paper-with-me

홈 › Papers

OpenStaxQA: A multilingual dataset based on open-source college textbooks

2025-10-03 · Pranav Gupta arxiv

We present OpenStaxQA, an evaluation benchmark specific to college-level educational applications based on 43 open-source college textbooks in English, Spanish, and Polish, available under a permissive Creative Commons license. We finetune and evaluate large language models (LLMs) with approximately 7 billion parameters on this dataset using quantized low rank adapters (QLoRa). Additionally we also perform a zero-shot evaluation on the AI2 reasoning challenge dev dataset in order to check if OpenStaxQA can lead to an improved performance on other tasks. We also discuss broader impacts relevant to datasets such as OpenStaxQA.

📄 PDF Abstract BibTeX arXiv:2510.06239

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The University of Maryland, College Park Submission to Large-Scale Multilingual Shared Task at WMT 2021

2021-11-01 · WMT (EMNLP) 2021 11 · Saptarashmi Bandyopadhyay, Tasnim Kabir, Zizhen Lian, Marine Carpuat

This paper describes the system submitted to Large-Scale Multilingual Shared Task (Small Task #2) at WMT 2021. It is based on the massively multilingual open-source model FLORES101_MM100 model, with selective fine-tuning…

Task 2

SCIMAT: Science and Mathematics Dataset

2021-09-30 · Neeraj Kollepara, Snehith Kumar Chatakonda, Pawan Kumar

In this work, we announce a comprehensive well curated and opensource dataset with millions of samples for pre-college and college level problems in mathematicsand science. A preliminary set of results using transformer …

Build Your Own Robot Friend: An Open-Source Learning Module for Accessible and Engaging AI Education

2024-01-06 · Zhonghao Shi, Allison O'Connell, Zongjian Li, SiQi Liu 외

As artificial intelligence (AI) is playing an increasingly important role in our society and global economy, AI education and literacy have become necessary components in college and K-12 education to prepare students fo…

Ethics

OpenDebateEvidence: A Massive-Scale Argument Mining and Summarization Dataset

2024-06-20 · Allen Roush, Yusuf Shabazz, Arvind Balaji, Peter Zhang 외

We introduce OpenDebateEvidence, a comprehensive dataset for argument mining and summarization sourced from the American Competitive Debate community. This dataset includes over 3.5 million documents with rich metadata, …

Abstractive Text SummarizationArgument Mining

Tagengo: A Multilingual Chat Dataset

2024-05-21 · Peter Devine

Open source large language models (LLMs) have shown great improvements in recent times. However, many of these models are focused solely on popular spoken languages. We present a high quality dataset of more than 70k pro…