paper-with-me

Papers Math Word Problem Solving

“Math Word Problem Solving” 태그가 달린 논문 107편 · 필터 해제

An Empirical Study of Data Ability Boundary in LLMs' Math Reasoning

2024-02-23 · Zui Chen, Yezeng Chen, Jiaqi Han, Zhijie Huang 외

Large language models (LLMs) are displaying emergent abilities for math reasoning tasks,and there is a growing attention on enhancing the ability of open-source LLMs through supervised fine-tuning (SFT).In this paper, we…

Arithmetic ReasoningAutomated Theorem ProvingMath Word Problem Solving

OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset

2024-02-15 · Shubham Toshniwal, Ivan Moshkov, Sean Narenthiran, Daria Gitman 외

Recent work has shown the immense potential of synthetically generated datasets for training large language models (LLMs), especially for acquiring targeted skills. Current large-scale math instruction tuning datasets su…

Arithmetic ReasoningGSM8KMathMath Word Problem Solving

DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

2024-02-05 · Zhihong Shao, Peiyi Wang, Qihao Zhu, Runxin Xu 외

Mathematical reasoning poses a significant challenge for language models due to its complex and structured nature. In this paper, we introduce DeepSeekMath 7B, which continues pre-training DeepSeek-Coder-Base-v1.5 7B wit…

Arithmetic ReasoningMathMathematical ReasoningMath Word Problem Solving

Augmenting Math Word Problems via Iterative Question Composing

2024-01-17 · Haoxiong Liu, Yifan Zhang, Yifan Luo, Andrew Chi-Chih Yao

Despite the advancements in large language models (LLMs) for mathematical reasoning, solving competition-level math problems remains a significant challenge, especially for open-source LLMs without external tools. We int…

MathMathematical ReasoningMath Word Problem Solving

Mixtral of Experts

2024-01-08 · Albert Q. Jiang, Alexandre Sablayrolles, Antoine Roux, Arthur Mensch 외

We introduce Mixtral 8x7B, a Sparse Mixture of Experts (SMoE) language model. Mixtral has the same architecture as Mistral 7B, with the difference that each layer is composed of 8 feedforward blocks (i.e. experts). For e…

Code GenerationCommon Sense ReasoningLanguage ModelingLanguage Modelling+4

Parameter-Efficient Sparsity Crafting from Dense to Mixture-of-Experts for Instruction Tuning on General Tasks

2024-01-05 · Haoyuan Wu, Haisheng Zheng, Zhuolun He, Bei Yu

Large language models (LLMs) have demonstrated considerable proficiency in general natural language processing (NLP) tasks. Instruction tuning, a successful paradigm, enhances the ability of LLMs to follow natural langua…

Arithmetic ReasoningCode GenerationCommon Sense ReasoningGPU+5

Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations

2023-12-14 · Peiyi Wang, Lei LI, Zhihong Shao, R. X. Xu 외

In this paper, we present an innovative process-oriented math process reward model called \textbf{Math-Shepherd}, which assigns a reward score to each step of math problem solutions. The training of Math-Shepherd is achi…

Arithmetic ReasoningGSM8KMathMathematical Reasoning+2

Frugal LMs Trained to Invoke Symbolic Solvers Achieve Parameter-Efficient Arithmetic Reasoning

2023-12-09 · Subhabrata Dutta, Joykirat Singh, Ishan Pandey, Sunny Manchanda 외

Large Language Models (LLM) exhibit zero-shot mathematical reasoning capacity as a behavior emergent with scale, commonly manifesting as chain-of-thoughts (CoT) reasoning. However, multiple empirical findings suggest tha…

Arithmetic ReasoningMathematical ReasoningMath Word Problem Solving

FinanceMath: Knowledge-Intensive Math Reasoning in Finance Domains

2023-11-16 · Yilun Zhao, Hongjun Liu, Yitao Long, Rui Zhang 외

We introduce FinanceMath, a novel benchmark designed to evaluate LLMs' capabilities in solving knowledge-intensive math reasoning problems. Compared to prior works, this study features three core advancements. First, Fin…

MathMath Word Problem SolvingRetrieval

VerityMath: Advancing Mathematical Reasoning by Self-Verification Through Unit Consistency

2023-11-13 · Vernon Toh Yan Han, Ratish Puduppully, Nancy F. Chen

Large Language Models (LLMs), combined with program-based solving techniques, are increasingly demonstrating proficiency in mathematical reasoning. For example, closed-source models such as OpenAI GPT-4 and Claude show e…

MathMathematical ReasoningMath Word Problem Solving

ATHENA: Mathematical Reasoning with Thought Expansion

2023-11-02 · EMNLP 2023 12 · JB. Kim, Hazel Kim, Joonghyuk Hahn, Yo-Sub Han

Solving math word problems depends on how to articulate the problems, the lens through which models view human linguistic expressions. Real-world settings count on such a method even more due to the diverse practices of …

MathMathematical ReasoningMath Word Problem Solving

An Expression Tree Decoding Strategy for Mathematical Equation Generation

2023-10-14 · Wenqi Zhang, Yongliang Shen, Qingpeng Nong, Zeqi Tan 외

Generating mathematical equations from natural language requires an accurate understanding of the relations among math expressions. Existing approaches can be broadly categorized into token-level and expression-level gen…

MathMathematical ReasoningMath Word Problem SolvingStructured Prediction

Mistral 7B

2023-10-10 · Albert Q. Jiang, Alexandre Sablayrolles, Arthur Mensch, Chris Bamford 외

We introduce Mistral 7B v0.1, a 7-billion-parameter language model engineered for superior performance and efficiency. Mistral 7B outperforms Llama 2 13B across all evaluated benchmarks, and Llama 1 34B in reasoning, mat…

answerability predictionArithmetic ReasoningChatbotCode Generation+11

MuggleMath: Assessing the Impact of Query and Response Augmentation on Math Reasoning

2023-10-09 · Chengpeng Li, Zheng Yuan, Hongyi Yuan, Guanting Dong 외

In math reasoning with large language models (LLMs), fine-tuning data augmentation by query evolution and diverse reasoning paths is empirically verified effective, profoundly narrowing the gap between open-sourced LLMs …

Arithmetic ReasoningData AugmentationGSM8KMath+2

MathCoder: Seamless Code Integration in LLMs for Enhanced Mathematical Reasoning

2023-10-05 · Ke Wang, Houxing Ren, Aojun Zhou, Zimu Lu 외

The recently released GPT-4 Code Interpreter has demonstrated remarkable proficiency in solving challenging math problems, primarily attributed to its ability to seamlessly reason with natural language, generate code, ex…

Arithmetic ReasoningGSM8KMathMathematical Reasoning+1

ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving

2023-09-29 · Zhibin Gou, Zhihong Shao, Yeyun Gong, Yelong Shen 외

Large language models have made significant progress in various language tasks, yet they still struggle with complex mathematics. In this paper, we propose ToRA a series of Tool-integrated Reasoning Agents designed to so…

Arithmetic ReasoningComputational EfficiencyImitation LearningMath+3

MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models

2023-09-21 · Longhui Yu, Weisen Jiang, Han Shi, Jincheng Yu 외

Large language models (LLMs) have pushed the limits of natural language understanding and exhibited excellent problem-solving ability. Despite the great success, most existing open-source LLMs (e.g., LLaMA-2) are still f…

Arithmetic ReasoningGSM8KLanguage ModelingLanguage Modelling+4

OpenChat: Advancing Open-source Language Models with Mixed-Quality Data

2023-09-20 · Guan Wang, Sijie Cheng, Xianyuan Zhan, Xiangang Li 외

Nowadays, open-source large language models like LLaMA have emerged. Recent developments have incorporated supervised fine-tuning (SFT) and reinforcement learning fine-tuning (RLFT) to align these models with human goals…

Arithmetic ReasoningCode GenerationMath Word Problem Solving

WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct

2023-08-18 · Haipeng Luo, Qingfeng Sun, Can Xu, Pu Zhao 외

Large language models (LLMs), such as GPT-4, have shown remarkable performance in natural language processing (NLP) tasks, including challenging mathematical reasoning. However, most existing open-source models are only …

Arithmetic ReasoningGSM8KMathMathematical Reasoning+1

Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

2023-08-15 · Aojun Zhou, Ke Wang, Zimu Lu, Weikang Shi 외

Recent progress in large language models (LLMs) like GPT-4 and PaLM-2 has brought significant advancements in addressing math reasoning problems. In particular, OpenAI's latest version of GPT-4, known as GPT-4 Code Inter…

Arithmetic ReasoningMathMathematical ReasoningMath Word Problem Solving
← 이전 21–40 / 107 다음 →