paper-with-me

Papers

SAAS: Solving Ability Amplification Strategy for Enhanced Mathematical Reasoning in Large Language Models

2024-04-05 · Hyeonwoo Kim, Gyoungjin Gim, Yungi Kim, Jihoo Kim, Byungju Kim, Wonseok Lee, Chanjun Park

This study presents a novel learning approach designed to enhance both mathematical reasoning and problem-solving abilities of Large Language Models (LLMs). We focus on integrating the Chain-of-Thought (CoT) and the Program-of-Thought (PoT) learning, hypothesizing that prioritizing the learning of mathematical reasoning ability is helpful for the amplification of problem-solving ability. Thus, the initial learning with CoT is essential for solving challenging mathematical problems. To this end, we propose a sequential learning approach, named SAAS (Solving Ability Amplification Strategy), which strategically transitions from CoT learning to PoT learning. Our empirical study, involving an extensive performance comparison using several benchmarks, demonstrates that our SAAS achieves state-of-the-art (SOTA) performance. The results underscore the effectiveness of our sequential learning approach, marking a significant advancement in the field of mathematical reasoning in LLMs.

📄 PDF Abstract BibTeX arXiv:2404.03887

Code (0)

등록된 구현이 없습니다.

Tasks

Mathematical Reasoning

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Segment Anything Across Shots: A Method and Benchmark

2025-11-17 · Hengrui Hu, Kaining Ying, Henghui Ding arxiv

This work focuses on multi-shot semi-supervised video object segmentation (MVOS), which aims at segmenting the target object indicated by an initial mask throughout a video with multiple shots. The existing VOS methods m…

Semi-Supervised Video Object SegmentationData Augmentation

SAAS: Self-Aware Reinforcement Learning for Over-Search Mitigation in Agentic Search

2026-05-28 · Yunbo Tang, Chengyi Yang, Shiyu Liu, Zhishang Xiang 외 arxiv

Agentic search enables LLMs to solve complex multi-hop questions through iterative reasoning and external search. Despite the effectiveness, these systems often suffer from a critical limitation in practice: agents fail …

Reinforcement Learning

From Static to Intelligent: Evolving SaaS Pricing with LLMs

2025-07-16 · Francisco Javier Cavero, Juan C. Alonso, Antonio Ruiz-Cortés arxiv

The SaaS paradigm has revolutionized software distribution by offering flexible pricing options to meet diverse customer needs. However, the rapid expansion of the SaaS market has introduced significant complexity for De…

SaaS: Speed as a Supervisor for Semi-supervised Learning

2018-05-02 · ECCV 2018 9 · Safa Cicek, Alhussein Fawzi, Stefano Soatto

We introduce the SaaS Algorithm for semi-supervised learning, which uses learning speed during stochastic gradient descent in a deep neural network to measure the quality of an iterative estimate of the posterior probabi…

Measuring directional bias amplification in image captions using predictability

2025-03-10 · Rahul Nair, Bhanu Tokas, Neel Shah, Hannah Kerner

When we train models on biased ML datasets, they not only learn these biases but can inflate them at test time - a phenomenon called bias amplification. To measure bias amplification in ML datasets, many co-occurrence-ba…

Image Captioningimage-classificationImage Classification