paper-with-me

Papers Small Language Model

“Small Language Model” 태그가 달린 논문 109편 · 필터 해제

Scaling Up RL: Unlocking Diverse Reasoning in LLMs via Prolonged Training

2025-07-16 · Mingjie Liu, Shizhe Diao, Jian Hu, Ximing Lu 외

Recent advancements in reasoning-focused language models such as OpenAI's O1 and DeepSeek-R1 have shown that scaling test-time computation-through chain-of-thought reasoning and iterative exploration-can yield substantia…

Code GenerationMathreinforcement-learningReinforcement Learning+2

Domain-Adaptive Small Language Models for Structured Tax Code Prediction

2025-07-15 · Souvik Nath, Sumit Wadhwa, Luiz Perez

Every day, multinational firms process thousands of transactions, each of which must adhere to tax regulations that vary by jurisdiction and are often nuanced. The determination of product and service tax codes, such as …

DecoderSmall Language Model

Towards Privacy-Preserving and Personalized Smart Homes via Tailored Small Language Models

2025-07-10 · Xinyu Huang, Leming Shen, Zijing Ma, Yuanqing Zheng

Large Language Models (LLMs) have showcased remarkable generalizability in language comprehension and hold significant potential to revolutionize human-computer interaction in smart homes. Existing LLM-based smart home a…

Privacy PreservingSmall Language Model

Counterfactual Influence as a Distributional Quantity

2025-06-25 · Matthieu Meeus, Igor Shilov, Georgios Kaissis, Yves-Alexandre de Montjoye

Machine learning models are known to memorize samples from their training data, raising concerns around privacy and generalization. Counterfactual self-influence is a popular metric to study memorization, quantifying how…

counterfactualimage-classificationImage ClassificationMemorization+1

Biomed-Enriched: A Biomedical Dataset Enriched with LLMs for Pretraining and Extracting Rare and Hidden Content

2025-06-25 · Rian Touchent, Nathan Godey, Eric de la Clergerie

We introduce Biomed-Enriched, a biomedical text dataset constructed from PubMed via a two-stage annotation process. In the first stage, a large language model annotates 400K paragraphs from PubMed scientific articles, as…

ArticlesContinual PretrainingLanguage ModelingLanguage Modelling+4

Distilling On-device Language Models for Robot Planning with Minimal Human Intervention

2025-06-20 · Zachary Ravichandran, Ignacio Hounie, Fernando Cladera, Alejandro Ribeiro 외

Large language models (LLMs) provide robots with powerful contextual reasoning abilities and a natural human interface. Yet, current LLM-enabled robots typically depend on cloud-hosted models, limiting their usability in…

Small Language Model

Lightweight Relevance Grader in RAG

2025-06-17 · Taehee Jeong

Retrieval-Augmented Generation (RAG) addresses limitations of large language models (LLMs) by leveraging a vector database to provide more accurate and up-to-date information. When a user submits a query, RAG executes a …

Language ModelingLanguage ModellingRAGRetrieval-augmented Generation+1

HypER: Literature-grounded Hypothesis Generation and Distillation with Provenance

2025-06-15 · Rosni Vasu, Chandrayee Basu, Bhavana Dalvi Mishra, Cristina Sarasua 외

Large Language models have demonstrated promising performance in research ideation across scientific domains. Hypothesis development, the process of generating a highly specific declarative statement connecting a researc…

Language ModelingLanguage ModellingSmall Language Modelvalid

Towards a Small Language Model Lifecycle Framework

2025-06-09 · Parsa Miraghaei, Sergio Moreschini, Antti Kolehmainen, David Hästbacka

Background: The growing demand for efficient and deployable language models has led to increased interest in Small Language Models (SLMs). However, existing research remains fragmented, lacking a unified lifecycle perspe…

Language ModelingLanguage ModellingmodelSmall Language Model

WhisQ: Cross-Modal Representation Learning for Text-to-Music MOS Prediction

2025-06-06 · Jakaria Islam Emon, Kazi Tamanna Alam, Md. Abu Salek

Mean Opinion Score (MOS) prediction for text to music systems requires evaluating both overall musical quality and text prompt alignment. This paper introduces WhisQ, a multimodal architecture that addresses this dual-as…

cross-modal alignmentLanguage ModelingLanguage ModellingRepresentation Learning+1

Prompt Candidates, then Distill: A Teacher-Student Framework for LLM-driven Data Annotation

2025-06-04 · Mingxuan Xia, Haobo Wang, Yixuan Li, Zewei Yu 외

Recently, Large Language Models (LLMs) have demonstrated significant potential for data annotation, markedly reducing the labor costs associated with downstream applications. However, existing methods mostly adopt an agg…

Small Language Modeltext-classificationText Classification

Adaptive Task Vectors for Large Language Models

2025-06-03 · Joonseong Kang, Soojeong Lee, Subeen Park, Sumin Park 외

In-Context Learning (ICL) enables Large Language Models (LLMs) to perform tasks without parameter updates by conditioning on a few demonstrations provided in the prompt. Despite its success, ICL suffers from several limi…

In-Context LearningSmall Language Model

Zero-Shot Vision Encoder Grafting via LLM Surrogates

2025-05-28 · Kaiyu Yue, Vasu Singla, Menglin Jia, John Kirchenbauer 외

Vision language models (VLMs) typically pair a modestly sized vision encoder with a large language model (LLM), e.g., Llama-70B, making the decoder the primary computational burden during training. To reduce costs, a pot…

DecoderLanguage ModelingLanguage ModellingLarge Language Model+1

A Lightweight Multi-Expert Generative Language Model System for Engineering Information and Knowledge Extraction

2025-05-27 · Bogdan Bogachov, Yaoyao Fiona Zhao

Despite recent advancements in domain adaptation techniques for large language models, these methods remain computationally intensive, and the resulting models can still exhibit hallucination issues. Most existing adapta…

Domain AdaptationHallucinationLanguage ModelingLanguage Modelling+1

Skip-Thinking: Chunk-wise Chain-of-Thought Distillation Enable Smaller Language Models to Reason Better and Faster

2025-05-24 · Xiao Chen, Sihang Zhou, Ke Liang, Xiaoyu Sun 외

Chain-of-thought (CoT) distillation allows a large language model (LLM) to guide a small language model (SLM) in reasoning tasks. Existing methods train the SLM to learn the long rationale in one iteration, resulting in …

Heuristic SearchLanguage ModelingLanguage ModellingLarge Language Model+1

Leveraging Online Data to Enhance Medical Knowledge in a Small Persian Language Model

2025-05-21 · Mehrdad ghassabi, Pedram Rostami, Hamidreza Baradaran Kashani, Amirhossein Poursina 외

The rapid advancement of language models has demonstrated the potential of artificial intelligence in the healthcare industry. However, small language models struggle with specialized domains in low-resource languages li…

Language ModelingLanguage ModellingMedical Question AnsweringPatient QA+2

TinyRS-R1: Compact Multimodal Language Model for Remote Sensing

2025-05-17 · Aybora Koksal, A. Aydin Alatan

Remote-sensing applications often run on edge hardware that cannot host today's 7B-parameter multimodal language models. This paper introduces TinyRS, the first 2B-parameter multimodal small language model (MSLM) optimiz…

Language ModelingLanguage ModellingOpen-Ended Question AnsweringQuestion Answering+4

Communication-Efficient Hybrid Language Model via Uncertainty-Aware Opportunistic and Compressed Transmission

2025-05-17 · Seungeun Oh, Jinhyuk Kim, Jihong Park, Seung-Woo Ko 외

To support emerging language-based applications using dispersed and heterogeneous computing resources, the hybrid language model (HLM) offers a promising architecture, where an on-device small language model (SLM) genera…

Language ModelingLanguage ModellingLarge Language ModelSmall Language Model

MilChat: Introducing Chain of Thought Reasoning and GRPO to a Multimodal Small Language Model for Remote Sensing

2025-05-12 · Aybora Koksal, A. Aydin Alatan

Remarkable capabilities in understanding and generating text-image content have been demonstrated by recent advancements in multimodal large language models (MLLMs). However, their effectiveness in specialized domains-pa…

Language ModelingLanguage ModellingSmall Language Model

Sadeed: Advancing Arabic Diacritization Through Small Language Model

2025-04-30 · Zeina Aldallal, Sara Chrouf, Khalil Hennara, Mohamed Motaism Hamed 외

Arabic text diacritization remains a persistent challenge in natural language processing due to the language's morphological richness. In this paper, we introduce Sadeed, a novel approach based on a fine-tuned decoder-on…

Arabic Text DiacritizationBenchmarkingDecoderLanguage Modeling+6
1–20 / 109 다음 →