paper-with-me

홈 › Papers

Are LLMs Effective Backbones for Fine-tuning? An Experimental Investigation of Supervised LLMs on Chinese Short Text Matching

2024-03-29 · Shulin Liu, Chengcheng Xu, Hao liu, TingHao Yu, Tao Yang

The recent success of Large Language Models (LLMs) has garnered significant attention in both academia and industry. Prior research on LLMs has primarily focused on enhancing or leveraging their generalization capabilities in zero- and few-shot settings. However, there has been limited investigation into effectively fine-tuning LLMs for a specific natural language understanding task in supervised settings. In this study, we conduct an experimental analysis by fine-tuning LLMs for the task of Chinese short text matching. We explore various factors that influence performance when fine-tuning LLMs, including task modeling methods, prompt formats, and output formats.

📄 PDF Abstract BibTeX arXiv:2403.19930

Code (0)

등록된 구현이 없습니다.

Tasks

Natural Language UnderstandingText Matching

Similar Papers 제목 키워드 기반

MathGLM-Vision: Solving Mathematical Problems with Multi-Modal Large Language Model

2024-09-10 · Zhen Yang, Jinhao Chen, Zhengxiao Du, Wenmeng Yu 외

Large language models (LLMs) have demonstrated significant capabilities in mathematical reasoning, particularly with text-based mathematical problems. However, current multi-modal large language models (MLLMs), especiall…

DiversityLanguage ModelingLanguage ModellingLarge Language Model+1

EAGLE: Expert-Augmented Attention Guidance for Tuning-Free Industrial Anomaly Detection in Multimodal Large Language Models

2026-02-19 · Xiaomeng Peng, Xilang Huang, Seon Han Choi arxiv

Multimodal large language models (MLLMs) can enrich industrial anomaly detection with semantic descriptions and anomaly reasoning, but they still lag specialist anomaly detectors in binary detection accuracy. Existing ap…

Anomaly Detection

Graph Representation Learning Augmented Model Manipulation on Federated Fine-Tuning of LLMs

2026-05-08 · Hanlin Cai, Kai Li, Houtianfu Wang, Haofan Dong 외 arxiv

Federated fine-tuning (FFT) has emerged as a privacy-preserving paradigm for collaboratively adapting large language models (LLMs). Built upon federated learning, FFT enables distributed agents to jointly refine a shared…

Graph Representation LearningFederated Learning

Mixture-of-Skills: Learning to Optimize Data Usage for Fine-Tuning Large Language Models

2024-06-13 · Minghao Wu, Thuy-Trang Vu, Lizhen Qu, Gholamreza Haffari

Large language models (LLMs) are typically fine-tuned on diverse and extensive datasets sourced from various origins to develop a comprehensive range of skills, such as writing, reasoning, chatting, coding, and more. Eac…

HBO: Hierarchical Balancing Optimization for Fine-Tuning Large Language Models

2025-05-18 · Weixuan Wang, Minghao Wu, Barry Haddow, Alexandra Birch

Fine-tuning large language models (LLMs) on a mixture of diverse datasets poses challenges due to data imbalance and heterogeneity. Existing methods often address these issues across datasets (globally) but overlook the …

Bilevel Optimization