paper-with-me

홈 › Papers

SOAEsV2-7B/72B: Full-Pipeline Optimization for State-Owned Enterprise LLMs via Continual Pre-Training, Domain-Progressive SFT and Distillation-Enhanced Speculative Decoding

2025-05-07 · Jingyang Deng, Ran Chen, Jo-Ku Cheng, Jinwen Ma

This study addresses key challenges in developing domain-specific large language models (LLMs) for Chinese state-owned assets and enterprises (SOAEs), where current approaches face three limitations: 1) constrained model capacity that limits knowledge integration and cross-task adaptability; 2) excessive reliance on domain-specific supervised fine-tuning (SFT) data, which neglects the broader applicability of general language patterns; and 3) inefficient inference acceleration for large models processing long contexts. In this work, we propose SOAEsV2-7B/72B, a specialized LLM series developed via a three-phase framework: 1) continual pre-training integrates domain knowledge while retaining base capabilities; 2) domain-progressive SFT employs curriculum-based learning strategy, transitioning from weakly relevant conversational data to expert-annotated SOAEs datasets to optimize domain-specific tasks; 3) distillation-enhanced speculative decoding accelerates inference via logit distillation between 72B target and 7B draft models, achieving 1.39-1.52$\times$ speedup without quality loss. Experimental results demonstrate that our domain-specific pre-training phase maintains 99.8% of original general language capabilities while significantly improving domain performance, resulting in a 1.08$\times$ improvement in Rouge-1 score and a 1.17$\times$ enhancement in BLEU-4 score. Ablation studies further show that domain-progressive SFT outperforms single-stage training, achieving 1.02$\times$ improvement in Rouge-1 and 1.06$\times$ in BLEU-4. Our work introduces a comprehensive, full-pipeline approach for optimizing SOAEs LLMs, bridging the gap between general language capabilities and domain-specific expertise.

📄 PDF Abstract BibTeX arXiv:2505.04723

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

SFT Shrink and Fine-Tune, or SFT, is a type of distillation that avoids explicit distillation by copying parameters to a student student model and then fine-tuning.…
BASE 설명 없음

Similar Papers 제목 키워드 기반

Winning Gold at IMO 2025 with a Model-Agnostic Verification-and-Refinement Pipeline

2025-07-21 · Yichen Huang, Lin F. Yang arxiv

The International Mathematical Olympiad (IMO) is widely regarded as the world championship of high-school mathematics. IMO problems are renowned for their difficulty and novelty, demanding deep insight, creativity, and r…

Super-efficiency of Listed Banks in China and Determinants Analysis (2006-2021)

2023-05-18 · Yun Liao, Ruihui Xu

This study employs the annual unbalanced panel data of 42 listed banks in China from 2006 to 2021, adopts the non-radial and non-oriented super-efficiency Data envelopment analysis (Super-SBM-UND-VRS based DEA) model con…

Market Crash Prediction Model for Markets in A Rational Bubble

2021-08-23 · HyeonJun Kim

Renowned method of log-periodic power law(LPPL) is one of the few ways that a financial market crash could be predicted. Alongside with LPPL, this paper propose a novel method of stock market crash using white box model …

Sensitivity

SovereignNegotiation-Bench: Evaluating User-Owned Personal Agents In Delegated Bargaining Under Privacy, Consent, Evidence, And Institutional Pressure

2026-07-02 · Dylan Zongmin Liu arxiv

Personal agents will increasingly negotiate on behalf of users: splitting costs with other personal agents, appealing platform decisions, escalating support disputes, requesting refunds, changing subscriptions, and negot…

Enhanced Facial Feature Extraction and Recignation Using Optimal Fully Dispersed Haar-like Filters

2024-04-16 · Zeinab Sedaghatjoo, Hossein Hosseinzadeh, Ahmad shirzadi

Haar-like filters are renowned for their simplicity, speed, and accuracy in various computer vision tasks. This paper proposes a novel algorithm to identify optimal fully dispersed Haar-like filters for enhanced facial f…

Face Detection