paper-with-me

홈 › Papers

PIKA: Expert-Level Synthetic Datasets for Post-Training Alignment from Scratch

2025-10-08 · Shangjian Yin, Shining Liang, Wenbiao Ding, Yuli Qian, Zhouxing Shi, Hongzhi Li, Yutao Xie arxiv

High-quality instruction data is critical for LLM alignment, yet existing open-source datasets often lack efficiency, requiring hundreds of thousands of examples to approach proprietary performance. In this work, we find that beyond the widely recognized importance of prompt-response quality, prompt difficulty itself plays a critical role in driving alignment gains. Motivated by this observation, we introduce PiKa, a data-efficient family of expert-level alignment datasets that concentrates supervision on high-difficulty instructions. The PiKa-SFT dataset contains only 30k examples, an order of magnitude fewer than state-of-the-art open datasets like Magpie-Pro. Despite its small size, fine-tuning Llama-3-8B-Base on PiKa-SFT even outperforms the official Llama-3-8B-Instruct model trained on over 10M proprietary examples on widely used benchmarks such as AlpacaEval 2.0 and Arena-Hard. We also validate the generalizability of PiKa across the Qwen2.5 series (0.5B-7B), consistently surpassing their official instruction-tuned counterparts. Additionally, we provide 30k high-quality preference optimization examples to further enhance alignment. Our results demonstrate that promising alignment is achievable with significantly reduced data, democratizing access for resource-constrained research. Our code and data will be available at https://github.com/SJY8460/PiKa.

📄 PDF Abstract BibTeX arXiv:2510.06670

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

PQA: Zero-shot Protein Question Answering for Free-form Scientific Enquiry with Large Language Models

2024-02-21 · Eli M Carrami, Sahand Sharifzadeh

Understanding protein structure and function is crucial in biology. However, current computational methods are often task-specific and resource-intensive. To address this, we propose zero-shot Protein Question Answering …

BenchmarkingFormQuestion Answering

Physics Informed Kolmogorov-Arnold Neural Networks for Dynamical Analysis via Efficent-KAN and WAV-KAN

2024-07-25 · Subhajit Patra, Sonali Panda, Bikram Keshari Parida, Mahima Arya 외

Physics-informed neural networks have proven to be a powerful tool for solving differential equations, leveraging the principles of physics to inform the learning process. However, traditional deep neural networks often …

PIKAN: Physics-Inspired Kolmogorov-Arnold Networks for Explainable UAV Channel Modelling

2025-10-07 · Kürşat Tekbıyık, Güneş Karabulut Kurt, Antoine Lesage-Landry arxiv

Unmanned aerial vehicle (UAV) communications demand accurate yet interpretable air-to-ground (A2G) channel models that can adapt to nonstationary propagation environments. While deterministic models offer interpretabilit…

Training Deep Physics-Informed Kolmogorov-Arnold Networks

2025-10-27 · Spyros Rigas, Fotios Anagnostopoulos, Michalis Papachristou, Georgios Alexandridis arxiv

Since their introduction, Kolmogorov-Arnold Networks (KANs) have been successfully applied across several domains, with physics-informed machine learning (PIML) emerging as one of the areas where they have thrived. In th…

Computational Efficiency

SpikACom: A Neuromorphic Computing Framework for Green Communications

2025-02-24 · Yanzhen Liu, Zhijin Qin, Yongxu Zhu, Geoffrey Ye Li

The ever-growing power consumption of wireless communication systems necessitates more energy-efficient algorithms. This paper introduces SpikACom ({Spik}ing {A}daptive {Com}munication), a neuromorphic computing-based fr…

Semantic Communication