paper-with-me

홈 › Papers

CoGraM: Context-sensitive granular optimization method with rollback for robust model fusion

2025-12-03 · Julius Lenz arxiv

Merging neural networks without retraining is central to federated and distributed learning. Common methods such as weight averaging or Fisher merging often lose accuracy and are unstable across seeds. CoGraM (Contextual Granular Merging) is a multi-stage, context-sensitive, loss-based, and iterative optimization method across layers, neurons, and weight levels that aligns decisions with loss differences and thresholds and prevents harmful updates through rollback. CoGraM is an optimization method that addresses the weaknesses of methods such as Fisher and can significantly improve the merged network.

📄 PDF Abstract BibTeX arXiv:2512.03610

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Word Midas Powered by StringNet: Discovering Lexicogrammatical Constructions in Situ

2016-12-01 · COLING 2016 12 · David Wible, Nai-Lung Tsao

Adult second language learners face the daunting but underappreciated task of mastering patterns of language use that are neither products of fully productive grammar rules nor frozen items to be memorized. Word Midas, a…

Language ModelingLanguage Modelling

DART: Semantic Recoverability for Structured Tool Agents

2026-05-22 · Ke Yang, Panpan Li, Zonghan Wu, Kejin Xu 외 arxiv

When a structured tool agent fails mid-execution, the runtime faces a dilemma: replaying the entire task is safe but wasteful, while restoring from a local checkpoint is efficient but can leave committed downstream work …

DiscoGraMS: Enhancing Movie Screen-Play Summarization using Movie Character-Aware Discourse Graph

2024-10-18 · Maitreya Prafulla Chitale, Uday Bindal, Rajakrishnan Rajkumar, Rahul Mishra

Summarizing movie screenplays presents a unique set of challenges compared to standard document summarization. Screenplays are not only lengthy, but also feature a complex interplay of characters, dialogues, and scenes, …

Document SummarizationQuestion Answering

Generator-Assistant Stepwise Rollback Framework for Large Language Model Agent

2025-03-04 · Xingzuo Li, Kehai Chen, Yunfei Long, Xuefeng Bai 외

Large language model (LLM) agents typically adopt a step-by-step reasoning framework, in which they interleave the processes of thinking and acting to accomplish the given task. However, this paradigm faces a deep-rooted…

Decision MakingLanguage ModelingLanguage ModellingLarge Language Model

Network Digital Untwinning: Towards Backward Optimization of Digital Twins

2026-04-30 · Zifan Zhang, Dianwei Chen, Anjun Gao, Manhua Wang 외 arxiv

Network digital twins (NDTs) are transforming network management by offering precise virtual replicas of physical network systems. However, their reliance on diverse and sensitive data introduces significant challenges r…