arxiv
PublishedApril 4, 2026 at 4:00 AM
▲bullish
Countering Catastrophic Forgetting of Large Language Models for Better Instruction Following via Weight-Space Model Merging
Publisher summary· verbatim
arXiv:2604.01538v1 Announce Type: new Abstract: Large language models have been adopted in the medical domain for clinical documentation to reduce clinician burden. However, studies have reported that LLMs often "forget" a significant amount of instruction-following ability when fine-tuned using a t
Models mentioned
01Related
04- arxivMay 28A Benchmark Construction and Evaluation Framework for Specialist Domains: Case Study on Defense-related Documents
- arxivMay 22GraphRAG on Consumer Hardware: Benchmarking Local LLMs for Healthcare EHR Schema Retrieval
- arxivMay 22DrugRAG: Enhancing Pharmacy LLM Performance Through A Novel Retrieval-Augmented Generation Pipeline
- arxivMay 8Measuring Evaluation-Context Divergence in Open-Weight LLMs: A Paired-Prompt Protocol with Pilot Evidence of Alignment-Pipeline-Specific Heterogeneity
Stay posted· Newsletter
A 5-min weekly brief — top movers, price watch, story of the week.
Discussion
No replies yet. Be first.
Related coverage
More from ARXIV
arxivADS-C: Antidistillation Sampling for Classification18harxivBeyond a Joke: Multi-Angle Reasoning for Detecting and Explaining Harmful Humor in Memes18harxivFrom Black Box to Executable Logic: Explainable Reinforcement Learning through Prolog Expert Systems18harxivA Formally Grounded ODRL Evaluator: Implementation and Comparison18hThe Bubble Brief
WEEKLYRead open-source insights every Tuesday — top movers, new releases, story of the week.
Originally published on arxiv ↗