·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
ExecRetrieval: Measuring the Functional-Correctness Gap in Code-Embedding Retrieval17m◆MasterControl Seventeen Every Time17m◆Speculative Macro Commit for Faster Tool-Using Agents17m◆Fresh Memory, Stale Plans: Dependency-Scoped Validation for Distributed LLM-Agent Memory17m◆A Prompt-Engineering Approach to Develop Scalable, Flexible, and Real-Time Hybrid Micro-Level Personalization in a General Purpose AI Teaching Assistant17m◆Caught in the Story: Narrative Captivity in Multi-turn LLMs Conversation17m◆Dude: A Dual-Detection Multi-Agent System for Paper-Code Discrepancy Detection17m◆DuplexSpeechBench-IFEval: Evaluating Implicit Instruction Following in Full-Duplex Voice Agents17m◆Do GUI Agents Know When Not to Act? Enabling Conflict-Aware Termination for Multimodal GUI Agents17m◆Beyond "Made with AI": Visualizing Provenance Density to Mitigate the Transparency Penalty17m◆AutoGraphForge: Towards Automated Graph Theory Discovery17m◆Making Every Tool Call Count: Necessary Tool-Evidence Path Rewards for Agentic Vision-Language Models17m◆GrowPage: On-Demand KV Budgeting for Efficient LLM Reasoning Serving17m◆PPO-STGNN: A Proximal Policy Optimization Approach with Spatio-Temporal Graph Neural Networks for DAG Task Scheduling in Cloud-Edge-End Computing17m◆What Matters for Aggressive Decoding-Time KV Eviction? Temporal Aggregation and Ranking Preservation17m◆CulturalMenuBench: Probing the Knowledge-Application Gap in Multimodal Culinary Reasoning17m◆NeoRed: A Knowledge-Logic-Alignment Multimodal Large Language Model for Neonatal Respiratory Disease Diagnosis17m◆Feature Reconfiguration With Visual Prior for Medical Lesion Segmentation17m◆Dalek: A Constructive Agent Machine17m◆GPS-Bench: A Governance Policy Benchmark for Automating Policy Analysis17m◆ExecRetrieval: Measuring the Functional-Correctness Gap in Code-Embedding Retrieval17m◆MasterControl Seventeen Every Time17m◆Speculative Macro Commit for Faster Tool-Using Agents17m◆Fresh Memory, Stale Plans: Dependency-Scoped Validation for Distributed LLM-Agent Memory17m◆A Prompt-Engineering Approach to Develop Scalable, Flexible, and Real-Time Hybrid Micro-Level Personalization in a General Purpose AI Teaching Assistant17m◆Caught in the Story: Narrative Captivity in Multi-turn LLMs Conversation17m◆Dude: A Dual-Detection Multi-Agent System for Paper-Code Discrepancy Detection17m◆DuplexSpeechBench-IFEval: Evaluating Implicit Instruction Following in Full-Duplex Voice Agents17m◆Do GUI Agents Know When Not to Act? Enabling Conflict-Aware Termination for Multimodal GUI Agents17m◆Beyond "Made with AI": Visualizing Provenance Density to Mitigate the Transparency Penalty17m◆AutoGraphForge: Towards Automated Graph Theory Discovery17m◆Making Every Tool Call Count: Necessary Tool-Evidence Path Rewards for Agentic Vision-Language Models17m◆GrowPage: On-Demand KV Budgeting for Efficient LLM Reasoning Serving17m◆PPO-STGNN: A Proximal Policy Optimization Approach with Spatio-Temporal Graph Neural Networks for DAG Task Scheduling in Cloud-Edge-End Computing17m◆What Matters for Aggressive Decoding-Time KV Eviction? Temporal Aggregation and Ranking Preservation17m◆CulturalMenuBench: Probing the Knowledge-Application Gap in Multimodal Culinary Reasoning17m◆NeoRed: A Knowledge-Logic-Alignment Multimodal Large Language Model for Neonatal Respiratory Disease Diagnosis17m◆Feature Reconfiguration With Visual Prior for Medical Lesion Segmentation17m◆Dalek: A Constructive Agent Machine17m◆GPS-Bench: A Governance Policy Benchmark for Automating Policy Analysis17m◆
News/MuCon: Clipped Muon Updates for LLM Training
arxiv
PublishedMay 27, 2026 at 4:00 AM

MuCon: Clipped Muon Updates for LLM Training

Source
arxiv.orgfull article ↗
Read on arxiv→
Publisher summary· verbatim

arXiv:2605.26459v1 Announce Type: new Abstract: Muon-style optimizers take a matrix-valued momentum or preconditioned update $B = U \operatorname{diag}(\sigma_1,\ldots,\sigma_r) V^\top$ and replace it with its canonical partial polar factor $\operatorname{Pol}(B) = U V^\top$. This maps every nonzero

Stay posted· Newsletter

A 5-min weekly brief — top movers, price watch, story of the week.

// no spam · unsubscribe one-click · free forever

Discussion
Source
↗
arxiv
Read original ↗All from arxiv →

No replies yet. Be first.

Source
↗
arxiv
Read original ↗All from arxiv →

Related coverage

More from ARXIV
arxivExecRetrieval: Measuring the Functional-Correctness Gap in Code-Embedding Retrieval17marxivMasterControl Seventeen Every Time17marxivSpeculative Macro Commit for Faster Tool-Using Agents17marxivFresh Memory, Stale Plans: Dependency-Scoped Validation for Distributed LLM-Agent Memory17m
The Bubble Brief
WEEKLY

Read AI insights every Tuesday — top movers, new releases, story of the week.

// no spam · unsubscribe one-click · free forever

Originally published on arxiv ↗
HomeModelsNews