·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
Hugging Face CEO calls for ‘radical transparency’ after ‘unprecedented’ OpenAI hack1h◆Monday.com is the latest tech company to blame AI for layoffs — here are 20 others16h◆Librarians are hosting viral ‘Avoiding AI’ workshops for people who are fed up with Big Tech1d◆One fallen power line exposed a growing AI data center problem. Here’s how to fix it.1d◆I tried out OpenAI’s new AI keypad — which will be fun for some coders and slightly mystifying to everyone else1d◆Prentis, new AI lab co-founded by Reid Hoffman, Mark Pincus in talks to raise $100M1d◆Midjourney bought the astrology app Co-Star1d◆Why Cognition bought Poke: AI personality is becoming a competitive advantage2d◆You can’t ignore Google Zero anymore2d◆Meta is making its AI chatbot more like an assistant2d◆Anthropic releases Opus 5 with ‘close’ to Fable 5’s capabilities2d◆Anthropic launches Opus 52d◆As US weighs response to Chinese AI, industry urges against broad open-weight restrictions2d◆Bluesky’s AI assistant Attie expands into an open social research tool2d◆Midjourney acquired the astrology app Co-Star2d◆The tech-broification of American science has officially begun2d◆‘AI communism’, rogue models, and the why Kimi K3 spooked Wall Street2d◆OpenAI’s new voice mode makes it to the ChatGPT desktop app2d◆SiGMA: Sign-Guided Merging and Adaptation for Multimodal Continual Instruction Tuning2d◆Neural Operator Surrogates for Two-Dimensional Neutron Flux Estimation2d◆Hugging Face CEO calls for ‘radical transparency’ after ‘unprecedented’ OpenAI hack1h◆Monday.com is the latest tech company to blame AI for layoffs — here are 20 others16h◆Librarians are hosting viral ‘Avoiding AI’ workshops for people who are fed up with Big Tech1d◆One fallen power line exposed a growing AI data center problem. Here’s how to fix it.1d◆I tried out OpenAI’s new AI keypad — which will be fun for some coders and slightly mystifying to everyone else1d◆Prentis, new AI lab co-founded by Reid Hoffman, Mark Pincus in talks to raise $100M1d◆Midjourney bought the astrology app Co-Star1d◆Why Cognition bought Poke: AI personality is becoming a competitive advantage2d◆You can’t ignore Google Zero anymore2d◆Meta is making its AI chatbot more like an assistant2d◆Anthropic releases Opus 5 with ‘close’ to Fable 5’s capabilities2d◆Anthropic launches Opus 52d◆As US weighs response to Chinese AI, industry urges against broad open-weight restrictions2d◆Bluesky’s AI assistant Attie expands into an open social research tool2d◆Midjourney acquired the astrology app Co-Star2d◆The tech-broification of American science has officially begun2d◆‘AI communism’, rogue models, and the why Kimi K3 spooked Wall Street2d◆OpenAI’s new voice mode makes it to the ChatGPT desktop app2d◆SiGMA: Sign-Guided Merging and Adaptation for Multimodal Continual Instruction Tuning2d◆Neural Operator Surrogates for Two-Dimensional Neutron Flux Estimation2d◆
News/How Much Orthogonalization Does Muon Need?
arxiv
PublishedJune 2, 2026 at 4:00 AM
—neutral

How Much Orthogonalization Does Muon Need?

Source
arxiv.orgfull article ↗
Read on arxiv→
Publisher summary· verbatim

arXiv:2606.00371v1 Announce Type: new Abstract: Muon optimizers improve neural-network training by replacing ill-conditioned momentum updates with approximately semi-orthogonal updates. This motivates a practical question: how much orthogonalization does Muon actually require? We study this question

Models mentioned
01
  • 01AI organization logo
    gpt2
    gpt2
Related
03
  • arxiv16d
    Efficient Long-Horizon Learning for Learned Optimization
  • arxivMay 1
    Making Logic a First-Class Citizen in Generative ML for Networking
  • arxivApr 9
    STQuant: Spatio-Temporal Adaptive Framework for Optimizer Quantization in Large Multimodal Model Training
Stay posted· Newsletter

A 5-min weekly brief — top movers, price watch, story of the week.

// no spam · unsubscribe one-click · free forever

Discussion
Mentioned models
04
  • 01
    NanoGPT
  • 02
    gpt2
    gpt2
  • 03
    Mamba
  • 04
    Muon
Source
↗
arxiv
Read original ↗All from arxiv →
Tags
04
#machine-learning#optimization#neural-networks#research

No replies yet. Be first.

Mentioned models
04
  • 01
    NanoGPT
  • 02
    gpt2
    gpt2
  • 03
    Mamba
  • 04
    Muon
Source
↗
arxiv
Read original ↗All from arxiv →
Tags
04
#machine-learning#optimization#neural-networks#research

Related coverage

More from ARXIV
arxivSiGMA: Sign-Guided Merging and Adaptation for Multimodal Continual Instruction Tuning2darxivNeural Operator Surrogates for Two-Dimensional Neutron Flux Estimation2d
The Bubble Brief
WEEKLY

Read machine-learning insights every Tuesday — top movers, new releases, story of the week.

// no spam · unsubscribe one-click · free forever

Originally published on arxiv ↗
HomeModelsNews