arxiv
PublishedJuly 31, 2026 at 4:00 AM
▲bullish
NorBERTo: A ModernBERT Model Trained for Portuguese with 331 Billion Tokens Corpus
Publisher summary· verbatim
arXiv:2605.00086v2 Announce Type: replace Abstract: High-quality corpora are essential for advancing Natural Language Processing (NLP) in Portuguese. Building on previous encoder-only models such as BERTimbau and Albertina PT-BR, we introduce NorBERTo, a modern encoder based on the ModernBERT archit
Models mentioned
01Related
04- arxivJul 10Efficient Long-Horizon Learning for Learned Optimization
- arxivJun 2How Much Orthogonalization Does Muon Need?
- arxivMay 1Making Logic a First-Class Citizen in Generative ML for Networking
- arxivApr 9STQuant: Spatio-Temporal Adaptive Framework for Optimizer Quantization in Large Multimodal Model Training
Stay posted· Newsletter
A 5-min weekly brief — top movers, price watch, story of the week.
Discussion
No replies yet. Be first.
The Bubble Brief
WEEKLYRead nlp insights every Tuesday — top movers, new releases, story of the week.
Originally published on arxiv ↗