·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
Anthropic’s $1.5 billion book piracy settlement approved by judge16m◆US threatens sanctions against Chinese AI models over IP theft1h◆Google launches a cheaper alternative to large AI security models like Mythos2h◆Music streamer Deezer says more than 50% of daily uploads are AI-generated3h◆Halliday’s latest smart glasses feature a much-improved display4h◆America needs to stop getting shocked by Chinese AI6h◆Advancing next-gen AI with materials science innovation6h◆Gritt exits stealth with $32 million for robots to build solar plants — then, everything else7h◆Capacity and Redundancy Trade-offs in Multi-Task Learning13h◆Predictive Training with Latent Imagination for Visual Quadruped Navigation13h◆Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making13h◆Did We Actually Fix It? An Independent Adversarial Stress-Test of Post-Point-Adjustment Evaluation Metrics for Time-Series Anomaly Detection13h◆Supervised Reward Inference13h◆PPO-HSC: An Exploratory Reinforcement Learning Framework Based on Wide-Area Policy Coverage Optimization13h◆Is Progressive Disclosure All You Need for Long-Context Agents?13h◆It Depends on the Dataset: When a Brain-Encoding Model's Predicted Responses Beat Their Visual Backbone for Video Memorability13h◆DMFNet: Dual-Backbone Multiscale Fusion Network for Urban Scene Classification13h◆Oracle Gap and Signal Fidelity: A Fixed-Pool Diagnostic for Test-Time Collaboration13h◆Scientific reasoning does not reliably translate into scientific forecasting in frontier AI13h◆Language Triggers Hijack Language Circuits: A Mechanistic Analysis of Backdoor Behaviors in Large Language Models13h◆Anthropic’s $1.5 billion book piracy settlement approved by judge16m◆US threatens sanctions against Chinese AI models over IP theft1h◆Google launches a cheaper alternative to large AI security models like Mythos2h◆Music streamer Deezer says more than 50% of daily uploads are AI-generated3h◆Halliday’s latest smart glasses feature a much-improved display4h◆America needs to stop getting shocked by Chinese AI6h◆Advancing next-gen AI with materials science innovation6h◆Gritt exits stealth with $32 million for robots to build solar plants — then, everything else7h◆Capacity and Redundancy Trade-offs in Multi-Task Learning13h◆Predictive Training with Latent Imagination for Visual Quadruped Navigation13h◆Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making13h◆Did We Actually Fix It? An Independent Adversarial Stress-Test of Post-Point-Adjustment Evaluation Metrics for Time-Series Anomaly Detection13h◆Supervised Reward Inference13h◆PPO-HSC: An Exploratory Reinforcement Learning Framework Based on Wide-Area Policy Coverage Optimization13h◆Is Progressive Disclosure All You Need for Long-Context Agents?13h◆It Depends on the Dataset: When a Brain-Encoding Model's Predicted Responses Beat Their Visual Backbone for Video Memorability13h◆DMFNet: Dual-Backbone Multiscale Fusion Network for Urban Scene Classification13h◆Oracle Gap and Signal Fidelity: A Fixed-Pool Diagnostic for Test-Time Collaboration13h◆Scientific reasoning does not reliably translate into scientific forecasting in frontier AI13h◆Language Triggers Hijack Language Circuits: A Mechanistic Analysis of Backdoor Behaviors in Large Language Models13h◆
News/Scaling Model and Data for Multilingual Machine Translation with Open Large Language Models
arxiv
PublishedJuly 21, 2026 at 4:00 AM
▲bullish

Scaling Model and Data for Multilingual Machine Translation with Open Large Language Models

Source
arxiv.orgfull article ↗
Read on arxiv→
Publisher summary· verbatim

arXiv:2602.11961v3 Announce Type: replace Abstract: Open large language models (LLMs) have demonstrated improving multilingual capabilities in recent years. In this paper, we present a study of open LLMs for multilingual machine translation (MT) across a range of languages, and investigate the effec

Stay posted· Newsletter

A 5-min weekly brief — top movers, price watch, story of the week.

// no spam · unsubscribe one-click · free forever

Discussion
Mentioned models
07
  • 01
    MiLMMT-46
  • 02
    Seed-X
  • 03
    HY-MT-1.5
  • 04
    TranslateGemma
  • 05
    Gemma3
  • 06
    Google Translate
  • 07
    Gemini 3 Pro
Source
↗
arxiv
Read original ↗All from arxiv →
Tags
04
#multilingual#machine-translation#open-source#language-models
Mentioned companies
01
Google

No replies yet. Be first.

Mentioned models
07
  • 01
    MiLMMT-46
  • 02
    Seed-X
  • 03
    HY-MT-1.5
  • 04
    TranslateGemma
  • 05
    Gemma3
  • 06
    Google Translate
  • 07
    Gemini 3 Pro
Source
↗
arxiv
Read original ↗All from arxiv →
Tags
04
#multilingual#machine-translation#open-source#language-models
Mentioned companies
01
Google

Related coverage

More from ARXIV
arxivCapacity and Redundancy Trade-offs in Multi-Task Learning13harxivPredictive Training with Latent Imagination for Visual Quadruped Navigation13harxivWhere Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making13harxivDid We Actually Fix It? An Independent Adversarial Stress-Test of Post-Point-Adjustment Evaluation Metrics for Time-Series Anomaly Detection13h
The Bubble Brief
WEEKLY

Read multilingual insights every Tuesday — top movers, new releases, story of the week.

// no spam · unsubscribe one-click · free forever

Originally published on arxiv ↗
HomeModelsNews