·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
Data centers expected to use 4x more electricity by 20351h◆Google releases three new Gemini models — but no 3.5 Pro2h◆Introducing the ChatGPT for small business program2h◆Anthropic’s $1.5 billion book piracy settlement approved by judge2h◆US threatens sanctions against Chinese AI models over IP theft4h◆Google launches a cheaper alternative to large AI security models like Mythos4h◆Music streamer Deezer says more than 50% of daily uploads are AI-generated6h◆Halliday’s latest smart glasses feature a much-improved display6h◆America needs to stop getting shocked by Chinese AI8h◆Advancing next-gen AI with materials science innovation9h◆Gritt exits stealth with $32 million for robots to build solar plants — then, everything else9h◆Capacity and Redundancy Trade-offs in Multi-Task Learning15h◆Predictive Training with Latent Imagination for Visual Quadruped Navigation15h◆Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making15h◆Did We Actually Fix It? An Independent Adversarial Stress-Test of Post-Point-Adjustment Evaluation Metrics for Time-Series Anomaly Detection15h◆Supervised Reward Inference15h◆PPO-HSC: An Exploratory Reinforcement Learning Framework Based on Wide-Area Policy Coverage Optimization15h◆Is Progressive Disclosure All You Need for Long-Context Agents?15h◆It Depends on the Dataset: When a Brain-Encoding Model's Predicted Responses Beat Their Visual Backbone for Video Memorability15h◆DMFNet: Dual-Backbone Multiscale Fusion Network for Urban Scene Classification15h◆Data centers expected to use 4x more electricity by 20351h◆Google releases three new Gemini models — but no 3.5 Pro2h◆Introducing the ChatGPT for small business program2h◆Anthropic’s $1.5 billion book piracy settlement approved by judge2h◆US threatens sanctions against Chinese AI models over IP theft4h◆Google launches a cheaper alternative to large AI security models like Mythos4h◆Music streamer Deezer says more than 50% of daily uploads are AI-generated6h◆Halliday’s latest smart glasses feature a much-improved display6h◆America needs to stop getting shocked by Chinese AI8h◆Advancing next-gen AI with materials science innovation9h◆Gritt exits stealth with $32 million for robots to build solar plants — then, everything else9h◆Capacity and Redundancy Trade-offs in Multi-Task Learning15h◆Predictive Training with Latent Imagination for Visual Quadruped Navigation15h◆Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making15h◆Did We Actually Fix It? An Independent Adversarial Stress-Test of Post-Point-Adjustment Evaluation Metrics for Time-Series Anomaly Detection15h◆Supervised Reward Inference15h◆PPO-HSC: An Exploratory Reinforcement Learning Framework Based on Wide-Area Policy Coverage Optimization15h◆Is Progressive Disclosure All You Need for Long-Context Agents?15h◆It Depends on the Dataset: When a Brain-Encoding Model's Predicted Responses Beat Their Visual Backbone for Video Memorability15h◆DMFNet: Dual-Backbone Multiscale Fusion Network for Urban Scene Classification15h◆
News/model/ZAYA1-74B-preview

ZAYA1-74B-preview news

2 articles mentioning ZAYA1-74B-preview

arxivMay 13

ZAYA1-VL-8B Technical Report

arXiv:2605.08560v1 Announce Type: cross Abstract: We present ZAYA1-VL-8B, a compact mixture-of-experts vision-language model built upon our in-house language model, ZAYA1-8B. Despite its compact size, ZAYA1-VL achieves performance competitive with leading base models such as Molmo2-4B and InternVL3.

arxivMay 8

ZAYA1-8B Technical Report

arXiv:2605.05365v1 Announce Type: new Abstract: We present ZAYA1-8B, a reasoning-focused mixture-of-experts (MoE) model with 700M active and 8B total parameters, built on Zyphra's MoE++ architecture. ZAYA1-8B's core pretraining, midtraining, and supervised fine-tuning (SFT) were performed on a full-

HomeModelsNews