·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
Learning to Discretize: Diffusion-Based Adaptive Mesh with Spectral Guidance8h◆Toward Trustworthy Autonomous Science: A Two-Year Community Roadmap8h◆ReflectWorld-MM: An Entity-Oriented Multimodal Memory System for Open-Ended Video Streams8h◆Mirror Horizon: Viable Path Entropy as a Measure of Bounded Reflection8h◆The Illusion of Robustness: Aggregate Accuracy Hides Prediction Flips under Task-Irrelevant Context8h◆What Makes a Representational Prior Work? Feature Families, Label-Free Invariances, and Critical Windows in Grokking8h◆Scalable Optimal Transport Algorithm for Network Alignment8h◆Mind the Gap: Promises and Pitfalls of Hierarchical Planning in LeWorldModel8h◆ISE: An Execution-Grounded Recipe for Multi-Turn OS-Agent Trajectories8h◆MAGE: Understanding Stability-Performance Trade-offs in Multi-component Prompt Optimization8h◆Evaluating Reliability in Machine Learning Models for Early Chronic Kidney Disease Prediction: A Systematic Review of Data Leakage and Predictor Stability8h◆Self-Evolving In-Context Learning for Direct Pilot-to-Beamformer Design in MU-MISO Systems8h◆Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs8h◆AutoTrace: From Patches to Triggers via Agentic Interprocedural Exploration8h◆The Benjamini--Hochberg Procedure Can Fail to Control the FDR for Correlated Two-Sided Gaussian Tests8h◆Does Topic Sentiment Cause Perceived Ideology? Comparing Human and LLM Annotations in Political News Articles8h◆Keep Policy Gradient in Charge: Sibling-Guided Credit Distillation for Long-Horizon Tool-Use Agents8h◆PLGSA-Transformer: Periocular Landmark-Guided Attention with Occlusion-Adaptive Cosine Thresholding for Cross-Modal Masked and Unmasked Face Recognition8h◆A Shared Subcircuit Lets LLMs Count Down Across Tasks8h◆Can LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment Reproduction8h◆Learning to Discretize: Diffusion-Based Adaptive Mesh with Spectral Guidance8h◆Toward Trustworthy Autonomous Science: A Two-Year Community Roadmap8h◆ReflectWorld-MM: An Entity-Oriented Multimodal Memory System for Open-Ended Video Streams8h◆Mirror Horizon: Viable Path Entropy as a Measure of Bounded Reflection8h◆The Illusion of Robustness: Aggregate Accuracy Hides Prediction Flips under Task-Irrelevant Context8h◆What Makes a Representational Prior Work? Feature Families, Label-Free Invariances, and Critical Windows in Grokking8h◆Scalable Optimal Transport Algorithm for Network Alignment8h◆Mind the Gap: Promises and Pitfalls of Hierarchical Planning in LeWorldModel8h◆ISE: An Execution-Grounded Recipe for Multi-Turn OS-Agent Trajectories8h◆MAGE: Understanding Stability-Performance Trade-offs in Multi-component Prompt Optimization8h◆Evaluating Reliability in Machine Learning Models for Early Chronic Kidney Disease Prediction: A Systematic Review of Data Leakage and Predictor Stability8h◆Self-Evolving In-Context Learning for Direct Pilot-to-Beamformer Design in MU-MISO Systems8h◆Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs8h◆AutoTrace: From Patches to Triggers via Agentic Interprocedural Exploration8h◆The Benjamini--Hochberg Procedure Can Fail to Control the FDR for Correlated Two-Sided Gaussian Tests8h◆Does Topic Sentiment Cause Perceived Ideology? Comparing Human and LLM Annotations in Political News Articles8h◆Keep Policy Gradient in Charge: Sibling-Guided Credit Distillation for Long-Horizon Tool-Use Agents8h◆PLGSA-Transformer: Periocular Landmark-Guided Attention with Occlusion-Adaptive Cosine Thresholding for Cross-Modal Masked and Unmasked Face Recognition8h◆A Shared Subcircuit Lets LLMs Count Down Across Tasks8h◆Can LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment Reproduction8h◆
News/ZAYA1-VL-8B Technical Report
arxiv
PublishedMay 13, 2026 at 4:00 AM
—neutral

ZAYA1-VL-8B Technical Report

Source
arxiv.orgfull article ↗
Read on arxiv→
Publisher summary· verbatim

arXiv:2605.08560v1 Announce Type: cross Abstract: We present ZAYA1-VL-8B, a compact mixture-of-experts vision-language model built upon our in-house language model, ZAYA1-8B. Despite its compact size, ZAYA1-VL achieves performance competitive with leading base models such as Molmo2-4B and InternVL3.

Stay posted· Newsletter

A 5-min weekly brief — top movers, price watch, story of the week.

// no spam · unsubscribe one-click · free forever

Discussion
Source
↗
arxiv
Read original ↗All from arxiv →

No replies yet. Be first.

Source
↗
arxiv
Read original ↗All from arxiv →

Related coverage

More from ARXIV
arxivLearning to Discretize: Diffusion-Based Adaptive Mesh with Spectral Guidance8harxivToward Trustworthy Autonomous Science: A Two-Year Community Roadmap8harxivReflectWorld-MM: An Entity-Oriented Multimodal Memory System for Open-Ended Video Streams8harxivMirror Horizon: Viable Path Entropy as a Measure of Bounded Reflection8h
The Bubble Brief
WEEKLY

Read AI insights every Tuesday — top movers, new releases, story of the week.

// no spam · unsubscribe one-click · free forever

Originally published on arxiv ↗
HomeModelsNews