·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
Capacity and Redundancy Trade-offs in Multi-Task Learning5h◆Predictive Training with Latent Imagination for Visual Quadruped Navigation5h◆Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making5h◆SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks?5h◆Did We Actually Fix It? An Independent Adversarial Stress-Test of Post-Point-Adjustment Evaluation Metrics for Time-Series Anomaly Detection5h◆Supervised Reward Inference5h◆Semi-Supervised Conditional Generative Learning through Stochastic Interpolation and Sufficient Representations5h◆PPO-HSC: An Exploratory Reinforcement Learning Framework Based on Wide-Area Policy Coverage Optimization5h◆A Survey on the Verification of Reinforcement Learning Policies5h◆RECON: Benchmarking Agent Memory for Compositional Reasoning over Long Contexts5h◆Lomekwi: Resource-Bounded Tool Discovery in LLM Agents5h◆Expected Free Energy as Belief-Dependent Utility for rho-POMDPs5h◆Constrained Path Reasoning: Measuring When Committed Stages Earn Their Cost5h◆Pailitao-MMSearch: Building Native E-Commerce Multimodal Search Foundation5h◆Is Progressive Disclosure All You Need for Long-Context Agents?5h◆Verify, Repair, Repeat, or Stop? Robust Stopping for Noisy Verify-Repair Loops in LLM Agents5h◆Hierarchical Wireless Foundation Model for Multi-Task Optimization5h◆Auto Research for Materials: Auditable AI-Scientist Workflows with Held-Out Transfer5h◆It Depends on the Dataset: When a Brain-Encoding Model's Predicted Responses Beat Their Visual Backbone for Video Memorability5h◆A${}^2$BM: Alignment-Aware Bridge Matching for Image-to-Image Translation5h◆Capacity and Redundancy Trade-offs in Multi-Task Learning5h◆Predictive Training with Latent Imagination for Visual Quadruped Navigation5h◆Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making5h◆SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks?5h◆Did We Actually Fix It? An Independent Adversarial Stress-Test of Post-Point-Adjustment Evaluation Metrics for Time-Series Anomaly Detection5h◆Supervised Reward Inference5h◆Semi-Supervised Conditional Generative Learning through Stochastic Interpolation and Sufficient Representations5h◆PPO-HSC: An Exploratory Reinforcement Learning Framework Based on Wide-Area Policy Coverage Optimization5h◆A Survey on the Verification of Reinforcement Learning Policies5h◆RECON: Benchmarking Agent Memory for Compositional Reasoning over Long Contexts5h◆Lomekwi: Resource-Bounded Tool Discovery in LLM Agents5h◆Expected Free Energy as Belief-Dependent Utility for rho-POMDPs5h◆Constrained Path Reasoning: Measuring When Committed Stages Earn Their Cost5h◆Pailitao-MMSearch: Building Native E-Commerce Multimodal Search Foundation5h◆Is Progressive Disclosure All You Need for Long-Context Agents?5h◆Verify, Repair, Repeat, or Stop? Robust Stopping for Noisy Verify-Repair Loops in LLM Agents5h◆Hierarchical Wireless Foundation Model for Multi-Task Optimization5h◆Auto Research for Materials: Auditable AI-Scientist Workflows with Held-Out Transfer5h◆It Depends on the Dataset: When a Brain-Encoding Model's Predicted Responses Beat Their Visual Backbone for Video Memorability5h◆A${}^2$BM: Alignment-Aware Bridge Matching for Image-to-Image Translation5h◆
News/Accelerating A/B-Tests with Counterfactual Estimation: Reducing Variance through Policy Overlap
arxiv
PublishedJuly 18, 2026 at 4:00 AM
—neutral

Accelerating A/B-Tests with Counterfactual Estimation: Reducing Variance through Policy Overlap

Source
arxiv.orgfull article ↗
Read on arxiv→
Publisher summary· verbatim

arXiv:2607.14604v1 Announce Type: new Abstract: Online controlled experiments are the gold standard for hypothesis testing in online platforms. Notwithstanding their ubiquity, they are notoriously expensive to run, and issues of variance hamper statistical power in assessing treatment effects. While

Stay posted· Newsletter

A 5-min weekly brief — top movers, price watch, story of the week.

// no spam · unsubscribe one-click · free forever

Discussion
Source
↗
arxiv
Read original ↗All from arxiv →

No replies yet. Be first.

Source
↗
arxiv
Read original ↗All from arxiv →

Related coverage

More from ARXIV
arxivCapacity and Redundancy Trade-offs in Multi-Task Learning5harxivPredictive Training with Latent Imagination for Visual Quadruped Navigation5harxivWhere Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making5harxivSEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks?5h
The Bubble Brief
WEEKLY

Read AI insights every Tuesday — top movers, new releases, story of the week.

// no spam · unsubscribe one-click · free forever

Originally published on arxiv ↗
HomeModelsNews