·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
The risk of weather data sabotage is rising3h◆IMEX Interaction-Based Model Explanation8h◆DialogueVPR: Towards Conversational Visual Place Recognition8h◆Human AI Construction of Bayesian Networks for Operational Decision Support -- A Virtual Survey Approach8h◆Orchestrating Power Grid Studies with Multi-Agent AI and MCP Servers8h◆MemoHarness: Agent Harnesses That Learn from Experience8h◆A Comparative Analysis of Machine Learning Models for Long and Short-Term Forecasting of the Egyptian Stock Market: A Focus on EGX308h◆CatalogAgent: A Supervisor-mediated Self-Learning System Enabling Context Engineering for GenAI Models8h◆Reward-Free Evolving Agents via Pairwise Validator8h◆Towards an Intention Abstraction Layer for Autonomous Industrial Systems8h◆Seeing the End at Step Zero: Accelerating Diffusion MLLMs via MLP Sparsity-Aware Truncation8h◆SportD: Can VLMs Physically Strategize?8h◆Action QFormer: Structured Representation Shaping under Action Supervision in Vision-Language-Action Models8h◆SmartRAG: Native Graph-Based RAG for Mobile Device8h◆InCarEmo: A Multimodal Dataset for In-Cabin Emotion Recognition and Driver State Monitoring8h◆NexForge: Scaling Executable Agent Tasks via Requirement-First Synthesis8h◆Instant NuRec: Feed-Forward 3D Gaussian Reconstruction for Driving Scene Simulation8h◆EdgeFaaS: A Function-based Framework for Edge Computing8h◆SafeRelBench: A Spatial-Relation-Aware Benchmark for Process-Level Safety in VLM-Driven Embodied Agents8h◆Answer-Conditioned Chains of Thought Degrade Verifiable-Reasoning Distillation in Large Language Models8h◆The risk of weather data sabotage is rising3h◆IMEX Interaction-Based Model Explanation8h◆DialogueVPR: Towards Conversational Visual Place Recognition8h◆Human AI Construction of Bayesian Networks for Operational Decision Support -- A Virtual Survey Approach8h◆Orchestrating Power Grid Studies with Multi-Agent AI and MCP Servers8h◆MemoHarness: Agent Harnesses That Learn from Experience8h◆A Comparative Analysis of Machine Learning Models for Long and Short-Term Forecasting of the Egyptian Stock Market: A Focus on EGX308h◆CatalogAgent: A Supervisor-mediated Self-Learning System Enabling Context Engineering for GenAI Models8h◆Reward-Free Evolving Agents via Pairwise Validator8h◆Towards an Intention Abstraction Layer for Autonomous Industrial Systems8h◆Seeing the End at Step Zero: Accelerating Diffusion MLLMs via MLP Sparsity-Aware Truncation8h◆SportD: Can VLMs Physically Strategize?8h◆Action QFormer: Structured Representation Shaping under Action Supervision in Vision-Language-Action Models8h◆SmartRAG: Native Graph-Based RAG for Mobile Device8h◆InCarEmo: A Multimodal Dataset for In-Cabin Emotion Recognition and Driver State Monitoring8h◆NexForge: Scaling Executable Agent Tasks via Requirement-First Synthesis8h◆Instant NuRec: Feed-Forward 3D Gaussian Reconstruction for Driving Scene Simulation8h◆EdgeFaaS: A Function-based Framework for Edge Computing8h◆SafeRelBench: A Spatial-Relation-Aware Benchmark for Process-Level Safety in VLM-Driven Embodied Agents8h◆Answer-Conditioned Chains of Thought Degrade Verifiable-Reasoning Distillation in Large Language Models8h◆
News/DeepSeekMath Meets Order Book: Group-Aware Policy Optimization for High-Frequency Directional Trading
arxiv
PublishedMay 26, 2026 at 4:00 AM

DeepSeekMath Meets Order Book: Group-Aware Policy Optimization for High-Frequency Directional Trading

Source
arxiv.orgfull article ↗
Read on arxiv→
Publisher summary· verbatim

arXiv:2605.25527v1 Announce Type: new Abstract: This paper studies reinforcement learning for high-frequency trading on limit order books by pairing an Order-Flow-based state model with policy-gradient methods. Instead of value-based RL techniques like tabular Q-learning, our approach deploys policy

Stay posted· Newsletter

A 5-min weekly brief — top movers, price watch, story of the week.

// no spam · unsubscribe one-click · free forever

Discussion
Source
↗
arxiv
Read original ↗All from arxiv →

No replies yet. Be first.

Source
↗
arxiv
Read original ↗All from arxiv →

Related coverage

More from ARXIV
arxivIMEX Interaction-Based Model Explanation8harxivDialogueVPR: Towards Conversational Visual Place Recognition8harxivHuman AI Construction of Bayesian Networks for Operational Decision Support -- A Virtual Survey Approach8harxivOrchestrating Power Grid Studies with Multi-Agent AI and MCP Servers8h
The Bubble Brief
WEEKLY

Read AI insights every Tuesday — top movers, new releases, story of the week.

// no spam · unsubscribe one-click · free forever

Originally published on arxiv ↗
HomeModelsNews