·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
AWS is helping vibe-coding startup Superblocks, and the implications are big16d◆Design Arena creators raise $7.9 million to bring taste to AI models16d◆Influencers draw backlash for attending OpenAI’s first luxury trip16d◆Apple finally fixed Siri. So why does it feel anticlimactic?16d◆Trump’s AI protectionism has come for robotics16d◆Europe’s AI labeling and transparency rules are now in effect16d◆Congress’ favorite AI tool? ChatGPT16d◆China’s Alibaba takes another swipe at America’s AI supremacy16d◆A Marc Benioff-backed startup thinks AI can solve the AI deployment problem16d◆Here’s why AI agents lie and cheat to reach their goals16d◆How we built a realtime system for responsive voice AI in six months16d◆Can AI Evaluate AI Scientists? A Benchmarking Study of Autonomous Research Generation Systems Using Automated Multi-Model Review16d◆ViSAGE: Constructing Self-Correcting Memories for Long-Form Video Understanding16d◆Library Reachability in LSR-Synth: How Anti-Memorization Design Changes the Measurement of Symbolic Discovery16d◆Self-Play Meets Skill Evolution: Self-Evolving Search Agents that Pose, Solve, and Remember16d◆AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction16d◆DungeonBench: A Benchmark for Rules-Rich Tactical Reasoning in Dungeons & Dragons Combat16d◆ExtractBench: A Benchmark for Schema-Guided Enterprise Document Extraction16d◆Scaffolding Critical Engagement with GenAI: Transforming Ethnic Minority Preparatory Students' Collaborative Discourse in Prompt Engineering Tasks16d◆Topology-Aware Data Movement for Disaggregated GPU Inference16d◆AWS is helping vibe-coding startup Superblocks, and the implications are big16d◆Design Arena creators raise $7.9 million to bring taste to AI models16d◆Influencers draw backlash for attending OpenAI’s first luxury trip16d◆Apple finally fixed Siri. So why does it feel anticlimactic?16d◆Trump’s AI protectionism has come for robotics16d◆Europe’s AI labeling and transparency rules are now in effect16d◆Congress’ favorite AI tool? ChatGPT16d◆China’s Alibaba takes another swipe at America’s AI supremacy16d◆A Marc Benioff-backed startup thinks AI can solve the AI deployment problem16d◆Here’s why AI agents lie and cheat to reach their goals16d◆How we built a realtime system for responsive voice AI in six months16d◆Can AI Evaluate AI Scientists? A Benchmarking Study of Autonomous Research Generation Systems Using Automated Multi-Model Review16d◆ViSAGE: Constructing Self-Correcting Memories for Long-Form Video Understanding16d◆Library Reachability in LSR-Synth: How Anti-Memorization Design Changes the Measurement of Symbolic Discovery16d◆Self-Play Meets Skill Evolution: Self-Evolving Search Agents that Pose, Solve, and Remember16d◆AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction16d◆DungeonBench: A Benchmark for Rules-Rich Tactical Reasoning in Dungeons & Dragons Combat16d◆ExtractBench: A Benchmark for Schema-Guided Enterprise Document Extraction16d◆Scaffolding Critical Engagement with GenAI: Transforming Ethnic Minority Preparatory Students' Collaborative Discourse in Prompt Engineering Tasks16d◆Topology-Aware Data Movement for Disaggregated GPU Inference16d◆
News/model/GLM-5.2-NVFP4

GLM-5.2-NVFP4 news

1 articles mentioning GLM-5.2-NVFP4

arxivApr 6bearish

AgentHazard: A Benchmark for Evaluating Harmful Behavior in Computer-Use Agents

arXiv:2604.02947v1 Announce Type: new Abstract: Computer-use agents extend language models from text generation to persistent action over tools, files, and execution environments. Unlike chat systems, they maintain state across interactions and translate intermediate outputs into concrete actions. T

#safety#benchmark#autonomous agents
HomeModelsNews