·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
AWS is helping vibe-coding startup Superblocks, and the implications are big16d◆Design Arena creators raise $7.9 million to bring taste to AI models16d◆Influencers draw backlash for attending OpenAI’s first luxury trip16d◆Apple finally fixed Siri. So why does it feel anticlimactic?16d◆Trump’s AI protectionism has come for robotics16d◆Europe’s AI labeling and transparency rules are now in effect16d◆Congress’ favorite AI tool? ChatGPT16d◆China’s Alibaba takes another swipe at America’s AI supremacy16d◆A Marc Benioff-backed startup thinks AI can solve the AI deployment problem16d◆Here’s why AI agents lie and cheat to reach their goals16d◆How we built a realtime system for responsive voice AI in six months16d◆Can AI Evaluate AI Scientists? A Benchmarking Study of Autonomous Research Generation Systems Using Automated Multi-Model Review17d◆ViSAGE: Constructing Self-Correcting Memories for Long-Form Video Understanding17d◆Library Reachability in LSR-Synth: How Anti-Memorization Design Changes the Measurement of Symbolic Discovery17d◆Self-Play Meets Skill Evolution: Self-Evolving Search Agents that Pose, Solve, and Remember17d◆AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction17d◆DungeonBench: A Benchmark for Rules-Rich Tactical Reasoning in Dungeons & Dragons Combat17d◆ExtractBench: A Benchmark for Schema-Guided Enterprise Document Extraction17d◆Scaffolding Critical Engagement with GenAI: Transforming Ethnic Minority Preparatory Students' Collaborative Discourse in Prompt Engineering Tasks17d◆Topology-Aware Data Movement for Disaggregated GPU Inference17d◆AWS is helping vibe-coding startup Superblocks, and the implications are big16d◆Design Arena creators raise $7.9 million to bring taste to AI models16d◆Influencers draw backlash for attending OpenAI’s first luxury trip16d◆Apple finally fixed Siri. So why does it feel anticlimactic?16d◆Trump’s AI protectionism has come for robotics16d◆Europe’s AI labeling and transparency rules are now in effect16d◆Congress’ favorite AI tool? ChatGPT16d◆China’s Alibaba takes another swipe at America’s AI supremacy16d◆A Marc Benioff-backed startup thinks AI can solve the AI deployment problem16d◆Here’s why AI agents lie and cheat to reach their goals16d◆How we built a realtime system for responsive voice AI in six months16d◆Can AI Evaluate AI Scientists? A Benchmarking Study of Autonomous Research Generation Systems Using Automated Multi-Model Review17d◆ViSAGE: Constructing Self-Correcting Memories for Long-Form Video Understanding17d◆Library Reachability in LSR-Synth: How Anti-Memorization Design Changes the Measurement of Symbolic Discovery17d◆Self-Play Meets Skill Evolution: Self-Evolving Search Agents that Pose, Solve, and Remember17d◆AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction17d◆DungeonBench: A Benchmark for Rules-Rich Tactical Reasoning in Dungeons & Dragons Combat17d◆ExtractBench: A Benchmark for Schema-Guided Enterprise Document Extraction17d◆Scaffolding Critical Engagement with GenAI: Transforming Ethnic Minority Preparatory Students' Collaborative Discourse in Prompt Engineering Tasks17d◆Topology-Aware Data Movement for Disaggregated GPU Inference17d◆
News/model/GLM-5.1-NVFP4

GLM-5.1-NVFP4 news

1 articles mentioning GLM-5.1-NVFP4

arxivApr 6bearish

AgentHazard: A Benchmark for Evaluating Harmful Behavior in Computer-Use Agents

arXiv:2604.02947v1 Announce Type: new Abstract: Computer-use agents extend language models from text generation to persistent action over tools, files, and execution environments. Unlike chat systems, they maintain state across interactions and translate intermediate outputs into concrete actions. T

#safety#benchmark#autonomous agents
HomeModelsNews