·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
Auto-FL-Research: Agentic Search for Federated Learning Algorithms21m◆The Wiola Architecture for Efficient Small Language Models21m◆When Should Service Agents Reconsider? Difficulty-Routed Control in Customer-Service Operations21m◆CreativityNeuro: Steering Language Model Weights to Improve Divergent Thinking and Reduce Mode Collapse21m◆Discrete Diffusion Language Models for Interactive Radiology Report Drafting21m◆Beyond Next-Token Prediction: An RLVR Proof of Concept for Tool-Use Agents on Atlassian Workflows21m◆World Feedback for Clinical Agents: Diagnosing RL in FHIR Environments21m◆Procedural Memory Distillation: Online Reflection for Self-Improving Language Models21m◆The Agentic Garden of Forking Paths21m◆Janus: a Playground for User-Involved Agentic Permission Management21m◆Revisiting Chain-of-Thought Reasoning under Limited Supervision: Semi-supervised Chain-of-Thought Learning21m◆OPINE-World: Programmatic World Modeling with Ontology-error-Prioritized Interactive Exploration21m◆Scaling Trends for Lie Detector Oversight in Preference Learning21m◆EO-Agents: A Three-Agent LLM Pipeline for Earth Observation Hypothesis Generation21m◆Hawk: Harnessing Hardware-Aware Knowledge for High-Performance NPU Kernel Generation21m◆Safe and Adaptive Cloud Healing: Verifying LLM-Generated Recovery Plans with a Neural-Symbolic World Model21m◆SemHash-LLM: A Multi-Granularity Semantic Hashing Framework for Document Deduplication21m◆Profit-Based Counterfactual Explanations for Product Improvement: A Case Study of Manga Sales in Japan21m◆Scaling with Confidence: Calibrating Confidence of LLMs for Adaptive Test Time Scaling21m◆Spatial Support Matters: Geometry-Aware Graph Fusion for Rainfall Field Reconstruction21m◆Auto-FL-Research: Agentic Search for Federated Learning Algorithms21m◆The Wiola Architecture for Efficient Small Language Models21m◆When Should Service Agents Reconsider? Difficulty-Routed Control in Customer-Service Operations21m◆CreativityNeuro: Steering Language Model Weights to Improve Divergent Thinking and Reduce Mode Collapse21m◆Discrete Diffusion Language Models for Interactive Radiology Report Drafting21m◆Beyond Next-Token Prediction: An RLVR Proof of Concept for Tool-Use Agents on Atlassian Workflows21m◆World Feedback for Clinical Agents: Diagnosing RL in FHIR Environments21m◆Procedural Memory Distillation: Online Reflection for Self-Improving Language Models21m◆The Agentic Garden of Forking Paths21m◆Janus: a Playground for User-Involved Agentic Permission Management21m◆Revisiting Chain-of-Thought Reasoning under Limited Supervision: Semi-supervised Chain-of-Thought Learning21m◆OPINE-World: Programmatic World Modeling with Ontology-error-Prioritized Interactive Exploration21m◆Scaling Trends for Lie Detector Oversight in Preference Learning21m◆EO-Agents: A Three-Agent LLM Pipeline for Earth Observation Hypothesis Generation21m◆Hawk: Harnessing Hardware-Aware Knowledge for High-Performance NPU Kernel Generation21m◆Safe and Adaptive Cloud Healing: Verifying LLM-Generated Recovery Plans with a Neural-Symbolic World Model21m◆SemHash-LLM: A Multi-Granularity Semantic Hashing Framework for Document Deduplication21m◆Profit-Based Counterfactual Explanations for Product Improvement: A Case Study of Manga Sales in Japan21m◆Scaling with Confidence: Calibrating Confidence of LLMs for Adaptive Test Time Scaling21m◆Spatial Support Matters: Geometry-Aware Graph Fusion for Rainfall Field Reconstruction21m◆
News/The Open Evaluation Standard: Benchmarking NVIDIA Nemotron 3 Nano with NeMo Evaluator
huggingface
PublishedDecember 17, 2025 at 1:22 PM

The Open Evaluation Standard: Benchmarking NVIDIA Nemotron 3 Nano with NeMo Evaluator

Source
huggingface.cofull article ↗
Read on huggingface→
Stay posted· Newsletter

A 5-min weekly brief — top movers, price watch, story of the week.

// no spam · unsubscribe one-click · free forever

Discussion
Source
↗
huggingface
Read original ↗All from huggingface →

No replies yet. Be first.

Source
↗
huggingface
Read original ↗All from huggingface →
The Bubble Brief
WEEKLY

Read AI insights every Tuesday — top movers, new releases, story of the week.

// no spam · unsubscribe one-click · free forever

Originally published on huggingface ↗
HomeModelsNews