·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
Bringing Value Models Back: Generative Critics for Value Modeling in LLM Reinforcement Learning7h◆Subagents vs Agent Skills: Executing Reusable Knowledge for Long-Horizon Agentic Tasks7h◆Distribution-Consistent Inference for Dynamic Sparse Mixture-of-Experts7h◆In RAG We Trust? Measuring Robustness of Retrieval-Augmented Generation Under Document Poisoning7h◆FrontierChallenge: Evaluating Scientific Workflow Completion7h◆IBIB: A Protocol for Measuring Enterprise AI Systems by Serving Route, Not Model Identifier7h◆AgenticGen: Reward-Guided Agentic Video Generation for Advertising7h◆Strangers to Themselves: What Language Models Say About Themselves Is Generic7h◆Omni Interaction Agent Technical Report7h◆CoGReV: A Confidence-Gated Post-Hoc Non-Monotonic Belief Revision Framework for Phishing Website Classification7h◆Bit-Flip Attacks on Vision-Language-Action Models: Action-Decoding Architecture Shapes the Vulnerability7h◆Can Foundation Models Moderate Online Content? Evaluating Instruction- vs. Example-Driven Policy Operationalization7h◆RiLM: Parameter-Efficient Language Modeling via Geodesic Decoding7h◆Tracing Computation Density in LLMs7h◆Cultural Binding Heads in Language Models7h◆Direct Diversity Optimization for Diverse Successful Trajectories in Preference Post-Training7h◆Left-Branching Transformers Excel at Right-Branching Languages: Data Shapes Word Order Preferences in Language Models7h◆Palmyra x6 Technical Report: An Agentic, Tool-Use Model Post-Trained via Anchored Supervised Fine-Tuning7h◆'Ghaib in Translation' aka Unseen Harm: Measuring Cross-Script Safety Inconsistency with 'Missed-in-Urdu' Scores in LLM Hate Speech Detection7h◆DexterSQL: Deep Schema Exploration and Rule-based Correction for Text-to-SQL Generation7h◆Bringing Value Models Back: Generative Critics for Value Modeling in LLM Reinforcement Learning7h◆Subagents vs Agent Skills: Executing Reusable Knowledge for Long-Horizon Agentic Tasks7h◆Distribution-Consistent Inference for Dynamic Sparse Mixture-of-Experts7h◆In RAG We Trust? Measuring Robustness of Retrieval-Augmented Generation Under Document Poisoning7h◆FrontierChallenge: Evaluating Scientific Workflow Completion7h◆IBIB: A Protocol for Measuring Enterprise AI Systems by Serving Route, Not Model Identifier7h◆AgenticGen: Reward-Guided Agentic Video Generation for Advertising7h◆Strangers to Themselves: What Language Models Say About Themselves Is Generic7h◆Omni Interaction Agent Technical Report7h◆CoGReV: A Confidence-Gated Post-Hoc Non-Monotonic Belief Revision Framework for Phishing Website Classification7h◆Bit-Flip Attacks on Vision-Language-Action Models: Action-Decoding Architecture Shapes the Vulnerability7h◆Can Foundation Models Moderate Online Content? Evaluating Instruction- vs. Example-Driven Policy Operationalization7h◆RiLM: Parameter-Efficient Language Modeling via Geodesic Decoding7h◆Tracing Computation Density in LLMs7h◆Cultural Binding Heads in Language Models7h◆Direct Diversity Optimization for Diverse Successful Trajectories in Preference Post-Training7h◆Left-Branching Transformers Excel at Right-Branching Languages: Data Shapes Word Order Preferences in Language Models7h◆Palmyra x6 Technical Report: An Agentic, Tool-Use Model Post-Trained via Anchored Supervised Fine-Tuning7h◆'Ghaib in Translation' aka Unseen Harm: Measuring Cross-Script Safety Inconsistency with 'Missed-in-Urdu' Scores in LLM Hate Speech Detection7h◆DexterSQL: Deep Schema Exploration and Rule-based Correction for Text-to-SQL Generation7h◆
DataBubble·

Model Detail

nvidia logo

NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4

—
Provider: NVIDIACategory: codePipeline: text-generationParameters: 120B
DB Score
0.8
Downloads
3.0M
Likes
409
Day
+0.0%
Week
+0.0%
Month
+0.0%
Overview

NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4 is a code generation model with 120B parameters released by NVIDIA. The model is registered under the text-generation pipeline tag on Hugging Face, distributed under a other license.

Technical

NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4 ships with 120B parameters. Total weight footprint is approximately 67.2 GB, which is the relevant figure when planning local-inference VRAM. Distribution is governed by the other license — review the exact terms before commercial deployment.

Use Cases

NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4 is best fit for code completion, repository-scale Q&A, and pair-programming integrations. It is a less obvious choice for one-shot generation of security-critical code without review. Treat this as a starting matrix rather than a benchmark verdict — the right deployment usually depends on the specific evaluation suite that mirrors your workload.

Download History
Research Paper
arXiv: 2512.20848→
Model Info
Licenseother
Citations56 (7 influential)
Recent newsView all news →
Related News
techcrunchneutral13h ago

Jensen Huang explains why Nvidia will grow an astounding 70% next year

Nvidia has its finger in every pie, and sees another year of plenty in its future, Jensen Huang says. But, he insists, its deals are not circular.

arxiv1d ago

Learning Exact NVIDIA SASS Encoders with $\mathbb{F}_2$ Linear Algebra

arXiv:2608.20532v2 Announce Type: replace Abstract: NVIDIA provides a SASS disassembler but no public SASS assembler for recent data-center GPUs, limiting controlled machine-code rewriting. We present F2Asm, which learns exact 128-bit SASS encoders from paired disassembly and original CUBIN instruct

techcrunchneutral6d ago

Apple’s Ternus era begins as Nvidia bets on the whole AI stack

It’s officially the Ternus era at Apple. Tim Cook stepped down as CEO this week, handing the company to former hardware chief John Ternus, whose first memo promised a “huge launch next week” — timing that puts Apple’s next iPhone event on his desk before he’s even settled in. Cook isn’t going far, t

Nvidia launches free tool that links idle computers into a personal AI data center
thevergeneutral7d ago

Nvidia launches free tool that links idle computers into a personal AI data center

Nvidia is announcing its new Personal AI Router (PAIR), a free tool that syncs up your home computers for tackling local AI inference tasks with tools like Ollama and LM Studio. Let's get the obvious thing out of the way, despite what its name might imply: PAIR is not a hardware router. It's open-so

techcrunchneutral7d ago

Nvidia confirms it will buy Hugging Face for $12.9 billion

Nvidia said Hugging Face hosts over 3 million models and is used by over 18 million developers.

Nvidia is buying Hugging Face for almost $13 billion
thevergeneutral7d ago

Nvidia is buying Hugging Face for almost $13 billion

Nvidia has agreed to buy Hugging Face for $12.93 billion, bringing one of the most popular hosting platforms for open-source AI models, datasets, and tools under the ownership of the world's biggest AI chipmaker. Hugging Face is an online platform founded in 2016 that gives AI developers a space to

Related Models
nvidia logo
Qwen3.6-35B-A3B-NVFP4
NVIDIA · 11.2M downloads
nvidia logo
Gemma-4-31B-IT-NVFP4
NVIDIA · 2.3M downloads
sentence-transformers logo
all-MiniLM-L6-v2
SBERT · 254.3M downloads
nomic-ai logo
nomic-embed-text-v1.5
nomic-ai · 17.1M downloads
HomeModelsNews