huggingface1d ago
Source
HuggingFace
100 articles indexed from HuggingFace
huggingface1d ago
IBM releases SOTA Granite Time Series PatchTST-FM-r2 model with commercial-friendly license
huggingface2d ago
Safety for Whom? Refusing the Right Subset of a Topic, Not the Whole Topic
huggingfaceSep 3
NeoMME: an efficient Multimodal-native and Multilingual Encoder
huggingfaceSep 3
Training a coding model to paint watercolours with TRL and OpenEnv
huggingfaceSep 3
Give Your Coding Agents a Memory You Own
huggingfaceSep 3
Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps
huggingfaceSep 2
Real-Time Intelligence with IBM Time Series Models on Confluent
huggingfaceSep 1
BenchMIRT: What are LLM benchmarks actually measuring?
huggingfaceSep 1
Introducing @huggingface/kernels: 200+ WebGPU Kernels for Local AI
huggingfaceAug 28
The Open ASR Leaderboard Adds Its First Global South Language
huggingfaceAug 26
Training and Finetuning Multi-Vector Embedding Models with Sentence Transformers
huggingfaceAug 25
Granite 4.2 LLMs: How They're Built
huggingfaceAug 25
Quantization-Aware Healing: a compressed, 4-bit model that outperforms its full-precision original
huggingfaceAug 25
Wire It, Run It, Deploy It: AI Workflows in Gradio
huggingfaceAug 21
How Hugging Face Inference Endpoints, Jobs, and Buckets Power Search on Papers with Code
huggingfaceAug 21
Measuring benchmark optimization in speech recognition
huggingfaceAug 20
Up to 3.2x Faster Inference with LFM2.5-DSpark
huggingfaceAug 18
How Much Memory Does Your Agent Actually Need?
huggingfaceAug 18
Multi-Vector (Late Interaction) Embedding Models with Sentence Transformers
huggingfaceAug 17
Same Cluster, 33 Points More Utilization: What Changed Was the Order
huggingfaceAug 14
State of Open Models: Summer 2026 Observations
huggingfaceAug 13
Record, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets
huggingfaceAug 13
What We Learned by Reproducing 2,200 papers from ICML
huggingfaceAug 12
Introducing OlmoEarth embeddings: Custom embedding exports from OlmoEarth Studio for downstream analysis
huggingfaceAug 11
Thinking of ACE? We Can Do It with Fewer Tokens
huggingfaceAug 10
Build Low-Latency Multilingual Voice Agents: Open Weights & Full Deployment Control with NVIDIA Magpie TTS
huggingfaceAug 10
Making Knowledge Distillation Cheap Enough to Run at Scale
huggingfaceAug 10
Meta is back with Muse Glimmer: local, agentic, multimodal, and open source
huggingfaceAug 6
Baseten on Hugging Face Inference Providers 🔥
huggingfaceJul 30
GPU Management: Why Idle GPUs Are the New Grounded Aircraft
huggingfaceJul 28
The OlmoEarth Platform: Geospatial inference at planetary scale
huggingfaceJul 28
LFM2.5-Encoders for Fast Long-Context Inference on CPU
huggingfaceJul 27
NVIDIA Cosmos-H-Dreams: Bringing Real-Time Generative Simulation to Surgical Robotics
huggingfaceJul 27
Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident
huggingfaceJul 23
Bringing Nunchaku 4-bit Diffusion Inference to Diffusers
huggingfaceJul 21
The State of Simulation for Physical AI: An Overview
huggingfaceJul 21
Grabette: an open system to record robot-manipulation data
huggingfaceJul 20
Introducing Cosmos 3 Edge
huggingfaceJul 17
Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 Diffusers
huggingfaceJul 16bullish
NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval
huggingfaceJul 16
Newer Models, Same Advantage
huggingfaceJul 16
Security incident disclosure — July 2026
huggingfaceJul 15
What building Shippy taught us about building agents
huggingfaceJul 15
Model Routing Is Simple. Until It Isn’t.
huggingfaceJul 15
Welcome Inkling by Thinking Machines
huggingfaceJul 15
Introducing Real World VoiceEQ: Measuring the human quality of voice AI
huggingfaceJul 10
Profiling in PyTorch (Part 3): Attention is all you profile
huggingfaceJul 8
Data for Agents
huggingfaceJul 8
Native-speed vLLM transformers modeling backend
huggingfaceJul 7
From Hugging Face to Amazon SageMaker Studio in one click
huggingfaceJul 7
Hugging Face Models on Foundry Managed Compute
huggingfaceJul 7
LeRobot v0.6.0: Imagine, Evaluate, Improve
huggingfaceJul 7
Run AI workloads on any cloud, store on Hugging Face: zero-egress storage with SkyPilot
huggingfaceJul 6
PRX Part 4: Our Data Strategy
huggingfaceJul 6
🤗 Kernels: Major Updates
huggingfaceJul 1
Hugging Face and Cerebras bring Gemma 4 to real-time voice AI
huggingfaceJun 30
ScarfBench: Benchmarking AI Agents for Enterprise Java Framework Migration
huggingfaceJun 30
Why Specialization Is Inevitable
huggingfaceJun 30
Featuring Every Eval Ever Results on Hugging Face Model Pages
huggingfaceJun 29
DiScoFormer: One transformer for density and score, across distributions
huggingfaceJun 26
Run a vLLM Server on HF Jobs in One Command
huggingfaceJun 25
Which tokens does a hybrid model predict better?
huggingfaceJun 24
Accelerating Transformers Fine-Tuning with NVIDIA NeMo AutoModel
huggingfaceJun 24
Introducing the FFASR Leaderboard: Benchmarking ASR in the Real World
huggingfaceJun 23
Build real agentic apps using CUGA: two dozen working examples on a lightweight harness
huggingfaceJun 23
Shipping huggingface_hub every week with AI, open tools, and a human in the loop
huggingfaceJun 23
Experimenting with the proposed Cross-Origin Storage API in Transformers.js
huggingfaceJun 22
PP-OCRv6 on Hugging Face: 50-Language OCR from 1.5M to 34.5M Parameters
huggingfaceJun 22
We got local models to triage the OpenClaw repo for FREE!*
huggingfaceJun 18
MosaicLeaks: Can your research agent keep a secret?
huggingfaceJun 18
Is it agentic enough? Benchmarking open models on your own tooling
huggingfaceJun 18
Beyond LoRA: Can you beat the most popular fine-tuning technique?
huggingfaceJun 17
MolmoMotion: Language-guided 3D motion forecasting
huggingfaceJun 17
From the Hugging Face Hub to robot hardware with Strands Agents and LeRobot
huggingfaceJun 17bullish
GLM-5.2: Built for Long-Horizon Tasks
huggingfaceJun 17
Agentic Resource Discovery: Let agents search
huggingfaceJun 12
olmo-eval: An evaluation workbench for the model development loop
huggingfaceJun 11
Profiling in PyTorch (Part 2): From nn.Linear to a Fused MLP
huggingfaceJun 9
Can Voice Agents Handle Bilingual Customers? Benchmarking Frontier ASR on Code-Switched Speech
huggingfaceJun 9
Introducing North Mini Code: Cohere’s First Model For Developers
huggingfaceJun 9
How an Agent Built a 3D Paris Gallery by Chaining Two Hugging Face Spaces
huggingfaceJun 9
Migrating Your GitHub CI to Hugging Face Jobs
huggingfaceJun 8
The Open Source Community is backing OpenEnv for Agentic RL
huggingfaceJun 6
Five labs, five minds: building a multi-model finance drama on small models
huggingfaceJun 6
Job Searcher
huggingfaceJun 6
Persona Atlas: Mapping How Famous Minds Think
huggingfaceJun 5
Thousand Token Wood: shipping a multi-agent economy on a 3B model
huggingfaceJun 4
Nemotron 3.5 Content Safety: Customizable Multimodal Safety for Global Enterprise AI
huggingfaceJun 4
How to Fine-Tune Nemotron 3.5 ASR for Your Language, Domain, or Accent
huggingfaceJun 4
EVA-Bench Data 2.0: 3 Domains, 121 Tools, 213 Scenarios
huggingfaceJun 4
Task-Seeded Synthetic Q&A Generation for Nemotron Pretraining
huggingfaceJun 4
Designing the hf CLI as an agent-optimized way to work with the Hub
huggingfaceJun 3
Direct Preference Optimization Beyond Chatbots
huggingfaceJun 3
Adding MCP Tools to Reachy Mini
huggingfaceJun 2
Holo3.1: Fast & Local Computer Use Agents
huggingfaceJun 1
Introducing Mellum2: A 12B Mixture-of-Experts Model by JetBrains
huggingfaceJun 1
Beyond LLMs: Why Scalable Enterprise AI Adoption Depends on Agent Logic
huggingfaceJun 1
Welcome NVIDIA Cosmos 3: The First Open Omni-model for Physical AI Reasoning and Action
huggingfaceMay 29