DataBubble
Home
Models
News
Compare
Boards
Pricing
About
Newsletter
Methodology
Contact
‹ collapse
Latest
China delivers a one-two punch to America’s AI dominance
5h
◆
AI is more likely than humans to form biases when hiring
6h
◆
Beyond a Single Direction: Chain-of-Thought Disrupts Simple Steering of Refusal
11h
◆
SkillCorpus: Consolidating and Evaluating the Open Skill Ecosystem for Real-World LLM Agents
11h
◆
PolyInterview: An LLM-based Platform for Immersive Mock Interview Practice with Comprehensive Multimodal Assessment
11h
◆
Latency-Response Theory Model: Evaluating Large Language Models via Response Accuracy and Chain-of-Thought Length
11h
◆
Label-Free Concept Drift Assessment for Reliable AI in Emerging Wireless Applications
11h
◆
DyneTrion: A Spatio-temporally Coherent Generative Emulator for Protein Dynamics Across Timescales
11h
◆
Improving Neural Network Training by Decoupling the Magnitude and Direction of Weight Vectors
11h
◆
CRAFT: Clustering Rubrics to Diagnose Weak LLM Capabilities and Generate Targeted Fine-Tuning Data
11h
◆
cGAP: Generalized Association Plots with HOMALS-Guided Heatmaps for Visualization of High-Dimensional Categorical Data
11h
◆
Do Coding Agents Need Executable World Models, Simplification, and Verification to Solve ARC-AGI-3?
11h
◆
Beyond a Joke: Multi-Angle Reasoning for Detecting and Explaining Harmful Humor in Memes
11h
◆
From Black Box to Executable Logic: Explainable Reinforcement Learning through Prolog Expert Systems
11h
◆
Closing the AI Trust Gap: The Case for Independent Certification for Trustworthy AI
11h
◆
The AI Fiction Paradox
11h
◆
Digital Pantheon: Simulating and Auditing Coalition Formation with LLM Agents
11h
◆
EpiNarrate: Agentic Generation of Grounded Narratives from Epidemiological Scenario Projections
11h
◆
Before the Action: Benchmarking LLMs on Prospective Hypothesis Discovery
11h
◆
How Much Human Label Variation Does Formal Semantic Structure Explain?: Group-Level Effects and Item-Level Ceilings in NLI
11h
◆
China delivers a one-two punch to America’s AI dominance
5h
◆
AI is more likely than humans to form biases when hiring
6h
◆
Beyond a Single Direction: Chain-of-Thought Disrupts Simple Steering of Refusal
11h
◆
SkillCorpus: Consolidating and Evaluating the Open Skill Ecosystem for Real-World LLM Agents
11h
◆
PolyInterview: An LLM-based Platform for Immersive Mock Interview Practice with Comprehensive Multimodal Assessment
11h
◆
Latency-Response Theory Model: Evaluating Large Language Models via Response Accuracy and Chain-of-Thought Length
11h
◆
Label-Free Concept Drift Assessment for Reliable AI in Emerging Wireless Applications
11h
◆
DyneTrion: A Spatio-temporally Coherent Generative Emulator for Protein Dynamics Across Timescales
11h
◆
Improving Neural Network Training by Decoupling the Magnitude and Direction of Weight Vectors
11h
◆
CRAFT: Clustering Rubrics to Diagnose Weak LLM Capabilities and Generate Targeted Fine-Tuning Data
11h
◆
cGAP: Generalized Association Plots with HOMALS-Guided Heatmaps for Visualization of High-Dimensional Categorical Data
11h
◆
Do Coding Agents Need Executable World Models, Simplification, and Verification to Solve ARC-AGI-3?
11h
◆
Beyond a Joke: Multi-Angle Reasoning for Detecting and Explaining Harmful Humor in Memes
11h
◆
From Black Box to Executable Logic: Explainable Reinforcement Learning through Prolog Expert Systems
11h
◆
Closing the AI Trust Gap: The Case for Independent Certification for Trustworthy AI
11h
◆
The AI Fiction Paradox
11h
◆
Digital Pantheon: Simulating and Auditing Coalition Formation with LLM Agents
11h
◆
EpiNarrate: Agentic Generation of Grounded Narratives from Epidemiological Scenario Projections
11h
◆
Before the Action: Benchmarking LLMs on Prospective Hypothesis Discovery
11h
◆
How Much Human Label Variation Does Formal Semantic Structure Explain?: Group-Level Effects and Item-Level Ceilings in NLI
11h
◆
DataBubble
·
Best Math Models
0 models
all
llm
image
audio
code
#
Model
Score
Downloads
Provider
Home
Models
News
More