·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
Evaluating Reliability in Machine Learning Models for Early Chronic Kidney Disease Prediction: A Systematic Review of Data Leakage and Predictor Stability5h◆Learning to Discretize: Diffusion-Based Adaptive Mesh with Spectral Guidance5h◆Toward Trustworthy Autonomous Science: A Two-Year Community Roadmap5h◆The Benjamini--Hochberg Procedure Can Fail to Control the FDR for Correlated Two-Sided Gaussian Tests5h◆Mind the Gap: Promises and Pitfalls of Hierarchical Planning in LeWorldModel5h◆What Does Goodness Measure? A Likelihood-Ratio Account of Forward-Forward Learning5h◆Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs5h◆AutoTrace: From Patches to Triggers via Agentic Interprocedural Exploration5h◆Traceback Translators Against Forgetting in Continual Fake Speech Detection5h◆Deep Learning-based Surrogate Modelling of the LOD Method for Multiscale Problems5h◆Explainable-by-Design Audio Deepfake Detection via Wiener-Hopf Linear Prediction5h◆Multi-Perspective Agentic Program Repair via Code Property Graphs and Temporal Execution Graphs5h◆DECO: Decoupled Multimodal Diffusion Transformer for Bimanual Dexterous Manipulation with a Plugin Tactile Adapter5h◆Neuro-Symbolic ODE Discovery with Latent Grammar Flow5h◆dMX: Differentiable Mixed-Precision Assignment for Low-Precision Floating-Point Formats5h◆Does Topic Sentiment Cause Perceived Ideology? Comparing Human and LLM Annotations in Political News Articles5h◆ISE: An Execution-Grounded Recipe for Multi-Turn OS-Agent Trajectories5h◆Atlas H&E-TME: Scalable AI-Based Tissue Profiling at Expert Pathologist-Level Accuracy5h◆Keep Policy Gradient in Charge: Sibling-Guided Credit Distillation for Long-Horizon Tool-Use Agents5h◆Prime Fourier Embeddings: A Principled Basis for Modular Arithmetic5h◆Evaluating Reliability in Machine Learning Models for Early Chronic Kidney Disease Prediction: A Systematic Review of Data Leakage and Predictor Stability5h◆Learning to Discretize: Diffusion-Based Adaptive Mesh with Spectral Guidance5h◆Toward Trustworthy Autonomous Science: A Two-Year Community Roadmap5h◆The Benjamini--Hochberg Procedure Can Fail to Control the FDR for Correlated Two-Sided Gaussian Tests5h◆Mind the Gap: Promises and Pitfalls of Hierarchical Planning in LeWorldModel5h◆What Does Goodness Measure? A Likelihood-Ratio Account of Forward-Forward Learning5h◆Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs5h◆AutoTrace: From Patches to Triggers via Agentic Interprocedural Exploration5h◆Traceback Translators Against Forgetting in Continual Fake Speech Detection5h◆Deep Learning-based Surrogate Modelling of the LOD Method for Multiscale Problems5h◆Explainable-by-Design Audio Deepfake Detection via Wiener-Hopf Linear Prediction5h◆Multi-Perspective Agentic Program Repair via Code Property Graphs and Temporal Execution Graphs5h◆DECO: Decoupled Multimodal Diffusion Transformer for Bimanual Dexterous Manipulation with a Plugin Tactile Adapter5h◆Neuro-Symbolic ODE Discovery with Latent Grammar Flow5h◆dMX: Differentiable Mixed-Precision Assignment for Low-Precision Floating-Point Formats5h◆Does Topic Sentiment Cause Perceived Ideology? Comparing Human and LLM Annotations in Political News Articles5h◆ISE: An Execution-Grounded Recipe for Multi-Turn OS-Agent Trajectories5h◆Atlas H&E-TME: Scalable AI-Based Tissue Profiling at Expert Pathologist-Level Accuracy5h◆Keep Policy Gradient in Charge: Sibling-Guided Credit Distillation for Long-Horizon Tool-Use Agents5h◆Prime Fourier Embeddings: A Principled Basis for Modular Arithmetic5h◆
News/Frame-Conditioned Moral Computation in LLaMA 3.1-8B-Instruct: A Mechanistic Interpretability Audit of Ethical Reasoning
arxiv
PublishedJune 16, 2026 at 4:00 AM
—neutral

Frame-Conditioned Moral Computation in LLaMA 3.1-8B-Instruct: A Mechanistic Interpretability Audit of Ethical Reasoning

Source
arxiv.orgfull article ↗
Read on arxiv→
Publisher summary· verbatim

arXiv:2606.15507v1 Announce Type: new Abstract: Behavioral audits of Large Language Models on moral prompts measure what the model says, not the internal computation producing it. We use Transluce, an AI-driven mechanistic-interpretability platform, to examine LLaMA 3.1-8B-Instruct on 54 moral promp

Stay posted· Newsletter

A 5-min weekly brief — top movers, price watch, story of the week.

// no spam · unsubscribe one-click · free forever

Discussion
Source
↗
arxiv
Read original ↗All from arxiv →

No replies yet. Be first.

Source
↗
arxiv
Read original ↗All from arxiv →

Related coverage

More from ARXIV
arxivEvaluating Reliability in Machine Learning Models for Early Chronic Kidney Disease Prediction: A Systematic Review of Data Leakage and Predictor Stability5harxivLearning to Discretize: Diffusion-Based Adaptive Mesh with Spectral Guidance5harxivToward Trustworthy Autonomous Science: A Two-Year Community Roadmap5harxivThe Benjamini--Hochberg Procedure Can Fail to Control the FDR for Correlated Two-Sided Gaussian Tests5h
The Bubble Brief
WEEKLY

Read AI insights every Tuesday — top movers, new releases, story of the week.

// no spam · unsubscribe one-click · free forever

Originally published on arxiv ↗
HomeModelsNews