Model Detail
cohere-transcribe-arabic-07-2026
—cohere-transcribe-arabic-07-2026 is an audio model with 1.0B parameters released by CohereLabs. The model is registered under the automatic-speech-recognition pipeline tag on Hugging Face, distributed under the permissive apache-2.0 license.
cohere-transcribe-arabic-07-2026 ships with 1.0B parameters. Total weight footprint is approximately 2.1 GB, which is the relevant figure when planning local-inference VRAM. The apache-2.0 license is permissive, allowing commercial deployment and derivative work without per-seat fees, though attribution requirements still apply.
cohere-transcribe-arabic-07-2026 is best fit for speech recognition, transcription, or speech synthesis depending on the task head. Treat this as a starting matrix rather than a benchmark verdict — the right deployment usually depends on the specific evaluation suite that mirrors your workload.
CORE-T: COherent REtrieval of Tables for Text-to-SQL
arXiv:2601.13111v3 Announce Type: replace-cross Abstract: Realistic text-to-SQL workflows often require joining multiple tables. As a result, accurately retrieving the relevant set of tables becomes a key bottleneck for end-to-end performance. We study an open-book setting where queries must be answ
Semantic Flow Regularization: Teaching LLMs to Generate Diverse Yet Coherent Responses
arXiv:2605.27971v2 Announce Type: replace-cross Abstract: When large language models are fine-tuned to generate persona- or tone-conditioned responses, their output diversity is severely limited--a failure we term Cross-Style Collapse. We trace this collapse to the cross-entropy objective, which und
An Explainable Coherence Score for Detecting Temporal Inconsistencies in Political News
arXiv:2608.29175v1 Announce Type: new Abstract: Temporal inconsistencies, such as mandates attributed outside their real interval, events presented as past before they occurred, or inverted causal sequences, are a form of political disinformation that evades style-based fake news detectors: a well-w
Fairness in multi-class multi-group classification problems via contextial coherent risk measures
arXiv:2608.30223v1 Announce Type: cross Abstract: We propose a new design of fair classifiers for multi-class classification problems in the presence of vector-valued sensitive attributes. In that scenario each sensitive attribute has multiple values and forms several groups relevant to the fairness
QAQ: Bidirectional Semantic Coherence for Selecting High-Quality Synthetic Code Instructions
arXiv:2603.12165v3 Announce Type: replace Abstract: Synthetic data has become essential for training code generation models, yet it introduces significant noise and hallucinations that are difficult to detect with current metrics. Existing data selection methods like Instruction-Following Difficulty
MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations
arXiv:2607.28956v1 Announce Type: new Abstract: Large language model agents are increasingly evaluated as autonomous tool users, yet most benchmarks focus on bounded tasks with immediate success criteria. Real-world deployments often require Long-Term Coherence, the capacity to preserve purposeful b