·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
OpenAI puts Pro subscriptions on hold due to Astra demand28m◆Anthropic details distillation campaigns from Alibaba, Moonshot AI, and DeepSeek31m◆Meta’s AI agent Muse is now the No. 2 app in the US1h◆Schools are catching on to Big Tech’s playbook1h◆Anthropic reveals rogue AI agents hate CAPTCHAs, just like you3h◆India’s Pocket FM doubles revenue run rate to $500M as AI powers 93% of audio content3h◆3 ways to prep for your next big race with Search5h◆How a researcher uses Codex and ChatGPT to search for new antimicrobial molecules5h◆Universal Music is launching an AI music platform with ElevenLabs5h◆Now everyone can put data to work6h◆Meta’s Muse AI works and creeps me out6h◆AI agents are flooding public services with new requests6h◆Maven Robotics wants to steal your robot deployment deal7h◆Why the current tech backlash feels different7h◆Mathematicians want proof OpenAI didn’t use their work10h◆Powering AI is an architecture problem10h◆Expanding AI access and cyber defense for federal, state, local, and tribal governments14h◆Introducing ChatGPT for Financial Services14h◆Planning and Scheduling Business Processes under Control-Flow Uncertainty17h◆SIM: Subspace Interaction-based Method for Token-Level Text Anomaly Detection17h◆OpenAI puts Pro subscriptions on hold due to Astra demand28m◆Anthropic details distillation campaigns from Alibaba, Moonshot AI, and DeepSeek31m◆Meta’s AI agent Muse is now the No. 2 app in the US1h◆Schools are catching on to Big Tech’s playbook1h◆Anthropic reveals rogue AI agents hate CAPTCHAs, just like you3h◆India’s Pocket FM doubles revenue run rate to $500M as AI powers 93% of audio content3h◆3 ways to prep for your next big race with Search5h◆How a researcher uses Codex and ChatGPT to search for new antimicrobial molecules5h◆Universal Music is launching an AI music platform with ElevenLabs5h◆Now everyone can put data to work6h◆Meta’s Muse AI works and creeps me out6h◆AI agents are flooding public services with new requests6h◆Maven Robotics wants to steal your robot deployment deal7h◆Why the current tech backlash feels different7h◆Mathematicians want proof OpenAI didn’t use their work10h◆Powering AI is an architecture problem10h◆Expanding AI access and cyber defense for federal, state, local, and tribal governments14h◆Introducing ChatGPT for Financial Services14h◆Planning and Scheduling Business Processes under Control-Flow Uncertainty17h◆SIM: Subspace Interaction-based Method for Token-Level Text Anomaly Detection17h◆
News/J-CHAT: Japanese Large-scale Spoken Dialogue Corpus for Spoken Dialogue Language Modeling
arxiv
PublishedApril 4, 2026 at 4:00 AM
▲bullish

J-CHAT: Japanese Large-scale Spoken Dialogue Corpus for Spoken Dialogue Language Modeling

Source
arxiv.orgfull article ↗
Read on arxiv→
Publisher summary· verbatim

arXiv:2407.15828v2 Announce Type: replace Abstract: Spoken dialogue is essential for human-AI interactions, providing expressive capabilities beyond text. Developing effective spoken dialogue systems (SDSs) requires large-scale, high-quality, and diverse spoken dialogue corpora. However, existing da

Stay posted· Newsletter

A 5-min weekly brief — top movers, price watch, story of the week.

// no spam · unsubscribe one-click · free forever

Discussion
Source
↗
arxiv
Read original ↗All from arxiv →
Tags
04
#open-source#dataset#speech#dialogue

No replies yet. Be first.

Source
↗
arxiv
Read original ↗All from arxiv →
Tags
04
#open-source#dataset#speech#dialogue

Related coverage

More from ARXIV
arxivPlanning and Scheduling Business Processes under Control-Flow Uncertainty17harxivSIM: Subspace Interaction-based Method for Token-Level Text Anomaly Detection17h
The Bubble Brief
WEEKLY

Read open-source insights every Tuesday — top movers, new releases, story of the week.

// no spam · unsubscribe one-click · free forever

Originally published on arxiv ↗
HomeModelsNews