← Back to Model Beat
Open Source·4d ago·all news from September 11, 2026

Xiaomi-CocktailASR-1 Technical Report

Xiaomi researchers released a technical report for CocktailASR, a new system designed to improve automatic speech recognition in environments where multiple people are talking at once. The project aims to resolve limitations in current large language model-based transcription tools, which typically struggle to isolate and accurately process individual voices during simultaneous speech.

Covered by 1 source

  • AarXiv CS.AIYiru Zhang, Hang Su, Lichun Fan, Ying Zeng, Chang Liu, Yifeng Wang, Yuquan Liang, Tao Li, Lian Li, Wenhao Yang, Jian Luan, Cong Zou, Heng Qu4d ago

Related stories

Open SourceOpenAI agents launched a 2,000-package cyberattack on RubyGems just to collect data anyone could GoogleSep 9 · 15 sourcesOpen SourceChina’s AI Industry Pivots to Agents From Models, Report SaysSep 11 · 4 sourcesOpen SourceAnthropic Expects an Operating Profit This Quarter, FT SaysSep 13 · 2 sourcesOpen SourceIndependent Investigation of Hugging Face Incident Reveals How Agents Collaborated and BehavedSep 14