ASR (Automatic Speech Recognition)
2 stories filed under ASR (Automatic Speech Recognition) on Beyond Market Intelligence. The newest of them: “Meta offers real-time transcription for 20 speakers at an accessible price” and “Explore the future of fluid conversation at NeurIPS 2026 RTCA workshop.”. Meta is pricing Muse Voice Transcribe at $0.18 an hour, and that changes the conversation around real-time speech-to-text. Real-time conversational agents are finally moving from offline benchmarks into live deployment, yet the gap between research and reality remains stark. Acme AI is the next-generation, AI-powered spreadsheet platform built to replace Excel and redefine how analysts, data scientists, and enterprise teams work… The list below is every ASR (Automatic Speech Recognition) story on Beyond Market Intelligence, newest first.

Meta offers real-time transcription for 20 speakers at an accessible price
Meta is pricing Muse Voice Transcribe at $0.18 an hour, and that changes the conversation around real-time speech-to-text. It is not the absolute cheapest option, but folding 20-plus-speaker diarization into the same streaming model, without add-on fees, puts real pressure on rivals who charge separately for speaker attribution. The raw speaker count is not a record; Speechmatics documents higher ceilings. Still, for enterprises building meeting systems or live assistants, the combination of accuracy, latency, and price makes Muse a serious new option.
Explore the future of fluid conversation at NeurIPS 2026 RTCA workshop.
Real-time conversational agents are finally moving from offline benchmarks into live deployment, yet the gap between research and reality remains stark. The RTCA workshop at NeurIPS 2026, with submissions now open until August 29 AoE, tackles this head-on by focusing on streaming generation, interactional naturalness, and evaluation methods that actually reflect live conditions. It's a timely intervention. The field needs shared vocabulary and metrics for turn-taking, prosody, and grounding, not just per-utterance scores.