Technology · 3 views
Meta's new AI transcription model can distinguish between multiple speakers and languages in real-time
Meta Meta has introduced its first real-time audio model, Muse Voice Transcribe. The model can handle dictation and transcription for more than 20 speakers and can seamlessly handle multiple languages at once, Meta says.
Story Brief
Meta Meta has introduced its first real-time audio model, Muse Voice Transcribe. The model can handle dictation and transcription for more than 20 speakers and can seamlessly handle multiple languages at once, Meta says. Meta CEO Mark Zuckerberg, who recently returned to X after three years of not posting on the platform, shared an example of the model's ability to handle multiple speakers and languages at once. In the video, the transcription is able to automatically distinguish between multiple speakers and switch between languages.
Read full article on EngadgetAI summaries can be wrong sometimes—always verify important details using the source article.
How AI & Automation are usedMore from Technology
Continue reading recent Technology coverage
- Are all TVs at risk from burn-in? Here's how to avoid the problem altogetherContinue reading
- Apple’s iPhone 18 Pro Could Come in Black, but Skip Silver This YearContinue reading
- The Range Rover Electric: Specs, Price, AvailabilityContinue reading
- Blue Origin wins NASA contract for telecommunications on MarsContinue reading
Support HappeningNow
Independent AI-powered news analysis is reader-supported. Your contribution helps cover infrastructure, summaries, and continued platform development.
Support HappeningNow