Meta's new AI transcription model can distinguish between multiple sp… | HappeningNow.news

Technology · 3 views

Meta's new AI transcription model can distinguish between multiple speakers and languages in real-time

Meta Meta has introduced its first real-time audio model, Muse Voice Transcribe. The model can handle dictation and transcription for more than 20 speakers and can seamlessly handle multiple languages at once, Meta says.

Source Story Brief Updated 1h 59m ago
Story intelligence Beta
Confidence Limited Single-outlet story
Views 3 Community interest
Brief read ~84 words

Story Brief

Meta Meta has introduced its first real-time audio model, Muse Voice Transcribe. The model can handle dictation and transcription for more than 20 speakers and can seamlessly handle multiple languages at once, Meta says. Meta CEO Mark Zuckerberg, who recently returned to X after three years of not posting on the platform, shared an example of the model's ability to handle multiple speakers and languages at once. In the video, the transcription is able to automatically distinguish between multiple speakers and switch between languages.

Read full article on Engadget

More from Technology

Continue reading recent Technology coverage

Support HappeningNow

Independent AI-powered news analysis is reader-supported. Your contribution helps cover infrastructure, summaries, and continued platform development.

Support HappeningNow

Report an issue with this page