Microsoft has introduced a preview version of MAI-Realtime, a bidirectional speech model built for real-time audio processing and more natural voice conversations. The company said the model supports multilingual use cases and can be integrated with tools such as Copilot. MAI-Realtime now joins MAI-Voice-1 and MAI-Transcribe-1 as part of Microsoft’s end-to-end speech AI lineup. According to CryptoBriefing, the move is aimed at reducing Microsoft’s reliance on OpenAI, while giving the company more control over its product roadmap and helping lower inference costs. The release adds another piece to Microsoft’s broader effort to expand its own AI infrastructure and product capabilities around speech.
Microsoft has rolled out a preview of MAI-Realtime, a bidirectional speech model designed for real-time audio processing and more natural voice interactions. The model supports multiple languages and can be integrated with tools including Copilot.
MAI-Realtime sits alongside the previously released MAI-Voice-1 and MAI-Transcribe-1, forming part of Microsoft’s end-to-end speech AI system. According to CryptoBriefing, the release is intended to reduce Microsoft’s dependence on OpenAI, give the company greater control over its product roadmap, and cut inference costs.
This article was originally published by Bit.Fan. For more cryptocurrency news and market insights, visit www.bit.fan. Disclaimer:
The market information, project data, and third-party content displayed on this platform are for industry information sharing only and do not constitute any form of investment advice or return commitment.
Cryptocurrency trading carries high risks. Users should fully assess their risk tolerance and make independent decisions. All profits, losses, and legal responsibilities are borne by the users themselves.