MAI-Voice-1

by Microsoft · Mai family · released Oct 7, 2025
Open weights Active

Microsoft's open-source neural text-to-speech and voice cloning model with strong speaker similarity. Released on Hugging Face and integrated into Azure AI Speech.

This model has no first-party API. Open-weight release: self-host from published weights.

Specifications

Model typeAudio gen
ArchitectureUnknown
ParametersNot disclosed
Context windowNot disclosed tokens
Max outputNot disclosed tokens
Knowledge cutoffNot disclosed
LicenseMIT License
Modalities Text input Audio input Audio output
Tags -
End of lifeNot disclosed
Last verifiedNot disclosed