Microsoft announced its most powerful visual artificial intelligence model, MAI-Image-2.5-Pro. Also, all the details about the new voice model MAI-Voice-2-Flash are in our news.
Technology giant Microsoft introduced MAI-Image-2.5-Pro, the most advanced visual artificial intelligence model developed within itself, to users. This model, the newest addition to the MAI-Image series that started in 2025, is defined as the company’s most ambitious work on visual production so far, with its technical infrastructure built on the previous versions of the series: MAI-Image-1, MAI-Image-2 and MAI-Image-2.5.
The new model draws attention with its ability to create high-resolution images as well as its ability to make precise edits on existing content. This technology, which is now available free of charge through the prestigious MAI Playground platform, ushers in a new era in artificial intelligence-supported visual design processes.
Continues to Improve Visual Modeling Skills
Microsoft’s MAI-Image series has gained a valuable place in the world of artificial intelligence with the progress it has made in a short time. This process, which started with MAI-Image-1, was always optimized with MAI-Image-2 and the subsequent version 2.5. MAI-Image-2.5-Pro, now released, makes a significant leap forward in visual editing and understanding complex commands, especially at a professional level. With this new model, users can create much more detailed and realistic images.
The new visual model takes the quality of content produced by artificial intelligence beyond standards.
A New Speed and Cost Standard Has Been Set in Audio Technologies
The company is making significant breakthroughs not only in visual artificial intelligence but also in audio technologies. The MAI-Voice-2-Flash model, first announced at the Build conference in June, is now ready for widespread use. Offering twice the speed of the main model, MAI-Voice-2, this new solution also stands out with its cost advantage. This technology, which is priced at $15 per 1 million characters, brings a humanoid quality to users in a much more economical way in text-to-speech conversion processes.
Artificial Intelligence Ecosystem Continues to Grow
These breakthroughs by Microsoft prove how rapidly artificial intelligence-supported vehicles are evolving in the technology world. These innovations, especially in both visual and audio fields, open new doors for content producers and developers. The company aims to improve its models by incorporating feedback from the community into its processes with these free testing opportunities offered through MAI Playground.
Microsoft’s new audio and video models make artificial intelligence integrations much more accessible.
What do you think about these new visual and audio artificial intelligence models from Microsoft? Which of the features you tried on MAI Playground attracted your attention the most? Share your opinions and experiences with us.