マイクロソフト、新たな音声・画像モデルでLLMを超える取り組みを発表
本文の状態
日本語全文を表示中
詳細モードで約1分の本文を読めます。
マイクロソフトが、自社開発のAIシステムを強化するため、新たな音声と画像のAIモデルを発表した。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
Microsoft、新たな音声・画像モデルでLLMsの先へ
新しいAIモデルは、Microsoft開発のAIシステムへのさらなる推進力を示すものです。
原文を表示
2 Min ReadMicrosoft on Thursday unveiled three new AI models, marking an expansion beyond typical large language models to multimodal, in-house capabilities.The models were introduced under the Microsoft AI (MAI) division.The release includes MAI-Transcribe-1, a new speech-to-text system, as well as voice generation and image models MAI-Voice-1 and MAI-Image-2. All three are the first models of their kind for Microsoft and are available on Microsoft Foundry and the MAI Playground.MAI-Transcribe-1 is Microsoft’s first dedicated transcription model, designed to convert audio into text across 25 languages. Potential applications include video captioning, meeting transcriptions and voice-enabled agents.According to Microsoft, the model can operate at speeds up to 2.5 times faster than its existing Azure Fast transcription model. MAI-Voice-1, meanwhile, is designed for high-quality speech generation. The model can generate up to a minute of audio in a single second, with an emphasis on natural, emotional tone and speaker personality.Related:The Real AI Shift Isn’t New Models. It’s Control.The third release, MAI-Image-2, represents the second generation of Microsoft’s in-house image model. The company says it offers at least twice the generation speed of its predecessor while providing more realistic details, such as skin tone, lighting and textures.The model is targeted for use in the creative industries, and is already being rolled out across Microsoft products, with integrations planned for the Bing search engine and PowerPoint.Early customers include marketing and communications firm WPP, Microsoft said.“MAI-Image-2 is a genuine game-changer,” Rob Reilly, global chief creative officer at WPP said in a MAI blog post on the launch. “It’s a platform that not only responds to the intricate nuance of creative direction, but deeply respects the sheer craft involved in generating real-world, campaign-ready images.”In the post, Microsoft said the updates come as it pursues a more "humanist" AI.“We have a distinct view when creating our AI models -- putting humans at the center, optimizing for how people actually communicate, training for practical use,” the company said.The launches also reflect a broader strategic shift as Microsoft looks to diversify its AI portfolio and reduce reliance on external partners such as OpenAI. It is also aiming to strengthen its competitive standing against rivals such as Google and Amazon, both of which have been investing heavily in proprietary AI stacks.About the AuthorContributing WriterScarlett Evans is a freelance writer with a focus on emerging technologies and the minerals industry. Previously, she served as assistant editor at IoT World Today, where she specialized in robotics and smart city technologies. Scarlett also has a background in the mining and resources sector, with experience at Mine Australia, Mine Technology and Power Technology. She joined Informa in April 2022 before transitioning to freelance work.
同じ出来事を5媒体で確認
同じ出来事を扱う別媒体の記事です。見出しと公開時刻を比較できます。
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み