• bitcoinBitcoin(BTC)$81,277.004.39%
  • ethereumEthereum(ETH)$2,524.025.00%
  • tetherTether(USDT)$1.000.04%
  • binancecoinBNB(BNB)$724.211.78%
  • rippleXRP(XRP)$1.455.96%
  • usd-coinUSDC(USDC)$1.000.01%
  • solanaSolana(SOL)$104.173.43%
  • tronTRON(TRX)$0.3289890.40%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.030.00%
  • HyperliquidHyperliquid(HYPE)$87.026.22%
  • zcashZcash(ZEC)$1,003.6919.09%
  • dogecoinDogecoin(DOGE)$0.0878895.65%
  • RainRain(RAIN)$0.0168951.84%
  • moneroMonero(XMR)$548.396.54%
  • USDSUSDS(USDS)$1.000.02%
  • chainlinkChainlink(LINK)$12.056.95%
  • whitebitWhiteBIT Coin(WBT)$74.124.01%
  • leo-tokenLEO Token(LEO)$9.30-1.58%
  • cardanoCardano(ADA)$0.2214596.97%
  • stellarStellar(XLM)$0.1847924.16%
  • bitcoin-cashBitcoin Cash(BCH)$259.823.95%
  • daiDai(DAI)$1.000.01%
  • CantonCanton(CC)$0.1109641.54%
  • Ethena USDeEthena USDe(USDE)$1.000.05%
  • USD1USD1(USD1)$1.000.04%
  • litecoinLitecoin(LTC)$51.251.29%
  • uniswapUniswap(UNI)$6.271.55%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.393.64%
  • hedera-hashgraphHedera(HBAR)$0.0789562.93%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • avalanche-2Avalanche(AVAX)$7.503.05%
  • suiSui(SUI)$0.781.16%
  • shiba-inuShiba Inu(SHIB)$0.0000052.69%
  • paypal-usdPayPal USD(PYUSD)$1.000.02%
  • crypto-com-chainCronos(CRO)$0.0576675.18%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,460.560.81%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • nearNEAR Protocol(NEAR)$2.026.50%
  • MemeCoreMemeCore(M)$1.07-0.89%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • okbOKB(OKB)$110.463.82%
  • BittensorBittensor(TAO)$229.813.21%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.36%
  • aaveAave(AAVE)$133.572.82%
  • AsterAster(ASTER)$0.752.96%
  • mantleMantle(MNT)$0.594.62%
  • pax-goldPAX Gold(PAXG)$4,467.820.72%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0574461.76%
  • OndoOndo(ONDO)$0.3675993.43%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

MAI-Transcribe-2 Tops FLEURS Benchmark Across 60 Languages, Microsoft Says – Unite.AI

September 3, 2026
in AI & Technology
Reading Time: 3 mins read
A A
MAI-Transcribe-2 Tops FLEURS Benchmark Across 60 Languages, Microsoft Says – Unite.AI
ShareShareShareShareShare

Microsoft AI released MAI-Transcribe-2 on September 3, 2026, a speech recognition model the lab said ranks first on the FLEURS benchmark across 60 languages with an average word error rate of 5.2%. In its announcement, Microsoft described the model as its most capable transcription system to date and priced it at $0.10 per hour of audio.

YOU MAY ALSO LIKE

Sam Altman Apologizes as GPT-6 Astra Staged Launch Denies Paid Access – Unite.AI

This Rugged Smartphone’s Camera Is A Removable Action Cam

Microsoft said MAI-Transcribe-2 adds speaker diarization, configurable transcription styles, and word-level timestamps, and outperforms competing models including Gemini 3.5 Transcribe, GPT-Transcribe, Whisper V3-Large, and ScribeV2 across a broader range of real-world audio. According to the announcement, the model defines the Pareto frontier for accuracy and latency on Artificial Analysis and ranks second on the Artificial Analysis word error rate leaderboard, improving on the results of earlier MAI-Transcribe versions.

Microsoft positioned the model for workloads including clinical note-taking, legal documentation, accessibility, and closed captioning. The company reported faster inference with substantially lower latency, particularly for long-form audio, at up to 10 times the processing speed of leading competitors. It also said the model maintains transcription quality in noisy conditions outside controlled recording environments.

Benchmark Results

Citing evaluations run by Artificial Analysis, Microsoft said MAI-Transcribe-2 is 10 times faster than OpenAI’s GPT-Transcribe, 7 times faster than ElevenLabs’ Scribe v2, and 5 times faster than Gemini 3.5 Transcribe while delivering higher accuracy. The company said the model sits alone in the most attractive quadrant of the benchmark’s accuracy-versus-speed chart, at a 2.0% error rate and a speed factor of 403.6, meaning an hour of audio returns in about ten seconds.

On the public multilingual FLEURS benchmark, Microsoft reported that MAI-Transcribe-2 holds a consistently high accuracy bar across all 60 tested languages. The company described the model as accurate across more languages than any other model and said developers transcribing across multiple languages can rely on a single model, reducing complexity and potentially saving GPU utilization. The announcement also states that the model’s speed and throughput allow Microsoft to offer what it called the most competitive price in the market. At launch, the $0.10 hourly rate is a limited-time offer running until the end of the year.

Availability and Developer Features

The Microsoft Learn documentation lists MAI-Transcribe-2 as available in Azure Speech in public preview, without a service-level agreement and not recommended for production workloads. The documentation describes MAI-Transcribe as a speech-to-text model built in-house by the Microsoft AI team, covering workloads such as video captioning, meetings, clinical notes, call center documentation, accessibility tools, content creation, and voice agents. It also lists MAI-Transcribe-2 alongside the earlier MAI-Transcribe-1.5 and MAI-Transcribe-1, the latter deprecated on August 20, 2026.

Requests route through the Fast Transcription API’s enhanced mode, with the model selected by setting the enhanced mode model property to MAI-Transcribe-2. Audio input is limited to files under 300 MB in WAV, MP3, or FLAC format, and use requires an Azure subscription and a Microsoft Foundry resource for Speech.

Optional parameters control the model’s feature set. Speaker diarization segments a recording by speaker and returns speaker-labelled segments with offset and duration metadata. Word-level timestamps return timing for every word, while a segment option returns timing per segment and a none option omits timing data. A phrase-list parameter biases recognition toward supplied terms such as domain-specific terminology, abbreviations, and proper nouns, with the documentation noting that terms act as hints rather than forced output.

The transcription style parameter defaults to verbatim, which captures speech exactly as spoken, including filler words and false starts, for compliance, QA, and analysis workloads. A clean setting removes fillers and auto-formats common speech patterns to produce more readable captions, notes, and published transcripts. Language selection is optional; by default the model automatically detects the spoken language, and the documentation advises forcing a specific language only when auto-detection fails. Code switching for blended language pairs such as Hinglish and Spanglish is handled automatically, and noise robustness for audio recorded outside controlled environments is inherent to the model.

The documentation also notes that MAI-Transcribe can provide input audio transcription in the Voice Live API through a session configuration field. Microsoft said the model is available to demo through Microsoft Foundry and the MAI Playground, and OpenRouter.

Credit: Source link

ShareTweetSendSharePin

Related Posts

Sam Altman Apologizes as GPT-6 Astra Staged Launch Denies Paid Access – Unite.AI
AI & Technology

Sam Altman Apologizes as GPT-6 Astra Staged Launch Denies Paid Access – Unite.AI

September 4, 2026
This Rugged Smartphone’s Camera Is A Removable Action Cam
AI & Technology

This Rugged Smartphone’s Camera Is A Removable Action Cam

September 4, 2026
The Ternus Era At Apple Begins, But Cook Isn’t Leaving
AI & Technology

The Ternus Era At Apple Begins, But Cook Isn’t Leaving

September 4, 2026
Mobile Games Designed to Be Addictive Get More Kid-Friendly
AI & Technology

Mobile Games Designed to Be Addictive Get More Kid-Friendly

September 4, 2026
Next Post
AI Spending Ripples Across Tech Stack; Nvidia Acquires Hugging Face | Bloomberg Tech 9/03/2026

AI Spending Ripples Across Tech Stack; Nvidia Acquires Hugging Face | Bloomberg Tech 9/03/2026

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Vercel AI Open-Sources vgpu: A TypeScript WebGPU Library for AI Agent Shaders

Vercel AI Open-Sources vgpu: A TypeScript WebGPU Library for AI Agent Shaders

August 28, 2026
FAA says it is investigating safety incident involving president’s helicopter

FAA says it is investigating safety incident involving president’s helicopter

August 29, 2026
Audacity’s New Look Is Finally Here, Along With Its Largest Feature Update In Years

Audacity’s New Look Is Finally Here, Along With Its Largest Feature Update In Years

September 3, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!