All models
Meta

Muse Voice Transcribe

Meta's streaming speech-recognition model for live captions, long audio, many speakers, code-switching, and domain-aware transcription.

Plain-English overview

What Muse Voice Transcribe actually is

Muse Voice Transcribe is designed for speech that arrives continuously rather than only as a finished upload. It can process small audio chunks, decide when an utterance has ended, label many speakers, and use language, keyword, and context hints to improve names or specialist vocabulary.

Meta advertises broad multilingual training while identifying a smaller set of specifically verified languages. That distinction is useful for procurement: test the exact accents, code-switching patterns, noise, and speaker overlap in your own calls before treating the wider training set as production coverage.

Good fit for

  • Live captions and conversational interfaces
  • Multi-speaker meetings and long recordings
  • Products that need keyword or context biasing

Category comparison

The facts that matter for transcription models

These are provider-published specifications, not Cody benchmark scores. Follow the linked sources for current limits and endpoint-specific exceptions.

Transcription price
$0.18 / audio hourCurrent provider price per audio hour or minute for the listed processing route.
Live or batch
Realtime streaming and long-form audioWhether the model handles realtime streams, uploaded recordings, or both.
Languages
Trained on 70+ languages; 25 specifically verifiedProvider-published language coverage, separating trained or advertised coverage from specifically verified languages where needed.
Speaker labels
Yes; more than 20 speakers advertisedWhether the model identifies who spoke and any published speaker limit.
Timestamps
Streaming transcript timing and endpointingAvailable word-, segment-, or utterance-level timing information.
Vocabulary control
Language, keyword, and context biasing; code-switchingKeyword boosting, custom spelling, context, prompting, or other ways to improve domain terms.
Where to use it
Meta Model API, Meta AI for Mac, Muse CodeDirect API, cloud catalog, application, or regional endpoint documented by the provider.

Pricing & comparisons

Estimate your cost

Set your usage. Your estimate updates as you type.

Uses the recording length in hours: 30 minutes = 0.5 hours. Extra features and minimum charges may change the bill.

Muse Voice Transcribe

Meta

Estimated total (USD)

No reviewed rate

For the usage above · USD · API pricing, not a subscription

How this estimate works

Estimates exclude taxes, tools, cache storage/writes, free allowances and custom discounts. Image estimates cover output only, not prompt or reference-image charges. Quality modes differ by model. Unlisted settings are not treated as free.

API and provider access

Where to get Muse Voice Transcribe

Availability

Regions and access stage

Available through Meta Model API, Meta AI for Mac, and Muse Code during the Model API public preview.

Meta does not enumerate one model-specific country or processing-region list on the public launch page; use the current developer-account terms.

Check live availability

Data and training

The route matters.

Meta's launch page describes privacy for its public demo, not a complete Model API retention or training-use commitment. Review current Model API terms and enterprise controls before sending sensitive audio.

This is a concise reading of the cited provider material, not legal advice. A third-party gateway can have different storage, routing, training, and residency terms from the model maker's direct API.

Read the provider policy

Frequently asked questions

Muse Voice Transcribe FAQ

What is Muse Voice Transcribe?

Meta's streaming speech-recognition model for live captions, long audio, many speakers, code-switching, and domain-aware transcription. Muse Voice Transcribe is designed for speech that arrives continuously rather than only as a finished upload. It can process small audio chunks, decide when an utterance has ended, label many speakers, and use language, keyword, and context hints to improve names or specialist vocabulary.

When was Muse Voice Transcribe released?

Muse Voice Transcribe was released on September 1, 2026 according to the cited provider materials.

Where can I access Muse Voice Transcribe?

Available through Meta Model API, Meta AI for Mac, and Muse Code during the Model API public preview. The access routes listed in this guide are Meta Model API and Meta.

How much does Muse Voice Transcribe cost?

$0.18 / audio hour. Meta lists this Model API rate on the reviewed catalog. Product allowances in Meta AI for Mac or Muse Code can differ.

Where is Muse Voice Transcribe available?

Available through Meta Model API, Meta AI for Mac, and Muse Code during the Model API public preview. Meta does not enumerate one model-specific country or processing-region list on the public launch page; use the current developer-account terms.

Is my Muse Voice Transcribe API data used for training?

Meta's launch page describes privacy for its public demo, not a complete Model API retention or training-use commitment. Review current Model API terms and enterprise controls before sending sensitive audio. The policy belongs to the provider route and account terms, so verify it again before production use.