Meta Model APIMeta
|
||||||
Related Products
|
||||||
About
Meta Model API is a new developer API for building with Muse Spark 1.1, Meta’s multimodal reasoning model built for agentic tasks, coding, tool use, computer use, and multimodal understanding. Now in public preview, it gives developers a way to access Muse Spark 1.1 through an OpenAI-compatible package, making it easier to point existing clients at the API, keep the same code structure, and set the model to muse-spark-1.1. Muse Spark 1.1 is designed for personal agentic tasks that require planning and orchestration across external apps and services, with the ability to generalize to new native tools, MCP servers, and custom skills. As a main agent, it can gather context, make a plan, and delegate execution across parallel subagents; as a subagent, it follows its role, understands available tools, and knows when to escalate back. The model can actively manage a 1 million-token context window, remember actions, retrieve information from much earlier work, and compact context.
|
About
Muse Voice Transcribe is Meta’s first real-time audio perception model, delivering streaming automatic speech recognition (ASR), diarization, and endpointing in real time. An autoregressive multimodal model from the Muse Spark family, it processes audio in 80 ms chunks and decides dynamically whether to continue listening or emit text. Its adaptive delay changes the amount of audio context used for each word based on difficulty, balancing transcription accuracy with latency. The model is trained on more than 70 languages, with 25 extensively verified at launch, and natively supports arbitrary code-switching both within and between sentences. Language, keyword, and context biasing can further improve recognition accuracy for specific names, places, contacts, or terminology. Streaming diarization identifies speaker changes and distinguishes more than 20 speakers, while endpointing detects when speech begins and when a user finishes speaking.
|
|||||
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
|||||
Audience
Developers and anyone needing a multimodal model API for coding agents, tool-using workflows, computer use, and long-context automation
|
Audience
Developers and AI researchers seeking to build real-time voice applications that transcribe multilingual speech, distinguish speakers, and detect conversational turns
|
|||||
Support
Phone Support
24/7 Live Support
Online
|
Support
Phone Support
24/7 Live Support
Online
|
|||||
API
Offers API
|
API
Offers API
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
$1.25 per 1M tokens
Free Version
Free Trial
|
Pricing
No information available.
Free Version
Free Trial
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Webinars
Live Online
In Person
|
Training
Documentation
Webinars
Live Online
In Person
|
|||||
Company InformationMeta
United States
developer.meta.com/ai/products/meta-model-api/
|
Company InformationMeta
Founded: 2004
United States
research.meta.ai/blog/introducing-muse-voice-transcribe
|
|||||
Alternatives |
Alternatives |
|||||
|
|
|
|||||
|
|
|
|||||
|
|
|
|||||
|
|
|
|||||
Categories |
Categories |
|||||
Integrations
Codex CLI
Continue
GitHub
Hugging Face
LangChain
Llama 3
Llama 4 Behemoth
Llama 4 Maverick
Llama 4 Scout
LlamaIndex
|
Integrations
Codex CLI
Continue
GitHub
Hugging Face
LangChain
Llama 3
Llama 4 Behemoth
Llama 4 Maverick
Llama 4 Scout
LlamaIndex
|
|||||
|
|
|