GPT‑Realtime‑WhisperOpenAI
|
NativBlaizzy
|
|||||
Related Products
|
||||||
About
GPT-Realtime-Whisper is OpenAI’s streaming transcription model built for low-latency speech-to-text experiences in live products. It transcribes audio as people speak, helping voice-enabled apps feel faster, more responsive, and more natural, from captions that appear in the moment to meeting notes that keep up with the conversation. It makes live speech usable inside business workflows as it happens, so teams can power captions for meetings, classrooms, broadcasts, and events, generate notes and summaries while conversations are still in progress, build voice agents that need to understand users continuously, and create faster follow-up workflows for high-volume spoken interactions. It is part of a new generation of real-time voice models in the API that can reason, translate, and transcribe as people speak, moving real-time audio beyond simple call-and-response toward voice interfaces that can listen, translate, transcribe, and take action as a conversation unfolds.
|
About
Nativ is a 100% open-source macOS app for running OpenAI models locally on Apple Silicon, putting frontier intelligence directly on your desk with no accounts or cloud required. It provides a clean chat interface with streaming responses, Markdown, code highlighting, image input, and per-message performance metrics, with every response generated locally. A curated model library includes open models from teams such as Google, Cohere, and Liquid AI, while Nativ recommends models suited to the hardware in your Mac. Built on MLX-VLM and tuned for M-series unified memory and Metal, it runs models without wrappers or translation layers. Live telemetry exposes tokens per second, memory pressure, thermal state, and time to first token so users can see what is actually happening during inference. Nativ supports language, vision, video, code, and audio workflows, including chatting with LLMs, captioning images, summarizing video, autocompleting code, transcribing audio, and generating speech.
|
|||||
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
|||||
Audience
Live events technology teams that need low-latency speech-to-text for real-time captions, transcripts, and post-event content workflows
|
Audience
Developers, researchers, hackers, and users seeking to run, inspect, customize, and connect open AI models locally for private multimodal and coding workflows
|
|||||
Support
Phone Support
24/7 Live Support
Online
|
Support
Phone Support
24/7 Live Support
Online
|
|||||
API
Offers API
|
API
Offers API
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
$0.017 per minute
Free Version
Free Trial
|
Pricing
Free
Free Version
Free Trial
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Webinars
Live Online
In Person
|
Training
Documentation
Webinars
Live Online
In Person
|
|||||
Company InformationOpenAI
Founded: 2015
United States
openai.com/index/advancing-voice-intelligence-with-new-models-in-the-api/
|
Company InformationBlaizzy
United States
blaizzy.github.io/nativ/
|
|||||
Alternatives |
Alternatives |
|||||
|
|
|
|||||
|
|
|
|||||
|
|
|
|||||
|
|
|
|||||
Categories |
Categories |
|||||
Integrations
OpenAI
Claude Code
Codex CLI
Cohere
Google
Hermes Agent
Liquid AI
Markdown
OpenAI Whisper
OpenCode
|
Integrations
OpenAI
Claude Code
Codex CLI
Cohere
Google
Hermes Agent
Liquid AI
Markdown
OpenAI Whisper
OpenCode
|
|||||
|
|
|