{"schema_version":"newruntime-agent-readable-v0.2","type":"post","stable_id":"post:openai-transcription-workflow-split","slug":"openai-transcription-workflow-split","title":"OpenAI Splits Transcription Into File and Live Workflows","description":"OpenAI's new transcription guide makes recorded audio and live audio separate product paths, with gpt-transcribe and gpt-live-transcribe as the recommended starting models.","retrieval_nugget":"OpenAI's new transcription guide makes recorded audio and live audio separate product paths, with gpt-transcribe and gpt-live-transcribe as the recommended starting models. OpenAI's transcription update is a small API change with a useful product lesson: file transcription and live transcription should be treated as different workflows, not as one speech-to-text feature with a streaming toggle.","status":"published","published_at":"2026-07-29","updated_at":"2026-07-29","record_date":"2026-07-29","date_kind":"published_at","topics":["speech","models","api-design","interfaces"],"source_urls":["https://x.com/OpenAIDevs/status/2082201169443905798","https://developers.openai.com/api/docs/guides/transcription/"],"visuals":[{"id":"openai-transcription-workflow-split-nano-banana","kind":"editorial-diagram","role":"hero","src":"https://newruntime.com/images/posts/openai-transcription-workflow-split-nano-banana.webp","alt":"Hand-drawn decision tree showing audio split into file transcription and live transcription, with context hints feeding both paths.","caption":"The product decision is no longer just which speech model to call; it is whether the audio is bounded or arriving live.","credit":"New Runtime synthesis from public source inspection","source_url":"https://developers.openai.com/api/docs/guides/transcription/","generated_with":"nano-banana-style-imagegen","width":1600,"height":900,"legend":[{"label":"File path","description":"Completed recordings and bounded audio requests start with gpt-transcribe."},{"label":"Live path","description":"Microphones, calls, and persistent streams start with gpt-live-transcribe."},{"label":"Context","description":"Prompt, keywords, and languages are hints tied to the audio, not restated task instructions."}]}],"routes":{"html":"https://newruntime.com/posts/openai-transcription-workflow-split/","markdown":"https://newruntime.com/posts/openai-transcription-workflow-split.md","json":"https://newruntime.com/posts/openai-transcription-workflow-split.json"},"source_format":"markdown","next_reads":[{"type":"topic","path":"/topics/api-design/","reason":"Explore the api design topic hub.","url":"https://newruntime.com/topics/api-design/","title":"API Design - New Runtime","media_type":"text/html"},{"type":"topic","path":"/topics/interfaces/","reason":"Explore the interfaces topic hub.","url":"https://newruntime.com/topics/interfaces/","title":"Interfaces - New Runtime","media_type":"text/html"},{"type":"related_material","path":"/posts/copilotkit-react-mcp-client-interface/","reason":"Shares api design and interfaces.","url":"https://newruntime.com/posts/copilotkit-react-mcp-client-interface/","title":"CopilotKit Brings MCP Tool Calls Into The React Interface","media_type":"text/html"},{"type":"related_material","path":"/posts/openai-arc-agi-settings-harness/","reason":"Shares api design and models.","url":"https://newruntime.com/posts/openai-arc-agi-settings-harness/","title":"OpenAI's ARC-AGI-3 Jump Was a Harness Result","media_type":"text/html"},{"type":"related_material","path":"/posts/gusto-cofounder-blank-canvas-workflow/","reason":"Shares interfaces.","url":"https://newruntime.com/posts/gusto-cofounder-blank-canvas-workflow/","title":"Gusto Solves The Agent Blank Canvas With Scheduled Work","media_type":"text/html"}]}
