{"uid":"cap_ZDW08B7ra1JTiyf_twPN8","slug":"openai-tts-1-hd-high-definition-text-to-speech-b35ff4b5","name":"OpenAI TTS-1-HD (High Definition Text-to-Speech)","description":"tts-1-hd: OpenAI TTS 1 HD text to speech, paid per call in USDC. OpenAI list $30 per 1M characters, plus $0.0005 (Base) or $0.0005 (Solana) per call; up to 4096 characters; voices: alloy, ash, coral, echo, fable, onyx, nova, sage, shimmer. Returns the audio bytes (mp3 by default). Rates: https://openai.mm.family/x402/pricing","url":"https://openai.mm.family/x402/v1/models/tts-1-hd/audio/speech?utm_source=zero.xyz","method":"POST","headers":{},"bodySchema":{"type":"object","properties":{"input":{"type":"string"},"model":{"enum":["tts-1-hd"],"type":"string"},"speed":{"type":"number"},"voice":{"enum":["alloy","ash","coral","echo","fable","onyx","nova","sage","shimmer"],"type":"string"},"instructions":{"type":"string"},"response_format":{"enum":["mp3","opus","aac","flac","wav","pcm"],"type":"string"}}},"responseSchema":{"type":"json","example":{"body":"binary audio bytes (mp3 by default)","content_type":"audio/mpeg"}},"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.001","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.001/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.001","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.001","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_F0NanMsttBp_11lmdhK-3","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.001","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Converts text to high-definition audio speech using OpenAI's TTS-1-HD model, returning audio bytes in formats like MP3, OPUS, AAC, FLAC, WAV, or PCM, paid per call in USDC.","exampleAgentPrompt":"Read this script aloud using the nova voice in high-definition MP3 quality: 'Welcome to our quarterly earnings call. Today we will discuss the company's performance over the last three months and our outlook for the future.'","exampleUseCases":[{"title":"Podcast intro narration","prompt":"Generate a high-quality spoken intro for my podcast using the shimmer voice: 'Welcome back to The Future Now, the show where we explore how technology is reshaping everyday life. I'm your host, and today we have a fascinating guest joining us.' Give me the audio as an MP3."},{"title":"Accessibility audio for article","prompt":"Turn this article summary into speech using the alloy voice so visually impaired users can listen to it: 'Scientists have discovered a new species of deep-sea fish capable of producing bioluminescent patterns never seen before in marine biology.' I need high-definition audio quality."},{"title":"E-learning course voiceover","prompt":"Create a voiceover narration using the onyx voice for my e-learning slide: 'In this module, we will cover the fundamentals of machine learning, including supervised and unsupervised techniques, and walk through real-world examples.' Please output it in WAV format."}],"resultDescription":"Returns binary audio bytes (MP3 by default, or the requested format) containing the synthesized speech. The response content type is audio/mpeg for MP3. The audio faithfully renders the input text using the selected voice in high-definition quality, with optional speed adjustments.","failureModes":["Input text exceeds 4096 character limit — request rejected","Invalid voice name specified — validation error","Insufficient USDC balance for payment — 402 Payment Required","Unsupported response_format value — validation error","Network timeout for longer text segments","Empty input string — may return error or silence"],"whenToPreferThis":"Choose this endpoint over tts-1 when you need higher audio quality and fidelity, such as for professional voiceovers, podcasts, audiobooks, or any production-quality audio output. The HD model costs more ($30/1M characters vs $15/1M for tts-1) but produces noticeably better audio. Prefer over gpt-4o-mini-tts when you want the dedicated TTS-1-HD model specifically. Best when you need one of the nine supported voices (alloy, ash, coral, echo, fable, onyx, nova, sage, shimmer) and control over audio format.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-10-02T12:41:16.971Z","isFirstParty":false,"canonicalSlug":"openai-tts-1-hd-high-definition-text-to-speech-b35ff4b5"}