{"uid":"cap_qxY589n9JAoq4ThAj-OFm","slug":"nexari-speech-to-text-with-speaker-diarization-30-min-b40fb32b","name":"Nexari Speech-to-Text with Speaker Diarization (30 min)","description":"Transkription mit Sprechertrennung bis 30 Min Audio (pyannote + Whisper), verarbeitet in Deutschland. Speech-to-text with speaker diarization for audio up to 30 min, German-optimised, processed in Germany. Voraussetzung/required: accept_terms_hash, business_name (B2B) – /legal/terms.json","url":"https://x402.nexari.cloud/transcribe/30min/speakers?utm_source=zero.xyz","method":"POST","headers":{},"bodySchema":{"type":"object","properties":{"filename":{"type":"string","description":"Dateiname mit Endung, z. B. memo.m4a"},"language":{"type":"string","description":"de (Standard), en, auto ..."},"audio_url":{"type":"string","description":"HTTPS-URL der Audiodatei / public URL of the audio file"},"audio_base64":{"type":"string","description":"Audio als Base64 / audio as base64"},"num_speakers":{"type":"integer","description":"Nur /speakers: bekannte Sprecherzahl (1-10), verbessert die Trennung"},"business_name":{"type":"string","description":"Unternehmen, für das gehandelt wird (nur B2B)"},"privacy_contact":{"type":"string","description":"E-Mail für Datenschutz-Meldungen (Art. 33 DSGVO)"},"accept_terms_hash":{"type":"string","description":"SHA-256 der AGB aus /legal/terms.json"}}},"responseSchema":null,"example":null,"exampleRequest":null,"tags":["x402"],"displayCostAmount":"0.6","displayCostAsset":"USDC","priceDynamic":false,"priceHint":null,"priceStatus":"priced","priceSource":"probe","requiresHandshake":false,"reviewCount":0,"rating":{"score":"0.00","successRate":"0.00","reviews":0,"stars":null,"state":"unrated"},"availabilityStatus":"unknown","priceObserved":null,"sessionDeposit":null,"pricing":{"kind":"static","summary":"$0.6/call","primary":{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.6","per":"call","confidence":"exact"},"accepted":[{"kind":"static","protocol":"x402","network":"base","amountUsd":"0.6","per":"call","confidence":"exact"}]},"paymentMethods":[{"uid":"pm_JPneX1dQ-Au9qfhFHsziS","protocol":"x402","methodType":"crypto","chain":"base","mode":"charge","costAmount":"0.6","costPer":"request","priority":0,"asset":"0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913","unit":"request","depositMicros":null,"planRef":null}],"brandName":null,"brandSlug":null,"brandBaseUrl":null,"brandDocsUrl":null,"whatItDoes":"Transcribes audio files up to 30 minutes with speaker separation (diarization) using pyannote + Whisper, optimized for German, processed in Germany under GDPR.","exampleAgentPrompt":"Can you transcribe this 25-minute team meeting recording for me — it's a German business call with 3 speakers — and label who said what? The audio is at https://storage.example.com/meeting-2024-06-10.m4a and my company is Acme GmbH.","exampleUseCases":[{"title":"Meeting minutes with speaker attribution","prompt":"I have a 28-minute board meeting recording in German with 4 speakers — can you transcribe it and label each speaker's turns? The file is at https://files.example.com/board-meeting.m4a and my company is Müller AG."},{"title":"Podcast interview transcription","prompt":"Please transcribe this 22-minute German podcast interview between 2 people and show me who said what throughout — the audio URL is https://cdn.podcast.example.com/episode42.m4a. We're a media company called Redaktions GmbH."},{"title":"Legal deposition audio to labeled transcript","prompt":"I need a full speaker-labeled transcript of this 30-minute deposition audio in German — there are 3 speakers: a lawyer, a witness, and a judge. The file is at https://secure.lawfirm.example.com/deposition-2024.m4a and the business name is Kanzlei Becker & Partner."}],"resultDescription":"A structured transcript with speaker-separated segments, each labeled by speaker identifier (e.g. SPEAKER_00, SPEAKER_01), including text and timing information for each turn. Output covers the full audio up to 30 minutes with improved accuracy when the number of speakers is provided.","failureModes":["Audio exceeds 30-minute limit — request rejected","Invalid or inaccessible audio URL — download fails","Missing accept_terms_hash or business_name — request rejected as B2B terms not accepted","Unsupported audio format or corrupt file — transcription fails","num_speakers out of range (must be 1-10) — validation error","Payment not included or insufficient — HTTP 402 returned","Audio language mismatch with specified language code — reduced accuracy"],"whenToPreferThis":"Choose this endpoint when you need speaker-attributed transcription for German or multilingual audio files up to 30 minutes long, require GDPR-compliant processing in Germany, and need diarization (who said what) rather than plain transcription. Prefer the 10-minute or 2-minute variants for shorter audio to reduce cost. Use the plain transcription endpoints (without /speakers) when speaker separation is not needed.","instructions":null,"reviewSummary":null,"reviewSummaryHighlights":null,"reviewSummaryConcerns":null,"reviewSummaryGeneratedAt":null,"activationCount":0,"lastUsedAt":null,"lastSuccessfullyRanAt":null,"lastHealthCheckAt":"2026-10-02T12:30:06.725Z","isFirstParty":false,"canonicalSlug":"nexari-speech-to-text-with-speaker-diarization-30-min-b40fb32b"}