ElevenLabs STT

Free.ai · stt · ~500 निशानियाँ प्रति minute

ऑडियो या वीडियो फ़ाइल को छोड़ें, या नीचे URL चिपकाएँ

~500 निशानियाँ प्रति minute
हमारे जीपीपी पर मुक्त चला जाता है. के लिए उन्नयन ElevenLabs STT →

ElevenLabs STT is a पाठ मॉडल. Routed through external models - ~500 tokens प्रति मिनट (50% markup over upstream cost).

एपीआई से प्रयोग करें

कृत्रिमताREST API खोलें. एक कुंजी बनाएँ और इस मॉडल को सेकेंड में कॉल करें.

curl -X POST https://api.free.ai/v1/stt/ \
  -H "Authorization: Bearer sk-free-..." \
  -H "Content-Type: application/json" \
  -d '{"model":"premium/elevenlabs/speech-to-text","audio_url":"https://..."}'
एपीआई प्रलेखन एपीआई कुंजी प्राप्त करें

बार बार पूछे जाने वाले प्रश्न

ElevenLabs STT transcribes spoken audio into text. Upload an MP3, WAV, M4A, or video file and ElevenLabs STT returns the full transcript plus optional SRT/VTT subtitles with timestamps.

ElevenLabs STT handles dozens of languages - Whisper-family models cover 90+, others vary. Pick "auto-detect" or specify the language for highest accuracy.

वर्ड-त्रुटि दर 5-10% शुद्ध अंग्रेजी ऑडियो पर है, 10-20% ध्वनि पर आवाज़ या उच्चारण पर. समान संरचना के बड़े अंतरों को मुश्किल मामलों पर अर्थपूर्ण रूप से बेहतर करते हैं जब ऑडियो को किसी तरह से कम किया जाता है.

हाँ, प्रत्येक खण्ड में प्रारंभ/ends शामिल हैं. एसआरटी या विट के रूप में निर्यात करें और समय चिह्न आपके वीडियो में सीधा.

ElevenLabs STT is a premium transcription engine. About ~500-1,500 tokens per minute of audio, shown live before you generate. Self-hosted transcription runs free on your 30,000-token daily pool; premium models are pay-as-you-go, with top-ups from $1.

MP3, WAV, M4A, FLAC, OGG, plus video (MP4, MOV, WebM) - we extract the audio. Max 500 MB per upload. Longer files? Split with /audio/cut/ or use /v1/stt/batch/.

Speaker diarization is a separate pass - toggle "diarize" on /transcribe/. ElevenLabs STT handles the transcription; diarization labels each segment with Speaker 1 / Speaker 2 / etc.

हां / / बेर्क ऑडियो फ़ाइलों के एक फ़ोल्डर को स्वीकार करता है. हर प्रदर्शन देश / tab/ tabbits मूल फ़ाइलनाम के साथ इतिहास. फ़ोल्डर संरक्षण के लिए Garaba.

Yes - POST your audio to /v1/stt/transcribe/ with model="ElevenLabs STT". Returns JSON with text + segments + word-level timestamps. /api/ has the full reference.

स्व-होड मॉडल हमारे जीपिप्टी पर ऑडियो रखते हैं; पूर्व-प्रानियम एक डीपी. से पार हो रहे हैं. साझा-24, 7 पर क्लिक करने के बाद (४-i). हम आपके इनपुट पर ट्रेन नहीं है.

जी हाँ, Free.ai नकल करने का कारोबार प्रदान करता है ।

Real-time factor is roughly 0.05-0.2× - a 60-minute podcast transcribes in 3-12 minutes. Premium models often finish faster. Use the queue button to close the tab.

Love Free.ai? Tell your friends!

इस पृष्ठ को दर करें