Distil-Whisper large-v3

Free.ai (self-hosted) · stt · ~500 निशानियाँ प्रति minute

ऑडियो या वीडियो फ़ाइल को छोड़ें, या नीचे URL चिपकाएँ

~500 निशानियाँ प्रति minute

Distil-Whisper large-v3 is a पाठ मॉडल built by HuggingFace. Strongest at Real-time transcription, large-volume batch STT.. Free.ai GPUs पर स्व-होस्ट - आपके दैनिक टोकन पूल (500 टोकन प्रति मिनट) के खिलाफ मुफ्त चलाता है. Released under MIT - commercial use permitted on Free.ai.

एपीआई से प्रयोग करें

कृत्रिमताREST API खोलें. एक कुंजी बनाएँ और इस मॉडल को सेकेंड में कॉल करें.

curl -X POST https://api.free.ai/v1/stt/ \
  -H "Authorization: Bearer sk-free-..." \
  -H "Content-Type: application/json" \
  -d '{"model":"distil-whisper-large-v3","audio_url":"https://..."}'
एपीआई प्रलेखन एपीआई कुंजी प्राप्त करें

बार बार पूछे जाने वाले प्रश्न

Distil-Whisper large-v3 transcribes spoken audio into text. Upload an MP3, WAV, M4A, or video file and Distil-Whisper large-v3 returns the full transcript plus optional SRT/VTT subtitles with timestamps.

Distil-Whisper large-v3 handles dozens of languages - Whisper-family models cover 90+, others vary. Pick "auto-detect" or specify the language for highest accuracy.

वर्ड-त्रुटि दर 5-10% शुद्ध अंग्रेजी ऑडियो पर है, 10-20% ध्वनि पर आवाज़ या उच्चारण पर. समान संरचना के बड़े अंतरों को मुश्किल मामलों पर अर्थपूर्ण रूप से बेहतर करते हैं जब ऑडियो को किसी तरह से कम किया जाता है.

हाँ, प्रत्येक खण्ड में प्रारंभ/ends शामिल हैं. एसआरटी या विट के रूप में निर्यात करें और समय चिह्न आपके वीडियो में सीधा.

Distil-Whisper large-v3 runs on our own GPUs against your daily free pool first; $5 → 200,000 paid tokens after that. About ~500 tokens per minute.

MP3, WAV, M4A, FLAC, OGG, plus video (MP4, MOV, WebM) - we extract the audio. Max 500 MB per upload. Longer files? Split with /audio/cut/ or use /v1/stt/batch/.

Speaker diarization is a separate pass - toggle "diarize" on /transcribe/. Distil-Whisper large-v3 handles the transcription; diarization labels each segment with Speaker 1 / Speaker 2 / etc.

हां / / बेर्क ऑडियो फ़ाइलों के एक फ़ोल्डर को स्वीकार करता है. हर प्रदर्शन देश / tab/ tabbits मूल फ़ाइलनाम के साथ इतिहास. फ़ोल्डर संरक्षण के लिए Garaba.

Yes - POST your audio to /v1/stt/transcribe/ with model="Distil-Whisper large-v3". Returns JSON with text + segments + word-level timestamps. /api/ has the full reference.

स्व-होड मॉडल हमारे जीपिप्टी पर ऑडियो रखते हैं; पूर्व-प्रानियम एक डीपी. से पार हो रहे हैं. साझा-24, 7 पर क्लिक करने के बाद (४-i). हम आपके इनपुट पर ट्रेन नहीं है.

जी हाँ, Free.ai नकल करने का कारोबार प्रदान करता है ।

Real-time factor is roughly 0.05-0.2× - a 60-minute podcast transcribes in 3-12 minutes. Premium models often finish faster. Use the queue button to close the tab.

Love Free.ai? Tell your friends!

इस पृष्ठ को दर करें