faster-whisper large-v3-turbo

Free.ai (self-hosted) · stt · ~500 निशानियाँ प्रति minute

ऑडियो या वीडियो फ़ाइल को छोड़ें, या नीचे URL चिपकाएँ

~500 निशानियाँ प्रति minute

faster-whisper large-v3-turbo is a पाठ मॉडल built by OpenAI / SYSTRAN. Strongest at Accurate transcription. Free.ai GPUs पर स्व-होस्ट - आपके दैनिक टोकन पूल (500 टोकन प्रति मिनट) के खिलाफ मुफ्त चलाता है. Released under MIT - commercial use permitted on Free.ai.

एपीआई से प्रयोग करें

कृत्रिमताREST API खोलें. एक कुंजी बनाएँ और इस मॉडल को सेकेंड में कॉल करें.

curl -X POST https://api.free.ai/v1/stt/ \
  -H "Authorization: Bearer sk-free-..." \
  -H "Content-Type: application/json" \
  -d '{"model":"whisper","audio_url":"https://..."}'
एपीआई प्रलेखन एपीआई कुंजी प्राप्त करें

बार बार पूछे जाने वाले प्रश्न

faster-whisper large-v3-turbo transcribes spoken audio into text. Upload an MP3, WAV, M4A, or video file and faster-whisper large-v3-turbo returns the full transcript plus optional SRT/VTT subtitles with timestamps.

faster-whisper large-v3-turbo handles dozens of languages - Whisper-family models cover 90+, others vary. Pick "auto-detect" or specify the language for highest accuracy.

वर्ड-त्रुटि दर 5-10% शुद्ध अंग्रेजी ऑडियो पर है, 10-20% ध्वनि पर आवाज़ या उच्चारण पर. समान संरचना के बड़े अंतरों को मुश्किल मामलों पर अर्थपूर्ण रूप से बेहतर करते हैं जब ऑडियो को किसी तरह से कम किया जाता है.

हाँ, प्रत्येक खण्ड में प्रारंभ/ends शामिल हैं. एसआरटी या विट के रूप में निर्यात करें और समय चिह्न आपके वीडियो में सीधा.

faster-whisper large-v3-turbo runs on our own GPUs against your daily free pool first; $5 → 200,000 paid tokens after that. About ~500 tokens per minute.

MP3, WAV, M4A, FLAC, OGG, plus video (MP4, MOV, WebM) - we extract the audio. Max 500 MB per upload. Longer files? Split with /audio/cut/ or use /v1/stt/batch/.

Speaker diarization is a separate pass - toggle "diarize" on /transcribe/. faster-whisper large-v3-turbo handles the transcription; diarization labels each segment with Speaker 1 / Speaker 2 / etc.

हां / / बेर्क ऑडियो फ़ाइलों के एक फ़ोल्डर को स्वीकार करता है. हर प्रदर्शन देश / tab/ tabbits मूल फ़ाइलनाम के साथ इतिहास. फ़ोल्डर संरक्षण के लिए Garaba.

Yes - POST your audio to /v1/stt/transcribe/ with model="faster-whisper large-v3-turbo". Returns JSON with text + segments + word-level timestamps. /api/ has the full reference.

स्व-होड मॉडल हमारे जीपिप्टी पर ऑडियो रखते हैं; पूर्व-प्रानियम एक डीपी. से पार हो रहे हैं. साझा-24, 7 पर क्लिक करने के बाद (४-i). हम आपके इनपुट पर ट्रेन नहीं है.

जी हाँ, Free.ai नकल करने का कारोबार प्रदान करता है ।

Real-time factor is roughly 0.05-0.2× - a 60-minute podcast transcribes in 3-12 minutes. Premium models often finish faster. Use the queue button to close the tab.

Love Free.ai? Tell your friends!

इस पृष्ठ को दर करें