Distil-Whisper large-v3

Free.ai (self-hosted) · stt · ~500 సూచనలు minute

ఆడియో లేదా వీడియో ఫైలును డ్రాప్‌చేయి లేదా ఒక URLను క్రిందకు అతికించుము

~500 సూచనలు minute

Distil-Whisper large-v3 is a పదాల నుండి పాఠ్యముకు నమూనా built by HuggingFace. Real-time transcription, large-volume batch STT. వద్ద బలప్రకాశం. Free.ai GPUs పై స్వయం-హోస్ట్ చేయబడినది — మీ రోజువారీ టోకెన్ పూల్ (500 టోకెన్లు నిమిషానికి) పై ఉచితంగా నడుస్తుంది. Released under MIT - commercial use permitted on Free.ai.

API ద్వారా వుపయోగించుము

OpI- సారూప్యమైన RAPI. కీ సృష్టించి ఈ మాదిరిని సెకనులనందు కాల్.

curl -X POST https://api.free.ai/v1/stt/ \
  -H "Authorization: Bearer sk-free-..." \
  -H "Content-Type: application/json" \
  -d '{"model":"distil-whisper-large-v3","audio_url":"https://..."}'
APIపత్రరచన API కీని పొందుము

ప్రశ్నలు

Distil-Whisper large-v3 transcribes spoken audio into text. Upload an MP3, WAV, M4A, or video file and Distil-Whisper large-v3 returns the full transcript plus optional SRT/VTT subtitles with timestamps.

Distil-Whisper large-v3 handles dozens of languages - Whisper-family models cover 90+, others vary. Pick "auto-detect" or specify the language for highest accuracy.

Word-error rate is 5-10% on clean English audio, 10-20% on noisy or accented audio. Large variants of the same architecture do meaningfully better on hard cases - pick larger when the audio is rough.

అవును —⁠ ప్రతి భాగము ప్రారంభం/ సెకనులు కలిగివుంటుంది. ఎస్ ఆర్టిటి లేదా VTT లా ఎగుమతి మరియు టైమ్ పటాలు మీ వీడియో లోకి నేరుగా చేరుస్తాయి.

Distil-Whisper large-v3 runs on our own GPUs against your daily free pool first; $5 → 200,000 paid tokens after that. About ~500 tokens per minute.

MP3, WAV, M4A, FLAC, OGG, plus video (MP4, MOV, WebM) - we extract the audio. Max 500 MB per upload. Longer files? Split with /audio/cut/ or use /v1/stt/batch/.

సేకరణ అనేది ప్రత్యేక స్ట్రీమర్ల ప్రొఫైల్‌ను సేకరణ చేసే / artanranch/. Distil-Whisper large-v3 న "diz" ను ప్రత్యేక పాస్‌ డైలారేషన్ అంటారు. డైలాజేషన్ ప్రతి భాగాన్నీ స్ట్రీమర్ 1 / 2 మరియు లీడర్ 2 ".

అవును / bache/soft/ అంగీకరించు. // cogam- లొని ప్రతి ఫారెన్ భూమిలు //accamia? Tab=tab వుత్ప్రత్యేక ఫైలుతో ఆకృతీకరించబడుతుంది. file- Talrip- Trees కొరకు API వుపయోగిస్తోంది.

Yes - POST your audio to /v1/stt/transcribe/ with model="Distil-Whisper large-v3". Returns JSON with text + segments + word-level timestamps. /api/ has the full reference.

Mygon-హోస్టు అయిన మాడ్యూస్ ఆడియోను GPUS న ఉంచుతుంది; A DPA ద్వారా దాటి పోతే. కరపత్రం స్వాగతం (24H ఆన్, 7d- సైన్సింగ్) తరువాత అది తొలగిపోతుంది. మేము మీ ఇన్పుట్సుల మీద శిక్షణ కాదు.

Yes - Free.ai grants commercial use of transcripts. You need rights to the audio you uploaded (your own recording, licensed material, or content with consent).

Real-time factor is roughly 0.05-0.2× - a 60-minute podcast transcribes in 3-12 minutes. Premium models often finish faster. Use the queue button to close the tab.

ప్రేమ Free.ai మీ స్నేహితులను చెప్పండి!

ఈ పేజీకి రేట్ చేయుము