Fal Speech-to-Text

Free.ai · stt · ~500 సూచనలు minute

ఆడియో లేదా వీడియో ఫైలును డ్రాప్‌చేయి లేదా ఒక URLను క్రిందకు అతికించుము

~500 సూచనలు minute
మా జిపిన్స్ ఉచిత అమలు. దీని కొరకు ఉన్నతీకరించు Fal Speech-to-Text →

Fal Speech-to-Text a పదాల నుండి పాఠ్యముకు నమూనా. బాహ్య మోడల్ ద్వారా మెసేజ్‌లు — ~500 చిహ్నాలు నిమిషానికి (50% రుద్దు అడ్రస్‌ వుటివ్‌ లొమాస్‌).

API ద్వారా వుపయోగించుము

OpI- సారూప్యమైన RAPI. కీ సృష్టించి ఈ మాదిరిని సెకనులనందు కాల్.

curl -X POST https://api.free.ai/v1/stt/ \
  -H "Authorization: Bearer sk-free-..." \
  -H "Content-Type: application/json" \
  -d '{"model":"premium/speech-to-text","audio_url":"https://..."}'
APIపత్రరచన API కీని పొందుము

ప్రశ్నలు

Fal Speech-to-Text transcribes spoken audio into text. Upload an MP3, WAV, M4A, or video file and Fal Speech-to-Text returns the full transcript plus optional SRT/VTT subtitles with timestamps.

Fal Speech-to-Text handles dozens of languages - Whisper-family models cover 90+, others vary. Pick "auto-detect" or specify the language for highest accuracy.

Word-error rate is 5-10% on clean English audio, 10-20% on noisy or accented audio. Large variants of the same architecture do meaningfully better on hard cases - pick larger when the audio is rough.

అవును —⁠ ప్రతి భాగము ప్రారంభం/ సెకనులు కలిగివుంటుంది. ఎస్ ఆర్టిటి లేదా VTT లా ఎగుమతి మరియు టైమ్ పటాలు మీ వీడియో లోకి నేరుగా చేరుస్తాయి.

Fal Speech-to-Text is a premium transcription engine. About ~500-1,500 tokens per minute of audio, shown live before you generate. Self-hosted transcription runs free on your 30,000-token daily pool; premium models are pay-as-you-go, with top-ups from $1.

MP3, WAV, M4A, FLAC, OGG, plus video (MP4, MOV, WebM) - we extract the audio. Max 500 MB per upload. Longer files? Split with /audio/cut/ or use /v1/stt/batch/.

సేకరణ అనేది ప్రత్యేక స్ట్రీమర్ల ప్రొఫైల్‌ను సేకరణ చేసే / artanranch/. Fal Speech-to-Text న "diz" ను ప్రత్యేక పాస్‌ డైలారేషన్ అంటారు. డైలాజేషన్ ప్రతి భాగాన్నీ స్ట్రీమర్ 1 / 2 మరియు లీడర్ 2 ".

అవును / bache/soft/ అంగీకరించు. // cogam- లొని ప్రతి ఫారెన్ భూమిలు //accamia? Tab=tab వుత్ప్రత్యేక ఫైలుతో ఆకృతీకరించబడుతుంది. file- Talrip- Trees కొరకు API వుపయోగిస్తోంది.

Yes - POST your audio to /v1/stt/transcribe/ with model="Fal Speech-to-Text". Returns JSON with text + segments + word-level timestamps. /api/ has the full reference.

Mygon-హోస్టు అయిన మాడ్యూస్ ఆడియోను GPUS న ఉంచుతుంది; A DPA ద్వారా దాటి పోతే. కరపత్రం స్వాగతం (24H ఆన్, 7d- సైన్సింగ్) తరువాత అది తొలగిపోతుంది. మేము మీ ఇన్పుట్సుల మీద శిక్షణ కాదు.

Yes - Free.ai grants commercial use of transcripts. You need rights to the audio you uploaded (your own recording, licensed material, or content with consent).

Real-time factor is roughly 0.05-0.2× - a 60-minute podcast transcribes in 3-12 minutes. Premium models often finish faster. Use the queue button to close the tab.

ప్రేమ Free.ai మీ స్నేహితులను చెప్పండి!

ఈ పేజీకి రేట్ చేయుము