Fal Speech-to-Text

Free.ai · stt · ~500 निशानियाँ प्रति minute

ऑडियो या वीडियो फ़ाइल को छोड़ें, या नीचे URL चिपकाएँ

~500 निशानियाँ प्रति minute
हमारे जीपीपी पर मुक्त चला जाता है. के लिए उन्नयन Fal Speech-to-Text →

Fal Speech-to-Text is a पाठ मॉडल. Routed through external models - ~500 tokens प्रति मिनट (50% markup over upstream cost).

एपीआई से प्रयोग करें

कृत्रिमताREST API खोलें. एक कुंजी बनाएँ और इस मॉडल को सेकेंड में कॉल करें.

curl -X POST https://api.free.ai/v1/stt/ \
  -H "Authorization: Bearer sk-free-..." \
  -H "Content-Type: application/json" \
  -d '{"model":"premium/speech-to-text","audio_url":"https://..."}'
एपीआई प्रलेखन एपीआई कुंजी प्राप्त करें

बार बार पूछे जाने वाले प्रश्न

Fal Speech-to-Text transcribes spoken audio into text. Upload an MP3, WAV, M4A, or video file and Fal Speech-to-Text returns the full transcript plus optional SRT/VTT subtitles with timestamps.

Fal Speech-to-Text handles dozens of languages - Whisper-family models cover 90+, others vary. Pick "auto-detect" or specify the language for highest accuracy.

वर्ड-त्रुटि दर 5-10% शुद्ध अंग्रेजी ऑडियो पर है, 10-20% ध्वनि पर आवाज़ या उच्चारण पर. समान संरचना के बड़े अंतरों को मुश्किल मामलों पर अर्थपूर्ण रूप से बेहतर करते हैं जब ऑडियो को किसी तरह से कम किया जाता है.

हाँ, प्रत्येक खण्ड में प्रारंभ/ends शामिल हैं. एसआरटी या विट के रूप में निर्यात करें और समय चिह्न आपके वीडियो में सीधा.

Fal Speech-to-Text is a premium transcription engine. About ~500-1,500 tokens per minute of audio, shown live before you generate. Self-hosted transcription runs free on your 30,000-token daily pool; premium models are pay-as-you-go, with top-ups from $1.

MP3, WAV, M4A, FLAC, OGG, plus video (MP4, MOV, WebM) - we extract the audio. Max 500 MB per upload. Longer files? Split with /audio/cut/ or use /v1/stt/batch/.

Speaker diarization is a separate pass - toggle "diarize" on /transcribe/. Fal Speech-to-Text handles the transcription; diarization labels each segment with Speaker 1 / Speaker 2 / etc.

हां / / बेर्क ऑडियो फ़ाइलों के एक फ़ोल्डर को स्वीकार करता है. हर प्रदर्शन देश / tab/ tabbits मूल फ़ाइलनाम के साथ इतिहास. फ़ोल्डर संरक्षण के लिए Garaba.

Yes - POST your audio to /v1/stt/transcribe/ with model="Fal Speech-to-Text". Returns JSON with text + segments + word-level timestamps. /api/ has the full reference.

स्व-होड मॉडल हमारे जीपिप्टी पर ऑडियो रखते हैं; पूर्व-प्रानियम एक डीपी. से पार हो रहे हैं. साझा-24, 7 पर क्लिक करने के बाद (४-i). हम आपके इनपुट पर ट्रेन नहीं है.

जी हाँ, Free.ai नकल करने का कारोबार प्रदान करता है ।

Real-time factor is roughly 0.05-0.2× - a 60-minute podcast transcribes in 3-12 minutes. Premium models often finish faster. Use the queue button to close the tab.

Love Free.ai? Tell your friends!

इस पृष्ठ को दर करें