Wizper (Whisper v3)

Free.ai · stt · ~500 ಇದಕ್ಕಾಗಿನ ಸೂಚನೆಗಳು minute

ಒಂದು ಆಡಿಯೊ ಅಥವ ವೀಡಿಯೊ ಕಡತವನ್ನು ಬೀಳಿಸು ಅಥವಾ ಒಂದು URL ಅನ್ನು ಕೆಳಗಿನಿಂದ ಅಂಟಿಸು

~500 ಇದಕ್ಕಾಗಿನ ಸೂಚನೆಗಳು minute

Wizper (Whisper v3) is a ಪದಗಳನ್ನು ಅವಲೋಕನದ ಮಾದರಿ. Routed through external models - ~500 tokens ನಿಮಿಷಕ್ಕೆ (50% markup over upstream cost).

API ಮೂಲಕ ಬಳಸು

OpI-ಹೊಂದಿದ RACT API. ಒಂದು ಕೀಲಿಯನ್ನು ನಿರ್ಮಿಸಿ ನಂತರ ಈ ಮಾದರಿಯನ್ನು ಸೆಕೆಂಡುಗಳಲ್ಲಿ ಕರೆಯಿರಿ.

curl -X POST https://api.free.ai/v1/stt/ \
  -H "Authorization: Bearer sk-free-..." \
  -H "Content-Type: application/json" \
  -d '{"model":"premium/wizper","audio_url":"https://..."}'
API ದಸ್ತಾವೇಜೀಕರಣ API ಕೀಲಿಯನ್ನು ಪಡೆದುಕೊಳ್ಳಿ

ಅನೇಕವೇಳೆ ಪ್ರಶ್ನೆಗಳು

Wizper (Whisper v3) transcribes spoken audio into text. Upload an MP3, WAV, M4A, or video file and Wizper (Whisper v3) returns the full transcript plus optional SRT/VTT subtitles with timestamps.

Wizper (Whisper v3) handles dozens of languages - Whisper-family models cover 90+, others vary. Pick "auto-detect" or specify the language for highest accuracy.

Word-error rate is 5-10% on clean English audio, 10-20% on noisy or accented audio. Large variants of the same architecture do meaningfully better on hard cases - pick larger when the audio is rough.

ಹೌದು, ಪ್ರತಿಯೊಂದು ಭಾಗವು ಆರಂಭ/ ಅಂತ್ಯದ ಸಮನಗಳನ್ನು ಒಳಗೊಂಡಿದೆ. ಆರೋಹಣೆ SRT ಅಥವಾ VTT ಎಂದು ರಫ್ತು ಮಾಡಿ ಹಾಗು ಸಮಯಗಳು ನಿಮ್ಮ ವೀಡಿಯೋನಲ್ಲಿ ನೆಟ್ಟಗೆ.

Wizper (Whisper v3) is a premium transcription engine. About ~500-1,500 tokens per minute of audio, shown live before you generate. Self-hosted transcription runs free on your 30,000-token daily pool; premium models are pay-as-you-go, with top-ups from $1.

MP3, WAV, M4A, FLAC, OGG, plus video (MP4, MOV, WebM) - we extract the audio. Max 500 MB per upload. Longer files? Split with /audio/cut/ or use /v1/stt/batch/.

Speaker diarization is a separate pass - toggle "diarize" on /transcribe/. Wizper (Whisper v3) handles the transcription; diarization labels each segment with Speaker 1 / Speaker 2 / etc.

ಹೌದು / bautch/sand files. ಪ್ರತಿಯೊಂದು ಉಪಾಯದ/account/? Tab=tab name ಎಂಬ ಹೆಸರಿನೊಂದಿಗೆ ಆಸ್ವಾತವಾದ. ಕಡತಕೋಶವನ್ನು ಉಳಿಸಲು API ಅನ್ನು ಬಳಸುತ್ತದೆ.

Yes - POST your audio to /v1/stt/transcribe/ with model="Wizper (Whisper v3)". Returns JSON with text + segments + word-level timestamps. /api/ has the full reference.

GPUS ನಲ್ಲಿ ಮಾತ್ರ ಆಡಿಯೊವನ್ನು ಇರಿಸಲಾಗುತ್ತದೆ. ಒಂದು DPA ಸಹಿಯೊಂದಿಗೆ ಹಾದು ಹೋಗುತ್ತದೆ. ಧ್ವನಿಗೂಡಿಸುವಿಕೆಯು ವಿದ್ಯುತ್ಧಕದ (24HA ಆನ್, 7d- ಅಂತಸ್ತು) ನಂತರ ಅಳಿಸಲ್ಪಟ್ಟಿದೆ. ನಾವು ನಿಮ್ಮ ಪ್ರದಾನಗಳಲ್ಲಿ ತರಬೇತಿ ಪಡೆಯುವುದಿಲ್ಲ.

ಹೌದು, Free.ai ಮಂದಿ ಈ ರೋಚಿಂಗ್‌ಗಳನ್ನು ವಾಣಿಜ್ಯ ವ್ಯವಹಾರವನ್ನು ನೀಡುತ್ತಾರೆ.

Real-time factor is roughly 0.05-0.2× - a 60-minute podcast transcribes in 3-12 minutes. Premium models often finish faster. Use the queue button to close the tab.

Like this tool? Share it!

ಈ ಪುಟಕ್ಕೆ ರೇಖಾರೂಪಿಸು