Moonshine Base

Free.ai (self-hosted) · stt · ~500 ಇದಕ್ಕಾಗಿನ ಸೂಚನೆಗಳು minute

ಒಂದು ಆಡಿಯೊ ಅಥವ ವೀಡಿಯೊ ಕಡತವನ್ನು ಬೀಳಿಸು ಅಥವಾ ಒಂದು URL ಅನ್ನು ಕೆಳಗಿನಿಂದ ಅಂಟಿಸು

~500 ಇದಕ್ಕಾಗಿನ ಸೂಚನೆಗಳು minute

Moonshine Base is a ಪದಗಳನ್ನು ಅವಲೋಕನದ ಮಾದರಿ built by Useful Sensors. Strongest at Low-latency live transcription, embedded devices.. Free.ai GPU ಗಳ ಮೇಲೆ ಸ್ವಯಂ- ಆತಿಥೇಯಗೊಂಡಿದೆ - ನಿಮ್ಮ ದಿನನಿತ್ಯದ ಟೋಕನ್ ಪೂಲ್ (500 ಟೋಕನ್ ನಿಮಿಷಕ್ಕೆ) ವಿರುದ್ಧ ಉಚಿತವಾಗಿ ಚಲಿಸುತ್ತದೆ. Released under MIT - commercial use permitted on Free.ai.

API ಮೂಲಕ ಬಳಸು

OpI-ಹೊಂದಿದ RACT API. ಒಂದು ಕೀಲಿಯನ್ನು ನಿರ್ಮಿಸಿ ನಂತರ ಈ ಮಾದರಿಯನ್ನು ಸೆಕೆಂಡುಗಳಲ್ಲಿ ಕರೆಯಿರಿ.

curl -X POST https://api.free.ai/v1/stt/ \
  -H "Authorization: Bearer sk-free-..." \
  -H "Content-Type: application/json" \
  -d '{"model":"moonshine-base","audio_url":"https://..."}'
API ದಸ್ತಾವೇಜೀಕರಣ API ಕೀಲಿಯನ್ನು ಪಡೆದುಕೊಳ್ಳಿ

ಅನೇಕವೇಳೆ ಪ್ರಶ್ನೆಗಳು

Moonshine Base transcribes spoken audio into text. Upload an MP3, WAV, M4A, or video file and Moonshine Base returns the full transcript plus optional SRT/VTT subtitles with timestamps.

Moonshine Base handles dozens of languages - Whisper-family models cover 90+, others vary. Pick "auto-detect" or specify the language for highest accuracy.

Word-error rate is 5-10% on clean English audio, 10-20% on noisy or accented audio. Large variants of the same architecture do meaningfully better on hard cases - pick larger when the audio is rough.

ಹೌದು, ಪ್ರತಿಯೊಂದು ಭಾಗವು ಆರಂಭ/ ಅಂತ್ಯದ ಸಮನಗಳನ್ನು ಒಳಗೊಂಡಿದೆ. ಆರೋಹಣೆ SRT ಅಥವಾ VTT ಎಂದು ರಫ್ತು ಮಾಡಿ ಹಾಗು ಸಮಯಗಳು ನಿಮ್ಮ ವೀಡಿಯೋನಲ್ಲಿ ನೆಟ್ಟಗೆ.

Moonshine Base runs on our own GPUs against your daily free pool first; $5 → 200,000 paid tokens after that. About ~500 tokens per minute.

MP3, WAV, M4A, FLAC, OGG, plus video (MP4, MOV, WebM) - we extract the audio. Max 500 MB per upload. Longer files? Split with /audio/cut/ or use /v1/stt/batch/.

Speaker diarization is a separate pass - toggle "diarize" on /transcribe/. Moonshine Base handles the transcription; diarization labels each segment with Speaker 1 / Speaker 2 / etc.

ಹೌದು / bautch/sand files. ಪ್ರತಿಯೊಂದು ಉಪಾಯದ/account/? Tab=tab name ಎಂಬ ಹೆಸರಿನೊಂದಿಗೆ ಆಸ್ವಾತವಾದ. ಕಡತಕೋಶವನ್ನು ಉಳಿಸಲು API ಅನ್ನು ಬಳಸುತ್ತದೆ.

Yes - POST your audio to /v1/stt/transcribe/ with model="Moonshine Base". Returns JSON with text + segments + word-level timestamps. /api/ has the full reference.

GPUS ನಲ್ಲಿ ಮಾತ್ರ ಆಡಿಯೊವನ್ನು ಇರಿಸಲಾಗುತ್ತದೆ. ಒಂದು DPA ಸಹಿಯೊಂದಿಗೆ ಹಾದು ಹೋಗುತ್ತದೆ. ಧ್ವನಿಗೂಡಿಸುವಿಕೆಯು ವಿದ್ಯುತ್ಧಕದ (24HA ಆನ್, 7d- ಅಂತಸ್ತು) ನಂತರ ಅಳಿಸಲ್ಪಟ್ಟಿದೆ. ನಾವು ನಿಮ್ಮ ಪ್ರದಾನಗಳಲ್ಲಿ ತರಬೇತಿ ಪಡೆಯುವುದಿಲ್ಲ.

ಹೌದು, Free.ai ಮಂದಿ ಈ ರೋಚಿಂಗ್‌ಗಳನ್ನು ವಾಣಿಜ್ಯ ವ್ಯವಹಾರವನ್ನು ನೀಡುತ್ತಾರೆ.

Real-time factor is roughly 0.05-0.2× - a 60-minute podcast transcribes in 3-12 minutes. Premium models often finish faster. Use the queue button to close the tab.

Like this tool? Share it!

ಈ ಪುಟಕ್ಕೆ ರೇಖಾರೂಪಿಸು