Moonshine Base

Free.ai (self-hosted) · stt · ~500 ഒരു സന്ദേശത്തിനും അടയാളങ്ങള്‍ നല്‍കുക minute

ഒരു ഓഡിയോ അല്ലെങ്കില്‍ വീഡിയോ ഫയല്‍ താഴെയിടുക അല്ലെങ്കില്‍ ഒരു യുആര്‍എല്‍ പകര്‍ത്തുക

~500 ഒരു സന്ദേശത്തിനും അടയാളങ്ങള്‍ നല്‍കുക minute

Moonshine Base is a സംസാരത്തിനുള്ള വാചക മാതൃക built by Useful Sensors. Strongest at Low-latency live transcription, embedded devices.. Free.ai GPUs-ല്‍ സ്വയം-ഹോസ്റ്റ് ചെയ്തിരിക്കുന്നു — നിങ്ങളുടെ ദിവസേനയുള്ള ടോക്കണ്‍ പൗളിനെതിരെ സ്വതന്ത്രമായി പ്രവര്‍ത്തിക്കുന്നു (500 ടോക്കണ്‍സ് ഒരു മിനിറ്റ്). Released under MIT - commercial use permitted on Free.ai.

_API വഴി ഉപയോഗിക്കുക

OpAI- യോജിപ്പുള്ള റേസ്റ്റ് API. ഒരു കീ ഉണ്ടാക്കൂ, ഈ മോഡലിനെ സെക്കന്‍ഡുകളില്‍ തന്നെ വിളിക്കുക.

curl -X POST https://api.free.ai/v1/stt/ \
  -H "Authorization: Bearer sk-free-..." \
  -H "Content-Type: application/json" \
  -d '{"model":"moonshine-base","audio_url":"https://..."}'
എപിഐ സഹായക്കുറിപ്പുകള്‍ API കീ ലഭ്യമാക്കുക

പലപ്പോഴും ചോദിക്കപ്പെടുന്ന ചോദ്യങ്ങൾ

Moonshine Base transcribes spoken audio into text. Upload an MP3, WAV, M4A, or video file and Moonshine Base returns the full transcript plus optional SRT/VTT subtitles with timestamps.

Moonshine Base handles dozens of languages - Whisper-family models cover 90+, others vary. Pick "auto-detect" or specify the language for highest accuracy.

Word-error rate is 5-10% on clean English audio, 10-20% on noisy or accented audio. Large variants of the same architecture do meaningfully better on hard cases - pick larger when the audio is rough.

അതെ, ഓരോ ഭാഗവും തുടങ്ങുക/ തുടങ്ങുക. SRT അല്ലെങ്കില്‍ VTT എന്നോ കാലോപകരണം നിങ്ങളുടെ വീഡിയോയില്‍ നേരിട്ട് ലഭ്യമാക്കുക.

Moonshine Base runs on our own GPUs against your daily free pool first; $5 → 200,000 paid tokens after that. About ~500 tokens per minute.

MP3, WAV, M4A, FLAC, OGG, plus video (MP4, MOV, WebM) - we extract the audio. Max 500 MB per upload. Longer files? Split with /audio/cut/ or use /v1/stt/batch/.

ശബ്ദകര്‍ത്താവ് (അടിസ്ഥാനങ്ങള്‍) എന്നറിയപ്പെടുന്ന ഒരു ഷീറ്ററിങ്ങ് ആണ് - /tranchannel /. Moonshine Base- ല്‍ "ഡിആറസി" എന്ന പരമ്പരയെ കൈകാര്യം ചെയ്യുന്നു; ഓരോ ഭാഗവും സ്പീഡര്‍ 1 / 2 / etc.

അതെ / bache/subject files. ഓരോ ഓഡിയോ- ശേഖരവും // കാഷ്/ സ്വീകരിയ്ക്കുക. /acam- config/? Tab=മുഴുവന്‍ പേരു്‌ കൂടെയുള്ളതാണു്. അറ- വൃക്ഷം സംരക്ഷിക്കുന്നതിനായി API ഉപയോഗിക്കുന്നു.

Yes - POST your audio to /v1/stt/transcribe/ with model="Moonshine Base". Returns JSON with text + segments + word-level timestamps. /api/ has the full reference.

GPUS-ല്‍ സ്വയമേയ മോഡല്‍സ് ഓഡിയോ സൂക്ഷിക്കുന്നു; ഒരു DPA വഴി കടന്നു പോകുന്നു. അതു് ഒരു പങ്കാളിത്ത-അന്‍-ആങ്കണത്തിനു ശേഷം (24HAn- 7d--ഇന്‍) നീക്കം ചെയ്യുന്നു. നിങ്ങളുടെ ഇന്‍പുട്ടുകളില്‍ ഞങ്ങള്‍ പരിശീലിക്കുന്നില്ല.

അതേ, Free.ai - ത്തോളം പേർക്ക് റെക്കോർഡ്‌ ചെയ്‌തിരിക്കുന്ന ഓഡിയോ - യുടെ (നിങ്ങളുടെ സ്വന്തം റെക്കോർഡ്‌, ലൈസന്‍സ്‌ ചെയ്‌തിരിക്കുന്ന വിവരങ്ങൾ അല്ലെങ്കിൽ സമ്മതത്തോടെ) വാണിജ്യ സാങ്കേതിക വിദ്യകൾ ലഭ്യമാണ്‌.

യഥാര്‍ത്ഥ- സമയം ആവശ്യത്തിനു് 0.05- 02× ആണ്. — 3-12 മിനുട്ടില്‍ 60-മിനിട പോസ്റ്റ് ട്രാന്‍ ബോര്‍ഡുകള്‍. പ്രീമിയം മോഡല്‍സ് പലപ്പോഴും വേഗത്തില്‍ പൂര്‍ത്തിയാക്കുന്നു. റെയിം ബട്ടണ്‍ ടാബ് അടയ്ക്കാന്‍ ഉപയോഗിയ്ക്കുക.

സ്നേഹം Free.ai, കൂട്ടുകാരോട് പറയൂ!

ഈ താള്‍ അനുബന്ധപ്പെടുത്തുക