Whisper local / open source
ਮੁਫ਼ਤ ਜੇ ਤੁਕੱਲ ਕੋਲ GPU ਅਤੇ ਇੱਕ afternoon ਹੈ। Speaker diarization box ਤੋਂ ਬਾਹਰ ਨਹੀਂ।
ਕਿਸੇ ਵੀ bitrate 'ਤੇ 64 ਤੋਂ 320 kbps ਤੱਕ MP3 ਫ਼ਾਈਲ ਡ੍ਰੌਪ ਕਰੋ। 99 ਭਾਸ਼ਾਵਾਂ ਵਿਚ timestamp ਕੀਤਾ, speaker-labeled transcript ਪ੍ਰਾਪਤ ਕਰੋ — ਕੋਈ format conversion, ਕੋਈ re-encoding, ਕੋਈ queue 'ਤੇ ਇੰਤਜ਼ਾਰ ਨਹੀਂ।
MP3 · WAV · M4A · MP4 · MOV · MKV · OGG · OPUS · FLAC · WEBM — up to 100 MB anonymously
YouTube · TikTok · Vimeo · Twitter · SoundCloud · Spotify · 50+ more
↓ ਵੇਖੋ ਕਿ ਕੀ ਨਿਕਲਦਾ ਹੈ
ਅਸੀਂ MP3 frame headers ਸਿੱਧੇ ਪੜ੍ਹਦੇ ਹਾਂ — VBR, CBR, joint-stereo, ਕੋਈ ਵੀ encoder (LAME, Fraunhofer, FFmpeg)। ਜੇ ਫ਼ਾਈਲ ਅਲੱਗ-ਅਲੱਗ ਚੈਨਲਾਂ 'ਤੇ ਸਪੀਕਰ ਦੇ ਨਾਲ true stereo ਹੈ, ਅਸੀਂ voices ਨੂੰ ਵੰਡਣ ਲਈ ਇਸ ਦੀ ਵਰਤੋਂ ਕਰਦੇ ਹਾਂ। Mono mix-down acoustic diarization 'ਤੇ ਵਾਪਸ ਜਾਂਦਾ ਹੈ।
ਤੁਸੀਂ ਪਹਿਲਾਂ ਕਦੋਂ ਸਮਝਿਆ ਕਿ ਆਰਕਾਈਵ ਅਧੂਰੀ ਸੀ?
ਸ਼ਾਇਦ 2019 ਦੇ ਆਸ-ਪਾਸ, ਜਦੋਂ ਮੇਂ reel-to-reels ਨੂੰ digitise ਕਰਨਾ ਸ਼ੁਰੂ ਕੀਤਾ।
ਅਤੇ ਗੁਆਪਤ ਟੇਪਾਂ — ਕੀ ਉਹ ਕਹੀਂ ਵੀ catalogue ਕੀਤੀਆਂ ਗਈਆਂ ਸਨ?
'78 ਤੋਂ ਵੀ ਇੱਕ paper index ਹੈ, ਪਰ ਅੱਧੀ ਪਾਣੀ ਦੀ ਨੁਕਸਾਨ ਹਾ।
↓ This is the dashboard
Same layout as the real dashboard — Summary, full Transcript, Speakers tab, Exports. Key points and action items extracted automatically. Auto-tags on every job.
Sample preview from a founder interview about post-call workflow. Real transcripts look exactly like this — same tabs, same summary block, same key-points / action-items split, same auto-tag chips.
ਤਿੰਨ ਅਸਲੀ ਵਿਕਲਪ · ਸ਼ੇਮ ਅਨੁਸਾਰ
ਤੁਸੀਂ ਮੁਫ਼ਤ ਵਿੱਚ ਆਪਣੇ laptop 'ਤੇ Whisper ਨੂੰ ਚਲਾ ਸਕਦੇ ਹੋ ਜੇ ਤੁਸੀਂ ਤਕਨੀਕੀ ਹੋ। Otter ਅਤੇ Sonix subscription dashboards ਵਿਚ MP3 uploads ਸਵੀਕਾਰ ਕਰਦੇ ਹਨ। ਅਸੀਂ ਫ਼ਾਈਲ ਲੈਂਦੇ ਹਾਂ, transcript ਵਾਪਸ ਕਰਦੇ ਹਾਂ, ਅਤੇ ਤੁਹਾਨੂੰ UI ਵਿਚ ਰਹਿਣ ਲਈ ਕਹਿ ਨਹੀਂ।
ਮੁਫ਼ਤ ਜੇ ਤੁਕੱਲ ਕੋਲ GPU ਅਤੇ ਇੱਕ afternoon ਹੈ। Speaker diarization box ਤੋਂ ਬਾਹਰ ਨਹੀਂ।
MP3 ਡ੍ਰੌਪ ਕਰੋ। Speaker-labeled ਲਿਖਤ ਵਾਪਸ ਪ੍ਰਾਪਤ ਕਰੋ roughly real-time × 0.025 ਵਿਚ।
Polished dashboard, monthly minutes cap, English-tuned। File upload ਪਾਸੇ ਦੀ feature ਲੱਗਦਾ ਹੈ।
May 2026 ਤੋਂ ਸਹੀ pricing ਅਤੇ ਫ਼ੀਚਿਅਰ availability। Whisper ਪ੍ਰਦਰਸ਼ਨ model ਸਾਈਜ਼ ਅਤੇ hardware ਅਨੁਸਾਰ ਵੱਖਵੱਖ।
MP3 ਨੂੰ Specific
MP3 ਇੱਕ format ਹੈ, ਰਿਕਾਰਡਿੰਗ style ਨਹੀਂ — ਜੋ ਮਤਲਬ ਹੈ failure modes encoder ਤੋਂ ਆਉਂਦੇ ਹਨ, ਬੋਲੀ ਤੋਂ ਨਹੀਂ।
Defaults ਜੋ ~80% MP3 ਫ਼ਾਈਲਾਂ ਨੂੰ fit ਕਰਦੀਆਂ ਹਨ। Form ਤੋਂ ਪ੍ਰਤੀ-job override।
Accuracy · real-world numbers
MP3 ਸ਼ੁੱਧਤਾ ਨੂੰ bounds ਜੋ encoder ਰੱਖਿਆ, ਸਾਨੂੰ ਨਹੀਂ। Perceptual compression ~96 kbps ਤੋਂ ਉੱਪਰ speech intelligibility ਨੂੰ ਬਹੁਤ ਚੰਗੀ ਤਰ੍ਹਾਂ ਸੁਰੱਖਿਅਤ ਰੱਖਦਾ ਹੈ; 64 kbps ਤੋਂ ਘੱਟ, sibilants ਅਤੇ consonants dissolve ਹੋਣ ਲਗਦੇ ਹਨ। ਹੇਠਲੀ ਸੰਖਿਆਵਾਂ production ਵਿਚ ਅਸਲੀ customer MP3s ਤੋਂ ਹਨ।
Near-lossless ਬੋਲੀ ਲਈ। Podcast masters, dictation app exports, professional interview rigs। Diarization ਸਾਫ਼ ਹੈ ਜੇ speakers ਅਲੱਗ channels 'ਤੇ ਹਨ।
ਸਭ ਤੋਂ ਆਮ bitrate spoken-word MP3s ਲਈ। Zoom exports, Riverside downloads, voice recorders default। Compression artifacts recognizer ਲਈ inaudible।
ਜ਼ਿ voice memo defaults ਆ ਜ਼ਿਭਪਟਭ phones 'ਤੇ। Acoustic diarization 2-4 speakers ਨੂੰ ਸੰਭਾਲਦੀ ਹੈ। ਨੰਬਾਂ ਅਤੇ proper nouns ਨੂੰ ਕਦੇ ਇੱਕ ਨਜ਼ਰ ਦੀ ਲੋੜ।
ਪੁਰਾਣੀ answering-machine rips, lecture archives, narrow-band ਸੋਤ। High-frequency consonants (f/s/sh) blur। ਫਿਰ ਵੀ legible — proofread ਦੀ ਯੋਜਨਾ ਬਣਾਓ।
Common ਸਵਾਲ
ਹਰ ਮਹੀਨਾ 30 ਮੁਫ਼ਤ ਮਿੰਟ। ਕਾਰਡ ਦੀ ਲੋੜ ਨਹੀਂ। Speaker ਲੇਬਲ, 99 ਭਾਸ਼ਾਵਾਂ, ਹਰ export format include।
ਮੁਫ਼ਤ ਸ਼ੁਰੂ ਕਰੋ