Whisper لوڪل / اوپن سورس
مفت جيڪڏهن توهان وٽ GPU ۽ ھڪ afternoon آهي. ڪوبه speaker diarization باڪس کان ٻاهر.
ھڪ MP3 فائل 64 کان 320 kbps پوري درجي ۾ ڪري. ھڪ timestamped، speaker-labeled transcript 99 ٻولين ۾ حاصل ڪريو — ڪوبه format conversion، ڪوبه re-encoding، ڪوبه queue تي انتظار.
MP3 · WAV · M4A · MP4 · MOV · MKV · OGG · OPUS · FLAC · WEBM — up to 100 MB anonymously
YouTube · TikTok · Vimeo · Twitter · SoundCloud · Spotify · 50+ more
↓ ڏسو ڇا نڪرندو آهي
اسان MP3 frame headers سڌو سڌو پڙهندا آهيون — VBR، CBR، joint-stereo، ڪوبه encoder (LAME، Fraunhofer، FFmpeg). جيڪڏهن فائل سچو stereo آهي speakers الڳ الڳ channels تي، اسان ان کي voice split ڪرڻ لاء استعمال ڪندا آهيون. Mono mix-down acoustic diarization تي واپس ڪندو آهي.
تر جڏهن توهان پهريون ڀيرو محسوس ڪيو ته آرڪائيو نامڪمل هو؟
شايد 2019 ڦيٹي، جڏهن اسان reel-to-reels digitise ڪرڻ شروع ڪيو.
۽ گمشدہ tapes — ڇا اهي ڪٿي catalog ٿيلا هئا؟
'78 وٽ ھڪ paper index آهي، پر اسان جو اڌ پاڻي نقصان آهي.
↓ This is the dashboard
Same layout as the real dashboard — Summary, full Transcript, Speakers tab, Exports. Key points and action items extracted automatically. Auto-tags on every job.
Sample preview from a founder interview about post-call workflow. Real transcripts look exactly like this — same tabs, same summary block, same key-points / action-items split, same auto-tag chips.
ٽي حقيقي اختيار · ايمان دار موازنو
توهان Whisper اپنے laptop تي مفت تي چلا سگهو ٿا جيڪڏهن توهان technical ئو. Otter ۽ Sonix subscription dashboards اندر MP3 uploads قبول ڪندا آهن. اسان فائل وٺندا آهيون، transcript واپس ڪندا آهيون، ۽ توهان کي UI اندر رهڻ لاء مجبور نٿا ڪندا.
مفت جيڪڏهن توهان وٽ GPU ۽ ھڪ afternoon آهي. ڪوبه speaker diarization باڪس کان ٻاهر.
MP3 ڪريو. Speaker-labeled متن واپس حاصل ڪريو تقريبن real-time × 0.025 ۾.
Polished dashboard، monthly منٹز cap، English-tuned. File upload ھڪ ضمني feature جهڙو محسوس ٿو.
قيمت ۽ feature دستيابي درست May 2026 جي طور تي. Whisper performance ماڊل سائز ۽ hardware جي لحاظ سان مختلف ٿو.
MP3 لاء مخصوص
MP3 ھڪ format آهي، نه ھڪ recording سٹائل — جيڪو مطلب آهي failure modes encoder کان آيندو آهين، نه speech کان.
Defaults جيڪي ~80% MP3 فائلز کي fit ٿين. فارم سان per-job override ڪريو.
Accuracy · real-world numbers
MP3 accuracy محدود جو encoder رکيو، اسان جي طرفان نه. Perceptual compression ~96 kbps کان مٿي speech intelligibility بہت سٺي محفوظ رکندو آهي; 64 kbps کان گهٹ، sibilants ۽ consonants ڦهلي گڏڻ شروع. ڊاٽا هيٺ حقيقي customer MP3s سان production ۾.
تقريبن lossless speech لاء. Podcast ماسٽرز، dictation app exports، professional interview rigs. Diarization صاف اگر speakers الڳ الڳ channels تي.
سب سے عام bitrate spoken-word MP3s لاء. Zoom exports، Riverside downloads، voice recorders default. Compression artifacts ناقابل شناخت recognizer لاء.
Voice memo defaults اڪثر فونز تي. Acoustic diarization 2-4 speakers سنڀاليندو ��هي. نمبر ۽ proper nouns بسا ھڪ نظر دي ضرورت.
پراڻا answering-machine rips، lecture archives، narrow-band ذرائع. High-frequency consonants (f/s/sh) ڦهليندا. اڃا تائين legible — ھڪ proofread منصوبو.
عام سوالات
30 مفت منٽز ھر ماہ. ڪوبه ڪارت ضروري نه. Speaker labels، 99 ٻولي، ھر export format شامل.
مفت شروع ڪريو