đď¸ Whisper ASR â Audio & Video Transcriber đ§ Ultraâfast, onâdevice transcription â up to 20Ă faster than realâtime 22 minutes transcribed in 58 seconds using the fastest GGML Whisper model (~20Ă realâtime) â GPU-enabled Experience fast, private, and GPUâaccelerated speechâtoâtext with a full Whisperâpowered transcription suite built for creators, developers, and power users. This app brings official OpenAI Whisper models directly to your desktop with seamless Hugging Face Hub integration, resumeâsupported downloads, and a clean GGML model browser. Run everything locally â no cloud, no data sharing, no latency. Choose from Englishâonly or multilingual models, including official FP16 Whisper releases and highâperformance GGML quantized variants. A builtâin GGML Model Server lets you route, manage, and benchmark more than 30 Whisper models across tiny, base, small, medium, and large tiers. Unlock maximum performance with NVIDIA GPU acceleration powered by CUDA. Monitor realâtime inference speed, decoding throughput, and GPU utilization with the integrated performance dashboard. Whether you're transcribing longâform audio, building AI workflows, or deploying Whisper in production, this app delivers a fast, secure, and professionalâgrade transcription experience. (Benchmark hardware: RTX 4050 (6GB VRAM) + AMD Ryzen 9 â performance varies by hardware) đ Licensing & Activation Whisper ASR is free to install from the Microsoft Store. To unlock full access, a license key is required. - License keys are available via Gumroad for a one-time purchase of $59 - Note: Whisper ASR license keys are not purchasable directly through the Microsoft Store đď¸ Purchase license key: https://productgenius.gumroad.com/l/whisper-asr đ Whatâs Included: ⢠â One-time purchase â lifetime access to all future updates ⢠đ¸ No subscriptions or monthly fees ⢠đĄ A transparent, costâeffective alternative to expensive cloudâbased transcription services â Core Features & Highlights ⢠âžď¸ Unlimited audio & video transcriptions ⢠đ Transcribe speech in 100 languages ⢠đď¸ Translate to English ⢠⥠Local GPU acceleration ⢠đŤ Zero perâminute cost ⢠đ No data leaves the device ⢠đ No API dependency ⢠đ No usage caps đ¤ Official OpenAI Whisper Models ⢠Hugging Face Hub â GGML Model Browser ⢠Direct Download (resume-supported) ⢠Full library, 33 GGML Whisper models ⢠11 official FP16 ⢠22 optimized quantized variants ⢠Tiny â Base â Small â Medium â Large (v1, v2, v3, turbo) đ Private, Local AI/ML Transcription ⢠100% onâdevice processing ⢠Englishâonly and multilingual models ⢠Official FP16 Whisper models + GGML quantized variants ⢠Built on the OpenAI Transformer architecture âłď¸ GPU Acceleration ⢠NVIDIA GPUâpowered highâperformance inference ⢠Optimized for HPC/AI workloads ⢠Whisper.cpp CUDA runtime for maximum throughput ⥠Real-time GPU Performance Monitor ⢠Live decoding + transcription telemetry ⢠GPU utilization, inference speed, and decoding performance ⢠Builtâin diagnostics for CUDAâaccelerated ASR đď¸ Audio & Video Transcription ⢠đ§ Transcribe MP3, WAV, FLAC, OGG (Vorbis), MP4, MOV, WEBM, and MKV ⢠âśď¸ YouTube support â download videos as MP4 and transcribe instantly ⢠đ¤ Builtâin Recorder â capture meetings, lectures, or conversations using laptop or external microphones đ¤ Transcript Tools + Export Options & Translation ⢠đ Copy to Clipboard for quick use in docs or productivity apps ⢠đ Raw SRT viewer with timestamp navigation and highlighted keyword search (original + translated SRT) ⢠đĽ Export in 17+ formats: TXT, VTT, SRT, HTML, LRC, CSV, DOCX, PDF, JSON, JSONâFULL, SBV, TTML/DFXP, ASS/SSA, RTTM, CTM ⢠đ HTML (Standalone Viewer) with dark/light themes and instant reâexport (JSON, VTT, CSV, TXT). Share anywhere â drop it in an email and your collaborators can open it immediately ⢠đ¤ Share transcripts via email for review or collaboration ⢠đ Translate transcriptions into 136 Languages (internet required). Note: Online translation may be temporarily unavailable during periods of limited server availability. đ§ Playback, Review & Sync Tools ⢠đź AudioâSynced LRC Transcript Player â autoâsynced timeline for audio & video ⢠Upload existing audio/video with .lrc for synchronized timeline playback ⢠đŚ Supports files & transcripts up to 200MB for stable, fast processing đĽď¸ GGML Model Server ⢠Route, manage, and brand 33 Whisper transcriber models (11 official + 22 quantized) ⢠Unified runtime across tiny â base â small â medium â large models (v1, v2, v3, turbo) ⢠Ideal for multiâmodel workflows and benchmarking âď¸ Performance & System Modes ⢠đĽď¸ GPU-accelerated for maximum speed or CPUâOnly for benchmarking ⢠đĽď¸ Run anywhere â from laptops to highâend NVIDIA GPUs ⢠đŚ Supports up to 1GB per transcription, enough for 10+ hours of audio. âď¸ System Guidance ⢠8â12âŻGB RAM: Best for short recordings and everyday transcription ⢠16â32âŻGB RAM: Ideal for creators and multiâhour audio ⢠64â128âŻGB RAM: Optimized for large models and highâvolume workloads ⢠This app supports audio & video files up to 1âŻGB, typically equal to 10+ hours of compressed formats like MP3, M4A(AAC/ALAC), or Opus. ⢠Performance and actual capacity vary based on your VRAM, RAM, available disk space, and the model size youâre using. â System Requirements Windows 10/11 đ Supported NVIDIA GPU RTX Generations RTX Generation | Architecture | Compute Capability ---------------|---------------|-------------------- RTX 20âseries | Turing | 7.5 RTX 30âseries | Ampere | 8.6 RTX 40âseries | Ada Lovelace | 8.9 NVIDIA TITAN Series (Various Architectures) NVIDIA Quadro & RTX AâSeries (Workstation GPUs) â ď¸ Compatibility Notice â RTX 50âseries | Blackwell | 12.0 - Currently not supported upstream. Whisper ASR supports NVIDIA GPUs with modern CUDA drivers. Older GPUs may not be compatible with the CUDA backend. â Not supported for CUDA acceleration: ⢠Fermiâgeneration GPUs (e.g., GeForce 300M / 400M / 500M) ⢠Maxwell GPUs (e.g., GTX 9xx, 940M) ⢠GTX 10âseries and GTX 16âseries (Pascal) ⢠GPUs are limited to legacy drivers that do not include modern CUDA features These older GPUs may fail to load the CUDA backend due to missing driver functions. đŻ Who This Is For ⢠Creators & podcasters ⢠Students & educators ⢠Professionals & teams ⢠Developers & AI builders ⢠Privacyâfocused users ⢠Accessibility & support roles đ Translate into 136 languages Major families supported: ⢠European: English, French, German, Spanish, Italian, Portuguese, Dutch, Swedish, Polish, Romanian, Czech, Greek ⢠Asian: Chinese (Simplified/Traditional), Japanese, Korean, Hindi, Bengali, Tamil, Telugu, Thai, Vietnamese ⢠Middle Eastern: Arabic, Hebrew, Persian, Kurdish (Sorani/Kurmanji), Turkish ⢠African: Swahili, Yoruba, Zulu, Hausa, Igbo, Amharic ⢠Pacific & Indigenous: Maori, Samoan, Hawaiian, Aymara, Quechua ⢠Plus 100+ additional languages Note: Online translation may be temporarily unavailable during periods of limited server availability. đ Visit @Generative Lab: https://genlab.i-mapz.com
Utilities ¡ Windows ¡ Free