Local speech to text Qwen3-ASR w/ OpenVINO (no API key)

Local offline ASR on Windows — no cloud, no API cost, full privacy. Qwen3-ASR 0.6B + Intel OpenVINO, GPU-accelerated inference. NETWORK: required for first-time setup (install deps + download 2 GB model); NOT required for inference. Auto-extracts audio from video files (mp4, mkv, webm, mov, avi) — just pass the video path. Also supports audio: mp3, wav, flac, m4a, ogg, aac, wma, opus. Single file, batch folder, or watch-mode continuous transcription with automatic txt/json archive. 30 languages + 22 Chinese dialects, auto language detection. One-line LLM API: from acoustic_pipeline import AcousticPipeline Local speech to text, transcribe audio, voice recognition, transcribe video, transcribe recording, convert speech to text, audio transcription, local ASR, offline speech recognition, dictation on Windows. 本地离线语音识别,零云端,支持视频音轨自动提取,支持批量转录和文件夹监听,自动保存转写文档。 网络说明:仅首次安装环境和下载模型时需要网络,推理转录完全离线。

Install

openclaw skills install @juan-oy/local-qwen3-asr-aipc