{"repo":"QuentinFuxa/WhisperLiveKit","free":true,"listed":false,"github":"https://github.com/QuentinFuxa/WhisperLiveKit","clone":"git clone https://github.com/QuentinFuxa/WhisperLiveKit.git","description":"Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.","language":"Python","stars":10617,"topics":["python","real-time","speaker-diarization","speech-recognition","speech-to-text","streaming","translation","websocket","whisper","automatic-speech-recognition"],"license":"Apache-2.0","category":"networking-infra","readme_excerpt":"WLK: Ultra-low-latency, self-hosted speech-to-text pipeline Powered by Leading Research: - Simul-Whisper/Streaming (SOTA 2025) - Ultra-low latency transcription using AlignAtt policy. - NLLW (2025), based on distilled NLLB (2022, 2024) - Simulatenous translation from & to 200 languages. - WhisperStreaming (SOTA 2023) - Low latency transcription using LocalAgreement policy - Streaming Sortformer (SOTA 2025) - Advanced real-time speaker diarization - Qwen3-ASR-causal (2026) - Causal streaming audio encoder for Qwen3-ASR: each audio block is encoded exactly once, constant compute per audio second, append-only transcripts. - AlignAtt4LLM (IWSLT 2026) - Simultaneous translation with decoder-only LLMs: attention-gated commits, append-only output. Why not just run a simple Whisper model on every audio batch? Whisper is designed for complete utterances, not real-time chunks. Processing small segments loses context, cuts off words mid-syllable, and produces poor transcription. WhisperLiveKit uses state-of-the-art simultaneous speech research for intelligent buffering and incremental processing. Architecture The backend supports multiple concurrent users. Voice Activity Detection reduces overhead when no voice is detected. Installation & Quick Start Quick Start API Compatibility WhisperLiveKit exposes compatibility-oriented subsets of popular APIs: Per-session WebSocket query parameters: param example effect --- --- --- language ?language=fr transcription language for this session (one","default_branch":null,"files":null,"tree":[],"storefront":"/r/QuentinFuxa","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/QuentinFuxa/WhisperLiveKit/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}