WhisperLiveKit — Real-Time Streaming Transcription and Diarization
Live captioning backend implementing Simul-Whisper AlignAtt and LocalAgreement policies with streaming speaker diarization for multi-user sessions.
Link
Related resources
- faster-whisper — CTranslate2 Whisper ReimplementationWhisper speech recognition reimplemented on CTranslate2 for roughly 4x faster transcription with lower memory, the standard backend for cap…
- FFmpeg 8.0 whisper Filter DocumentationReference for FFmpeg 8.0's built-in Whisper ASR filter: model selection, VAD, queueing, and SRT/JSON transcript output destinations.
- COVER — Comprehensive Video Quality EvaluatorCVPRW 2024 winner of the AIS 2024 video quality assessment challenge; tri-branch Swin/ConvNet/CLIP architecture doing real-time 1080p no-re…
- YoMo — dockerized active QoE measurement for streamingContainerized active measurement harness from Uni Würzburg that automates browser playback sessions and extracts QoE metrics (startup delay…
- 2026 NFL Final: OTT Streaming Latency Beats Traditional Broadcast at ScaleThird-party measurement report comparing OTT vs DTH/DTV glass-to-glass latency across providers during a major live sports event, including…
- PyNvVideoCodec API Programming GuideNVIDIA's Python bindings over the Video Codec SDK for GPU-accelerated decode/encode pipelines feeding ML and analytics workloads.
- hdr10plus_tool — Codec Wiki Usage GuideCommunity reference with flags and FFmpeg pipe recipes for extracting, injecting and inspecting HDR10+ dynamic metadata in HEVC/MKV.
- Shutter Encoder DocumentationDocs for the professional FFmpeg-based conversion GUI, covering its encode/decode functions, subtitle generation via Whisper, upscaling, an…
- UVQ — Universal Video Quality (Google)Google's no-reference perceptual video quality model for user-generated content, scoring compression, content and distortion without a pris…
- ColorVideoVDP — reference implementationPyTorch implementation of the ColorVideoVDP full-reference metric (SIGGRAPH 2024), usable as a CLI quality tool or as a differentiable perc…
- ColorVideoVDP — color spatiotemporal visible difference predictor (Cambridge Rainbow Group)Project page for ColorVideoVDP, a full-reference perceptual video quality metric modeling chromatic and spatiotemporal contrast sensitivity…
- QTGMC — AviSynth Wiki ReferenceCanonical parameter reference for QTGMC, the highest-quality motion-compensated deinterlacer, including double-rate output and FPSDivisor s…