Ross ROSS = Recommend OSS · open-source software intelligence for agents

collabora/WhisperLive

A nearly-live implementation of OpenAI's Whisper. observed · 2026-08-28

github.com/collabora/WhisperLive · Python · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

86/100

  • Activity 96
  • Release rhythm 74
  • Longevity 86
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: 76
  • age_days: 1217
  • days_rel: 92
  • days_push: 29
  • n_releases_24m: 8

Full methodology

Adoption not part of the score

4241 stars · 582 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

WhisperLive is a nearly-live speech-to-text application built on OpenAI's Whisper, serving real-time transcription over WebSocket and REST from microphone or file audio. It supports faster_whisper, TensorRT-LLM, and OpenVINO backends, plus features like word-level timestamps, hotwords, and speaker diarization.

Use cases

  • transcribe live microphone audio in real time
  • transcribe pre-recorded audio files
  • add live captions to an OBS stream
  • dictate text hands-free
  • translate spoken audio to another language
  • run a self-hosted speech-to-text server with GPU acceleration
  • diarize speakers in a meeting recording

When to choose

  • you need low-latency streaming transcription with a client-server setup
  • you want GPU-optimized Whisper inference via TensorRT or OpenVINO
  • you need speaker diarization or word-level timestamps alongside transcription

When to avoid

  • you only need simple offline batch transcription without a server
  • you cannot install GPU toolchains or PortAudio dependencies
  • you need a fully managed cloud speech API

Facets

application · maturity active

speech-recognition http-server websocket llm-inference speech-processing artificial-intelligence python cross-platform whisper transcription real-time dictation tensorrt openvino speaker-diarization translation natural-language-processing docker gpu web-server

1 source

Member repositories

RepositoryRoleHealth v2
collabora/WhisperLivemain86

For agents

markdown · JSON · MCP: product_card(name="collabora/WhisperLive")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem