EvolvingLMMs-Lab/lmms-eval
One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks observed · 2026-08-28
Health v2 · maintenance only
85/100
- Activity 99
- Release rhythm 78
- Longevity 64
Flags: no_license
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: 40
- age_days: 909
- days_rel: 70
- days_push: 7
- n_releases_24m: 16
Adoption not part of the score
4377 stars · 647 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded
lmms-eval is a unified Python framework for evaluating large multimodal models across text, image, video, and audio tasks, with 100+ benchmarks and 30+ model integrations. It emphasizes reproducible, efficient, and trustworthy evaluation pipelines for frontier model research.
Use cases
- evaluate a vision-language model on standard benchmarks
- compare multimodal model scores reproducibly across teams
- benchmark video understanding models
- run audio evaluation on large multimodal models
- measure LMM performance across image, video, and audio tasks
- set up a unified eval pipeline for model releases
When to choose
- you need reproducible multimodal benchmark numbers
- you want one toolkit covering text, image, video, and audio evals
- you are a research lab evaluating frontier LMMs at scale
When to avoid
- you only need text-only LLM evaluation (lm-eval-harness may suffice)
- you need a lightweight ad-hoc metric script rather than a full harness
- you require a commercial benchmarking service with hosted infrastructure
Facets
framework · maturity active
benchmarking machine-learning llm-inference testing large-language-models machine-learning computer-vision artificial-intelligence python cross-platform multimodal-evaluation vision-language-models vlm llm-evaluation benchmarks video-understanding audio-evaluation audio video gpu linux
4 sources
- readme: https://github.com/EvolvingLMMs-Lab/lmms-eval · fetched 2026-08-28 · 1a0c2b0a87ae
- homepage: https://www.lmms-lab.com · fetched 2026-08-29 · 77429eb7a494
- site_page: https://www.lmms-lab.com/about · fetched 2026-08-29 · 14cba1c6f595
- registry_pypi: https://pypi.org/pypi/lmms-eval/json · fetched 2026-08-29 · 83285aaeb88b
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| EvolvingLMMs-Lab/lmms-eval | main | 85 |
For agents
markdown · JSON · MCP: product_card(name="EvolvingLMMs-Lab/lmms-eval")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem