Ross ROSS = Recommend OSS · open-source software intelligence for agents

HuangOwen/Awesome-LLM-Compression resource

Awesome LLM compression research papers and tools. observed · 2026-08-28

github.com/HuangOwen/Awesome-LLM-Compression · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

74/100

  • Activity 99
  • Release rhythm 35
  • Longevity 85

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 1191
  • days_rel: n/a
  • days_push: 7
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1865 stars · 130 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A curated awesome-list of research papers and tools on large language model compression, covering quantization, pruning and sparsity, distillation, efficient prompting, and KV cache compression. Papers are organized by category and year, with a tools section for accelerating LLM training and inference.

Use cases

  • find recent papers on LLM quantization
  • research KV cache compression techniques
  • learn about pruning and sparsity for large language models
  • discover knowledge distillation methods for LLMs
  • find tools to compress and accelerate LLM inference
  • survey efficient prompting techniques
  • keep up with LLM efficiency research by year

When to choose

  • you need a curated, categorized reading list on LLM compression
  • you want to track the latest efficiency papers grouped by year
  • you are looking for both papers and practical compression tools in one place

When to avoid

  • you need runnable compression software rather than a paper index
  • you want tutorials or courses rather than research paper links
  • you need compression methods for non-LLM models

Facets

learning-resource · maturity active

llm-inference llm-training machine-learning large-language-models machine-learning artificial-intelligence awesome-lists tutorials awesome-list model-compression quantization pruning knowledge-distillation kv-cache research-papers efficiency web-server

1 source

Member repositories

RepositoryRoleHealth v2
HuangOwen/Awesome-LLM-Compressionmain74

For agents

markdown · JSON · MCP: product_card(name="HuangOwen/Awesome-LLM-Compression")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem