Ross ROSS = Recommend OSS · open-source software intelligence for agents

horseee/Awesome-Efficient-LLM resource

A curated list for Efficient Large Language Models observed · 2026-08-28

github.com/horseee/Awesome-Efficient-LLM · Python observed · 2026-08-28

Health v2 · maintenance only

41/100

  • Activity 27
  • Release rhythm 35
  • Longevity 85

Flags: no_releases no_license

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 1199
  • days_rel: n/a
  • days_push: 443
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

2036 stars · 169 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A curated awesome-list of research papers and resources on efficient large language models, covering pruning, quantization, knowledge distillation, KV cache compression, efficient MoE, and serving systems. It is organized into topic subdirectories with regularly updated recent papers and recommended reads.

Use cases

  • find papers on LLM quantization
  • research KV cache compression techniques
  • survey efficient fine-tuning methods
  • learn about pruning and sparsity for transformers
  • find inference acceleration research
  • keep up with efficient MoE architectures
  • find surveys and benchmarks on efficient LLMs

When to choose

  • starting research on efficient LLM techniques
  • looking for a broad, organized survey of model compression literature
  • tracking recent papers in LLM efficiency

When to avoid

  • need runnable software or a library rather than a paper list
  • need production tooling for model compression
  • require a licensed or citable software artifact

Facets

learning-resource · maturity active

llm-inference llm-training machine-learning documentation large-language-models machine-learning deep-learning awesome-lists tutorials python cross-platform awesome-list model-compression quantization knowledge-distillation pruning kv-cache moe inference-acceleration curated-papers

1 source

Member repositories

RepositoryRoleHealth v2
horseee/Awesome-Efficient-LLMmain41

For agents

markdown · JSON · MCP: product_card(name="horseee/Awesome-Efficient-LLM")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem