Ross ROSS = Recommend OSS · open-source software intelligence for agents

hemingkx/SpeculativeDecodingPapers resource

📰 Must-read papers and blogs on Speculative Decoding ⚡️ observed · 2026-08-28

github.com/hemingkx/SpeculativeDecodingPapers · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

68/100

  • Activity 89
  • Release rhythm 35
  • Longevity 77

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 1078
  • days_rel: n/a
  • days_push: 68
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1291 stars · 80 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A regularly updated curated reading list (awesome list) of must-read papers, blogs, and tutorials on speculative decoding for efficient large language model inference. It accompanies an ACL 2024 Findings survey and organizes research by drafting method, application area, and features such as multi-token prediction, MoE, and long-context decoding.

Use cases

  • learn about speculative decoding for LLM inference
  • find research papers on faster LLM token generation
  • catch up on multi-token prediction and draft-model techniques
  • prepare a related-work section on efficient LLM decoding
  • find tutorials and blogs explaining speculative decoding
  • track new publications on LLM inference acceleration

When to choose

  • you are starting research or a literature review on speculative decoding
  • you need a curated, categorized, regularly updated bibliography with venue and method tags
  • you want tutorials, slides, and blogs alongside academic papers

When to avoid

  • you need runnable inference acceleration code or a serving library rather than a reading list
  • you need production LLM serving or quantization tooling
  • you need benchmark results you can run rather than references to benchmark papers

Facets

learning-resource · maturity active

llm-inference machine-learning deep-learning nlp large-language-models artificial-intelligence awesome-lists tutorials performance awesome-list speculative-decoding paper-list survey efficient-inference multi-token-prediction draft-models reading-list research natural-language-processing

1 source

Member repositories

RepositoryRoleHealth v2
hemingkx/SpeculativeDecodingPapersmain68

For agents

markdown · JSON · MCP: product_card(name="hemingkx/SpeculativeDecodingPapers")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem