neuralmagic/deepsparse
Sparsity-aware deep learning inference runtime for CPUs observed · 2026-08-28
Health v2 · maintenance only
10/100
- Activity 24
- Release rhythm 8
- Longevity 100
Flags: archived no_license
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 2088
- days_rel: 457
- days_push: 457
- n_releases_24m: 1
Adoption not part of the score
3158 stars · 193 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded
DeepSparse is a sparsity-aware deep learning inference runtime that delivers GPU-class performance on x86 CPUs for ONNX models, covering computer vision, NLP, and LLM workloads. Development ceased and the project was deprecated on June 2, 2025 following Neural Magic's acquisition by Red Hat.
Use cases
- run ONNX models fast on CPUs without a GPU
- deploy sparse pruned and quantized models for inference
- serve object detection models on x86 servers
- run LLM inference on CPU hardware
- integrate ML inference into Python applications
When to choose
- you need fast CPU-only inference for ONNX models and are comfortable with an unmaintained runtime
- you already have sparsified models from Neural Magic's tooling and accept no future updates
When to avoid
- you need ongoing support, updates, or security patches
- you are starting a new project - use vLLM or other actively maintained runtimes instead
- you need GPU or non-x86 (e.g. ARM) inference
Facets
library · maturity abandoned
llm-inference machine-learning computer-vision nlp machine-learning deep-learning large-language-models computer-vision python cpp onnx inference-runtime sparsity cpu-optimization deprecated pruning quantization natural-language-processing linux
4 sources
- readme: https://github.com/neuralmagic/deepsparse · fetched 2026-08-28 · f0800317b21f
- homepage: https://neuralmagic.com/deepsparse/ · fetched 2026-08-29 · 8c6922ebda3f
- site_page: https://www.redhat.com/en/about/press-releases/red-hat-ai-factory-nvidia-accelerates-path-scalable-production-ai · fetched 2026-08-29 · e10bc40cf2db
- registry_pypi: https://pypi.org/pypi/deepsparse/json · fetched 2026-08-29 · ea18c11fae32
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| neuralmagic/deepsparse | main | 10 |
For agents
markdown · JSON · MCP: product_card(name="neuralmagic/deepsparse")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem