Ross ROSS = Recommend OSS · open-source software intelligence for agents

awslabs/multi-model-server

Multi Model Server is a tool for serving neural net models for inference observed · 2026-08-28

github.com/awslabs/multi-model-server · Java · Apache-2.0 (permissive) · archived observed · 2026-08-28

Health v2 · maintenance only

10/100

  • Activity 0
  • Release rhythm 8
  • Longevity 100

Flags: archived

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 3255
  • days_rel: n/a
  • days_push: 835
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1024 stars · 230 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

Multi Model Server (MMS) is a tool for serving deep learning model inference over HTTP endpoints, supporting models from any ML/DL framework including MXNet and ONNX. It provides a server CLI and pre-configured Docker images to deploy models as inference services.

Use cases

  • serve deep learning models for inference over http
  • deploy neural network models as a rest api
  • host onnx model inference server
  • serve mxnet models in production
  • package and serve pytorch or mxnet model archives
  • run a multi-model inference server in docker

When to choose

  • you need HTTP endpoints for model inference requests
  • you want to serve models from multiple DL frameworks
  • you want pre-configured Docker images for model serving
  • you are on Linux or macOS with Java 8 and Python available

When to avoid

  • you need a modern, actively developed serving stack (consider TorchServe, its successor)
  • you need first-class Windows support
  • you want GPU-optimized LLM serving with batching and quantization
  • you need a framework-free client-side inference library

Facets

service · maturity maintenance

http-server machine-learning llm-inference deep-learning machine-learning apis self-hosted python jvm cross-platform model-serving inference-server mxnet onnx deep-learning-serving http-endpoints linux macos docker

1 source

Member repositories

RepositoryRoleHealth v2
awslabs/multi-model-servermain10

For agents

markdown · JSON · MCP: product_card(name="awslabs/multi-model-server")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem