Ross ROSS = Recommend OSS · open-source software intelligence for agents

getumbrel/llama-gpt

A self-hosted, offline, ChatGPT-like chatbot. Powered by Llama 2. 100% private, with no data leaving your device. New: Code Llama support! observed · 2026-08-28

github.com/getumbrel/llama-gpt · homepage · TypeScript · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

28/100

  • Activity 0
  • Release rhythm 35
  • Longevity 81

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 1138
  • days_rel: n/a
  • days_push: 862
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

10940 stars · 706 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

LlamaGPT is a self-hosted, offline ChatGPT-like chatbot powered by Llama 2 and Code Llama models via llama.cpp, with all data staying on the user's device. It ships as a Docker-deployable app with an OpenAI-compatible API and one-click install on umbrelOS home servers.

Use cases

  • run a private chatgpt alternative on my own hardware
  • self-host a local llm chatbot with no data leaving my device
  • chat with llama 2 offline
  • serve an openai-compatible api from a local model
  • run code llama locally for coding help
  • install a chatbot on my umbrel home server

When to choose

  • you want a fully private, offline ChatGPT-like experience on your own hardware
  • you run an umbrelOS home server and want one-click local AI
  • you need an OpenAI-compatible API backed by a local Llama 2 or Code Llama model
  • you have an M1/M2 Mac or Nvidia GPU and enough RAM for 7B-70B quantized models

When to avoid

  • you need the latest models, custom model support, or frequent updates - the project appears in maintenance with limited recent activity
  • you lack the RAM/GPU required for the quantized models
  • you want multi-user or cloud-scale serving rather than a personal assistant
  • you prefer actively developed alternatives like Ollama or Open WebUI

Facets

application · maturity maintenance

chatbot llm-inference self-hosted chat-interface api-framework large-language-models chatbots artificial-intelligence self-hosted privacy self-hosted llama-2 code-llama llama-cpp offline-ai local-llm openai-compatible-api umbrel docker macos linux web-server kubernetes gpu

3 sources

Member repositories

RepositoryRoleHealth v2
getumbrel/llama-gptmain28

For agents

markdown · JSON · MCP: product_card(name="getumbrel/llama-gpt")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem