# nobodywho-ooo/nobodywho

NobodyWho is an inference engine that lets you run LLMs locally and efficiently on any device.

Repository: https://github.com/nobodywho-ooo/nobodywho
Canonical: https://ross.abutalabs.com/products/nobodywho
Homepage: https://nobodywho.ai
Language: Rust
License: EUPL-1.2
License Family: copyleft
Topics: godot, godot-engine, godot-plugin, godot4, llm, inference-engine, python, python-llm, slm, flutter, on-device-ai, ai, react-native, kotlin, swift, stt, tts, expo
Last push: 2026-09-02T15:50:35+00:00

## Health v2 (maintenance only)
Score: 86/100 (v2, computed 2026-09-03T02:39:23.370411+00:00)
- activity 100, release rhythm 87, longevity 52
- inputs: {"age_days": 729, "days_push": 0, "days_rel": 9, "gap_med": 1.0, "n_releases_24m": 147}
- flags: none
- formula: round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10)

## Adoption (not part of the score)
Stars 1086, forks 77 (observed 2026-09-03T02:15:15.558742+00:00)

## What it is
NobodyWho is an open-source (EUPL-1.2) on-device LLM inference engine written in Rust, built on llama.cpp, with SDKs for Kotlin, Swift, Python, Flutter, React Native, and Godot. It supports any GGUF chat model, GPU acceleration via Vulkan/Metal, tool calling, multimodal input, plus Whisper speech-to-text and local text-to-speech.

## Use cases
- run LLMs locally on mobile apps without a server
- add offline chatbot AI to a Godot game
- embed on-device AI in a Flutter or React Native app
- transcribe audio to text locally with Whisper
- synthesize speech from text on-device
- run GGUF models from Hugging Face with GPU acceleration
- add type-safe tool calling to a local LLM

## When to choose
- you need offline, privacy-preserving LLM inference on mobile or desktop
- you want to avoid API costs and servers by running models on-device
- you're building with Godot, Flutter, React Native, Kotlin, Swift, or Python
- you need speech-to-text or text-to-speech bundled with LLM inference

## When to avoid
- you need large models that exceed on-device memory and require cloud GPUs
- you need a managed inference service with autoscaling
- your stack isn't among the supported SDKs

## Facets
- artifact type: library
- maturity: active
- function: llm-inference, speech-recognition, tts, sdk, machine-learning
- domain: large-language-models, artificial-intelligence, mobile-development, speech-processing, cross-platform
- platform: cross-platform, python, rust, jvm, game-engine
- tags: on-device-ai, gguf, llama.cpp, godot-plugin, flutter, react-native, offline-inference, tool-calling, multimodal, game-development, android, ios, swift, desktop

## Member repositories
- nobodywho-ooo/nobodywho (main) score 86

## Provenance
- Observed fields: from GitHub, fetched 2026-09-03T02:15:15.558742+00:00.
- Health v2: computed from the inputs above; adoption is never an input.
- Inferred fields (summary, facets, guidance): AI-extracted, prompt v1, taxonomy v1, on 2026-08-30T06:53:11.236876+00:00, confidence not recorded.
  - readme: https://github.com/nobodywho-ooo/nobodywho (fetched 2026-09-03T02:15:15.558742+00:00, sha 42a79d45000e)
  - homepage: https://nobodywho.ai (fetched 2026-08-29T12:54:18.038624+00:00, sha e0686e683bb9)
  - site_page: https://www.nobodywho.ai/about (fetched 2026-08-29T12:54:18.041782+00:00, sha d47834bed1f7)
- Data as of 2026-08-30T08:39:29.467469+00:00.
