fpgaminer/joycaption
JoyCaption is an image captioning Visual Language Model (VLM) being built from the ground up as a free, open, and uncensored model for the community to use in training Diffusion models. observed · 2026-08-28
Health v2 · maintenance only
53/100
- Activity 69
- Release rhythm 35
- Longevity 49
Flags: no_releases
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 690
- days_rel: n/a
- days_push: 190
- n_releases_24m: 0
Adoption not part of the score
1244 stars · 69 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
JoyCaption is an open, free, and uncensored image captioning Visual Language Model (VLM) with released weights and training scripts. It generates descriptive captions for images, primarily to support training and finetuning of diffusion models.
Use cases
- generate captions for images to train diffusion models
- automatically caption datasets for stable diffusion finetuning
- caption NSFW and SFW images without censorship
- describe anime, furry, and digital art images
- replace paid captioning services like ChatGPT for image description
- run an uncensored vision language model locally
When to choose
- you need descriptive image captions for text-to-image model training
- you want an open, uncensored alternative to GPT-4o for captioning
- you need coverage of diverse image styles including anime and digital art
- you want to inspect or reproduce the model training pipeline
When to avoid
- you need a general-purpose chat or reasoning VLM
- you require a small CPU-only model for edge devices
- you need a fully moderated, safety-filtered captioning service
- you want a hosted API rather than self-hosted inference
Facets
library · maturity active
machine-learning image-processing nlp llm-inference machine-learning artificial-intelligence image-processing large-language-models python cross-platform image-captioning vision-language-model vlm diffusion-model-training uncensored stable-diffusion llava gpu
1 source
- readme: https://github.com/fpgaminer/joycaption · fetched 2026-08-28 · 47f5631c9840
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| fpgaminer/joycaption | main | 53 |
For agents
markdown · JSON · MCP: product_card(name="fpgaminer/joycaption")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem