Ross ROSS = Recommend OSS · open-source software intelligence for agents

fpgaminer/joycaption

JoyCaption is an image captioning Visual Language Model (VLM) being built from the ground up as a free, open, and uncensored model for the community to use in training Diffusion models. observed · 2026-08-28

github.com/fpgaminer/joycaption · Jupyter Notebook · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

53/100

  • Activity 69
  • Release rhythm 35
  • Longevity 49

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 690
  • days_rel: n/a
  • days_push: 190
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1244 stars · 69 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

JoyCaption is an open, free, and uncensored image captioning Visual Language Model (VLM) with released weights and training scripts. It generates descriptive captions for images, primarily to support training and finetuning of diffusion models.

Use cases

  • generate captions for images to train diffusion models
  • automatically caption datasets for stable diffusion finetuning
  • caption NSFW and SFW images without censorship
  • describe anime, furry, and digital art images
  • replace paid captioning services like ChatGPT for image description
  • run an uncensored vision language model locally

When to choose

  • you need descriptive image captions for text-to-image model training
  • you want an open, uncensored alternative to GPT-4o for captioning
  • you need coverage of diverse image styles including anime and digital art
  • you want to inspect or reproduce the model training pipeline

When to avoid

  • you need a general-purpose chat or reasoning VLM
  • you require a small CPU-only model for edge devices
  • you need a fully moderated, safety-filtered captioning service
  • you want a hosted API rather than self-hosted inference

Facets

library · maturity active

machine-learning image-processing nlp llm-inference machine-learning artificial-intelligence image-processing large-language-models python cross-platform image-captioning vision-language-model vlm diffusion-model-training uncensored stable-diffusion llava gpu

1 source

Member repositories

RepositoryRoleHealth v2
fpgaminer/joycaptionmain53

For agents

markdown · JSON · MCP: product_card(name="fpgaminer/joycaption")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem