Ross ROSS = Recommend OSS · open-source software intelligence for agents

LianjiaTech/BELLE resource

BELLE: Be Everyone's Large Language model Engine(开源中文对话大模型) observed · 2026-08-28

github.com/LianjiaTech/BELLE · HTML · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

21/100

  • Activity 0
  • Release rhythm 8
  • Longevity 90
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 1265
  • days_rel: n/a
  • days_push: 686
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

8275 stars · 756 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

BELLE is an open-source project providing Chinese-optimized instruction-tuned large language models, training data, and fine-tuning code built on open pretrained backbones like LLaMA and BLOOM. It also releases enhanced Chinese speech recognition models (fine-tuned Whisper) and multimodal variants, along with technical reports on training techniques.

Use cases

  • fine-tune a Chinese instruction-following chatbot on an open LLM
  • download Chinese instruction tuning datasets generated from ChatGPT
  • train models with LoRA or RLHF (PPO/DPO) for Chinese dialogue
  • improve Whisper speech recognition accuracy for Chinese audio
  • explore multimodal Chinese vision-language models
  • reduce the barrier to building custom Chinese LLMs

When to choose

  • you need Chinese-language instruction-tuned models or training data
  • you want open recipes for LoRA, full fine-tuning, or RLHF on Chinese LLMs
  • you need Chinese-optimized speech recognition models
  • you are researching how training data quality affects LLM performance

When to avoid

  • you need a production-ready hosted chat service rather than models and training code
  • your focus is English-only or multilingual models beyond Chinese optimization
  • you need actively cutting-edge releases, as updates have slowed since late 2024

Facets

learning-resource · maturity maintenance

llm-training machine-learning speech-recognition rag large-language-models machine-learning speech-processing python chinese-nlp instruction-tuning lora rlhf whisper multimodal open-models chatglm llama natural-language-processing gpu linux

1 source

Member repositories

RepositoryRoleHealth v2
LianjiaTech/BELLEmain21

For agents

markdown · JSON · MCP: product_card(name="LianjiaTech/BELLE")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem