yangjianxin1/Firefly
Firefly: 大模型训练工具,支持训练Qwen2.5、Qwen2、Yi1.5、Phi-3、Llama3、Gemma、MiniCPM、Yi、Deepseek、Orion、Xverse、Mixtral-8x7B、Zephyr、Mistral、Baichuan2、Llma2、Llama、Qwen、Baichuan、ChatGLM2、InternLM、Ziya2、Vicuna、Bloom等大模型 observed · 2026-08-28
Health v2 · maintenance only
21/100
- Activity 0
- Release rhythm 8
- Longevity 89
Flags: no_license
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 1249
- days_rel: n/a
- days_push: 679
- n_releases_24m: 0
Adoption not part of the score
6653 stars · 582 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded
Firefly is an open-source one-stop training tool for large language models, supporting pretraining, instruction fine-tuning (SFT), and DPO with full-parameter, LoRA, and QLoRA methods. It supports many mainstream open-source LLMs such as Qwen2, Llama3, Gemma, Mistral, Mixtral, ChatGLM, and InternLM, with chat-template alignment and Unsloth acceleration.
Use cases
- fine-tune llama3 with lora on a single gpu
- qlora instruction tuning for qwen2.5
- train a chat model with dpo
- pretrain a large language model on custom data
- fine-tune chatglm2 or internlm with sft
- reduce gpu memory usage when training llms with unsloth
- align training templates with official chat models
When to choose
- you want to fine-tune or DPO-train mainstream open-source LLMs with LoRA/QLoRA on limited GPU resources
- you need config-file-driven training that matches each model's official chat template
- you want memory- and time-efficient training via Unsloth integration
When to avoid
- you need a managed no-code training platform rather than Python scripts and config files
- you only want inference or serving of models rather than training
- you require a permissively licensed project - the repository has no declared license
Facets
framework · maturity active
llm-training machine-learning deep-learning large-language-models machine-learning deep-learning python fine-tuning lora qlora sft dpo pretraining unsloth peft chinese-llm gpu linux
1 source
- readme: https://github.com/yangjianxin1/Firefly · fetched 2026-08-28 · baba99829c20
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| yangjianxin1/Firefly | main | 21 |
For agents
markdown · JSON · MCP: product_card(name="yangjianxin1/Firefly")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem