Ross ROSS = Recommend OSS · open-source software intelligence for agents

THUDM/AgentTuning resource

AgentTuning: Enabling Generalized Agent Abilities for LLMs observed · 2026-08-28

github.com/THUDM/AgentTuning · homepage · Python observed · 2026-08-28

Health v2 · maintenance only

27/100

  • Activity 0
  • Release rhythm 35
  • Longevity 75

Flags: no_releases no_license

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 1050
  • days_rel: n/a
  • days_push: 1037
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1504 stars · 104 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

AgentTuning is a research project from Tsinghua University that instruction-tunes LLMs on multi-task agent interaction trajectories to improve their agent abilities. It releases the AgentInstruct dataset (1,866 filtered GPT-4 trajectories across 6 tasks) and the AgentLM-7B/13B/70B models fine-tuned from Llama 2.

Use cases

  • fine-tune an open-source LLM to act as an agent
  • get training data of agent interaction trajectories
  • improve tool-use and planning abilities of Llama 2 models
  • find an open alternative to GPT-3.5 for agent tasks
  • download AgentInstruct dataset for agent instruction tuning
  • evaluate open LLMs on held-in and held-out agent benchmarks

When to choose

  • you need an open-weight LLM with stronger agent/tool-use capabilities
  • you want a curated instruction-tuning dataset of agent trajectories
  • you are researching how instruction tuning affects agent generalization

When to avoid

  • you need a production-ready agent framework rather than models and data
  • you cannot host or fine-tune large models on GPUs
  • you need a permissively licensed artifact - the repo has no license

Facets

dataset · maturity maintenance

llm-training agent-framework machine-learning data-generation large-language-models machine-learning python instruction-tuning agent-trajectories llama2 huggingface research agentlm agentinstruct fine-tuning ai-agents natural-language-processing gpu linux

2 sources

Member repositories

RepositoryRoleHealth v2
THUDM/AgentTuningmain27

For agents

markdown · JSON · MCP: product_card(name="THUDM/AgentTuning")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem