resource: rag
229 resources, primary matches first, then adoption-weighted; health v2 shown.
| Resource | Health v2 | Stars | Maturity |
|---|---|---|---|
| langchain-ai/agents-from-scratch A tutorial repository and companion course from LangChain Academy that teaches how to build AI agents from scratch using LangGraph, culmina… | 63 | 2156 | active |
| LianjiaTech/BELLE BELLE is an open-source project providing Chinese-optimized instruction-tuned large language models, training data, and fine-tuning code bu… | 21 | 8275 | maintenance |
| huggingface/transformers.js-examples A collection of demos and example applications showcasing Hugging Face Transformers.js, which runs transformer models directly in the brows… | 53 | 2087 | active |
| safe-graph/graph-fraud-detection-papers A curated awesome-list of academic papers and resources on graph- and transformer-based fraud, anomaly, and outlier detection, organized by… | 73 | 1889 | active |
| tau-bench τ-Bench (tau2-bench) is a Python benchmark for evaluating LLM agents on tool-agent-user interaction in real-world domains like retail, airl… | 79 | 1882 | active |
| vstorm-co/full-stack-ai-agent-template A production-ready full-stack project generator that scaffolds FastAPI + Next.js 15 applications with AI agents, RAG pipelines, WebSocket s… | 80 | 1856 | active |
| thinkwee/AgentsMeetRL A curated awesome list of open-source repositories for training LLM agents with reinforcement learning, organized by taxonomy categories li… | 62 | 1799 | active |
| datawhalechina/deepagents-in-action An open-source Chinese-language course ('Deep Agents 实战') that teaches building production-grade AI agents with the LangChain/LangGraph Dee… | 58 | 1790 | active |
| neural-maze/ava-whatsapp-agent-course Ava is a course project and codebase for building a WhatsApp AI agent that chats realistically, understands voice and images, and replies w… | 43 | 1673 | active |
| tencent-ailab/persona-hub PersonaHub is Tencent AI Lab's collection of 1 billion diverse personas plus code for persona-driven synthetic data creation with LLMs. It … | 27 | 1642 | active |
| modelscope/modelscope-classroom ModelScope Classroom is a collection of deep learning and large language model tutorials from the ModelScope community, delivered as Jupyte… | 61 | 1482 | active |
| elicit/machine-learning-list A curated, tiered reading list and curriculum for learning machine learning with a focus on language models and foundation models, from fun… | 49 | 1464 | active |
| groq/groq-api-cookbook A collection of tutorials, sample code, and guidelines for building applications with the Groq API, maintained as Jupyter Notebook examples… | 69 | 1418 | active |
| no-magic-ai/no-magic A curated collection of 48 single-file, zero-dependency Python implementations of core AI algorithms, from GPT and LSTM to LoRA, quantizati… | 67 | 1410 | active |
| lvgalvao/data-engineering-roadmap The official repository of a professional training program in Data Engineering and AI (university extension) by Jornada de Dados. It contai… | 63 | 1394 | active |
| philschmid/deep-learning-pytorch-huggingface A collection of Jupyter notebook tutorials and examples for deep learning with PyTorch and Hugging Face libraries like Transformers, Datase… | 35 | 1392 | active |
| philschmid/gemini-samples A collection of Jupyter notebook samples, snippets, and guides demonstrating how to use Google DeepMind Gemini models. It covers function c… | 52 | 1372 | active |
| OpenDriveLab/DriveLM DriveLM is a research project and dataset for autonomous driving built around Graph Visual Question Answering (GVQA), where QA pairs about … | 41 | 1338 | active |
| whwangovo/pyre-code A self-hosted coding practice platform with ~76 ML implementation problems covering attention, training, RLHF, diffusion, and GNNs, graded … | 51 | 1253 | active |
| CrazyBoyM/llama3-Chinese-chat A community repository releasing Chinese post-trained (SFT and DPO) versions of Llama3/Llama3.1, including open model weights, quantized GG… | 55 | 4146 | maintenance |
| unipds-engenharia-de-ia-aplicada/engenharia-de-software-com-ia-aplicada A companion repository of code examples, demos, and reference links for the UNIPDS postgraduate program 'Engenharia de Software com IA Apli… | 60 | 1084 | active |
| wdndev/tiny-llm-zh An educational project that implements a small-parameter Chinese large language model from scratch, covering the full pipeline: tokenizer t… | 25 | 1078 | active |
| Kent0n-Li/ChatDoctor ChatDoctor is a medical chat model fine-tuned on LLaMA using real patient-doctor conversations, with training code, datasets (HealthCareMag… | 30 | 3631 | maintenance |
| data-science-on-aws/data-science-on-aws Companion Jupyter Notebook repository for the O'Reilly book 'Data Science on AWS', containing end-to-end AI/ML pipeline examples using Amaz… | 32 | 3433 | maintenance |
| CVI-SZU/Linly Linly is a project releasing Chinese-adapted open large language models, including Chinese-LLaMA 1&2, Chinese-Falcon, Linly-OpenLLaMA base … | 21 | 3044 | maintenance |
| HarderThenHarder/transformers_tasks A collection of Jupyter Notebook-based NLP algorithm implementations built on the Hugging Face transformers library, covering text classifi… | 32 | 2420 | maintenance |
| anthropics/cwc-workshops Workshop materials from Anthropic's 'Code with Claude' events, covering building and evaluating agents with Claude Managed Agents, Skills, … | 58 | 2026 | maintenance |
| fastai/lm-hackers A Jupyter notebook and accompanying video guide from fast.ai explaining how language models work, including tokenization, base models, inst… | 27 | 1869 | maintenance |
| openai/summarize-from-feedback Research code and human feedback dataset from OpenAI's paper 'Learning to Summarize from human feedback', including a supervised baseline, … | 10 | 1062 | abandoned |
← prev page 3 / 3