function: rag
1112 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| sdyckjq-lab/llm-wiki-skill A personal knowledge base Skill based on Andrej Karpathy's llm-wiki methodology, installable into AI CLI agents like Claude Code, Codex, Op… | 57 | 2378 | active |
| kevinho/clawfeed ClawFeed is an AI-powered news digest application that curates sources like Twitter, RSS, HackerNews, Reddit, and GitHub Trending into stru… | 64 | 2376 | active |
| Natively-AI-assistant/natively-cluely-ai-assistant Natively is a free, source-available desktop AI meeting assistant and interview copilot that provides real-time transcription, AI-generated… | 82 | 2354 | active |
| 1517005260/graph-rag-agent A Python framework that integrates GraphRAG, LightRAG, and Neo4j-based graph building to construct knowledge graphs and perform explainable… | 43 | 2326 | active |
| severian42/GraphRAG-Local-UI A local-first GraphRAG application adapted from Microsoft's GraphRAG, providing a FastAPI backend plus Gradio-based UIs for indexing, promp… | 14 | 2321 | active |
| trypromptly/LLMStack LLMStack is a no-code platform for building generative AI agents, workflows, and chatbots by chaining multiple LLMs and connecting them to … | 30 | 2311 | active |
| 0xMassi/webclaw webclaw is a Rust-based web extraction toolkit that turns any URL into clean, LLM-ready markdown, JSON, or token-optimized text, including … | 77 | 2305 | active |
| togethercomputer/OpenChatKit OpenChatKit is an open-source toolkit from Together, LAION, and Ontocord.ai for building and fine-tuning chat language models, including in… | 30 | 8984 | maintenance |
| HKUDS/FastCode FastCode is a token-efficient framework for repository-scale code understanding and reasoning, using structural scouting over a semantic ma… | 58 | 2292 | active |
| lioensky/VCPToolBox VCPToolBox is a Node.js middleware layer deployed between LLM APIs and frontend applications, providing a unified Variable & Command Protoc… | 79 | 2265 | active |
| qhjqhj00/MemoRAG MemoRAG is a Python RAG framework that uses a super-long memory model to build a global understanding of an entire corpus (up to ~1M tokens… | 43 | 2265 | active |
| PrismML-Eng/Bonsai-demo A demo repository for running PrismML's Bonsai family of 1-bit and ternary-quantized language models locally via llama.cpp and MLX. It prov… | 59 | 2252 | active |
| guy-hartstein/company-research-agent A multi-agent company research application built with LangGraph and Tavily that generates comprehensive due-diligence reports on any compan… | 75 | 2250 | active |
| coleam00/mcp-crawl4ai-rag An MCP server that combines Crawl4AI web crawling with RAG capabilities backed by a Supabase vector database, exposing tools for AI agents … | 34 | 2245 | active |
| datapizza-labs/datapizza-ai Datapizza AI is a Python framework for building production-ready generative AI applications with agents, LLM clients, and RAG pipelines. It… | 68 | 2239 | active |
| codejunkie99/agentic-stack A Python-based framework providing a portable .agent/ folder of memory, skills, and protocols that plugs into many coding-agent harnesses (… | 79 | 2238 | active |
| vasu-devs/JustHireMe JustHireMe is a local-first desktop workbench (Tauri frontend, Python backend) that scrapes job postings, ranks role fit against your profi… | 78 | 2236 | stable |
| apconw/Aix-DB Aix-DB is an AI-powered data analysis system (ChatBI) built on LangChain/LangGraph with an MCP Skills multi-agent architecture, converting … | 72 | 2234 | active |
| baturyilmaz/wordpecker-app WordPecker is a personalized language-learning web application that combines Duolingo-style lessons with user-curated vocabulary lists. It … | 36 | 2229 | active |
| OpenSPG/openspg OpenSPG is a knowledge graph engine built on the SPG (Semantic-enhanced Programmable Graph) framework, developed by Ant Group with OpenKG. … | 38 | 2217 | active |
| LigphiDonk/academic-figure-generator A self-hosted AI-powered platform that generates high-quality academic paper figures: users upload a paper (PDF/DOCX/TXT), Claude analyzes … | 62 | 2190 | active |
| MinishLab/model2vec Model2Vec is a Python library that distills any sentence transformer into a tiny, fast static embedding model by computing one fixed vector… | 90 | 2186 | active |
| 1sdv/TripStar TripStar is an AI-powered travel planning application built on the HelloAgents multi-agent framework, using LLMs to generate personalized t… | 59 | 2181 | active |
| intel/intel-extension-for-transformers Intel's toolkit for accelerating transformer-based GenAI/LLM workloads on Intel platforms, offering state-of-the-art compression (e.g., INT… | 10 | 2174 | active |
| alfredfrancis/ai-chatbot-framework AI Chatbot Framework is an open-source, self-hosted Python platform for building AI-powered chatbots with a low-code admin dashboard. It co… | 79 | 2167 | active |
| withcatai/node-llama-cpp node-llama-cpp is a Node.js library providing bindings to llama.cpp for running LLMs locally, with pre-built binaries and automatic GPU sup… | 93 | 2162 | active |
| btahir/open-deep-research An open-source web application that replicates Gemini's Deep Research by searching the web, extracting page content, and generating AI-writ… | 10 | 2141 | active |
| satellitecomponent/Neurite Neurite is an open-source fractal-based mind-mapping workspace that combines the Mandelbrot set interface with graph-of-thought reasoning f… | 60 | 2120 | active |
| ModelEngine-Group/fit-framework FIT is an enterprise-grade AI development framework for the Java ecosystem, combining a polyglot function engine (Java/Python/C++), a strea… | 66 | 2117 | active |
| zilliztech/GPTCache GPTCache is a Python library for building a semantic cache that stores and retrieves LLM responses based on embedding similarity, reducing … | 35 | 8172 | maintenance |
| LC1332/Chat-Haruhi-Suzumiya An open-source role-playing chatbot project that revives anime characters like Haruhi Suzumiya using large language models, mimicking their… | 29 | 2107 | active |
| tsingyuai/scientify Scientify is an OpenClaw plugin that automates end-to-end scientific research workflows, continuously ingesting new papers, evolving hypoth… | 67 | 2094 | active |
| raphaelmansuy/edgequake EdgeQuake is a high-performance Graph-RAG framework written in Rust, inspired by LightRAG, that transforms documents (PDFs, markdown, text)… | 79 | 2078 | active |
| agentset-ai/agentset Agentset is an open-source RAG-as-a-service platform that handles document ingestion, chunking, embeddings, retrieval, and agentic search w… | 61 | 2075 | active |
| Nutlope/llamatutor LlamaTutor is an open-source AI personal tutor web application powered by Meta's Llama 3.1 70B model via Together AI inference. It is built… | 65 | 2049 | active |
| Weizhena/Deep-Research-skills A structured deep research skill (prompt/workflow package) for Claude Code, OpenCode, and Codex that runs a two-phase research process: out… | 60 | 2028 | active |
| JuneYaooo/nihaisha-nishi-tcm An Agent Skill for Claude Code / Codex / OpenClaw that organizes Ni Haixia's traditional Chinese medicine course materials into a searchabl… | 57 | 2019 | active |
| QwenLM/Qwen3-Embedding Qwen3-Embedding is a series of text embedding and reranking models (0.6B, 4B, 8B) built on Qwen3 foundation models, with a Python repositor… | 38 | 2016 | active |
| watercrawl/WaterCrawl WaterCrawl is a self-hostable web application (Python/Django/Scrapy/Celery) that crawls websites and transforms web content into LLM-ready … | 82 | 2010 | active |
| HKUDS/MiniRAG MiniRAG is an extremely simple retrieval-augmented generation framework designed to work with small, open-source language models. It uses s… | 39 | 2008 | active |
| shcherbak-ai/contextgem ContextGem is an open-source Python LLM framework for extracting structured data and insights from documents with minimal code. It provides… | 87 | 1993 | active |
| patterns-ai-core/langchainrb Langchain.rb is a Ruby library for building LLM-powered applications, providing a unified interface to many LLM providers (OpenAI, Anthropi… | 76 | 1991 | active |
| forloopcodes/contextplus Context+ is an MCP server that gives AI coding agents semantic understanding of large codebases by combining RAG embeddings, Tree-sitter AS… | 56 | 1976 | active |
| run-llama/notebookllama NotebookLlaMa is an open-source, Python-based alternative to Google's NotebookLM, backed by LlamaCloud for document ingestion and retrieval… | 52 | 1967 | active |
| SaiAkhil066/CORTEX-AI-SUPER-RAG CORTEX RAG is a local-first, agentic retrieval-augmented generation application that lets users upload PDFs and ask questions with cited an… | 61 | 1962 | active |
| mudler/LocalAGI LocalAGI is a self-hostable, privacy-focused AI agent platform written in Go that lets users create no-code agents, automations, and chatbo… | 90 | 1961 | active |
| firecrawl/fireplexity Fireplexity is an open-source, Perplexity-style AI search engine built with TypeScript that delivers answers with real-time citations, stre… | 36 | 1958 | active |
| kenforthewin/atomic Atomic is a self-hosted, local-first personal knowledge base that turns markdown notes into a semantically-connected, AI-augmented knowledg… | 77 | 1928 | active |
| microsoft/sample-app-aoai-chatGPT A Microsoft sample web chat application that connects to Azure OpenAI chat models, including support for the 'On Your Data' feature with da… | 74 | 1926 | active |
| alexpinel/Dot Dot is a standalone Electron desktop application for chatting with your documents using fully local LLMs and Retrieval Augmented Generation… | 16 | 1911 | active |
| doobidoo/mcp-memory-service A self-hosted persistent memory service for AI agents and Claude, exposing storage and semantic retrieval via REST API, MCP, OAuth, CLI, an… | 84 | 1907 | active |
| melih-unsal/DemoGPT DemoGPT is a Python framework and autonomous agent that generates LangChain-based LLM agent applications from natural-language prompts, bun… | 53 | 1905 | active |
| bhaskatripathi/pdfGPT pdfGPT is an open-source Python application that lets users chat with the contents of PDF files using GPT capabilities. It implements a sim… | 62 | 7167 | maintenance |
| cortexkit/magic-context Magic Context is a TypeScript library that gives coding agents self-managing, unbounded memory so a single session can last indefinitely. I… | 77 | 1892 | active |
| CaviraOSS/PageLM PageLM is an open-source, AI-powered education platform inspired by NotebookLM that transforms study materials like PDFs and notes into int… | 62 | 1884 | active |
| netease-youdao/BCEmbedding BCEmbedding is NetEase Youdao's open-source library of bilingual and crosslingual embedding and reranker models for English and Chinese, bu… | 44 | 1883 | active |
| BidingCC/BuildingAI BuildingAI is an open-source, enterprise-grade platform for visually assembling AI applications without code, described as the 'WordPress o… | 86 | 1870 | active |
| e-p-armstrong/augmentoolkit Augmentoolkit is a Python application that generates domain-expert fine-tuning datasets from uploaded documents and trains custom LLMs on t… | 60 | 1864 | active |
| Kav-K/GPTDiscord GPTDiscord is a self-hosted Discord bot providing an all-in-one GPT interface with ChatGPT-style conversations, DALL-E image generation, AI… | 62 | 1855 | active |
| moorcheh-ai/memanto Memanto is a companion 'memory agent' that manages long-term semantic memory for AI agents — deciding what to keep, resolving contradiction… | 81 | 1841 | active |
| NVIDIA-AI-Blueprints/video-search-and-summarization NVIDIA's GPU-accelerated AI Blueprint reference architecture for building video analytics agents that search, summarize, and reason over li… | 83 | 1824 | active |
| jacoblee93/fully-local-pdf-chatbot A Next.js web application that lets users chat with uploaded PDF documents entirely locally, performing chunking, embedding, vector storage… | 52 | 1817 | active |
| generative-computing/mellea Mellea is a Python library for writing generative programs, where LLM calls are first-class operations with type-annotated outputs, verifia… | 83 | 1799 | active |
| AI-Citizen/SolidGPT SolidGPT is an AI-powered semantic search assistant for developers that lets you chat with your codebase and Notion workspace. It ships as … | 19 | 1796 | active |
| jordan-gibbs/hyperresearch Hyperresearch is a Python CLI harness that turns Claude Code into a deep research agent running a 16-step adversarially-audited pipeline. I… | 78 | 1795 | active |
| onestardao/WFGY WFGY is an open-source ecosystem of protocols and tools for debugging and improving AI reasoning, RAG pipelines, and agent workflows, curre… | 80 | 1785 | active |
| hitsz-ids/airda airda (Air Data Agent) is a Python-based multi-agent system for data analysis that understands natural language data requirements and gener… | 17 | 1785 | active |
| SmartFlowAI/EmoLLM EmoLLM is a series of open-source large language models fine-tuned for mental health understanding and support, built on models like Intern… | 59 | 1780 | active |
| neuml/paperai paperai is an AI application for medical and scientific papers that runs bulk LLM inference and RAG pipelines over article repositories to … | 67 | 1779 | active |
| xhluca/bm25s BM25S is a pure Python library implementing BM25 lexical ranking using Numpy/Scipy sparse matrices with an optional Numba backend, achievin… | 92 | 1774 | active |
| OpenBMB/StaffDeck StaffDeck is an open-source enterprise platform for building and managing AI-powered digital employees that codify professional experience,… | 80 | 1763 | active |
| run-llama/rags A Streamlit web app that lets users build a RAG (retrieval-augmented generation) pipeline over their data using natural language instructio… | 27 | 6549 | maintenance |
| GoogleCloudPlatform/agent-starter-pack A Python package from Google Cloud Platform that provides production-ready templates for building and deploying GenAI agents on Google Clou… | 74 | 6545 | maintenance |
| Feather-2/Burner-X Paper Burner X is a browser-based AI workstation for processing, translating, and analyzing academic documents like PDFs, DOCX, PPTX, and E… | 51 | 1750 | active |
| parthsarthi03/raptor The official Python implementation of RAPTOR (Recursive Abstractive Processing for Tree-Organized Retrieval), a retrieval-augmented generat… | 25 | 1750 | active |
| trpc-group/trpc-agent-go tRPC-Agent-Go is a Go-native framework for building production LLM agent systems, offering agents, graph workflows (LangGraph-equivalent), … | 81 | 1733 | active |
| lifan0127/ai-research-assistant Aria is a Zotero plugin that brings GPT-powered AI assistance to reference management, letting researchers chat with their library items, P… | 38 | 1726 | active |
| McGill-NLP/llm2vec LLM2Vec is a Python library that converts decoder-only large language models into powerful text encoders via bidirectional attention, maske… | 49 | 1712 | active |
| LLPhant/LLPhant LLPhant is a comprehensive PHP generative AI framework inspired by LangChain and LlamaIndex, providing tools for working with LLMs, embeddi… | 92 | 1708 | active |
| glidea/zenfeed zenfeed is a self-hosted AI-powered RSS reader and information hub built in Go. It aggregates RSS feeds, uses LLMs to filter, summarize, an… | 68 | 1707 | active |
| undreamai/LLMUnity LLMUnity is an open-source Unity plugin that runs large language models locally in Unity games via llama.cpp, enabling AI-driven characters… | 75 | 1701 | active |
| arabold/docs-mcp-server A self-hosted MCP server that indexes documentation from websites, GitHub, npm, PyPI, and local files so AI coding assistants can query up-… | 82 | 1692 | active |
| facebookresearch/coconut Official PyTorch implementation of Coconut, a method for training large language models to reason in a continuous latent space instead of e… | 61 | 1689 | active |
| xerj-org/xerj XERJ is a Rust-based, single-binary search engine built for AI agents, offering BM25, kNN vector, and hybrid search with Elasticsearch-comp… | 76 | 1670 | active |
| deepsense-ai/ragbits Ragbits is a Python framework of modular building blocks for rapidly developing generative AI applications, covering LLM interaction, promp… | 75 | 1668 | active |
| lotus-data/lotus LOTUS is a Python library providing a Pandas-like API of LLM-powered semantic operators (map, filter, extract, aggregate, top-k) for bulk p… | 83 | 1661 | active |
| Nutlope/turboseek TurboSeek is an open-source AI search engine web application inspired by Perplexity, built with Next.js and TypeScript. It searches the web… | 65 | 1655 | active |
| AgentEra/Agently Agently is a Python framework for building production-grade GenAI applications with structured outputs, observable actions, MCP/tool integr… | 94 | 1644 | active |
| langchain-ai/langmem LangMem is a Python library that helps AI agents learn and adapt from their interactions over time by extracting information from conversat… | 64 | 1626 | active |
| yincongcyincong/MuseBot MuseBot is a self-hosted Go chatbot application that connects messaging platforms (Telegram, Discord, Slack, Lark/Feishu, DingTalk, WeCom, … | 78 | 1623 | active |
| a16z-infra/companion-app A tutorial starter stack for building and hosting AI companions with personality, backstory, and conversational memory. It combines Next.js… | 29 | 5977 | maintenance |
| stanford-oval/WikiChat WikiChat is a retrieval-augmented generation (RAG) framework that grounds LLM chatbot responses in a corpus (Wikipedia by default) to reduc… | 47 | 1614 | active |
| Lynpoint/CyberVerse CyberVerse is an open-source, self-hosted framework for building real-time, voice-first AI agents with optional digital-human video (talkin… | 63 | 1612 | active |
| clusterzx/paperless-ai A self-hosted web application that automatically analyzes, tags, and classifies documents in Paperless-ngx using LLMs via OpenAI-compatible… | 73 | 5910 | maintenance |
| THUDM/WebGLM WebGLM is an efficient web-enhanced question answering system (KDD 2023) that combines a large language model with web search retrieval and… | 35 | 1602 | active |
| qnguyen3/chat-with-mlx A chat playground application for running LLMs locally on Apple Silicon Macs using Apple's MLX framework. It provides a chat UI with suppor… | 25 | 1593 | active |
| D-Star-AI/dsRAG dsRAG is a high-performance retrieval engine for unstructured data, implemented as a Python library for building RAG pipelines. It improves… | 48 | 1589 | active |
| AkariAsai/OpenScholar OpenScholar is a retrieval-augmented language model system from the Allen Institute for AI that answers scientific questions by searching t… | 38 | 1584 | active |
| ShenSeanChen/waku-agent Waku Agent is a local-first AI agent harness in Python that exposes the full agent stack — harness, loop, memory, and eval/LLM-Ops — as sma… | 78 | 1580 | active |
| MLSysOps/MLE-agent MLE-Agent is an LLM-powered CLI agent that acts as a pairing assistant for machine learning engineers and researchers. It automates ML task… | 67 | 1568 | active |