开源精选 · 开源大模型
最好的开源大模型
每日更新。60 个项目,按 star 排序,并核对 fork 与维护活跃度。
开源大模型 镇场榜
成熟且仍在更新的项目。GitHub 上标注了 large-language-models, local-llm。
-
Langflow is a powerful tool for building and deploying AI-powered agents and workflows.
-
Implement a ChatGPT-like LLM in PyTorch from scratch, step by step
-
Course to get into Large Language Models (LLMs) with roadmaps and Colab notebooks.
-
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
-
为GPT/GLM等LLM大语言模型提供实用化交互接口,特别优化论文阅读/润色/写作体验,模块化设计,支持自定义快捷按钮&函数插件,支持Python和C++等项目剖析&自译解功能,PDF/LaTex论文翻译&总结功能,支持并行问询多种LLM模型,支持chatglm3等本地模型。接入通义千问, deepseekcoder, 讯飞星火, 文心一言, llama2, rwkv, claude2, moss等。
-
《深入理解 AI Agent:设计原理与工程实践》(李博杰 著)开源主仓库:全书正文、编译版 PDF 与按章配套代码
-
Ultra-lightweight, open-source, self-hosted personal AI agent framework in Python with WebUI, tools, memory, MCP, multi-agent workflows, automation, and chat apps
-
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
-
[EMNLP2025] LightRAG: Simple and Fast Retrieval-Augmented Generation
-
A Python library for extracting structured information from unstructured text using LLMs with precise source grounding and interactive visualization.
-
🪢 Open source agent evals & observability: Trace, evaluate, and improve LLM applications with one open platform.
-
Build and run agents you can see, understand and trust.
-
A one stop repository for generative AI research updates, interview resources, notebooks and much more!
-
Official code repo for the O'Reilly Book - "Hands-On Large Language Models"
-
Open-source AI orchestration framework for building context-engineered, production-ready LLM applications. Design modular pipelines and agent workflows with explicit control over retrieval, routing, memory, and generation. Built for scalable agents, RAG, multimodal applications, semantic search, and conversational systems.
-
Distribute and run LLMs with a single file.
-
The official repo of Qwen (通义千问) chat & pretrained large language model proposed by Alibaba Cloud.
-
FinGPT: Open-Source Financial Large Language Models! Revolutionize 🔥 We release the trained model on HuggingFace.
-
Code for Machine Learning for Trading, 3rd edition - from data sourcing to live execution.
-
Machine Learning Engineering Open Book
-
中文LLaMA&Alpaca大语言模型+本地CPU/GPU训练部署 (Chinese LLaMA & Alpaca LLMs)
-
Ongoing research training transformer models at scale
-
:sparkles::sparkles:Latest Advances on Multimodal Large Language Models
-
🐫 CAMEL: The first and the best multi-agent framework. Finding the Scaling Law of Agents. https://www.camel-ai.org
-
Sample code and notebooks for Generative AI on Google Cloud, with Gemini Enterprise Agent Platform
-
Automated Penetration Testing Agentic Framework Powered by Large Language Models
-
20+ high-performance LLMs with recipes to pretrain, finetune and deploy at scale.
-
💡 All-in-one AI framework for semantic search, LLM orchestration and language model workflows
-
A curated list of modern Generative Artificial Intelligence projects and services
-
Pocket Flow: Codebase to Tutorial
-
Hierarchical Reasoning Model Official Release
-
An open-source, tool-augmented conversational language model from Fudan University
-
A straightforward method for training your LLM, from downloading data to generating text.
-
Pocket Flow: 100-line LLM framework. Let Agents build Agents!
-
A curated list of 120+ LLM libraries category wise.
- Mooler0410/LLMsPracticalGuide ★ 10,208
A curated list of practical guide resources of LLMs (LLMs Tree, Examples, Papers)
-
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
-
High-speed Large Language Model Serving for Local Deployment
-
Anomaly detection related books, papers, videos, and toolboxes. Last update late 2025 for LLM and VLM works!
-
~95% on SimpleQA (e.g. Qwen3.6-27B on a 3090). Supports all local and cloud LLMs (llama.cpp, Ollama, Google, ...). 10+ search engines - arXiv, PubMed, your private documents. Everything Local & Encrypted.
开源大模型 新秀榜
近 90 天新建、涨势最猛的项目。
-
Serve large Qwen models fast on the GPUs you actually own. Qwen3.8-27B on a single 24 GB card with vLLM: 127 tok/s single-user (381 when the answer quotes the prompt), ~1,035 tok/s at 64 concurrent, 150k-262k context. vLLM patches, requant pipeline, benchmarks.
-
NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness.
-
大模型(LLM)全栈学习路线与中文教程🔥:覆盖 Prompt Engineering、RAG、AI Agent、MCP、微调、模型部署、Transformer、AI 编程与大厂面试,从入门到生产实践。
-
The fastest way to run Qwen3.8-Flash-Next on Strix Halo (gfx1151)
-
Run full Kimi K3 on a single device. And an OpenAI-compatible API server for local chat and coding agents.
-
Halo is an open-source framework built by White Circle for training large language and multimodal models
-
Swiftlet is a Swift and Metal runtime that runs large Qwen Mixture-of-Experts models locally on Apple devices by streaming expert weights from storage, enabling 35B and 80B models to run with low RAM, including on iPhone.
-
Peer-to-peer LLM inference in the browser: pool your devices to run big open models, for chat and coding agents.
-
A clinical decision-evidence benchmark: pull the one fact an answer rests on, and see whether the assistant's action moves with the evidence. Built on HealthBench.
-
AI Engineering Course - A free and complete AI Engineering Course to learn AI Engineering step by step - from Machine Learning, Neural Networks, and Transformers to LLMs, Fine-Tuning, RAG, AI Agents, LLM Inference, Evaluation, AI Safety, and AI System Design.
-
Run a 105 GB AI model on a Mac that can't hold it. Slotstream streams Qwen3.8-Flash-Next (125B mixture of experts) from your SSD and caches the busiest experts in memory, so it runs on Macs with 16 to 64 GB. One native Swift binary on MLX and Metal, no Python, offline. Works with Claude Code, Codex and Ollama or OpenAI clients.
-
Your iPhone helps your Mac run a 27B model: faster prompt reading and more context over a USB-C cable
-
本地优先的全栈 AI 创作工作站:AI对话 · AI漫画 · 漫剧创作 · 写作台 · 知识学习五大模块;vLLM / llama.cpp / Transformers 多引擎本地推理,显存智能调度、舱壁隔离自愈、仅供学习使用,欢迎交流指正。(最新进度(除插件系统)受到协议限制,暂时无法提交到仓库)
-
An intentionally vulnerable OWASP LLM Top 10 training platform for AI Security, Prompt Injection, RAG Security, Agent Security, and GenAI penetration testing.
-
Learn LLM Inference Engineering step by step - from KV cache, PagedAttention, and continuous batching to vLLM, SGLang, and GPUs.
-
Your car as a chat-room agent: Raspberry Pi 5 + dashcam + local AI. CodeWatch's sibling for the garage.
-
Local Qwen 3.8 27B uncensored Q4_K_M + harvested SYSTEM pack. Official 3.8 weights, not a 3.6 retitle.
-
TLive-Omni: An Omni-Modal Understanding Model for E-Commerce Live Streaming
-
One CLI for all your robots. Connect them, command them, and let them work together, each with an LLM for a brain (Claude, OpenAI, Gemini, Grok, or local via Ollama or vLLM), VLAs for arms and decision LLMs (Jev, Laya, Kev). Drives Microduck, Open Duck Mini, LeRobot, XLeRobot, AlohaMini, ToddlerBot, or any ROS 2 base from a laptop, offboard.
-
Jacobian-Brainwash : A manual alignment tool for large language models built on Anthropic's Jacobian Lens. Results are exportable.
这份榜单怎么来的
候选来自 GitHub 话题标签 large-language-models, local-llm。star 数与周增长取自 GitHub 官方的 star 历史接口,因此数字与 GitHub 自己的口径一致。榜单收录活跃、话题数量至多 20 个的仓库。某个项目的 star 数相对 fork 数明显偏高时会被标注,可结合仓库活动进一步了解原因。