#llm-agents (30 Repositories)
Ranked open-source repositories tagged with #llm-agents, scored by pull request acceptance likelihood and maintainer engagement velocity.
32.4%
25.9h
30 repositories tagged #llm-agents
oliver-zehentleitner/keep-the-why
Keep the Why: a repo-native convention and agent skill that preserves the reasoning behind a codebase as a byproduct of working with your agent — so it stops re-suggesting rejected approaches, gives better answers, speeds up onboarding, and makes legacy projects tractable again.
azalio/map-framework
Plan-then-build AI coding for Claude Code & Codex CLI — you approve the plan before the model writes a line of code. SPEC → PLAN → TEST → CODE → REVIEW → LEARN
arcee-ai/nac
Give AI agents ambitious work without losing the plot. nac is an open-source harness for long-running tasks, using a central orchestrator, threads, and structured episodes to stay aligned with your intent.
openlegion-ai/openlegion
Secure autonomous AI agent framework and platform. Build AI teams by describing what you want. Orchestrate agents that can do everything a human can do.
strukto-ai/mirage
The World's First Unified Virtual Filesystem For AI Agents
Human-Agent-Society/CORAL
Open-source autoresearch powered by autonomous coding agents. Run Claude Code, OpenCode, and Codex with grading, shared knowledge, and multi-agent evolution. Accepted at COLM 2026.
KimGLee/Cambium
Governance standard and reference toolset for LLM-maintained knowledge corpora
r-uby-dev/llm
cruby's capable AI runtime
IBM/AssetOpsBench
AssetOpsBench - Industry 4.0: A unified benchmark and framework for building, orchestrating, and evaluating domain-specific AI agents for Industry 4.0 asset operations and maintenance, with 460+ scenarios, 5 specialist agents (IoT, FMSR, TSFM, Work Order,...), and multi-agent orchestration blueprints (MetaAgent, AgentHive) over MCP.
HKUDS/nanobot
Ultra-lightweight, open-source, self-hosted personal AI agent framework in Python with WebUI, tools, memory, MCP, multi-agent workflows, automation, and chat apps
WenyuChiou/awesome-agentic-ai-zh
A trilingual (繁中 / English / 简中) learning roadmap for agentic AI: from LLM basics to multi-agent systems, with 240+ curated resources and hands-on examples. 中文 AI agent 學習地圖。
a-streetcoder/agent-deck
Agent Deck
stripe/ai
One-stop shop for building AI-powered products and businesses with Stripe.
aiming-lab/AutoResearchClaw
Fully autonomous & self-evolving research from idea to paper. Chat an Idea. Get a Paper. 🦞
alexisfox7/PRO-LONG
Programmatic memory for long-horizon LLM agents: the harness appends everything to one log, and the agent searches it with code. 97.4% on ARC-AGI-3 (arXiv:2607.20064)
maze-agent/Maze
A distributed framework for LLM agents
yzhao062/awesome-auditable-ai
Auditing AI agents: a curated list of papers, tools, datasets, benchmarks, and standards covering reliability, monitoring, failure attribution, and decision records.
A-EVO-Lab/a-evolve
The official repository of "Position: Agentic Evolution is the Path to Evolving LLMs".
svd-ai-lab/sim-cli
CLI-first runtime for Codex, Claude Code, and AI agents to operate CAE solvers via plugins: COMSOL, Abaqus, Ansys.
OpenRaiser/PaperFlow
📚 PaperFlow: Dynamic personalized scientific-paper recommendation, reading, and reporting
Text2SqlAgent/text2sql-framework
Agentic text-to-SQL SDK: hand the LLM one execute_sql tool and let it explore the schema, test queries, and self-correct — no RAG, no semantic layer. 20/20 on an 80-table Spider run. Built-in tracing and scenario support.
smileformylove/XScientist
Turn ideas into autonomous research with Git-like evidence histories—inspectable, reproducible, and reversible.
TIGER-AI-Lab/TheoremExplainAgent
Official Repo for "TheoremExplainAgent: Towards Video-based Multimodal Explanations for LLM Theorem Understanding" [ACL 2025 oral]
stanford-iris-lab/meta-harness
Reference code for the Meta-Harness paper.
camel-ai/oasis
🏝️ OASIS: Open Agent Social Interaction Simulations with One Million Agents.
EverMind-AI/SkillCorpus
Open-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.
Birfy/agentdescent
Gradient descent, but the parameters are agents — a parallel, asynchronous framework for self-evolving agents (skills, prompts, harnesses). Diffs are the gradients; the aggregator is the optimizer.
mll-lab-nu/RAGEN
Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics
vonzosten/awesome-LangGraph
An index of the LangChain + LangGraph ecosystem: concepts, projects, tools, templates, and guides for LLM & multi-agent apps.
Purewhiter/mobilegym
MobileGym: A Verifiable and Highly Parallel Simulation Platform for Mobile GUI Agent Research · 浏览器里运行的安卓模拟器 · Browser-hosted Android Simulator · Verifiable Evaluation · Scalable Online RL Training (EMNLP 2026)