Category
AI Agents
Autonomous and stateful agents for research, coding, and workflows.
| Project name | Stars | Category | Language | Tags | Summary | View Project |
|---|---|---|---|---|---|---|
| ajean | ★ 103 | AI Agents | Go | local-ai · self-hosted · llama.cpp · gguf · mcp · agent · browser-control · computer-use · persistent-memory · encryption-at-rest · openai-compatible · web-search · cuda · privacy · multimodal · single-binary | Single Go binary wrapping llama.cpp into a self-hosted AI assistant: chat UI, persistent Markdown memory, terminal/file/web/browser tools, MCP servers, encryption at rest and encrypted remote access. Handles GPU detection, model presets and an OpenAI-compatible endpoint. | View Project → |
| Micro-Agent | ★ 101 | AI Agents | Python | AI agent · ReAct · vertical domains · RAG · MCP · skills · subagents · agent memory · tool use · FastAPI · SSE | A Python framework for building domain-specific ReAct agents with configurable LLMs, tools, memory, skills, RAG, and MCP integration. Includes a FastAPI service with streaming output and example task workflows. | View Project → |
| P-ai | ★ 101 | AI Agents | Rust | desktop-ai-assistant · multi-agent · long-term-memory · mcp · agent-skills · tool-review · local-first · computer-use · tauri · rust · vue · workspace-automation · remote-im · personas · self-growing · privacy | Rust/Tauri desktop AI work system with multi-persona agents, long-term compressed memory, MCP and skill management, reversible tool execution with AI review, and remote IM bridges. Local-first, no intermediary servers; aimed at long-running tasks rather than chat. | View Project → |
| SourceWeft | ★ 100 | AI Agents | TypeScript | multi-agent · self-hosted · RAG · MCP · skills · knowledge-base · citations · sandbox · BYOK · AI workstation | A self-hostable AI workstation for multi-agent tasks, cited knowledge search, skills, and MCP-connected tools. Supports multiple model providers, sandboxed execution, and web and desktop clients. | View Project → |
| ENZO | ★ 100 | AI Agents | TypeScript | AI agents · self-hosted · BYOK · multi-provider · agent builder · agent skills · deep research · code generation · chatbot · Docker · model catalog | Self-hosted AI workspace for chatting with models, building agents, deep research, and code generation. Connect provider API keys directly; agents can use bundled skills and tools, including Gmail and Calendar integrations. | View Project → |
| ScienceBuddy | ★ 100 | AI Agents | Python | scientific-agents · interactive-agent · recursive-self-improvement · agent-harness · llm-agents · reinforcement-learning · grpo · agent-evaluation · biomedicine · tool-use · research-automation · trajectory-inspection · verifier · python | Research code and preview for ScienceBuddy, an interactive scientific agent workspace plus a double-recursive self-improvement experiment that alternates harness refinement with SkyRL GRPO model training on frozen scientific tasks. | View Project → |
| End-to-End-Agentic-Ai-Automation-Lab | ★ 100 | AI Agents | Jupyter Notebook | agentic-ai · multi-agent · langgraph · autogen · mcp · rag · langchain · n8n · vllm · memory · human-in-the-loop · guardrails · fine-tuning · deployment · docker · browser-automation | A large hands-on lab of Jupyter notebooks and Python projects covering LangGraph and AutoGen agent systems, MCP servers, production RAG with reranking, mem0 memory, n8n automation, plus fine-tuning and vLLM deployment. Best for developers learning end-to-end agentic AI by example. | View Project → |
| AMA-Bench | ★ 82 | AI Agents | Python | agent-memory · long-horizon · long-context · benchmark · evaluation · llm-as-judge · memory-retrieval · embeddings · bm25 · vllm · leaderboard · agent-trajectories · icml · python | AMA-Bench is an ICML 2026 evaluation framework for agentic memory: methods build memory from long agent trajectories, retrieve evidence, and answer QA scored by LLM-as-judge. Includes vLLM/API pipelines, cross-judge validation, and a HF leaderboard. | View Project → |
| MemoryArena | ★ 64 | AI Agents | Python | agent-memory · benchmark · multi-session · llm-agents · evaluation · memory-systems · long-context · tool-use · environments · research | MemoryArena is a research framework and benchmark for agent memory in interdependent multi-session agentic tasks. It wires pluggable memory backends (long-context, mem0, Letta, Mirix, GraphRAG, MemoRAG, BM25) into task agents and step-based environments for web shopping, travel, search, and formal reasoning. | View Project → |
| OneVOneJev | ★ 38 | AI Agents | TypeScript | AI opponent · game agent · FPS · quickscope · real-time · WebSocket · spectator mode · heuristic fallback | A browser-based 1v1 quickscope FPS where players fight Jev, an AI opponent driven by the TypeSafe SDK. The server runs the match simulation and uses a deterministic heuristic fallback when the AI service is unavailable. | View Project → |
| agency-agents | ★ 25 | AI Agents | Shell | ai-agents · subagents · agent-personas · claude-code · cursor · codex · gemini-cli · prompt-engineering · multi-agent · agent-roster · developer-tools · copilot · windsurf · aider · markdown-prompts · workflow-automation | A MIT-licensed roster of 200+ markdown AI agent personas (engineering, design, marketing, sales, security, GIS, game-dev) installable as subagents into Claude Code, Cursor, Codex, Gemini CLI and other agentic tools via shell scripts or a desktop app. | View Project → |
| jev-trader | ★ 20 | AI Agents | Python | AI trading · market making · autonomous trading · risk management · model calibration · paper trading · deterministic state | A Python market-making system that uses Jev for selected market judgments while deterministic code handles state, quote policy, execution, and hard risk vetoes. Includes paper trading, fallback behavior, and decision calibration. | View Project → |
| askgrokwallet | ★ 8 | AI Agents | JavaScript | agent-governance · agentic-commerce · ai-agents · human-in-the-loop · policy-engine · verifiable-receipts · wallet · ed25519 · x402 · erc-8196 · erc-8126 · mcp · approval-workflow · audit-log · base · smart-contracts | Rules-and-receipts layer for AI agents that spend money: plain-English policy compiles to allow/ask/deny, risky actions hit a human approval inbox, and every outcome produces an Ed25519-signed receipt anchored on Base. Contracts are unaudited; mainnet settlement not yet demonstrated. | View Project → |
| customer-service-agent | ★ 1 | AI Agents | Python | customer-service · langgraph · rag · booking · appointment-scheduling · ragflow · fastapi · stripe-payments · sse-streaming · voice · whisper-stt · tts · postgres · human-handoff · ticketing · single-tenant | ServiceEmma is a LangGraph customer-service agent for appointment businesses: it answers FAQs from RAG knowledge, books or cancels appointments, holds slots via Stripe Checkout, and escalates to support tickets. Ships with FastAPI backend, SSE chat widget, Postgres state, voice, and an owner dashboard. | View Project → |