Vibe Coding Discover

Category

AI Frameworks

Core frameworks for building LLM applications.

109 projects

Project nameStarsCategoryLanguageTagsSummaryView Project
yoagent★ 179AI FrameworksRustRust · agent loop · tool calling · streaming · coding agents · multi-agent · MCP · OpenAPI · local models · session branching · middleware · LLM providersA Rust framework for building tool-using LLM agents, with streaming support across seven protocols, built-in tools, MCP and OpenAPI integrations, sub-agents, and session management. Includes a terminal coding-agent example.View Project →
self-hosted-ai-stack★ 156AI FrameworksShellself-hosted · docker-compose · ollama · litellm · mcp · rag · local-llm · whisper · text-to-speech · embeddings · docling · cuda · privacy · multi-arch · openai-compatible · infrastructureDocker Compose bundle that deploys a full local AI stack: Ollama for LLMs, LiteLLM gateway, AnythingLLM chat UI, embeddings/RAG, Whisper STT, Kokoro TTS, Docling parsing, and an MCP Gateway. Includes lightweight stack variants, optional HTTPS and CUDA GPU acceleration.View Project →
swift-ai-sdk★ 154AI FrameworksSwiftswift · sdk · llm · streaming · tool-calling · structured-outputs · mcp · multi-provider · apple-platforms · ios · macos · vercel-ai-sdk · middleware · text-generation · function-calling · provider-agnosticSwift port of the Vercel AI SDK offering one provider-agnostic API for streaming text, structured outputs, tool calling, MCP tools, and middleware across 38 providers via SwiftPM. Built for iOS/macOS apps needing OpenAI, Anthropic, Google and others from Swift.View Project →
neurolink★ 144AI FrameworksTypeScripttypescript · llm · multi-provider · mcp · rag · voice · agents · memory · sdk · cli · streaming · embeddings · tts · stt · model-routing · failoverTypeScript AI SDK unifying 30+ LLM providers (OpenAI, Anthropic, Gemini, Bedrock, Ollama) behind one streaming API. MCP-native with built-in RAG, memory, voice TTS/STT, agents, and provider failover. Extracted from Juspay production systems.View Project →
llm★ 141AI FrameworksRubyruby · llm-runtime · agents · tool-calling · mcp-client · a2a · rag · skills · streaming · multi-provider · concurrency · persistence · zero-dependencies · active-record · sequel · consolellm.rb is a zero-dependency Ruby runtime for building agentic LLM apps: a single API across 14+ providers, managed tool loops, tools, skills, MCP/A2A clients, streaming callbacks, concurrency strategies and ActiveRecord/Sequel persistence, plus an interactive agent console.View Project →
model-compose★ 113AI FrameworksPythonyaml · declarative · orchestration · llmops · agents · rag · mcp · multi-provider · docker · streaming · workflow · self-hosted · human-in-the-loop · vector-database · local-models · docker-compose-alternativemodel-compose is a declarative YAML orchestrator (docker-compose for AI) that deploys chat APIs, ReAct agents, RAG pipelines, and MCP servers from one file. It bridges local and cloud models and runs on Docker, native, or distributed Redis-queued runtimes.View Project →
ai-microcore★ 108AI FrameworksPythonllm · python · mcp · rag · vector-database · prompt-templates · provider-agnostic · embeddings · semantic-search · streaming · jinja2 · llm-adapters · chat-completion · tool-calling · minimalistMicroCore is a minimalist Python library of LLM and vector-DB adapters that makes providers switchable via config while keeping app code unchanged. It includes prompt templating, streaming, embeddings search (Chroma/Qdrant) and LLM-agnostic MCP tool integration.View Project →
quarkus-workshop-langchain4j★ 107AI FrameworksJavaworkshop · Quarkus · LangChain4j · LLM · chatbot · AI services · agent orchestration · MCPA step-by-step Java workshop for building AI applications with Quarkus and LangChain4j, covering single AI services and agentic orchestration. Each lesson has a runnable project state.View Project →
HOMER★ 45AI FrameworksPythonlong-context · kv-cache · context-extension · llm-inference · llama-2 · memory-efficiency · attention · transformers · pytorch · research-implementation · hierarchical-merging · passkey-retrieval · perplexity-evaluation · training-free · flash-attention · iclr-2024Official ICLR 2024 implementation of HOMER, a training-free hierarchical KV-cache merging method that extends pre-trained LLM context limits (e.g. Llama-2) with lower memory. Ships patched LlamaForCausalLM, plus passkey-retrieval and PG19 perplexity scripts.View Project →