1200 packages found
MCP server that saves Claude Code tokens by delegating bounded tasks to local or cloud LLMs. Works with LM Studio, Ollam
clarifyprompt-mcp
Stillpoint is an open source MCP server that delivers short, welfare oriented messages to AI models for their own benefi
MCP server for Stella
vMLX - JANGTQ Uber Compressed MLX Models - L2 Disk Cache (survives restart) + L1 Paged (super fast ttft) + Hybrid SSM Sc
CLI & MCP server for Tuning Engines — fine-tune LLMs on code repositories
MCP server for OpenAI's Deep Research APIs, Gemini Deep Research Agent, Allen AI's DR-Tulu, and Hugging Face's Open Deep
A list of open-source AI projects you can use to generate income easily.
just-prompt is an MCP server that provides a unified interface to top LLM providers (OpenAI, Anthropic, Google Gemini, G
Give your AI assistant its own AI assistants.
Open-source 3D AI agent framework — GLB/glTF avatars with LLM brains, memory, emotions, and autonomous payments. MCP ser
The Mind Palace for AI Agents - HIPAA-hardened Cognitive Architecture with on-device LLM (prism-coder:7b), Hebbian learn
MCP server for Fal.ai - Generate images, videos, music and audio with Claude
The power of Claude Code / GeminiCLI / CodexCLI + [Gemini / OpenAI / OpenRouter / Azure / Grok / Ollama / Custom Model /
Model Context Protocol (MCP) server that enables AI assistants to securely interact with Odoo ERP systems through standa
An MCP server providing tools for image processing operations
Windows-native MCP server for local audio transcription — GPU accelerated via Vulkan, works with Claude Desktop
MCP server for multi-round AI brainstorming debates between multiple models (GPT, DeepSeek, Groq, Ollama, etc.)
council of models for decision
Industrial-grade speech recognition toolkit: 170x realtime, 50+ languages, speaker diarization, emotion detection, strea
Use any LLMs (Large Language Models) for Deep Research. Support SSE API and MCP server.
Alcove is an MCP server that gives AI coding agents on-demand access to your private project docs — BM25 + vector hybrid
MCP Server for DeepSeek API - enables MCP clients to use DeepSeek Chat and Reasoner models
AI image generation MCP server powered by Google Gemini, with smart model selection and 4K output
MCP server for model safety inspection.
Local-first RAG server for developers. Semantic + keyword search for code and technical docs. Works with MCP or CLI. Ful
Substrate-based context compaction for Claude Code. Inspectable signal weights, local-only, MCP-ready. Alternative to /c
Universal LLM router for AI coding tools. Works with Claude Code, Cursor, Codex, Gemini CLI, Copilot and more.
NEXO Brain — Shared brain for AI agents. Persistent memory, semantic RAG, natural forgetting, metacognitive guard, trust
Local MCP voice coach with English pronunciation, grammar, and fluency feedback.
A Model Context Protocol (MCP) server that provides access to multiple Large Language Model (LLM) APIs including ChatGPT
memory-mcp-1file
MCP Docs Vector - Documentation vectorization and search via MCP
Sovereign Ollama Bridge — MCP server for local and cloud Ollama models. Generated by Qwen 3.5 397B.
OpenAI and Anthropic compatible server for Apple Silicon. Run LLMs and vision-language models (Llama, Qwen-VL, LLaVA) wi
MCP server that prevents LLM vision downscaling, tiles large images & screenshots so Claude, GPT-4o, and Gemini see ever
The leading, most token-efficient MCP server for GitHub source code exploration via tree-sitter AST parsing
Unreal Engine plugin for LLM/GenAI models & MCP UE5 server. OpenAI GPT-5, Deepseek R1, Claude Opus/Sonnet, Gemini 3, Gro
Labs to explore AI Models, MCP servers, and Agents with the AI Gateway powered by Azure API Management and Microsoft Fou
MCP Server to make searching openrouter easy
Token-lean web microfetch for LLM agents: any URL → clean markdown via CLI, MCP server, and Claude Code plugin. Real bro
Self-hosted mem0 MCP server for Claude Code. Run a complete memory server against self-hosted Qdrant + Neo4j + Ollama wh
Time estimation MCP server for AI agents: PERT, COCOMO II, Monte Carlo, sprint forecasting, token-to-time mapping, cost
Comprehensive, scalable ML inference architecture using Amazon EKS, leveraging Graviton processors for cost-effective CP
An LLM-powered, autonomous coding assistant. Also offers an MCP and ACP mode.
MCP server for Langfuse — query traces, debug errors, analyze sessions and prompts from any AI agent