33 packages found
One place to manage & connect to all your MCP servers
MCP server that saves Claude Code tokens by delegating bounded tasks to local or cloud LLMs. Works with LM Studio, Ollam
Comprehensive open-source library of AI research and engineering skills for any AI model. Package the skills and your cl
A Model Context Protocol (MCP) server that provides access to multiple Large Language Model (LLM) APIs including ChatGPT
MCP Server to Use HuggingFace spaces, easy configuration and Claude Desktop mode.
Give your AI assistant its own AI assistants.
Context Rot Detection & Healing MCP Service — gives AI agents self-awareness about their cognitive state
The Mind Palace for AI Agents - HIPAA-hardened Cognitive Architecture with on-device LLM (prism-coder:7b), Hebbian learn
MCP server for LLM quantization. Compress any model to GGUF/GPTQ/AWQ in one tool call. First MCP server for model compre
The unified web layer for AI agents. Search (8 engines), stealth browse, auth, and act on 24 platforms. One npm install,
Deterministic research MCP server on FastMCP 3 — 5-engine web search, 9-platform social search, 6 academic DBs, news agg
Local-first RAG server for developers. Semantic + keyword search for code and technical docs. Works with MCP or CLI. Ful
memory-mcp-1file
Python SDK, Proxy Server (AI Gateway) to call 100+ LLM APIs in OpenAI (or native) format, with cost tracking, guardrails
vMLX - JANGTQ Uber Compressed MLX Models - L2 Disk Cache (survives restart) + L1 Paged (super fast ttft) + Hybrid SSM Sc
Single-binary, Chrome-free headless browser for LLM agents. Cheap layer of a two-engine scraping stack — handles 80% of
Token optimizer for Claude Code Gemini CLI & Qwen Code — compresses shell outputs via PreToolUse hook, tracks USD saving
Optimize Claude Code token usage — 22 skills, 8 hooks, 6 rules. One npm install. Zero manual steps. Cut costs by 67%.
Precomputed reasoning cache for AI infrastructure decisions — 220K+ pages tracking the open-source AI ecosystem
Lightweight Long-Term Memory for LLM Agents.
Optimize your AI prompts for maximum efficiency. Reduce tokens, improve clarity, and learn better prompting techniques.
AlgoTune is a NeurIPS 2025 benchmark made up of 154 math, physics, and computer science problems. The goal is write code
syftr is an agent optimizer that helps you find the best agentic workflows for your budget.
Industrial-grade speech recognition toolkit: 170x realtime, 50+ languages, speaker diarization, emotion detection, strea
🪢 Open source LLM engineering platform: LLM Observability, metrics, evals, prompt management, playground, datasets. Int
CLI & MCP server for Tuning Engines — fine-tune LLMs on code repositories
Universal LLM router for AI coding tools. Works with Claude Code, Cursor, Codex, Gemini CLI, Copilot and more.
Your AI assistants don't know who you are. mcp-me fixes that: a local MCP server that gives any AI a full picture of who
An MCP server providing tools for image processing operations
Creative AI MCP server — 32 tools (18 free): image generation, upscaling, bg removal, mockups, CMYK, print-ready PDF, ve
wishfinity-mcp-plusw
Comprehensive, scalable ML inference architecture using Amazon EKS, leveraging Graviton processors for cost-effective CP