5 packages found
Offline MCP server that predicts LLM call cost and recommends the cheapest capable model before you call it. No API keys
A lossless token-optimization codec for LLM agents — cut input tokens up to 82%, nothing dropped, every fact kept. Drop-
a local-first, multi-provider tool that captures LLM API spend and exposes it to coding agents via the Model Context Pro
Fastest enterprise AI gateway (50x faster than LiteLLM) with adaptive load balancer, cluster mode, guardrails, 1000+ mod
Real-time cost observability for Model Context Protocol (MCP) tool calls. Wraps any MCP server, attributes spend per too