2 packages found
MCP server for LLM quantization. Compress any model to GGUF/GPTQ/AWQ in one tool call. First MCP server for model compre
The highest-scoring AI memory system ever benchmarked that isn't reliant on LLM reranking. And it's free & burns less to