7 packages found
LLM-as-judge evaluation framework for assessing AI agent output quality
let your LLM drive bitbucket ci/cd
Manage LLM prompts with collaborative version control — official Langfuse
LLM prompt engineering — chain of thought, few-shot, system prompts, evaluation
Give Claude Code memory that evolves with your codebase via hooks and LLM-compiled knowledge
MCP server for real-time Google, Bing, Yandex, and DuckDuckGo search results with sub-second latency.
Prompt engineering — crafting, testing, optimizing prompts for LLMs