4 packages found
LLM-as-judge evaluation framework for assessing AI agent output quality
Manage LLM prompts with collaborative version control — official Langfuse
Give Claude Code memory that evolves with your codebase via hooks and LLM-compiled knowledge
Prompt engineering — crafting, testing, optimizing prompts for LLMs