5 packages found
LLM-as-judge evaluation framework for assessing AI agent output quality
Manage LLM prompts with collaborative version control — official Langfuse
LLM prompt engineering — chain of thought, few-shot, system prompts, evaluation
Give Claude Code memory that evolves with your codebase via hooks and LLM-compiled knowledge
Prompt engineering — crafting, testing, optimizing prompts for LLMs