2 packages found
The benchmark tasks and evaluation harness for "PhysicianBench: Evaluating LLM Agents in Real-World EHR Environments".
📓 Claude Code skill and agent for automatic project documentation in Obsidian