13 packages found
Regression testing for AI agents. Snapshot behavior,diff tool calls,catch regressions in CI. Works with LangGraph, CrewA
A Claude Code skill that adds a rubric-based eval layer to any agent project. Framework-agnostic — generates rubric, tes
Build Claude Code–style deep agents in Python: tool-calling, sandboxed execution, multi-agent teams, skills, checkpoints
Agentic QA boilerplate built on Playwright + KATA + TypeScript. Multi-agent skills following the agentskills.io spec, or
Claude Code skills for medical research — literature search, reporting guidelines, statistical analysis, publication fig
Claude Code plugin marketplace: 145 framework-aware code-review skills plus AI-writing detection, doc and test-plan gene
4-stage evaluation framework for testing Claude Code plugin component triggering. Validates skills, agents, and commands
A curated list of developer tools, SDKs, libraries, and testing utilities for Model Context Protocol (MCP) server develo
A curated system of production-ready Claude Code skills with quantitative evaluation reports, golden test fixtures, and
欢迎大家提issue。 AI Test Agent 是一个基于人工智能的自动化测试平台,利用大语言模型(LLM)和浏览器自动化技术,实现测试用例的智能生成、自动执行、Bug 分析和报告生成。 Enterprise_AI_QA_Agent
An MCP server that autonomously evaluates web applications.
AI 測試大師 — MCP server driving pytest / Jest / Cypress / Go / Maestro. Analyze, generate, run, advise. Web + Mobile (iOS/A
Production-ready Claude Code skills for building AI agents with Pydantic AI. Includes dependency injection, tools, valid