42 packages found
CI for Claude Skills — lint, eval, and regression-test SKILL.md files across a model matrix, with a self-growing eval lo
Monocle is a framework for tracing GenAI app code. This repo contains implementation of Monocle for GenAI apps written i
How real engineers run Claude Code and Codex: spec-driven planning, enforced TDD, persistent memory, and quality enforce
Automated Claude Code QA gate for AI-assisted development — Codex code review + Playwright browser smoke testing before
Claude Skills for Governance, Risk, & Compliance (GRC): Expert-level compliance guidance for ISO 27001, SOC 2, FedRAMP,
Elixir implementation of a LangChain style framework that lets Elixir projects integrate with and leverage LLMs.
50 QA and test-automation skills for Claude Code, Codex, Cursor, and any Agent Skills Standard runtime.
Playwright skills for writing, reviewing, and debugging reliable end-to-end tests.
Run Codex and Claude Code reliably in the background on macOS—with a native dashboard, persistent tmux sessions, local c
Frontier models sell confidence. FABULA ships proof — an agent harness where any model is a swappable chip and every fin
Repackages the official Claude Desktop .deb installer for other distros
An experimental game engine for robots (and their humans), made by robots (and their humans)
Cover your Mac screen with a hotkey while AI agents keep running. Dog/cat mascot, native Settings, Touch ID unlock.
Production-grade Go SDK for building AI agents with long-term memory, knowledge retrieval, and voice — runnable as a lib
Multi-agent sprint orchestration for Claude Code — /tickets plans, /sprint executes
🌊 The leading agent orchestration platform for Claude. Deploy intelligent multi-agent swarms, coordinate autonomous wor
🚀 One-command requirement development flow for Claude Code - Complete workflow system with sub-agents, quality gates, a
AI-driven SDLC harness for Claude Code — multi-agent TDD workflow (plan → tests → code → review → PR) with Azure DevOps,
Communicate with an LLM provider using a single interface
Give your AI agent eyes and hands on your real Chrome browser — your tabs, your logins, your page state. 42 commands, ze
Personal AI assistant supporting Claude Code and Gemini CLI - productivity automation, AppleScript skills, and custom wo
Reusable Claude Code skills: deck-builder (polished .pptx decks), review-contrib (safe PR/fork review), autocli-password
⌥ AI Coding agent for the terminal — hash-anchored edits, optimized tool harness, LSP, Python, browser, subagents, and
AI-powered QA pipeline: Playwright tests + Claude API agent that classifies failures, auto-heals broken locators via PR,
A self-hostable LLM router / gateway: OpenAI- and Anthropic-compatible, explicit-first routing with fallbacks, spend lim
Autonomous Claude Code agent — 14 specialized agents, 10-parallel execution, project-specific skill learning, mechanical
Unofficial Kotlin multiplatform variant of the Anthropic SDK
The operating system an executive runs their company from — research, communications, CRM, content, and operations. Clau
Security audit tool for Claude Desktop and Claude Code on macOS — single-command visibility into MCP servers, extensions
PHP SDK for Claude - Provides complete 1-for-1 functionality of the Official Python SDK
Autonomous B2B prospecting engine. The AI sources, researches and writes your outreach every evening, you review your qu
Local multi-agent coding orchestrator with 22 pipeline roles, TDD enforcement, SonarQube integration, and automated code
🧠 Knowledge-Driven Development (KDD): ready-to-use template combining OKF (knowledge as linked markdown nodes) + CCDD (
The pretty much "official" DSPy framework for Typescript
Custom AWS Transform agent that migrates PyTorch/Triton kernels to AWS Trainium NKI (@nki.jit) and compiles, numerically
YantrikDB memory provider for NousResearch/hermes-agent — self-maintaining memory with canonicalization, contradiction t
Universal, model-agnostic operating harness for AI agents (Claude, Codex, Gemini, …) — a lean core + work-type profiles
A map of what AI tokens actually cost, and where they're wasted vs. well spent. Tools, research, practices, and copy-pas
Multi-agent autonomous SDLC framework. Spec to deployed app. PRD, GitHub issue, OpenAPI/JSON/YAML, or one-line brief. 5
Agentic Theorem Prover for Rocq for Program Verification