118 packages found
Deterministic quality scorer for AI agent instruction files — 8-dimension scoring with security, multi-format (SKILL.md,
A map of what AI tokens actually cost, and where they're wasted vs. well spent. Tools, research, practices, and copy-pas
Eco mode for Claude Code. /eco: -31% to -73% output tokens with critical findings intact; /eco-max: up to -75% with lowe
Awesome papers involving LLMs in Social Science.
Claude Skills for Governance, Risk, & Compliance (GRC): Expert-level compliance guidance for ISO 27001, SOC 2, FedRAMP,
Source-backed GPT-5.6 use cases for coding, agents, creative work, integrations, benchmarks, and practical limits.
Give your AI agent eyes and hands on your real Chrome browser — your tabs, your logins, your page state. 42 commands, ze
Local code navigation for coding agents: deterministic symbol graph, semantic search, compact briefings, and byte-exact
It is a comprehensive resource hub compiling all LLM papers accepted at the International Conference on Learning Represe
Agent skills for Claude Code where every entry ships with receipts: accuracy-gated benchmarks vs baseline AND placebo. R
AI writes code. This automates everything else · 24 plugins · 49 agents · 44 skills · for Claude Code, OpenCode, Codex,
AI builds faster than anyone, with no skin in the game and no memory of yesterday. How do you govern that? With X2, deci
Official companion repository for our survey "A Survey of the OpenClaw Ecosystem: From Platform Extensibility to Constra
Engineering decisions engine that know when they're stale. Frame, compare, decide — with evidence decay and parity enfo
Find out where your coding agent starts degrading. Personal context-rot analytics from your own sessions - 100% local, z
Claude Code usage governor: compact professional output, context slimming, tool-output filtering, telemetry, and drift g
An agent that takes a dead research repo and turns it into a callable pipeline component.
A repo lists papers related to LLM based agent
Ultimate Claude Fable 5 Guide 2026: Use Cases, Integrations & Benchmarks
Frontier models sell confidence. FABULA ships proof — an agent harness where any model is a swappable chip and every fin
Run Claude in self-improving loops to optimize measurable goals.
Agent memory for LLMs: 30 runnable Jupyter notebooks covering conversation buffers, vector stores, knowledge graphs, epi
[🏆 CHI26 Best Paper] CoBRA: Reproducible control of LLM agent behavior via classic social science experiments
历年ICLR论文和开源项目合集,包含ICLR2021、ICLR2022、ICLR2023、ICLR2024、ICLR2025.
Source-backed Claude Opus 5 use cases, workflows, model comparisons, benchmarks, integrations, costs, and limitations.
Multi-model orchestration layer for Claude Code — the frontier model plans, cheaper models execute, verification guards
AI agent platform for building multi-agent systems with orchestration, memory, RAG, workflows, and enterprise observabi
Token Cost Parity: Multilingual LLM Efficiency Analysis 2026
Multi-Agent Code Review Cost Benchmark 2026: Librarian vs Prompt Cache
Multi-agent code review mesh — orchestrates AI agents from multiple providers to review code in parallel, cross-review e
HumanStudy-Bench: Towards AI Agent Design for Participant Simulation
Parameterized multi-agent orchestration framework for Claude Code and Gemini
A curated list of Generative AI tools, works, models, and references
🌊 The leading agent orchestration platform for Claude. Deploy intelligent multi-agent swarms, coordinate autonomous wor
Open-source CLI that measures AI coding-assistant impact (Claude Code, Copilot, Cursor, Aider) across Bitbucket+Jira, Gi
Your AI forgets. This remembers. Spec-driven coding harness for vibecoders, product owners, CEOs and real builders — sel
macOS menu bar AI usage monitor for multiple Claude and Codex accounts: quota, rate limits, and a smart launcher for Cla
Self-hosted Python job-search pipeline for H-1B & visa holders. You run it on your own laptop — no server, no sign-up, n
🔴 VERY LARGE AI TOOL LIST! 🔴 Curated list of AI Tools - Updated 2026
Save 30% token costs when using Claude Code, Codex, OpenCode for free - with open source, local semantic search. Works f
Multi-agent autonomous SDLC framework. Spec to deployed app. PRD, GitHub issue, OpenAPI/JSON/YAML, or one-line brief. 5
A Next.js starter showing how to run Claude Code / Codex in safe, verified loops — plan → build → report cycles with loc
ConcoLLMic: the first language- and theory-agonistic concolic execution engine via LLM agents
Live DevTools for your running React + TanStack app — driven from the terminal by your AI agent, through one CLI
Ultra-Fast Python Solutions: 110x Speed Optimization Guide 2026
How real engineers run Claude Code and Codex: spec-driven planning, enforced TDD, persistent memory, and quality enforce
Custom AWS Transform agent that migrates PyTorch/Triton kernels to AWS Trainium NKI (@nki.jit) and compiles, numerically
Run Codex and Claude Code reliably in the background on macOS—with a native dashboard, persistent tmux sessions, local c