82 packages found
AI agent security scanner. Detect vulnerabilities in agent configurations, MCP servers, and tool permissions. Available
A structural X-ray for the codebases AI agents are writing.
Claude Skills for Governance, Risk, & Compliance (GRC): Expert-level compliance guidance for ISO 27001, SOC 2, FedRAMP,
Official Implementation of UA^{2}-Agent and other baseline algorithms of "Towards Unified Alignment Between Agents, Huma
Deterministic, resumable, tournament-based orchestrator for LLM-driven software development. Turns Claude Code or Cursor
🔥 An autonomous AI agent that runs your deep learning experiments 24/7 while you sleep. Zero-cost monitoring, Leader-Wo
It is a comprehensive resource hub compiling all LLM papers accepted at the International Conference on Learning Represe
AAAI24(Oral) ProAgent: Building Proactive Cooperative Agents with Large Language Models
AI writes code. This automates everything else · 24 plugins · 49 agents · 44 skills · for Claude Code, OpenCode, Codex,
Lightweight, auditable Python code agent (~1500 LOC) — ReAct + Planner + Reflexion + Hybrid RAG, with SWE-bench Lite e
Open-source CLI that measures AI coding-assistant impact (Claude Code, Copilot, Cursor, Aider) across Bitbucket+Jira, Gi
Local multi-agent coding orchestrator with 22 pipeline roles, TDD enforcement, SonarQube integration, and automated code
AI-Powered Predictive Maintenance & Fault Diagnosis through Model Context Protocol. An open-source framework for integra
Multi-agent sprint orchestration for Claude Code — /tickets plans, /sprint executes
Official Implementation for the Paper [AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent](https://ar
Governed local runtime for AI coding agents: task lifecycle, mandatory gates, reviews, doc-impact checks, and auditable
[ICLR'25] OpenRCA: Can Large Language Models Locate the Root Cause of Software Failures?
Hierarchical Expert Prompt for Large-Language-Models: An Approch Defeat Elite AI in TextStarCraft-II for the First Time
🌊 The leading agent orchestration platform for Claude. Deploy intelligent multi-agent swarms, coordinate autonomous wor
Engineering decisions engine that know when they're stale. Frame, compare, decide — with evidence decay and parity enfo
MASSW is a comprehensive text dataset on Multi-Aspect Summarization of Scientific Workflows. MASSW includes more than 15
A general purpose scientific writer
Tree-of-Debate converts scientific papers into LLM personas that debate their respective novelties. To emphasize structu
Lightweight Claude Code statusLine: 5h/7d rate-limit usage, reset countdowns, model + context window, prompt-cache age —
Fully autonomous AI Agents system capable of performing complex penetration testing tasks
How real engineers run Claude Code and Codex: spec-driven planning, enforced TDD, persistent memory, and quality enforce
Solana meme-coin auto-trading bot with MiniMax M2.7 LLM exit advisor, Jupiter Ultra swaps, Telegram control, and a Pepe-
Windows Agent Arena (WAA) 🪟 is a scalable OS platform for testing and benchmarking of multi-modal AI agents.
Multi agens orchestration setup for Github Copilot, Cursor, Claude Code, OpenCode, Windsurf, Codex and Antigravity.
"DeepCode: Open Agentic Coding (Paper2Code & Text2Web & Text2Backend)"
Awesome LLM Papers and repos on very comprehensive topics.
历年ICLR论文和开源项目合集,包含ICLR2021、ICLR2022、ICLR2023、ICLR2024、ICLR2025.
Multi-agent orchestration on top of Claude Code. Supervisor delegates, specialists execute, you gate. Browser-based IDE;
Most AI agents forget you the moment the tab closes. Constellation Engine gives them a hippocampus — a living star map w
Exploring Zep & Knowledge Graphs with Open-Source models as an alternative to OpenAI's Assistant Threads
DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents
Hivemind turns your traces into reusable skills across agents
Open Source Generative Process Automation (i.e. Generative RPA). AI-First Process Automation with Large ([Language (LLMs
Your agents now know when the ground shifted under their feet
Run Claude Desktop’s Cowork mode natively on Linux — no macOS or VM required
Open Android AI agent runtime for phone control, app automation, VLM screen reading, skill routing, mini apps, and Mihom
Multi-provider routing for Claude Code CLI. Use your Copilot subscription, Ollama offline, or Anthropic Direct.
Claude Multi-Agent Project Management Framework - AI-driven orchestration with LangGraph and OpenAI integration
[ICLR 2026] Meta-RL Induces Exploration in Language Agents
Open survey and evidence map for AI agent evolution, self-evolving agents, memory, skills, harnesses, benchmarks, and ag
Build mods for Claude Code: Hook any request, modify any response, /model "with-your-custom-model", intelligent model ro
PFI: Prompt Flow Integrity to Prevent Privilege Escalation in LLM Agents
A lightweight, agent-style framework for fact-checking atomic claims using iterative retrieval and verification. Reduces
HumanStudy-Bench: Towards AI Agent Design for Participant Simulation
A repo lists papers related to LLM based agent