825 packages found
Source-backed GPT-5.6 use cases for coding, agents, creative work, integrations, benchmarks, and practical limits.
Claude Code skill for benchmark research. Survey papers to find datasets, metrics, and evaluation protocols used in a re
历年ICLR论文和开源项目合集,包含ICLR2021、ICLR2022、ICLR2023、ICLR2024、ICLR2025.
AINL helps turn AI from "a smart conversation" into "a structured worker." It is designed for teams building AI workflo
SIGNAL — Agent Skills: terse structured output (tiers, templates, checkpoints), git workflow skills (commit, push, PR, r
Open survey and evidence map for AI agent evolution, self-evolving agents, memory, skills, harnesses, benchmarks, and ag
A repo lists papers related to LLM based agent
It is a comprehensive resource hub compiling all LLM papers accepted at the International Conference on Learning Represe
Official companion repository for our survey "A Survey of the OpenClaw Ecosystem: From Platform Extensibility to Constra
🌊 The leading agent orchestration platform for Claude. Deploy intelligent multi-agent swarms, coordinate autonomous wor
A curated list of tools, papers, and datasets for applying AI to cybersecurity tasks. This list primarily focuses on mod
Persistent memory for Claude Code & Codex CLI. Auto-extracted knowledge graph, multi-representation embeddings, 3D WebGL
Context optimization for AI agents — 97 domain knowledge graphs, MCP-native traversal. 4× F1 of RAG, 11× fewer tokens pe
Comprehensive MCP server exposing dozens of capabilities to AI agents: multi-provider LLM delegation, browser automation
A framework for designing App Store-compliant subscription paywalls. 4 layers: AI skill (Claude/GPT/Cursor) + knowledge
💻 A curated list of papers and resources for multi-modal Graphical User Interface (GUI) agents.
Save 30% token costs when using Claude Code, Codex, OpenCode for free - with open source, local semantic search. Works f
A Systematic Analysis and Discussion of Claude Code for Designing Today's and Future AI Agent Systems
Lightweight, auditable Python code agent (~1500 LOC) — ReAct + Planner + Reflexion + Hybrid RAG, with SWE-bench Lite e
Drawdown-first portfolio tool with a read-only MCP addon for Claude — a deterministic core computes every number; the AI
A curated list of Generative AI tools, works, models, and references
Hook-based token compressor for 5 AI CLI hosts (Claude Code, Copilot CLI, OpenCode, Gemini CLI, Codex CLI). Up to 95% ba
Windows Agent Arena (WAA) 🪟 is a scalable OS platform for testing and benchmarking of multi-modal AI agents.
Universal documentation knowledge-graph MCP server with hybrid full-text + vector search.
A map of what AI tokens actually cost, and where they're wasted vs. well spent. Tools, research, practices, and copy-pas
Claude Code execution playbook with 3 pilot modes: cost-first (Haiku), quality-first (Sonnet), ceiling-elevation (Opus).
非线智能 NoneLinear - ReLE评测:中文AI大模型能力评测(持续更新):目前已囊括374个大模型,覆盖chatgpt、gpt-5.4、谷歌gemini-3.1-pro、Claude-4.6、文心ERNIE-X1.1、ERNIE
The most complete SEO + GEO + AEO skill for Claude Code. 20 phases, 0 to 100 of 100. Benchmarked against 61 top SaaS and
~95% on SimpleQA (e.g. Qwen3.6-27B on a 3090). Supports all local and cloud LLMs (llama.cpp, Ollama, Google, ...). 10+
🛰️ A CLI tool for tracking token usage from OpenCode, Claude Code, 🦞OpenClaw (Clawdbot/Moltbot), Pi, Codex, Gemini, Cu
[Up-to-date] A curated list of resources on graph-empowered agents and agent-facilitated graph learning (Graphs Meet Age
[NeurIPS 2024 D&B] GTA: A Benchmark for General Tool Agents & [arXiv 2026] GTA-2
A Systematic Survey of Deep Research
Local-first MCP memory layer for AI coding assistants; persistent cross-session memory, stored entirely on your machine
Persistent AI memory for Claude Code, OpenClaw, and any MCP-compatible agent. BM25F + vector hybrid, governance-aware, l
Give your AI agent eyes and hands on your real Chrome browser — your tabs, your logins, your page state. 42 commands, ze
HealthFlow: Automating electronic health record analysis via a strategically self-evolving multi-agent framework
Benchmarked agent execution runtime for Python. Sub-10ms cold starts, real-time streaming, time-travel debugging, and se
AI Agent plugin for Autoresearch with AI (Claude, OpenClaw, etc) to improve anything!
Vault-native, accountable memory for Claude Code and MCP clients. Markdown is the source of truth, no LLM on the Stop pa
Fast and Accurate Code Search for Agents. Uses ~98% fewer tokens than grep+read
Comprehensive paid advertising audit & optimization skill for Claude Code. 250+ checks across Google, Meta, YouTube, Lin
A LangGraph-powered multi-agent deep research system featuring task planning, human-in-the-loop review, multi-source ret
LLM agents that write machine-checked cryptographic proofs in EasyCrypt (arXiv:2607.02847)
[ICML2025 Oral] LLM-SRBench: A New Benchmark for Scientific Equation Discovery with Large Language Models
A comprehensive best-practices wiki for Claude Code - setup, CLAUDE.md templates, workflows, multi-agent patterns, and c
The Mind Palace for AI Agents - HIPAA-hardened Cognitive Architecture with on-device LLM (prism-coder:7b), Hebbian learn
Local-first, auditable memory layer for AI apps and coding agents — Codex, Claude Code, MCP, HTTP, TypeScript, and Pytho