50 packages found
An unofficial C#/.NET SDK for accessing the Anthropic Claude API. This package is not affiliated with, endorsed by, or s
A research agent system deeply rooted in your Zotero library.
A general purpose scientific writer
🔴 VERY LARGE AI TOOL LIST! 🔴 Curated list of AI Tools - Updated 2026
A curated list of Generative AI tools, works, models, and references
ML-Dev-Bench is a benchmark for evaluating AI agents against various ML development tasks.
RepairAgent is an autonomous LLM-based agent for software repair.
Code and Data for "MIRAI: Evaluating LLM Agents for Event Forecasting"
A Systematic Survey of Deep Research
Towards Large Multimodal Models as Visual Foundation Agents
A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
ICML 2026 · Plug-and-play long-term memory for LLM agents
A simple yet versatile context engineered for scalable online data collection
[Up-to-date] A curated list of resources on graph-empowered agents and agent-facilitated graph learning (Graphs Meet Age
AgentHER: Hindsight Experience Replay for LLM Agents
A powerful AI assistant integrated into KOReader.
Official Repo for ICML 2024 paper "Executable Code Actions Elicit Better LLM Agents" by Xingyao Wang, Yangyi Chen, Lifan
[ICML 2024] LLMCompiler: An LLM Compiler for Parallel Function Calling
[CVPR2024 Highlight] Editable Scene Simulation for Autonomous Driving via LLM-Agent Collaboration
[ICLR 2025 Oral] This is the official repo for the paper "LLM-SR" on Scientific Equation Discovery and Symbolic Regressi
SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution
[NeurIPS 2024 D&B] GTA: A Benchmark for General Tool Agents & [arXiv 2026] GTA-2
[ICML2025 Oral] LLM-SRBench: A New Benchmark for Scientific Equation Discovery with Large Language Models
[🏆 CHI26 Best Paper] CoBRA: Reproducible control of LLM agent behavior via classic social science experiments
OrcaLoca: An LLM Agent Framework for Software Issue Localization [ICML 25]
[ICLR2026] The official repository for the CodeGym project: "Generalizable End-to-End Tool-Use RL with Synthetic CodeGym
LLM Agent that leverages cheminformatics tools to provide informed responses.
A curated list of awesome things related to Anthropic Claude
Yunjue Agent: A Fully Reproducible, Zero-Start In-Situ Self-Evolving Agent System for Open-Ended Tasks
Ask the oracle when you're stuck. Invoke GPT-5 Pro with a custom context and files.
[ICLR'25] OpenRCA: Can Large Language Models Locate the Root Cause of Software Failures?
Text2Sim MCP Server is a conversational simulation engine that transforms natural language into working simulation model
irresponsible innovation. Try now at https://chat.dev/
"DeepCode: Open Agentic Coding (Paper2Code & Text2Web & Text2Backend)"
[CVPR 2025 🔥]A Large Multimodal Model for Pixel-Level Visual Grounding in Videos
The benchmark tasks and evaluation harness for "PhysicianBench: Evaluating LLM Agents in Real-World EHR Environments".
MR. Video: MapReduce is the Principle for Long Video Understanding
DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents
The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
A lightweight nanobot/OpenClaw variant in pure Julia
Asynchronous LLM Agent playing games of Mafia against human players
Agentic Theorem Prover for Rocq for Program Verification
[ICML 2026] This is the official implementation for paper HiPER: Hierarchical Reinforcement Learning with Explicit Credi
MCP Server to manage a Wordpress CMS system.
The official repository of "SmartAgent: Chain-of-User-Thought for Embodied Personalized Agent in Cyber World".
Hypernetworks that update LLMs to remember factual information
[CVPR 2024 🔥] Grounding Large Multimodal Model (GLaMM), the first-of-its-kind model capable of generating natural langu
Odyssey: Empowering Minecraft Agents with Open-World Skills