42 packages found
Awesome LLM Papers and repos on very comprehensive topics.
A Systematic Survey of Deep Research
[Up-to-date] A curated list of resources on graph-empowered agents and agent-facilitated graph learning (Graphs Meet Age
Odyssey: Empowering Minecraft Agents with Open-World Skills
Transform any arXiv papers into slides using LLMs
xLAM: A Family of Large Action Models to Empower AI Agent Systems
A curated list of Generative AI tools, works, models, and references
[NeurIPS 2024 D&B] GTA: A Benchmark for General Tool Agents & [arXiv 2026] GTA-2
A simple yet versatile context engineered for scalable online data collection
MR. Video: MapReduce is the Principle for Long Video Understanding
LLM Agent that leverages cheminformatics tools to provide informed responses.
🔴 VERY LARGE AI TOOL LIST! 🔴 Curated list of AI Tools - Updated 2026
[🏆 CHI26 Best Paper] CoBRA: Reproducible control of LLM agent behavior via classic social science experiments
[ICML2025 Oral] LLM-SRBench: A New Benchmark for Scientific Equation Discovery with Large Language Models
SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution
[ACL 2024 Findings] MedAgents: Large Language Models as Collaborators for Zero-shot Medical Reasoning https://arxiv.org/
Retrieval Augmented Generation for youtube videos with a BRAD agent
DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents
Official Implementation of UA^{2}-Agent and other baseline algorithms of "Towards Unified Alignment Between Agents, Huma
[ICML 2026] This is the official implementation for paper HiPER: Hierarchical Reinforcement Learning with Explicit Credi
The official repository of "SmartAgent: Chain-of-User-Thought for Embodied Personalized Agent in Cyber World".
A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
[CVPR2024 Highlight] Editable Scene Simulation for Autonomous Driving via LLM-Agent Collaboration
[ICLR 2025 Oral] This is the official repo for the paper "LLM-SR" on Scientific Equation Discovery and Symbolic Regressi
Towards Large Multimodal Models as Visual Foundation Agents
Code and Data for "MIRAI: Evaluating LLM Agents for Event Forecasting"
[ICLR2026] The official repository for the CodeGym project: "Generalizable End-to-End Tool-Use RL with Synthetic CodeGym
Awesome papers involving LLMs in Social Science.
Yunjue Agent: A Fully Reproducible, Zero-Start In-Situ Self-Evolving Agent System for Open-Ended Tasks
ML-Dev-Bench is a benchmark for evaluating AI agents against various ML development tasks.
[CVPR 2025 🔥]A Large Multimodal Model for Pixel-Level Visual Grounding in Videos
Code repo for the paper: Attacking Vision-Language Computer Agents via Pop-ups
The benchmark tasks and evaluation harness for "PhysicianBench: Evaluating LLM Agents in Real-World EHR Environments".
[ICML 2024] LLMCompiler: An LLM Compiler for Parallel Function Calling
[CVPR 2024 🔥] Grounding Large Multimodal Model (GLaMM), the first-of-its-kind model capable of generating natural langu
[NeurIPS 2024] Official implementation for "AgentPoison: Red-teaming LLM Agents via Memory or Knowledge Base Backdoor Po
The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
"DeepCode: Open Agentic Coding (Paper2Code & Text2Web & Text2Backend)"
ICML 2026 · Plug-and-play long-term memory for LLM agents
ALICE and its prior work, Voice2Action: Language Models as Agent for Efficient Real-Time Interaction in Virtual Reality
OrcaLoca: An LLM Agent Framework for Software Issue Localization [ICML 25]