85 packages found
Awesome LLM Papers and repos on very comprehensive topics.
💻 A curated list of papers and resources for multi-modal Graphical User Interface (GUI) agents.
A Systematic Survey of Deep Research
[Up-to-date] A curated list of resources on graph-empowered agents and agent-facilitated graph learning (Graphs Meet Age
✨✨Latest Advances on Neuro-Symbolic Learning in the era of Large Language Models
Odyssey: Empowering Minecraft Agents with Open-World Skills
Transform any arXiv papers into slides using LLMs
~95% on SimpleQA (e.g. Qwen3.6-27B on a 3090). Supports all local and cloud LLMs (llama.cpp, Ollama, Google, ...). 10+
A curated list of Generative AI tools, works, models, and references
All-in-one Web Agent framework for post-training. Start building with a few clicks!
xLAM: A Family of Large Action Models to Empower AI Agent Systems
An automated AI research-paper writer based off Google's PaperOrchestra paper's implementation through a skills - bench
🔥 An autonomous AI agent that runs your deep learning experiments 24/7 while you sleep. Zero-cost monitoring, Leader-Wo
[NeurIPS 2024 D&B] GTA: A Benchmark for General Tool Agents & [arXiv 2026] GTA-2
A simple yet versatile context engineered for scalable online data collection
⌥ AI Coding agent for the terminal — hash-anchored edits, optimized tool harness, LSP, Python, browser, subagents, and
A collection of Summoner clients and agents featuring example implementations and reusable templates
Open-source relational AI framework with identity persistence, memory, and MCP integration. Build relationship-aware AI
MR. Video: MapReduce is the Principle for Long Video Understanding
LLM Agent that leverages cheminformatics tools to provide informed responses.
A LangGraph-powered multi-agent deep research system featuring task planning, human-in-the-loop review, multi-source ret
SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution
[🏆 CHI26 Best Paper] CoBRA: Reproducible control of LLM agent behavior via classic social science experiments
🔴 VERY LARGE AI TOOL LIST! 🔴 Curated list of AI Tools - Updated 2026
[ICML2025 Oral] LLM-SRBench: A New Benchmark for Scientific Equation Discovery with Large Language Models
Framework and toolkits for building and evaluating collaborative agents that can work together with humans.
[ACL 2024 Findings] MedAgents: Large Language Models as Collaborators for Zero-shot Medical Reasoning https://arxiv.org/
Retrieval Augmented Generation for youtube videos with a BRAD agent
PFI: Prompt Flow Integrity to Prevent Privilege Escalation in LLM Agents
DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents
Tree-of-Debate converts scientific papers into LLM personas that debate their respective novelties. To emphasize structu
MASSW is a comprehensive text dataset on Multi-Aspect Summarization of Scientific Workflows. MASSW includes more than 15
Official Implementation of UA^{2}-Agent and other baseline algorithms of "Towards Unified Alignment Between Agents, Huma
[ICML 2026] This is the official implementation for paper HiPER: Hierarchical Reinforcement Learning with Explicit Credi
The official repository of "SmartAgent: Chain-of-User-Thought for Embodied Personalized Agent in Cyber World".
A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
Self-evolving agent: grows skill tree from 3.3K-line seed, achieving full system control with 6x less token consumption
RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios
The GEP-powered self-evolving engine for AI agents. Auditable evolution with Genes, Capsules, and Events. | evomap.ai
[CVPR2024 Highlight] Editable Scene Simulation for Autonomous Driving via LLM-Agent Collaboration
[ICLR 2025 Oral] This is the official repo for the paper "LLM-SR" on Scientific Equation Discovery and Symbolic Regressi
Towards Large Multimodal Models as Visual Foundation Agents
CivAgent is an LLM-based Human-like Agent acting as a Digital Player within the Strategy Game Unciv.
High-value AI skills repository for Codex, Claude Code, OpenClaw, agents, prompts, and automation workflows. 高价值的 AI 技能
MobileUse: an open-source mobile GUI agent for Android phone automation, AndroidWorld/AndroidLab evaluation, hierarchica
[NAACL2025] LiteWebAgent: The Open-Source Suite for VLM-Based Web-Agent Applications
DialOp: Decision-oriented dialogue environments for collaborative language agents
Code and Data for "MIRAI: Evaluating LLM Agents for Event Forecasting"
AlgoTune is a NeurIPS 2025 benchmark made up of 154 math, physics, and computer science problems. The goal is write code
Official Implementation for the Paper [AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent](https://ar