Agent Instructs Large Language Models to be General Zero-Shot Reasoners (arXiv 2023) [paper] [code]
OKR-Agent: An Object and Key Results Driven Agent System with Hierarchical Self-Collaboration and Self-Evaluation (arXiv 2023) [paper]
MAGDi: Structured Distillation of Multi-Agent Interaction Graphs Improves Reasoning in Smaller Language Models (arXiv 2024) [paper] [code]
Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration (NAACL 2024) [paper] [code]
MechAgents: Large language model multi-agent collaborations can solve mechanics problems, generate new data, and integrate knowledge (EML 2024) [paper] [code]
Exploring Collaboration Mechanisms for LLM Agents: A Social Psychology View (arXiv 2023) [paper] [code]
SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents (ICLR 2024) [paper] [code]
ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate (ICLR 2024) [paper] [code]
Dynamic LLM-Agent Network: An LLM-agent Collaboration Framework with Agent Team Optimization (arXiv 2023) [paper] [code]
Playing repeated games with Large Language Models (arXiv 2023) [paper]
Communicative Agents for Software Development (ICLR 2024) [paper] [code]
AutoAgents: A Framework for Automatic Agent Generation (arXiv 2024) [paper] [code]
Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents (arXiv 2024) [paper]
Player-Driven Emergence in LLM-Driven Game Narrative (arXiv 2024) [paper]
Enhancing the Efficiency and Accuracy of Underlying Asset Reviews in Structured Finance: The Application of Multi-agent Framework (arXiv 2024) [paper] [code]
MapCoder: Multi-Agent Code Generation for Competitive Problem Solving (arXiv 2024) [paper] [code]
Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration (NAACL 2024) [paper] [code]
AgentCF: Collaborative Learning with Autonomous Language Agents for Recommender Systems (WWW 2024) [paper] [code](不存在)
On Generative Agents in Recommendation (SIGIR 2024) [paper] [code]
评测
LLM-Coordination: Evaluating and Analyzing Multi-agent Coordination Abilities in Large Language Models (arXiv 2024) [paper] [code]
MAgIC: Investigation of Large Language Model Powered Multi-Agent in Cognition, Adaptability, Rationality and Collaboration (arXiv 2023) [paper] [code]
LLMArena: Assessing Capabilities of Large Language Models in Dynamic Multi-Agent Environments (arXiv 2024) [paper] [code]
SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents (ICLR 2024) [paper] [code]
ALYMPICS: LLM Agents Meet Game Theory -- Exploring Strategic Decision-Making with AI Agents (arXiv 2024) [paper] [code]
Large Language Models as Rational Players in Competitive Economics Game (arXiv 2024) [paper]
AgentBench: Evaluating LLMs as Agents (ICLR 2024) [paper] [code]
MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework (ICLR 2024) [paper] [code]
LARGE LANGUAGE MODEL EVALUATION VIA MULTI AI AGENTS: PRELIMINARY RESULTS (ICLR 2024) [paper]
MobileAgentBench: An Efficient and User-Friendly Benchmark for Mobile LLM Agents (arXiv 2024) [paper] [code]
综述
Understanding the planning of LLM agents: A survey (arXiv 2024) [paper]
A Survey on Large Language Model based Autonomous Agents (Front. Comput. Sci. 2024) [paper] [code]
AGENT AI: SURVEYING THE HORIZONS OF MULTIMODAL INTERACTION (arXiv 2024) [paper]
The Rise and Potential of Large Language Model Based Agents: A Survey (ICML 2024) [paper] [code]
Large Language Models Empowered Agent-based Modeling and Simulation: A Survey and Perspectives (arXiv 2023) [paper]
LLM-based Multi-Agent Reinforcement Learning: Current and Future Directions (arXiv 2024) [paper]
Large Language Model based Multi-Agents: A Survey of Progress and Challenges (arXiv 2024) [paper] [code]
Agent Instructs Large Language Models to be General Zero-Shot Reasoners (arXiv 2023) [paper] [code]
OKR-Agent: An Object and Key Results Driven Agent System with Hierarchical Self-Collaboration and Self-Evaluation (arXiv 2023) [paper]
MAGDi: Structured Distillation of Multi-Agent Interaction Graphs Improves Reasoning in Smaller Language Models (arXiv 2024) [paper] [code]
Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration (NAACL 2024) [paper] [code]
MechAgents: Large language model multi-agent collaborations can solve mechanics problems, generate new data, and integrate knowledge (EML 2024) [paper] [code]
Exploring Collaboration Mechanisms for LLM Agents: A Social Psychology View (arXiv 2023) [paper] [code]
SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents (ICLR 2024) [paper] [code]
ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate (ICLR 2024) [paper] [code]
Dynamic LLM-Agent Network: An LLM-agent Collaboration Framework with Agent Team Optimization (arXiv 2023) [paper] [code]
Playing repeated games with Large Language Models (arXiv 2023) [paper]
Communicative Agents for Software Development (ICLR 2024) [paper] [code]
AutoAgents: A Framework for Automatic Agent Generation (arXiv 2024) [paper] [code]
Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents (arXiv 2024) [paper]
Player-Driven Emergence in LLM-Driven Game Narrative (arXiv 2024) [paper]
Enhancing the Efficiency and Accuracy of Underlying Asset Reviews in Structured Finance: The Application of Multi-agent Framework (arXiv 2024) [paper] [code]
MapCoder: Multi-Agent Code Generation for Competitive Problem Solving (arXiv 2024) [paper] [code]
Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration (NAACL 2024) [paper] [code]
AgentCF: Collaborative Learning with Autonomous Language Agents for Recommender Systems (WWW 2024) [paper] [code](不存在)
On Generative Agents in Recommendation (SIGIR 2024) [paper] [code]
评测
LLM-Coordination: Evaluating and Analyzing Multi-agent Coordination Abilities in Large Language Models (arXiv 2024) [paper] [code]
MAgIC: Investigation of Large Language Model Powered Multi-Agent in Cognition, Adaptability, Rationality and Collaboration (arXiv 2023) [paper] [code]
LLMArena: Assessing Capabilities of Large Language Models in Dynamic Multi-Agent Environments (arXiv 2024) [paper] [code]
SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents (ICLR 2024) [paper] [code]
ALYMPICS: LLM Agents Meet Game Theory -- Exploring Strategic Decision-Making with AI Agents (arXiv 2024) [paper] [code]
Large Language Models as Rational Players in Competitive Economics Game (arXiv 2024) [paper]
AgentBench: Evaluating LLMs as Agents (ICLR 2024) [paper] [code]
MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework (ICLR 2024) [paper] [code]
LARGE LANGUAGE MODEL EVALUATION VIA MULTI AI AGENTS: PRELIMINARY RESULTS (ICLR 2024) [paper]
MobileAgentBench: An Efficient and User-Friendly Benchmark for Mobile LLM Agents (arXiv 2024) [paper] [code]
综述
Understanding the planning of LLM agents: A survey (arXiv 2024) [paper]
A Survey on Large Language Model based Autonomous Agents (Front. Comput. Sci. 2024) [paper] [code]
AGENT AI: SURVEYING THE HORIZONS OF MULTIMODAL INTERACTION (arXiv 2024) [paper]
The Rise and Potential of Large Language Model Based Agents: A Survey (ICML 2024) [paper] [code]
Large Language Models Empowered Agent-based Modeling and Simulation: A Survey and Perspectives (arXiv 2023) [paper]
LLM-based Multi-Agent Reinforcement Learning: Current and Future Directions (arXiv 2024) [paper]
Large Language Model based Multi-Agents: A Survey of Progress and Challenges (arXiv 2024) [paper] [code]