lucianma05-create/Awesome-Social-AI

A curated collection of paper summaries for Social AI research.

Python

20

118 commits

updated Sep 16, 2026

See the code

README

🌐 Awesome Social AI

Awesome License Papers Summaries Views

我们精心收集整理社交人工智能(Social AI)领域的研究论文,并持续更新论文的中文摘要,从而支持快速了解该领域的代表性工作。仓库将不断更新,追踪社交 AI 前沿。欢迎 Follow 和 Star!⭐

🎯 范畴与结构

Social AI 是研究与构建具备社会智能的 AI 系统的领域。社会智能指智能体在社交情境中的三组能力:社交理解(感知情绪、意图、信念、关系、规范与多方动态)、社交推理(心智建模、归因与策略规划)、社交行动(说服、谈判、共情支持、建立信任)。

本仓库按四层结构组织:

不收录:计算社会科学 / "AI for social science" 类工作。方法层的通用强化学习与对齐技术,以"支撑社交智能体的基础方法"身份收录;与社交无关、亦不服务于社交智能体的技术不收录。

参考论文:

  1. Towards Social AI: A Survey on Understanding Social Interactions
  2. Advancing Social Intelligence in AI Agents: Technical Challenges and Open Questions

📑 目录

📚 论文列表

🗣️ Persuasion & Negotiation · 41 篇
年份会议/期刊论文链接摘要代码
2008PERSUASIVEA Systematic Framework for Designing and Evaluating Persuasive Systems查看摘要-
2008-Measure Of Belief Change as an Evaluation of Persuasion查看摘要-
2014COLINGReinforcement Learning of Cooperative Persuasive Dialogue Policies using Framing查看摘要-
2017HCIPersuasive Argumentation and Emotions- An Empirical Evaluation with Users查看摘要-
2023ArgCompStrategic argumentation dialogues for persuasion- Framework and experiments based on modelling the beliefs and concerns of the persuadee查看摘要-
2023arXivImproving Language Model Negotiation with Self-Play and In-Context Learning from AI Feedback查看摘要-
2024ICLRPlug-and-Play Policy Planner for LLM-Powered Dialogue Agents查看摘要代码
2024ACLHow Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs查看-代码
2024ACLMeasuring Bargaining Abilities of LLMs: A Benchmark and A Buyer-Enhancement Method查看-代码
2024EMNLPNegotiationToM: A Benchmark for Stress-testing Machine Theory of Mind on Negotiation Surrounding查看-代码
2024ICMLDebating with More Persuasive LLMs Leads to More Truthful Answers查看-代码
2024NeurIPSCooperation, Competition, and Maliciousness: LLM-Stakeholders Interactive Negotiation查看-代码
2025ScienceDurably reducing conspiracy beliefs through dialogues with AI查看摘要-
2025ScienceThe Levers of Political Persuasion查看摘要代码
2025ACLBattling against Tough Resister- Strategy Planning with Adversarial Game for Non-collaborative Dialogues查看摘要-
2025TACLHuman Choice Prediction in Language-based Persuasion Games- Simulation-based Off-Policy Evaluation查看摘要代码
2025EMNLPEnhancing LLM-Based Persuasion Simulations with Cultural and Speaker-Specific Information查看摘要代码
2025EMNLPEnhancing Persuasive Dialogue Agents by Synthesizing Cross-Disciplinary Communication Strategies查看摘要-
2025NAACLTeaching models to balance resisting and accepting persuasion查看-代码
2025arXivLLM Can be a Dangerous Persuader- Empirical Study of Persuasion Safety in Large Language Models查看摘要代码
2025EMNLPPRINCIPLES- Synthetic Strategy Memory for Proactive Dialogue Agents查看摘要代码
2025NeurIPSPersuade Me if You Can- A Framework for Evaluating Persuasion Effectiveness and Susceptibility Among Large Language Models查看摘要代码
2025AAAISimulation-free hierarchical latent policy planning for proactive dialogues查看摘要-
2025arXivToMAP: Training Opponent-Aware LLM Persuaders with Theory of Mind查看摘要代码
2025ACLEPO: Explicit Policy Optimization for Strategic Reasoning in LLMs via RL查看摘要-
2025arXivDisagreements in Reasoning: How a Model's Thinking Process Dictates Persuasion in Multi-Agent Systems查看摘要-
2025arXivEvoEmo: Evolved Emotional Policies for Adversarial LLM Agents in Multi-Turn Price Negotiation查看摘要-
2025arXivFrom Simulation to Strategy: Automating Personalized Interaction Planning for Conversational Agents查看摘要-
2025arXivPersuasion Should be Double-Blind: A Multi-Domain Dialogue Dataset With Faithfulness Based on Causal Theory of Mind查看摘要-
2025arXivVerbalized Bayesian Persuasion查看摘要-
2025EMNLPPersuasion-Dynamics in LLMs: Investigating Robustness and Adaptability in Knowledge and Safety with DuET-PD查看摘要-
2025EMNLPProfiling LLM Copyright Infringement Risks under Adversarial Persuasive Prompting查看摘要代码
2025CSURPersuasive Conversational Agents for Environmental Sustainability: A Survey查看摘要-
2025ACLASTRO: Automatic Strategy Optimization For Non-Cooperative Dialogues查看摘要代码
2026arXivOne Model, All Roles- Multi-Turn, Multi-Agent Self-Play Reinforcement Learning for Conversational Social Intelligence查看--
2026arXivPersonality-Aware Reinforcement Learning for Persuasive Dialogue with LLM-Driven Simulation查看摘要-
2026CSURA Comprehensive Survey of Computational Persuasion查看摘要代码
2026ICLRRebuttalAgent: Strategic Persuasion in Academic Rebuttal via Theory of Mind查看摘要代码
2026ICLRTowards Strategic Persuasion with Language Models查看摘要-
2026arXivMETRO: Towards Strategy Induction from Expert Dialogue Transcripts for Non-collaborative Dialogues查看摘要代码
2026ICLRStrategic Planning and Rationalizing on Trees Make LLMs Better Debaters查看-代码

❤️‍🩹 Empathy & Emotional Support · 19 篇
年份会议/期刊论文链接摘要代码
2021ACLTowards Emotional Support Dialog Systems查看-代码
2023arXivCharacterChat- Learning towards Conversational AI with Personalized Social Support查看摘要代码
2023EMNLPSoulChat- Improving LLMs’ Empathy, Listening, and Comfort Abilities through Fine-tuning with Multi-turn Empathy Conversations查看摘要代码
2023ACLTransESC: Smoothing Emotional Support Conversation via Turn-Level State Transition查看-代码
2024ACLEmoBench- Evaluating the Emotional Intelligence of Large Language Models查看摘要代码
2024ACLCan Large Language Models be Good Emotional Supporter? Mitigating Preference Bias on Emotional Support Conversation查看-代码
2024ACLESCoT: Towards Interpretable Emotional Support Dialogue Systems查看-代码
2025arXivEcho-N1- Affective RL Frontier查看摘要-
2025ACLPsyDT- Using LLMs to Construct the Digital Twin of Psychological Counselor with Personalized Counseling Style for Psychological Counseling查看摘要代码
2025ACLPsyDial- A Large-scale Long-term Conversational Dataset for Mental Health Support查看摘要代码
2025arXivReinforcement Learning with Verifiable Emotion Rewards for Empathetic Agents查看摘要代码
2025arXivSAGE- Steering and Refining Dialog Generation with State-Action Augmentation查看摘要代码
2025ACLBeyond Verbal Cues: Emotional Contagion Graph Network for Causal Emotion Entailment查看-代码
2025EMNLPChain of Strategy Optimization Makes Large Language Models Better Emotional Supporter查看摘要代码
2025CHICustomizing Emotional Support: How Do Individuals Construct and Interact With LLM-Powered Chatbots查看--
2025NAACLEmoDynamiX: Emotional Support Dialogue Strategy Prediction by Modelling MiXed Emotions and Discourse Dynamics查看-代码
2026arXivAffective Flow Language Model for Emotional Support Conversation查看摘要代码
2026arXivEMPA: Evaluating Persona-Aligned Empathy as a Process查看摘要代码
2026ACLYou Never Know a Person, You Only Know Their Defenses- Detecting Levels of Psychological Defense Mechanisms in Supportive Conversations查看--

🛍️ Conversational Recommendation · 24 篇
年份会议/期刊论文链接摘要代码
2020ACLTowards Conversational Recommendation over Multi-Type Dialogs查看摘要代码
2021ACLRevCore- Review-augmented Conversational Recommendation查看摘要代码
2023KDDImproving conversational recommendation systems via counterfactual data simulation查看摘要代码
2023EMNLPRethinking the Evaluation for Conversational Recommendation in the Era of Large Language Models查看摘要代码
2023CIKMLarge Language Models as Zero-Shot Conversational Recommenders查看-代码
2024EMNLPBeyond Persuasion: Towards Conversational Recommender System with Credible Explanations查看摘要代码
2024WWWHow Reliable is Your Simulator- Analysis on the Limitations of Current LLM-based User Simulators for Conversational Recommendation查看摘要代码
2024ACLLLM-REDIAL: A Large-Scale Dataset for Conversational Recommender Systems Created from User Behaviors with LLMs查看摘要代码
2024arXivReindex-Then-Adapt- Improving Large Language Models for Conversational Recommendation查看摘要-
2024ACLPearl: A Review-driven Persona-Knowledge Grounded Conversational Recommendation Dataset查看-代码
2024EMNLPMitigating Matthew Effect: Multi-Hypergraph Boosted Multi-Interest Self-Supervised Learning for Conversational Recommendation查看-代码
2024SIGIRBroadening the View: Demonstration-augmented Prompt Learning for Conversational Recommendation查看--
2025TKDEA Causal-Based Attribute Selection Strategy for Conversational Recommender Systems查看摘要-
2025WWWBridging Conversational and Collaborative Signals for Conversational Recommendation查看摘要-
2025WWWCollaborative Retrieval for Large Language Model-based Conversational Recommender Systems查看摘要代码
2025WWWTowards Efficient Conversational Recommendations- Expected Value of Information Meets Bandit Learning查看摘要-
2025arXivA Framework for Generating Conversational Recommendation Datasets from Behavioral Interactions查看摘要-
2025EMNLPLLM-based Conversational Recommendation Agents with Collaborative Verbalized Experience查看摘要代码
2025EMNLPTowards Personalized Conversational Sales Agents查看摘要-
2025NAACLEmpowering Retrieval-based Conversational Recommendation with Contrasting User Preferences查看-代码
2026WWWNot All Information Brings Benefits- Personalization-Driven Agent Debate for Conversational Recommendation查看摘要-
2026WWWOptimizing Multi-Turn Interactive Recommendation Agents via Generative Intrinsic Motivation查看摘要代码
2026arXivUser Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation查看摘要-
2026ICLRRank-GRPO: Training LLM-based Conversational Recommender Systems with Reinforcement Learning查看-代码

🤝 Cooperation & Collaboration · 11 篇
年份会议/期刊论文链接摘要代码
2022ScienceHuman-level play in the game of Diplomacy by combining language models with strategic reasoning查看--
2022CHIAI Chains: Transparent and Controllable Human-AI Interaction by Chaining Large Language Model Prompts查看--
2024ICLRBuilding Cooperative Embodied Agents Modularly with Large Language Models查看--
2024TACLDecision-Oriented Dialogue for Human-AI Collaboration查看--
2024ICMLShould we be going MAD- A Look at Multi-Agent Debate Strategies for LLMs查看-代码
2024ACLYour Co-Workers Matter- Evaluating Collaborative Capabilities of Language Models in Blocks World查看--
2024ACLExploring Collaboration Mechanisms for LLM Agents: A Social Psychology View查看-代码
2024ICLRMetaGPT: Meta Programming for A Multi-Agent Collaborative Framework查看-代码
2024NeurIPSCooperate or Collapse: Emergence of Sustainable Cooperation in a Society of LLM Agents查看-代码
2024ScienceAI can help humans find common ground in democratic deliberation查看-代码
2025Nature Human BehaviourPlaying repeated games with large language models查看-代码

💭 Theory of Mind · 19 篇
年份会议/期刊论文链接摘要代码
2011CogSciBayesian Theory of Mind- Modeling Joint Belief-Desire Attribution查看摘要-
2019COBSTheory of Mind as Inverse Reinforcement Learning查看摘要-
2024ACLThink Twice: Perspective-Taking Improves Large Language Models' Theory-of-Mind Capabilities查看-代码
2025ACLMachine Theory of Mind Needs Machine Validation查看摘要-
2025ACLTheory of Mind in Large Language Models- Assessment and Enhancement查看摘要-
2025arXivMINDGAMES: Do Large Language Models Have a Planning Theory of Mind?查看摘要代码
2025arXivModeling the Mental World for Embodied AI- A Comprehensive Review查看摘要-
2025arXivToM-agent: Large Language Models as Theory of Mind Aware Generative Agents with Counterfactual Reflection查看摘要-
2025arXivToM-RL: Reinforcement Learning Unlocks Theory of Mind in Small LLMs查看摘要代码
2025ICMLOvercoming Multi-step Complexity in Multimodal Theory-of-Mind Reasoning- A Scalable Bayesian Planner查看摘要-
2025NeurIPSAutoToM- Scaling Model-based Mental Inference via Automated Agent Modeling查看摘要-
2025NeurIPSMetaMind- Modeling Human Social Thoughts with Metacognitive Multi-Agent Systems查看摘要代码
2026AAAIReality vs Counterfactual- Multi-World Contrastive Reinforcement Learning for Enhancing MLLM’s Theory of Mind in Egocentric Videos查看摘要-
2026arXivInfusing Theory of Mind into Socially Intelligent LLM Agents查看摘要代码
2026arXivMetaMind- General and Cognitive World Models in Multi-Agent Systems by Meta-Theory of Mind查看摘要-
2026arXivMindClaw- Closed-Loop Embodied Mental-State Reasoning for Precision Intervention查看摘要-
2026arXivUserHarness- Harnessing User Minds for Stronger Agent Theory-of-Mind查看摘要-
2026CVPRVideo-Only ToM- Enhancing Theory of Mind in Multimodal Large Language Models查看摘要代码
2026EACLLet's Put Ourselves in Sally's Shoes: Shoes of Others Prefilling Improves Theory of Mind in LLMs查看摘要-

🎭 Emotion Understanding · 10 篇
年份会议/期刊论文链接摘要代码
2019AAAIDialogueRNN- An Attentive RNN for Emotion Detection in Conversations查看--
2019ACLMELD- A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations查看--
2021AAAICOSMIC- COmmonSense knowledge for eMotion Identification in Conversations查看--
2023arXivInstructERC- Reforming Emotion Recognition in Conversation with Multi-task Retrieval-Augmented Large Language Models查看--
2023AAAIKnowledge-Bridged Causal Interaction Network for Causal Emotion Entailment查看-代码
2024KDDEmoLLMs: A Series of Emotional Large Language Models and Annotation Tools for Comprehensive Affective Analysis查看-代码
2024NAACLTelME: Teacher-leading Multimodal Fusion Network for Emotion Recognition in Conversation查看-代码
2024NeurIPSEmotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning查看-代码
2025arXivDo LLMs Feel- Teaching Emotion Recognition with Prompts, Retrieval, and Curriculum Learning查看--
2025ACLCoE: A Clue of Emotion Framework for Emotion Recognition in Conversations查看--

⚖️ Social Norms & Morality · 12 篇
年份会议/期刊论文链接摘要代码
2020EMNLPSocial Chemistry 101- Learning to Reason about Social and Moral Norms查看--
2021arXivDelphi- Towards Machine Ethics and Norms查看--
2021ICLRAligning AI With Shared Human Values查看-代码
2022ACLThe Moral Integrity Corpus- A Benchmark for Ethical Dialogue Systems查看-代码
2022NeurIPSWhen to Make Exceptions: Exploring Language Models as Accounts of Human Moral Judgment查看-代码
2023ICMLDo the Rewards Justify the Means- Measuring Trade-Offs Between Rewards and Ethical Behavior in the MACHIAVELLI Benchmark查看--
2023ACLNormBank- A Knowledge Bank of Situational Social Norms查看-代码
2023NeurIPSEvaluating the Moral Beliefs Encoded in LLMs查看-代码
2024arXivCultureBank- An Online Community-Driven Knowledge Base Towards Culturally Aware Language Technologies查看-代码
2024AAAIValue Kaleidoscope: Engaging AI with Pluralistic Human Values, Rights, and Duties查看-代码
2025NAACLNormAd: A Framework for Measuring the Cultural Adaptability of Large Language Models查看-代码
2025Science AdvancesEmergent social conventions and collective bias in LLM populations查看-代码

💾 Agent Memory · 11 篇
年份会议/期刊论文链接摘要代码
2023UISTGenerative Agents- Interactive Simulacra of Human Behavior查看摘要代码
2023arXivMemoryBank- Enhancing Large Language Models with Long-Term Memory查看-代码
2023NeurIPSReflexion: Language Agents with Verbal Reinforcement Learning查看-代码
2024AAAIExpeL: LLM Agents Are Experiential Learners查看-代码
2024NeurIPSHippoRAG: Neurobiologically Inspired Long-Term Memory for Large Language Models查看-代码
2025arXivA-Mem: Agentic Memory for LLM Agents查看摘要代码
2025arXivEvo-Memory- Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory查看摘要代码
2025arXivRemember Me, Refine Me- A Dynamic Procedural Memory Framework for Experience-Driven Agent Evolution查看摘要代码
2025ICLRHuman-inspired Episodic Memory for Infinite Context LLMs查看-代码
2026ICLRMemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent查看摘要-
2026ICLRReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory查看摘要代码

🤖 Reinforcement Learning & Alignment · 18 篇
年份会议/期刊论文链接摘要代码
2022NeurIPSThe Surprising Effectiveness of PPO in Cooperative Multi-Agent Games查看-代码
2022NeurIPSTraining language models to follow instructions with human feedback查看--
2023NeurIPSDirect Preference Optimization- Your Language Model is Secretly a Reward Model查看摘要-
2024ICMLRLAIF vs. RLHF- Scaling Reinforcement Learning from Human Feedback with AI Feed查看摘要-
2024ICLRLet's Verify Step by Step查看-代码
2024ICMLArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL查看-代码
2024ICMLKTO: Model Alignment as Prospect Theoretic Optimization查看-代码
2025arXivDCPO- Dynamic Clipping Policy Optimization查看摘要代码
2025arXivGroup Sequence Policy Optimization查看摘要-
2025COLINGMCA-Model-Based Causal RL for Efficient Dialogue Policy查看摘要-
2025NeurIPSWorld Models Should Prioritize the Unification of Physical and Social Dynamics查看摘要-
2025EMNLPDream to Chat: Model-based Reinforcement Learning on Dialogues with User Belief Modeling查看摘要-
2025arXivEnhancing User Engagement in Socially-Driven Dialogue through Interactive LLM Alignments查看摘要-
2025arXivMAPO: Mixed Advantage Policy Optimization for Long-Horizon Multi-Turn Dialogue查看摘要-
2025ICMLVinePPO: Refining Credit Assignment in RL Training of LLMs查看-代码
2026ICLRToward Evaluative Thinking: Meta-Policy Optimization with Evolving Reward Models查看摘要-
2026arXivBetter LLM Reasoning via Dual-Play查看摘要代码
2026ICLRTreeSearch for LLM Agent Reinforcement Learning查看摘要代码

🕹️ User Simulation & Interactive Environments · 16 篇
年份会议/期刊论文链接摘要代码
2006KERA Survey of Statistical User Simulation Techniques for RL Dialogue Management查看摘要-
2023TOISMetaphorical User Simulators for Evaluating Task-oriented Dialogue Systems查看摘要代码; 代码
2024ICLRSOTOPIA- Interactive Evaluation for Social Intelligence in Language Agents查看摘要-
2024WWWAn In-depth Investigation of User Response Simulation for Conversational Search查看摘要代码
2024AAAIAdversarial Socialbots Modeling Based on Structural Information Principles查看摘要代码
2024arXivStrength Lies in Differences! Improving Strategy Planning for Non-collaborative Dialogues via Diversified User Simulation查看摘要-
2024ACLPlatoLM: Teaching LLMs in Multi-Round Dialogue via a User Simulator查看-代码
2024arXivLLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals查看-代码
2024arXivOASIS: Open Agent Social Interaction Simulations with One Million Agents查看-代码
2025SIGDIALGenerating Diverse Personas for User Simulators to Test Interview Dialogue Systems查看摘要-
2025SIGIRSimulating Before Planning- Constructing Intrinsic User World Model for User-Tailored Dialogue Policy Planning查看摘要-
2025SIGIRTheory and Toolkits for User Simulation in the Era of Generative AI- User Modeling, Synthetic Data Generation, and System Evaluation查看摘要-
2025WWWA LLM-based Controllable, Scalable, Human-Involved User Simulator Framework for Conversational Recommender Systems查看摘要代码
2025NeurIPSGoal Alignment in LLM-Based User Simulators for Conversational AI查看摘要代码
2025ICLRτ-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains查看-代码
2026CSURFrom Individual to Society: A Survey on Social Simulation Driven by Large Language Model-based Agents查看-代码

📊 Benchmark & Evaluation · 33 篇
年份会议/期刊论文链接摘要代码
2019CVPRSocial-IQ- A Question Answering Benchmark for Artificial Social Intelligence查看--
2023ICCVSocial-IQ 2.0 Challenge- Benchmarking Multimodal Social Understanding查看-代码
2023EMNLPFANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions查看-代码
2023NeurIPSJudging LLM-as-a-Judge with MT-Bench and Chatbot Arena查看-代码
2023NeurIPSUnderstanding Social Reasoning in Language Models with Language Models查看-代码
2024ACLEvaluating Intention Detection Capability of Large Language Models in Persuasive Dialogues查看摘要代码
2024ICMLAgent-as-a-Judge- Evaluate Agents with Agents查看摘要代码
2024ACLEvaluating Very Long-Term Conversational Memory of LLM Agents查看-代码
2024ACLMM-SOC: Benchmarking Multimodal Large Language Models in Social Media Platforms查看-代码
2024ACLMMToM-QA: Multimodal Theory of Mind Question Answering查看-代码
2024ACLOpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models查看-代码
2024ACLSocialBench: Sociality Evaluation of Role-Playing Conversational Agents查看-代码
2024ACLToMBench: Benchmarking Theory of Mind in Large Language Models查看-代码
2024ICLRWho is ChatGPT? Benchmarking LLMs' Psychological Portrayal Using PsychoBench查看-代码
2025AAAIMuMA-ToM- Multi-modal Multi-Agent Theory of Mind查看摘要代码
2025AAAIToMATO: Verbalizing the Mental States of Role-Playing LLMs for Benchmarking Theory of Mind查看摘要代码
2025ACLTowards Dynamic Theory of Mind- Evaluating LLM Adaptation to Temporal Evolution of Human States查看摘要代码
2025EMNLPMOMENT S- A Comprehensive Multimodal Benchmark for Theory of Mind查看摘要代码
2025ICLRExplore theory of mind: program-guided adversarial data generation for theory of mind reasoning查看摘要代码
2025NAACLCommunication Makes Perfect: Persuasion Dataset Construction via Multi-LLM Communication查看摘要代码
2025ACLDICE-BENCH- Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues查看-代码
2025ACLIn Search of the Lost Arch in Dialogue- A Dependency Dialogue Acts Corpus for Multi-Party Dialogues查看--
2025NAACLWHoW- A Cross-domain Approach for Analysing Conversation Moderation查看--
2025arXivYou need to MIMIC to get FAME- Solving Meeting Transcript Scarcity with Multi-Agent Conversations查看--
2025EMNLPPersonaGym: Evaluating Persona Agents and LLMs查看-代码
2025ICLRLongMemEval: Benchmarking Chat Assistants on Long-Term Interactive Memory查看-代码
2026AAAIRecToM: A Benchmark for Evaluating Machine Theory of Mind in LLM-based Conversational Recommender Systems查看摘要代码
2026arXivEnactToM- An Evolving Benchmark for Functional Theory of Mind in Embodied Agents查看摘要代码
2026arXivLarge language model psychometrics- A systematic review of evaluation, validation, and enhancement查看摘要代码
2026WWWES-MemEval- Benchmarking Conversational Agents on Personalized Long-Term Emotional Support查看摘要代码
2026arXivPIVOTSBench- Evaluating Fine-Grained Interpersonal Relationship Reasoning in Multimodal Large Language Models查看--
2026arXivTIDES- A Longitudinal Bilingual Dataset for Modeling Multi-Party Social Dynamics查看--
2026ICLRMME-Emotion: A Holistic Evaluation Benchmark for Emotional Intelligence in Multimodal Large Language Models查看-代码

📁 仓库结构

Awesome-Social-AI/
├── image/              # 存放论文相关的图片、图表等
├── paper/              # 论文摘要(main,按 8 个方向前缀命名)
├── Example.md          # 论文摘要的撰写范例
└── README.md           # 本说明文件

✍️ 如何贡献

我们鼓励所有成员积极贡献自己阅读的论文摘要,请严格遵循以下步骤:

1. Fork & Clone

Fork 本仓库到你的 GitHub 账户,然后克隆到本地。

git clone https://github.com/你的用户名/Awesome-Social-AI.git && cd Awesome-Social-AI

2. 添加原始仓库为 upstream

git remote add upstream https://github.com/lucianma05-create/Awesome-Social-AI.git

3. 每次准备撰写新摘要前,请先同步主分支并创建一个独立分支:

  1. 切换回主分支并拉取上游最新代码
git checkout main
git pull upstream main
  1. 创建并切换到一个新分支 (分支名建议反映论文内容)
git checkout -b 分支名

4. 命名规范

\paper 下的文件名请严格按照以下格式命名,以便于检索和管理:
[方向]-[会议/期刊名]-[年份]-[论文名(完整的名字而不是缩写)].md
示例: Memory-NeurIPS-2023-Retrieval-Augmented-Generation.md
研究方向请从当前 8 个方向中选择最接近的归类;确需新增方向时,请先在组内讨论后再添加。
当前方向前缀:PD、ED、Recommend、Coop、ToM、Emotion、Norms、Memory、RLHF、US、Data。
\image 下的文件命名为 [年]-[月]-[日]-[编号]-[姓名缩写].png
示例:2024010101mmh.png

5. 填写内容

a. 参照 Example.md 中的模板,填写论文的各项信息,确保内容精炼、准确。

b. 或者可以使用我们专门开发的 Auto-Summary, 请在自动化生成后进行必要的人工校对和修改。

6. 修改 README.md

参照 README.md 中的表格,增加新论文的链接、摘要和代码;可以用以下提示词提示codex或copilot进行自动化填充:

请根据 paper 目录新增的 .md 文件,按 README.md 里现有表格格式补全相应方向的行,填 链接、摘要、代码 字段,缺失用 - 

或运行仓库自带的同步脚本,自动把 paper/ 目录的新论文填入表格:

python scripts/update_readme.py --dry-run   # 预览改动
python scripts/update_readme.py             # 应用改动

7. 提交、同步与推送

  1. 暂存并提交本地更改
git add .
git commit -m "你的提交信息,例如:Add summary for [论文名]"
  1. 推送当前功能分支到你个人的 GitHub (origin)
git push origin 分支名

8. 发起 Pull Request

向本仓库的主分支发起一个 Pull Request (PR),并等待审核合并。

🧠 关于我们

本仓库由 NWPU Crowd-HMT-Lab Social-AI-Group 维护。

lucianma05-create/Awesome-Social-AI

A curated collection of paper summaries for Social AI research.

Python

20

118 commits

updated Sep 16, 2026

See the code

README

🌐 Awesome Social AI

Awesome License Papers Summaries Views

我们精心收集整理社交人工智能(Social AI)领域的研究论文,并持续更新论文的中文摘要,从而支持快速了解该领域的代表性工作。仓库将不断更新,追踪社交 AI 前沿。欢迎 Follow 和 Star!⭐

🎯 范畴与结构

Social AI 是研究与构建具备社会智能的 AI 系统的领域。社会智能指智能体在社交情境中的三组能力:社交理解(感知情绪、意图、信念、关系、规范与多方动态)、社交推理(心智建模、归因与策略规划)、社交行动(说服、谈判、共情支持、建立信任)。

本仓库按四层结构组织:

不收录:计算社会科学 / "AI for social science" 类工作。方法层的通用强化学习与对齐技术,以"支撑社交智能体的基础方法"身份收录;与社交无关、亦不服务于社交智能体的技术不收录。

参考论文:

  1. Towards Social AI: A Survey on Understanding Social Interactions
  2. Advancing Social Intelligence in AI Agents: Technical Challenges and Open Questions

📑 目录

📚 论文列表

🗣️ Persuasion & Negotiation · 41 篇
年份会议/期刊论文链接摘要代码
2008PERSUASIVEA Systematic Framework for Designing and Evaluating Persuasive Systems查看摘要-
2008-Measure Of Belief Change as an Evaluation of Persuasion查看摘要-
2014COLINGReinforcement Learning of Cooperative Persuasive Dialogue Policies using Framing查看摘要-
2017HCIPersuasive Argumentation and Emotions- An Empirical Evaluation with Users查看摘要-
2023ArgCompStrategic argumentation dialogues for persuasion- Framework and experiments based on modelling the beliefs and concerns of the persuadee查看摘要-
2023arXivImproving Language Model Negotiation with Self-Play and In-Context Learning from AI Feedback查看摘要-
2024ICLRPlug-and-Play Policy Planner for LLM-Powered Dialogue Agents查看摘要代码
2024ACLHow Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs查看-代码
2024ACLMeasuring Bargaining Abilities of LLMs: A Benchmark and A Buyer-Enhancement Method查看-代码
2024EMNLPNegotiationToM: A Benchmark for Stress-testing Machine Theory of Mind on Negotiation Surrounding查看-代码
2024ICMLDebating with More Persuasive LLMs Leads to More Truthful Answers查看-代码
2024NeurIPSCooperation, Competition, and Maliciousness: LLM-Stakeholders Interactive Negotiation查看-代码
2025ScienceDurably reducing conspiracy beliefs through dialogues with AI查看摘要-
2025ScienceThe Levers of Political Persuasion查看摘要代码
2025ACLBattling against Tough Resister- Strategy Planning with Adversarial Game for Non-collaborative Dialogues查看摘要-
2025TACLHuman Choice Prediction in Language-based Persuasion Games- Simulation-based Off-Policy Evaluation查看摘要代码
2025EMNLPEnhancing LLM-Based Persuasion Simulations with Cultural and Speaker-Specific Information查看摘要代码
2025EMNLPEnhancing Persuasive Dialogue Agents by Synthesizing Cross-Disciplinary Communication Strategies查看摘要-
2025NAACLTeaching models to balance resisting and accepting persuasion查看-代码
2025arXivLLM Can be a Dangerous Persuader- Empirical Study of Persuasion Safety in Large Language Models查看摘要代码
2025EMNLPPRINCIPLES- Synthetic Strategy Memory for Proactive Dialogue Agents查看摘要代码
2025NeurIPSPersuade Me if You Can- A Framework for Evaluating Persuasion Effectiveness and Susceptibility Among Large Language Models查看摘要代码
2025AAAISimulation-free hierarchical latent policy planning for proactive dialogues查看摘要-
2025arXivToMAP: Training Opponent-Aware LLM Persuaders with Theory of Mind查看摘要代码
2025ACLEPO: Explicit Policy Optimization for Strategic Reasoning in LLMs via RL查看摘要-
2025arXivDisagreements in Reasoning: How a Model's Thinking Process Dictates Persuasion in Multi-Agent Systems查看摘要-
2025arXivEvoEmo: Evolved Emotional Policies for Adversarial LLM Agents in Multi-Turn Price Negotiation查看摘要-
2025arXivFrom Simulation to Strategy: Automating Personalized Interaction Planning for Conversational Agents查看摘要-
2025arXivPersuasion Should be Double-Blind: A Multi-Domain Dialogue Dataset With Faithfulness Based on Causal Theory of Mind查看摘要-
2025arXivVerbalized Bayesian Persuasion查看摘要-
2025EMNLPPersuasion-Dynamics in LLMs: Investigating Robustness and Adaptability in Knowledge and Safety with DuET-PD查看摘要-
2025EMNLPProfiling LLM Copyright Infringement Risks under Adversarial Persuasive Prompting查看摘要代码
2025CSURPersuasive Conversational Agents for Environmental Sustainability: A Survey查看摘要-
2025ACLASTRO: Automatic Strategy Optimization For Non-Cooperative Dialogues查看摘要代码
2026arXivOne Model, All Roles- Multi-Turn, Multi-Agent Self-Play Reinforcement Learning for Conversational Social Intelligence查看--
2026arXivPersonality-Aware Reinforcement Learning for Persuasive Dialogue with LLM-Driven Simulation查看摘要-
2026CSURA Comprehensive Survey of Computational Persuasion查看摘要代码
2026ICLRRebuttalAgent: Strategic Persuasion in Academic Rebuttal via Theory of Mind查看摘要代码
2026ICLRTowards Strategic Persuasion with Language Models查看摘要-
2026arXivMETRO: Towards Strategy Induction from Expert Dialogue Transcripts for Non-collaborative Dialogues查看摘要代码
2026ICLRStrategic Planning and Rationalizing on Trees Make LLMs Better Debaters查看-代码

❤️‍🩹 Empathy & Emotional Support · 19 篇
年份会议/期刊论文链接摘要代码
2021ACLTowards Emotional Support Dialog Systems查看-代码
2023arXivCharacterChat- Learning towards Conversational AI with Personalized Social Support查看摘要代码
2023EMNLPSoulChat- Improving LLMs’ Empathy, Listening, and Comfort Abilities through Fine-tuning with Multi-turn Empathy Conversations查看摘要代码
2023ACLTransESC: Smoothing Emotional Support Conversation via Turn-Level State Transition查看-代码
2024ACLEmoBench- Evaluating the Emotional Intelligence of Large Language Models查看摘要代码
2024ACLCan Large Language Models be Good Emotional Supporter? Mitigating Preference Bias on Emotional Support Conversation查看-代码
2024ACLESCoT: Towards Interpretable Emotional Support Dialogue Systems查看-代码
2025arXivEcho-N1- Affective RL Frontier查看摘要-
2025ACLPsyDT- Using LLMs to Construct the Digital Twin of Psychological Counselor with Personalized Counseling Style for Psychological Counseling查看摘要代码
2025ACLPsyDial- A Large-scale Long-term Conversational Dataset for Mental Health Support查看摘要代码
2025arXivReinforcement Learning with Verifiable Emotion Rewards for Empathetic Agents查看摘要代码
2025arXivSAGE- Steering and Refining Dialog Generation with State-Action Augmentation查看摘要代码
2025ACLBeyond Verbal Cues: Emotional Contagion Graph Network for Causal Emotion Entailment查看-代码
2025EMNLPChain of Strategy Optimization Makes Large Language Models Better Emotional Supporter查看摘要代码
2025CHICustomizing Emotional Support: How Do Individuals Construct and Interact With LLM-Powered Chatbots查看--
2025NAACLEmoDynamiX: Emotional Support Dialogue Strategy Prediction by Modelling MiXed Emotions and Discourse Dynamics查看-代码
2026arXivAffective Flow Language Model for Emotional Support Conversation查看摘要代码
2026arXivEMPA: Evaluating Persona-Aligned Empathy as a Process查看摘要代码
2026ACLYou Never Know a Person, You Only Know Their Defenses- Detecting Levels of Psychological Defense Mechanisms in Supportive Conversations查看--

🛍️ Conversational Recommendation · 24 篇
年份会议/期刊论文链接摘要代码
2020ACLTowards Conversational Recommendation over Multi-Type Dialogs查看摘要代码
2021ACLRevCore- Review-augmented Conversational Recommendation查看摘要代码
2023KDDImproving conversational recommendation systems via counterfactual data simulation查看摘要代码
2023EMNLPRethinking the Evaluation for Conversational Recommendation in the Era of Large Language Models查看摘要代码
2023CIKMLarge Language Models as Zero-Shot Conversational Recommenders查看-代码
2024EMNLPBeyond Persuasion: Towards Conversational Recommender System with Credible Explanations查看摘要代码
2024WWWHow Reliable is Your Simulator- Analysis on the Limitations of Current LLM-based User Simulators for Conversational Recommendation查看摘要代码
2024ACLLLM-REDIAL: A Large-Scale Dataset for Conversational Recommender Systems Created from User Behaviors with LLMs查看摘要代码
2024arXivReindex-Then-Adapt- Improving Large Language Models for Conversational Recommendation查看摘要-
2024ACLPearl: A Review-driven Persona-Knowledge Grounded Conversational Recommendation Dataset查看-代码
2024EMNLPMitigating Matthew Effect: Multi-Hypergraph Boosted Multi-Interest Self-Supervised Learning for Conversational Recommendation查看-代码
2024SIGIRBroadening the View: Demonstration-augmented Prompt Learning for Conversational Recommendation查看--
2025TKDEA Causal-Based Attribute Selection Strategy for Conversational Recommender Systems查看摘要-
2025WWWBridging Conversational and Collaborative Signals for Conversational Recommendation查看摘要-
2025WWWCollaborative Retrieval for Large Language Model-based Conversational Recommender Systems查看摘要代码
2025WWWTowards Efficient Conversational Recommendations- Expected Value of Information Meets Bandit Learning查看摘要-
2025arXivA Framework for Generating Conversational Recommendation Datasets from Behavioral Interactions查看摘要-
2025EMNLPLLM-based Conversational Recommendation Agents with Collaborative Verbalized Experience查看摘要代码
2025EMNLPTowards Personalized Conversational Sales Agents查看摘要-
2025NAACLEmpowering Retrieval-based Conversational Recommendation with Contrasting User Preferences查看-代码
2026WWWNot All Information Brings Benefits- Personalization-Driven Agent Debate for Conversational Recommendation查看摘要-
2026WWWOptimizing Multi-Turn Interactive Recommendation Agents via Generative Intrinsic Motivation查看摘要代码
2026arXivUser Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation查看摘要-
2026ICLRRank-GRPO: Training LLM-based Conversational Recommender Systems with Reinforcement Learning查看-代码

🤝 Cooperation & Collaboration · 11 篇
年份会议/期刊论文链接摘要代码
2022ScienceHuman-level play in the game of Diplomacy by combining language models with strategic reasoning查看--
2022CHIAI Chains: Transparent and Controllable Human-AI Interaction by Chaining Large Language Model Prompts查看--
2024ICLRBuilding Cooperative Embodied Agents Modularly with Large Language Models查看--
2024TACLDecision-Oriented Dialogue for Human-AI Collaboration查看--
2024ICMLShould we be going MAD- A Look at Multi-Agent Debate Strategies for LLMs查看-代码
2024ACLYour Co-Workers Matter- Evaluating Collaborative Capabilities of Language Models in Blocks World查看--
2024ACLExploring Collaboration Mechanisms for LLM Agents: A Social Psychology View查看-代码
2024ICLRMetaGPT: Meta Programming for A Multi-Agent Collaborative Framework查看-代码
2024NeurIPSCooperate or Collapse: Emergence of Sustainable Cooperation in a Society of LLM Agents查看-代码
2024ScienceAI can help humans find common ground in democratic deliberation查看-代码
2025Nature Human BehaviourPlaying repeated games with large language models查看-代码

💭 Theory of Mind · 19 篇
年份会议/期刊论文链接摘要代码
2011CogSciBayesian Theory of Mind- Modeling Joint Belief-Desire Attribution查看摘要-
2019COBSTheory of Mind as Inverse Reinforcement Learning查看摘要-
2024ACLThink Twice: Perspective-Taking Improves Large Language Models' Theory-of-Mind Capabilities查看-代码
2025ACLMachine Theory of Mind Needs Machine Validation查看摘要-
2025ACLTheory of Mind in Large Language Models- Assessment and Enhancement查看摘要-
2025arXivMINDGAMES: Do Large Language Models Have a Planning Theory of Mind?查看摘要代码
2025arXivModeling the Mental World for Embodied AI- A Comprehensive Review查看摘要-
2025arXivToM-agent: Large Language Models as Theory of Mind Aware Generative Agents with Counterfactual Reflection查看摘要-
2025arXivToM-RL: Reinforcement Learning Unlocks Theory of Mind in Small LLMs查看摘要代码
2025ICMLOvercoming Multi-step Complexity in Multimodal Theory-of-Mind Reasoning- A Scalable Bayesian Planner查看摘要-
2025NeurIPSAutoToM- Scaling Model-based Mental Inference via Automated Agent Modeling查看摘要-
2025NeurIPSMetaMind- Modeling Human Social Thoughts with Metacognitive Multi-Agent Systems查看摘要代码
2026AAAIReality vs Counterfactual- Multi-World Contrastive Reinforcement Learning for Enhancing MLLM’s Theory of Mind in Egocentric Videos查看摘要-
2026arXivInfusing Theory of Mind into Socially Intelligent LLM Agents查看摘要代码
2026arXivMetaMind- General and Cognitive World Models in Multi-Agent Systems by Meta-Theory of Mind查看摘要-
2026arXivMindClaw- Closed-Loop Embodied Mental-State Reasoning for Precision Intervention查看摘要-
2026arXivUserHarness- Harnessing User Minds for Stronger Agent Theory-of-Mind查看摘要-
2026CVPRVideo-Only ToM- Enhancing Theory of Mind in Multimodal Large Language Models查看摘要代码
2026EACLLet's Put Ourselves in Sally's Shoes: Shoes of Others Prefilling Improves Theory of Mind in LLMs查看摘要-

🎭 Emotion Understanding · 10 篇
年份会议/期刊论文链接摘要代码
2019AAAIDialogueRNN- An Attentive RNN for Emotion Detection in Conversations查看--
2019ACLMELD- A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations查看--
2021AAAICOSMIC- COmmonSense knowledge for eMotion Identification in Conversations查看--
2023arXivInstructERC- Reforming Emotion Recognition in Conversation with Multi-task Retrieval-Augmented Large Language Models查看--
2023AAAIKnowledge-Bridged Causal Interaction Network for Causal Emotion Entailment查看-代码
2024KDDEmoLLMs: A Series of Emotional Large Language Models and Annotation Tools for Comprehensive Affective Analysis查看-代码
2024NAACLTelME: Teacher-leading Multimodal Fusion Network for Emotion Recognition in Conversation查看-代码
2024NeurIPSEmotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning查看-代码
2025arXivDo LLMs Feel- Teaching Emotion Recognition with Prompts, Retrieval, and Curriculum Learning查看--
2025ACLCoE: A Clue of Emotion Framework for Emotion Recognition in Conversations查看--

⚖️ Social Norms & Morality · 12 篇
年份会议/期刊论文链接摘要代码
2020EMNLPSocial Chemistry 101- Learning to Reason about Social and Moral Norms查看--
2021arXivDelphi- Towards Machine Ethics and Norms查看--
2021ICLRAligning AI With Shared Human Values查看-代码
2022ACLThe Moral Integrity Corpus- A Benchmark for Ethical Dialogue Systems查看-代码
2022NeurIPSWhen to Make Exceptions: Exploring Language Models as Accounts of Human Moral Judgment查看-代码
2023ICMLDo the Rewards Justify the Means- Measuring Trade-Offs Between Rewards and Ethical Behavior in the MACHIAVELLI Benchmark查看--
2023ACLNormBank- A Knowledge Bank of Situational Social Norms查看-代码
2023NeurIPSEvaluating the Moral Beliefs Encoded in LLMs查看-代码
2024arXivCultureBank- An Online Community-Driven Knowledge Base Towards Culturally Aware Language Technologies查看-代码
2024AAAIValue Kaleidoscope: Engaging AI with Pluralistic Human Values, Rights, and Duties查看-代码
2025NAACLNormAd: A Framework for Measuring the Cultural Adaptability of Large Language Models查看-代码
2025Science AdvancesEmergent social conventions and collective bias in LLM populations查看-代码

💾 Agent Memory · 11 篇
年份会议/期刊论文链接摘要代码
2023UISTGenerative Agents- Interactive Simulacra of Human Behavior查看摘要代码
2023arXivMemoryBank- Enhancing Large Language Models with Long-Term Memory查看-代码
2023NeurIPSReflexion: Language Agents with Verbal Reinforcement Learning查看-代码
2024AAAIExpeL: LLM Agents Are Experiential Learners查看-代码
2024NeurIPSHippoRAG: Neurobiologically Inspired Long-Term Memory for Large Language Models查看-代码
2025arXivA-Mem: Agentic Memory for LLM Agents查看摘要代码
2025arXivEvo-Memory- Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory查看摘要代码
2025arXivRemember Me, Refine Me- A Dynamic Procedural Memory Framework for Experience-Driven Agent Evolution查看摘要代码
2025ICLRHuman-inspired Episodic Memory for Infinite Context LLMs查看-代码
2026ICLRMemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent查看摘要-
2026ICLRReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory查看摘要代码

🤖 Reinforcement Learning & Alignment · 18 篇
年份会议/期刊论文链接摘要代码
2022NeurIPSThe Surprising Effectiveness of PPO in Cooperative Multi-Agent Games查看-代码
2022NeurIPSTraining language models to follow instructions with human feedback查看--
2023NeurIPSDirect Preference Optimization- Your Language Model is Secretly a Reward Model查看摘要-
2024ICMLRLAIF vs. RLHF- Scaling Reinforcement Learning from Human Feedback with AI Feed查看摘要-
2024ICLRLet's Verify Step by Step查看-代码
2024ICMLArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL查看-代码
2024ICMLKTO: Model Alignment as Prospect Theoretic Optimization查看-代码
2025arXivDCPO- Dynamic Clipping Policy Optimization查看摘要代码
2025arXivGroup Sequence Policy Optimization查看摘要-
2025COLINGMCA-Model-Based Causal RL for Efficient Dialogue Policy查看摘要-
2025NeurIPSWorld Models Should Prioritize the Unification of Physical and Social Dynamics查看摘要-
2025EMNLPDream to Chat: Model-based Reinforcement Learning on Dialogues with User Belief Modeling查看摘要-
2025arXivEnhancing User Engagement in Socially-Driven Dialogue through Interactive LLM Alignments查看摘要-
2025arXivMAPO: Mixed Advantage Policy Optimization for Long-Horizon Multi-Turn Dialogue查看摘要-
2025ICMLVinePPO: Refining Credit Assignment in RL Training of LLMs查看-代码
2026ICLRToward Evaluative Thinking: Meta-Policy Optimization with Evolving Reward Models查看摘要-
2026arXivBetter LLM Reasoning via Dual-Play查看摘要代码
2026ICLRTreeSearch for LLM Agent Reinforcement Learning查看摘要代码

🕹️ User Simulation & Interactive Environments · 16 篇
年份会议/期刊论文链接摘要代码
2006KERA Survey of Statistical User Simulation Techniques for RL Dialogue Management查看摘要-
2023TOISMetaphorical User Simulators for Evaluating Task-oriented Dialogue Systems查看摘要代码; 代码
2024ICLRSOTOPIA- Interactive Evaluation for Social Intelligence in Language Agents查看摘要-
2024WWWAn In-depth Investigation of User Response Simulation for Conversational Search查看摘要代码
2024AAAIAdversarial Socialbots Modeling Based on Structural Information Principles查看摘要代码
2024arXivStrength Lies in Differences! Improving Strategy Planning for Non-collaborative Dialogues via Diversified User Simulation查看摘要-
2024ACLPlatoLM: Teaching LLMs in Multi-Round Dialogue via a User Simulator查看-代码
2024arXivLLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals查看-代码
2024arXivOASIS: Open Agent Social Interaction Simulations with One Million Agents查看-代码
2025SIGDIALGenerating Diverse Personas for User Simulators to Test Interview Dialogue Systems查看摘要-
2025SIGIRSimulating Before Planning- Constructing Intrinsic User World Model for User-Tailored Dialogue Policy Planning查看摘要-
2025SIGIRTheory and Toolkits for User Simulation in the Era of Generative AI- User Modeling, Synthetic Data Generation, and System Evaluation查看摘要-
2025WWWA LLM-based Controllable, Scalable, Human-Involved User Simulator Framework for Conversational Recommender Systems查看摘要代码
2025NeurIPSGoal Alignment in LLM-Based User Simulators for Conversational AI查看摘要代码
2025ICLRτ-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains查看-代码
2026CSURFrom Individual to Society: A Survey on Social Simulation Driven by Large Language Model-based Agents查看-代码

📊 Benchmark & Evaluation · 33 篇
年份会议/期刊论文链接摘要代码
2019CVPRSocial-IQ- A Question Answering Benchmark for Artificial Social Intelligence查看--
2023ICCVSocial-IQ 2.0 Challenge- Benchmarking Multimodal Social Understanding查看-代码
2023EMNLPFANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions查看-代码
2023NeurIPSJudging LLM-as-a-Judge with MT-Bench and Chatbot Arena查看-代码
2023NeurIPSUnderstanding Social Reasoning in Language Models with Language Models查看-代码
2024ACLEvaluating Intention Detection Capability of Large Language Models in Persuasive Dialogues查看摘要代码
2024ICMLAgent-as-a-Judge- Evaluate Agents with Agents查看摘要代码
2024ACLEvaluating Very Long-Term Conversational Memory of LLM Agents查看-代码
2024ACLMM-SOC: Benchmarking Multimodal Large Language Models in Social Media Platforms查看-代码
2024ACLMMToM-QA: Multimodal Theory of Mind Question Answering查看-代码
2024ACLOpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models查看-代码
2024ACLSocialBench: Sociality Evaluation of Role-Playing Conversational Agents查看-代码
2024ACLToMBench: Benchmarking Theory of Mind in Large Language Models查看-代码
2024ICLRWho is ChatGPT? Benchmarking LLMs' Psychological Portrayal Using PsychoBench查看-代码
2025AAAIMuMA-ToM- Multi-modal Multi-Agent Theory of Mind查看摘要代码
2025AAAIToMATO: Verbalizing the Mental States of Role-Playing LLMs for Benchmarking Theory of Mind查看摘要代码
2025ACLTowards Dynamic Theory of Mind- Evaluating LLM Adaptation to Temporal Evolution of Human States查看摘要代码
2025EMNLPMOMENT S- A Comprehensive Multimodal Benchmark for Theory of Mind查看摘要代码
2025ICLRExplore theory of mind: program-guided adversarial data generation for theory of mind reasoning查看摘要代码
2025NAACLCommunication Makes Perfect: Persuasion Dataset Construction via Multi-LLM Communication查看摘要代码
2025ACLDICE-BENCH- Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues查看-代码
2025ACLIn Search of the Lost Arch in Dialogue- A Dependency Dialogue Acts Corpus for Multi-Party Dialogues查看--
2025NAACLWHoW- A Cross-domain Approach for Analysing Conversation Moderation查看--
2025arXivYou need to MIMIC to get FAME- Solving Meeting Transcript Scarcity with Multi-Agent Conversations查看--
2025EMNLPPersonaGym: Evaluating Persona Agents and LLMs查看-代码
2025ICLRLongMemEval: Benchmarking Chat Assistants on Long-Term Interactive Memory查看-代码
2026AAAIRecToM: A Benchmark for Evaluating Machine Theory of Mind in LLM-based Conversational Recommender Systems查看摘要代码
2026arXivEnactToM- An Evolving Benchmark for Functional Theory of Mind in Embodied Agents查看摘要代码
2026arXivLarge language model psychometrics- A systematic review of evaluation, validation, and enhancement查看摘要代码
2026WWWES-MemEval- Benchmarking Conversational Agents on Personalized Long-Term Emotional Support查看摘要代码
2026arXivPIVOTSBench- Evaluating Fine-Grained Interpersonal Relationship Reasoning in Multimodal Large Language Models查看--
2026arXivTIDES- A Longitudinal Bilingual Dataset for Modeling Multi-Party Social Dynamics查看--
2026ICLRMME-Emotion: A Holistic Evaluation Benchmark for Emotional Intelligence in Multimodal Large Language Models查看-代码

📁 仓库结构

Awesome-Social-AI/
├── image/              # 存放论文相关的图片、图表等
├── paper/              # 论文摘要(main,按 8 个方向前缀命名)
├── Example.md          # 论文摘要的撰写范例
└── README.md           # 本说明文件

✍️ 如何贡献

我们鼓励所有成员积极贡献自己阅读的论文摘要,请严格遵循以下步骤:

1. Fork & Clone

Fork 本仓库到你的 GitHub 账户,然后克隆到本地。

git clone https://github.com/你的用户名/Awesome-Social-AI.git && cd Awesome-Social-AI

2. 添加原始仓库为 upstream

git remote add upstream https://github.com/lucianma05-create/Awesome-Social-AI.git

3. 每次准备撰写新摘要前,请先同步主分支并创建一个独立分支:

  1. 切换回主分支并拉取上游最新代码
git checkout main
git pull upstream main
  1. 创建并切换到一个新分支 (分支名建议反映论文内容)
git checkout -b 分支名

4. 命名规范

\paper 下的文件名请严格按照以下格式命名,以便于检索和管理:
[方向]-[会议/期刊名]-[年份]-[论文名(完整的名字而不是缩写)].md
示例: Memory-NeurIPS-2023-Retrieval-Augmented-Generation.md
研究方向请从当前 8 个方向中选择最接近的归类;确需新增方向时,请先在组内讨论后再添加。
当前方向前缀:PD、ED、Recommend、Coop、ToM、Emotion、Norms、Memory、RLHF、US、Data。
\image 下的文件命名为 [年]-[月]-[日]-[编号]-[姓名缩写].png
示例:2024010101mmh.png

5. 填写内容

a. 参照 Example.md 中的模板,填写论文的各项信息,确保内容精炼、准确。

b. 或者可以使用我们专门开发的 Auto-Summary, 请在自动化生成后进行必要的人工校对和修改。

6. 修改 README.md

参照 README.md 中的表格,增加新论文的链接、摘要和代码;可以用以下提示词提示codex或copilot进行自动化填充:

请根据 paper 目录新增的 .md 文件,按 README.md 里现有表格格式补全相应方向的行,填 链接、摘要、代码 字段,缺失用 - 

或运行仓库自带的同步脚本,自动把 paper/ 目录的新论文填入表格:

python scripts/update_readme.py --dry-run   # 预览改动
python scripts/update_readme.py             # 应用改动

7. 提交、同步与推送

  1. 暂存并提交本地更改
git add .
git commit -m "你的提交信息,例如:Add summary for [论文名]"
  1. 推送当前功能分支到你个人的 GitHub (origin)
git push origin 分支名

8. 发起 Pull Request

向本仓库的主分支发起一个 Pull Request (PR),并等待审核合并。

🧠 关于我们

本仓库由 NWPU Crowd-HMT-Lab Social-AI-Group 维护。

Languages

Python

100.0%