Curated embodied AI list: surveys, VLA models, datasets, simulators, humanoids, robot learning, and safety resources.
Python
248
18 commits
updated Sep 24, 2026
A curated, actively maintained list of surveys, papers, datasets, simulators, benchmarks, toolkits, and project pages for embodied AI, robot learning, vision-language-action models, humanoids, and safety.

430+ curated resources across 10 major research and tooling tracks.See CONTRIBUTING.md to add a paper, fix a link, or propose a new section. If this repo is useful, please star it and cite it.
Cheng Yin, Chenyu Yang, Zhiwen Hu, Yunxiang Mi, Weichen Lin, Yimeng Wang.
2026-06-14: added WAM representation/alignment, tactile foresight, humanoid loco-manipulation/navigation, VLA social safety, and embodied benchmark automation papers.2026-06-13: added recent WAM/VLA memory, real-time execution, force-aware manipulation, humanoid recovery, robot-learning safety, and in-context execution papers.2026-06-12: refreshed VLA/world-model manipulation, tactile VLA, embodied planners, humanoid self-modeling, embodied safety, simulators, and datasets.2026-03-30: added a dedicated Safety section with representative papers across perception, cognition, planning, interaction, and agentic systems.2025-11-05: expanded robotic code-as-policy and robotic in-context learning coverage.2025-09-07: refreshed surveys, perception, brain models, VLA models, and embodied RL entries.Functional Cache Grafting: Robust and Rapid Code-Policy Synthesis for Embodied Agents [Paper Link] [2026]
CodeDiffuser: Attention-Enhanced Diffusion Policy via VLM-Generated Code for Instruction Ambiguity [Paper Link] [Project Link] [2025]
Embodied large language models enable robots to complete complex tasks in unpredictable environments [Paper Link] [Project Link] [2025]
Maestro: Orchestrating Robotics Modules with Vision-Language Models for Zero-Shot Generalist Robots [Paper Link] [2025]
Code as Policies: Language Model Programs for Embodied Control [Paper Link] [Project Link] [2023]
Manipulate-Anything: Automating Real-World Robots using Vision-Language Models [Paper Link] [Project Link] [2024]
MotionWAM: Towards Foundation World Action Models for Real-Time Humanoid Loco-Manipulation [Paper Link] [2026]
GuideWalk: Learning Unified Autonomous Navigation and Locomotion for Humanoid Robots across Versatile Terrains [Paper Link] [2026]
RoboNaldo: Accurate, Stable and Powerful Humanoid Soccer Shooting via Motion-Guided Curriculum Reinforcement Learning [Paper Link] [Project Link] [2026]
GenHOI: Contact-Aware Humanoid-Object Interaction by Imitating Generated Videos without Task-Specific Training [Paper Link] [2026]
Stubborn: A Streamlined and Unified Reinforcement Learning Framework for Robust Motion Tracking and Fall Recovery for Humanoids [Paper Link] [Project Link] [2026]
Proprioceptive-visual correspondence enables self-other distinction in humanoid robots [Paper Link] [Project Link] [2026]
Learning Whole-Body Humanoid Locomotion via Motion Generation and Motion Tracking [Paper Link] [Project Link] [2026]
Scalable and General Whole-Body Control for Cross-Humanoid Locomotion [Paper Link] [Project Link] [2026]
Humanoid Manipulation Interface: Humanoid Whole-Body Manipulation from Robot-Free Demonstrations [Paper Link] [Project Link] [2026]
ULTRA: Unified Multimodal Control for Autonomous Humanoid Whole-Body Loco-Manipulation [Paper Link] [Project Link] [2026]
SUGAR: A Scalable Human-Video-Driven Generalizable Humanoid Loco-Manipulation Learning Framework [Paper Link] [Project Link] [2026]
OASIS: From Simulation Data Collection to Real-World Humanoid Loco-Manipulation [Paper Link] [Project Link] [2026]
HumanoidMimicGen: Data Generation for Loco-Manipulation via Whole-Body Planning [Paper Link] [Project Link] [2026]
Learning to Learn Faster from Human Feedback with Language Model Predictive Control [Paper Link] [Project Link] [2024]
ELEGNT: Expressive and Functional Movement Design for Non-anthropomorphic Robot [Paper Link] [Project Link] [2025]
Generative Expressive Robot Behaviors using Large Language Models [Paper Link] [Project Link] [2024]
A Generative Model to Embed Human Expressivity into Robot Motions [Paper Link] [2024]
Exploring the Design Space of Extra-Linguistic Expression for Robots [Paper Link] [2023]
Collection of Metaphors for Human-Robot Interaction [Paper Link] [2021]
RHINO: Learning Real-Time Humanoid-Human-Object Interaction from Human Demonstrations [Paper Link] [Project Link] [2025]
ASAP: Aligning Simulation and Real-World Physics for Learning Agile Humanoid Whole-Body Skills [Paper Link] [Project Link] [2025]
ExBody2: Advanced Expressive Humanoid Whole-Body Control [Paper Link] [Project Link] [2024]
Expressive Whole-Body Control for Humanoid Robots [Paper Link] [Project Link] [2024]
HOVER: Versatile Neural Whole-Body Controller for Humanoid Robots [Paper Link] [Project Link] [2024]
OmniH2O: Universal and Dexterous Human-to-Humanoid Whole-Body Teleoperation and Learning [Paper Link] [Project Link] [2024]
Learning Human-to-Humanoid Real-Time Whole-Body Teleoperation [Paper Link] [Project Link] [2024]
Learning from Massive Human Videos for Universal Humanoid Pose Control [Paper Link] [Project Link] [2024]
Mobile-TeleVision: Predictive Motion Priors for Humanoid Whole-Body Control [Paper Link] [Project Link] [2024]
HumanPlus: Humanoid Shadowing and Imitation from Humans [Paper Link] [Project Link] [2024]
Humanoid-VLA: Towards Universal Humanoid Control with Visual Integration [Paper Link] [2025]
XBG: End-to-End Imitation Learning for Autonomous Behaviour in Human-Robot Interaction and Collaboration [Paper Link] [2024]
EMOTION: Expressive Motion Sequence Generation for Humanoid Robots with In-Context Learning [Paper Link] [Project Link] [2024]
HARMON: Whole-Body Motion Generation of Humanoid Robots from Language Descriptions [Paper Link] [Project Link] [2024]
ImitationNet: Unsupervised Human-to-Robot Motion Retargeting via Shared Latent Space [Paper Link] [Project Link] [2023]
FABG : End-to-end Imitation Learning for Embodied Affective Human-Robot Interaction [Paper Link] [Project Link] [2025]
HAPI: A Model for Learning Robot Facial Expressions from Human Preferences [Paper Link] [2025]
Human-robot facial coexpression [Paper Link] [2024]
Unlocking Human-Like Facial Expressions in Humanoid Robots: A Novel Approach for Action Unit Driven Facial Expression Disentangled Synthesis [Paper Link] [2024]
UGotMe: An Embodied System for Affective Human-Robot Interaction [Paper Link] [Project Link] [2024]
Knowing Where to Look: A Planning-based Architecture to Automate the Gaze Behavior of Social Robots* [Paper Link] [2022]
Naturalistic Head Motion Generation from Speech [Paper Link] [2022]
Transitioning to Human Interaction with AI Systems: New Challenges and Opportunities for HCI Professionals to Enable Human-Centered AI [Paper Link] [2023]
Roots and Requirements for Collaborative AI [Paper Link] [2023]
From Human-Computer Interaction to Human-AI Interaction:New Challenges and Opportunities for Enabling Human-Centered AI [Paper Link] [2021]
From explainable to interactive AI: A literature review on current trends in human-AI interaction [Paper Link] [2024]
Treat robots as humans? Perspective choice in human-human and human-robot spatial language interaction [Paper Link] [2023]
Advances in Large Language Models for Robotics [Paper Link] [2024]
Grounding Language to Natural Human-Robot Interaction in Robot Navigation Tasks [Paper Link] [2021]
Multi-modal interaction with transformers: bridging robots and human with natural language [Paper Link] [2024]
Robot Control Platform for Multimodal Interactions with Humans Based on ChatGPT [Paper Link] [2024]
Multi-Grained Multimodal Interaction Network for Sentiment Analysis [Paper Link] [2024]
Vision-Language Navigation with Embodied Intelligence: A Survey [Paper Link] [2024]
SweepMM: A High-Quality Multimodal Dataset for Sweeping Robots in Home Scenarios for Vision-Language Model [Paper Link] [2024]
Recent advancements in multimodal human–robot interaction [Paper Link] [2023]
Multi-Modal Data Fusion in Enhancing Human-Machine Interaction for Robotic Applications: A Survey [Paper Link] [2022]
LaMI: Large Language Models for Multi-Modal Human-Robot Interaction [Paper Link] [2024]
"Help Me Help the AI": Understanding How Explainability Can Support Human-AI Interaction [Paper Link] [2022]
Employing Co-Learning to Evaluate the Explainability of Multimodal Sentiment Analysis [Paper Link] [2024]
Towards Responsible AI: Developing Explanations to Increase Human-AI Collaboration [Paper Link] [2023]
Toward Affective XAI: Facial Affect Analysis for Understanding Explainable Human-AI Interactions [Paper Link] [2021]
As embodied AI systems are deployed in safety-critical environments (autonomous driving, healthcare, household robotics), ensuring their safety becomes technically challenging and socially indispensable. This section highlights representative works on attacks and defenses across five safety layers. We intentionally select ~80 representative papers rather than the full 500+ to avoid overwhelming this repo -- for the complete collection, see Awesome-Embodied-AI-Safety.
Visual Perception — adversarial attacks and backdoors on visual recognition, detection, and tracking:
Auditory Perception — voice command injection, audio adversarial examples, and defenses:
Spatial Perception — LiDAR spoofing, point cloud attacks, and 3D perception robustness:
Motion Perception — IMU/GPS/radar sensor spoofing and drone attacks:
Cross-Modal Perception — attacks exploiting multi-sensor fusion inconsistencies:
Instruction Understanding — attacks on embodied instruction following and VQA:
World Model — hallucination, robustness, and safety in learned world models:
Reasoning — jailbreaking chain-of-thought and embodied reasoning:
Task Planning — jailbreaking LLM planners and backdooring robotic task plans:
Trajectory Planning — adversarial scenarios for autonomous driving trajectory prediction:
Multi-Agent Planning — Byzantine resilience and adversarial communication in swarms:
Robot Control — adversarial RL, backdoors in policies, and safe VLA models:
Human-Agent Interaction — perceived safety and psychological risks:
Multi-Agent Collaboration — inter-agent infection and collusion:
Tool Use — prompt injection and skill poisoning in tool-using agents:
Memory — memory poisoning, privacy leakage, and prompt extraction:
Self-Evolving — risks from self-improving and hallucinating agents:
Cascading Risks — cross-layer failures, supply chain attacks, and system-level vulnerabilities:
OmniSim [Project Link] [2026]
SIMPLE: Simulation-Based Policy Learning and Evaluation for Humanoid Loco-manipulation [Paper Link] [Project Link] [2026]
An Embodied Simulation Platform, Benchmark, and Data-Efficient Augmentation Framework for Wet-Lab Robotics [Paper Link] [2026]
Embodied-BenchClaw: An Autonomous Multi-Agent System for Embodied Spatial Intelligence Benchmark Construction [Paper Link] [2026]
ORBIT: A Unified Simulation Framework for Interactive Robot Learning Environments [Paper Link] [Project Link] [2023]
Gazebo [Paper Link] [Project Link] [2004]
Pybullet, a python module for physics simulation for games, robotics and machine learning [Project Link] [2021]
Mujoco: A physics engine for model-based control [Paper Link] [Project Link] [2012]
V-REP: A versatile and scalable robot simulation framework [Project Link] [2013]
AI2-THOR: An Interactive 3D Environment for Visual AI [Paper Link] [Project Link] [2017]
CLIPORT: What and Where Pathways for Robotic Manipulation [Paper Link] [Project Link] [2021]
BEHAVIOR-1K: A Human-Centered, Embodied AI Benchmark with 1,000 Everyday Activities and Realistic Simulation [Paper Link] [Project Link] [2024]
RLBench: The Robot Learning Benchmark & Learning Environment [Paper Link] [Project Link] [2019]
MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations [Paper Link] [Project Link] [2023]
CALVIN: A Benchmark for Language-Conditioned Policy Learning for Long-Horizon Robot Manipulation Tasks [Paper Link] [Project Link] [2022]
Meta-World: A Benchmark and Evaluation for Multi-Task and Meta Reinforcement Learning [Paper Link] [Project Link] [2019]
ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI [Paper Link] [Project Link] [2024]
HomeRobot: Open-Vocabulary Mobile Manipulation [Paper Link] [Project Link] [2023]
ARNOLD: A Benchmark for Language-Grounded Task Learning With Continuous States in Realistic 3D Scenes [Paper Link] [Project Link] [2023]
Habitat 3.0: A Co-Habitat for Humans, Avatars and Robots [Paper Link] [Project Link] [2023]
InfiniteWorld: A Unified Scalable Simulation Framework for General Visual-Language Robot Interaction [Paper Link] [Project Link] [2024]
ProcTHOR: Large-Scale Embodied AI Using Procedural Generation [Paper Link] [Project Link] [2022]
Holodeck: Language Guided Generation of 3D Embodied AI Environments [Paper Link] [Project Link] [2023]
PhyScene: Physically Interactable 3D Scene Synthesis for Embodied AI [Paper Link] [Project Link] [2024]
RoboGen: Towards Unleashing Infinite Data for Automated Robot Learning via Generative Simulation [Paper Link] [Project Link] [2023]
Genesis: A Universal and Generative Physics Engine for Robotics and Beyond [Project Link] [2025]
Webots: open-source robot simulator [Paper Link] [Project Link] [2018]
Unity: A General Platform for Intelligent Agents [Paper Link] [Project Link] [2020]
ThreeDWorld: A Platform for Interactive Multi-Modal Physical Simulation [Paper Link] [Project Link] [2021]
iGibson 1.0: A Simulation Environment for Interactive Tasks in Large Realistic Scenes [Paper Link] [Project Link] [2021]
SAPIEN: A SimulAted Part-based Interactive ENvironment [Paper Link] [Project Link] [2020]
VirtualHome: Simulating Household Activities via Programs [Paper Link] [Project Link] [2018]
Modular Open Robots Simulation Engine: MORSE [Paper Link] [Project Link] [2011]
VRKitchen: an Interactive 3D Virtual Environment for Task-oriented Learning [Paper Link] [Project Link] [2019]
CHALET: Cornell House Agent Learning Environment [Paper Link] [Project Link] [2018]
Habitat: A Platform for Embodied AI Research [Paper Link] [Project Link] [2019]
MineDojo: Building Open-Ended Embodied Agents with Internet-Scale Knowledge [Paper Link] [Project Link] [2022]
ALFRED: A Benchmark for Interpreting Grounded Instructions for Everyday Tasks [Paper Link] [Project Link] [2019]
BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning [Paper Link] [Project Link] [2019]
Gibson Env: Real-World Perception for Embodied Agents [Paper Link] [Project Link] [2018]
iGibson 2.0: Object-Centric Simulation for Robot Learning of Everyday Household Tasks [Paper Link] [Project Link] [2021]
RoboTHOR: An Open Simulation-to-Real Embodied AI Platform [Paper Link] [Project Link] [2020]
LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning [Paper Link] [Project Link] [2023]
robosuite: A Modular Simulation Framework and Benchmark for Robot Learning [Paper Link] [Project Link] [2020]
Demonstrating HumanTHOR: A Simulation Platform and Benchmark for Human-Robot Collaboration in a Shared Workspace [Paper Link] [Project Link] [2024]
Robomimic: What Matters in Learning from Offline Human Demonstrations for Robot Manipulation [Paper Link] [Project Link] [2021]
Adroit: Manipulators and Manipulation in high dimensional spaces [Paper Link] [Project Link] [2016]
Gymnasium-Robotics [Paper Link] [Project Link] [2024]
RoboHive: A Unified Framework for Robot Learning [Paper Link] [Project Link] [2024]
If this repo helps your work, please use the metadata in CITATION.cff or cite it as:
@misc{yin2025awesomeembodiedai,
title = {Awesome-Embodied-AI},
author = {Cheng Yin and Chenyu Yang and Zhiwen Hu and Yunxiang Mi and Weichen Lin and Yimeng Wang},
year = {2025},
howpublished = {\url{https://github.com/wadeKeith/Awesome-Embodied-AI}},
note = {Curated repository of embodied AI resources}
}
This repo builds on and cross-links with several strong community collections:
Python
100.0%
Curated embodied AI list: surveys, VLA models, datasets, simulators, humanoids, robot learning, and safety resources.
Python
248
18 commits
updated Sep 24, 2026
A curated, actively maintained list of surveys, papers, datasets, simulators, benchmarks, toolkits, and project pages for embodied AI, robot learning, vision-language-action models, humanoids, and safety.

430+ curated resources across 10 major research and tooling tracks.See CONTRIBUTING.md to add a paper, fix a link, or propose a new section. If this repo is useful, please star it and cite it.
Cheng Yin, Chenyu Yang, Zhiwen Hu, Yunxiang Mi, Weichen Lin, Yimeng Wang.
2026-06-14: added WAM representation/alignment, tactile foresight, humanoid loco-manipulation/navigation, VLA social safety, and embodied benchmark automation papers.2026-06-13: added recent WAM/VLA memory, real-time execution, force-aware manipulation, humanoid recovery, robot-learning safety, and in-context execution papers.2026-06-12: refreshed VLA/world-model manipulation, tactile VLA, embodied planners, humanoid self-modeling, embodied safety, simulators, and datasets.2026-03-30: added a dedicated Safety section with representative papers across perception, cognition, planning, interaction, and agentic systems.2025-11-05: expanded robotic code-as-policy and robotic in-context learning coverage.2025-09-07: refreshed surveys, perception, brain models, VLA models, and embodied RL entries.Functional Cache Grafting: Robust and Rapid Code-Policy Synthesis for Embodied Agents [Paper Link] [2026]
CodeDiffuser: Attention-Enhanced Diffusion Policy via VLM-Generated Code for Instruction Ambiguity [Paper Link] [Project Link] [2025]
Embodied large language models enable robots to complete complex tasks in unpredictable environments [Paper Link] [Project Link] [2025]
Maestro: Orchestrating Robotics Modules with Vision-Language Models for Zero-Shot Generalist Robots [Paper Link] [2025]
Code as Policies: Language Model Programs for Embodied Control [Paper Link] [Project Link] [2023]
Manipulate-Anything: Automating Real-World Robots using Vision-Language Models [Paper Link] [Project Link] [2024]
MotionWAM: Towards Foundation World Action Models for Real-Time Humanoid Loco-Manipulation [Paper Link] [2026]
GuideWalk: Learning Unified Autonomous Navigation and Locomotion for Humanoid Robots across Versatile Terrains [Paper Link] [2026]
RoboNaldo: Accurate, Stable and Powerful Humanoid Soccer Shooting via Motion-Guided Curriculum Reinforcement Learning [Paper Link] [Project Link] [2026]
GenHOI: Contact-Aware Humanoid-Object Interaction by Imitating Generated Videos without Task-Specific Training [Paper Link] [2026]
Stubborn: A Streamlined and Unified Reinforcement Learning Framework for Robust Motion Tracking and Fall Recovery for Humanoids [Paper Link] [Project Link] [2026]
Proprioceptive-visual correspondence enables self-other distinction in humanoid robots [Paper Link] [Project Link] [2026]
Learning Whole-Body Humanoid Locomotion via Motion Generation and Motion Tracking [Paper Link] [Project Link] [2026]
Scalable and General Whole-Body Control for Cross-Humanoid Locomotion [Paper Link] [Project Link] [2026]
Humanoid Manipulation Interface: Humanoid Whole-Body Manipulation from Robot-Free Demonstrations [Paper Link] [Project Link] [2026]
ULTRA: Unified Multimodal Control for Autonomous Humanoid Whole-Body Loco-Manipulation [Paper Link] [Project Link] [2026]
SUGAR: A Scalable Human-Video-Driven Generalizable Humanoid Loco-Manipulation Learning Framework [Paper Link] [Project Link] [2026]
OASIS: From Simulation Data Collection to Real-World Humanoid Loco-Manipulation [Paper Link] [Project Link] [2026]
HumanoidMimicGen: Data Generation for Loco-Manipulation via Whole-Body Planning [Paper Link] [Project Link] [2026]
Learning to Learn Faster from Human Feedback with Language Model Predictive Control [Paper Link] [Project Link] [2024]
ELEGNT: Expressive and Functional Movement Design for Non-anthropomorphic Robot [Paper Link] [Project Link] [2025]
Generative Expressive Robot Behaviors using Large Language Models [Paper Link] [Project Link] [2024]
A Generative Model to Embed Human Expressivity into Robot Motions [Paper Link] [2024]
Exploring the Design Space of Extra-Linguistic Expression for Robots [Paper Link] [2023]
Collection of Metaphors for Human-Robot Interaction [Paper Link] [2021]
RHINO: Learning Real-Time Humanoid-Human-Object Interaction from Human Demonstrations [Paper Link] [Project Link] [2025]
ASAP: Aligning Simulation and Real-World Physics for Learning Agile Humanoid Whole-Body Skills [Paper Link] [Project Link] [2025]
ExBody2: Advanced Expressive Humanoid Whole-Body Control [Paper Link] [Project Link] [2024]
Expressive Whole-Body Control for Humanoid Robots [Paper Link] [Project Link] [2024]
HOVER: Versatile Neural Whole-Body Controller for Humanoid Robots [Paper Link] [Project Link] [2024]
OmniH2O: Universal and Dexterous Human-to-Humanoid Whole-Body Teleoperation and Learning [Paper Link] [Project Link] [2024]
Learning Human-to-Humanoid Real-Time Whole-Body Teleoperation [Paper Link] [Project Link] [2024]
Learning from Massive Human Videos for Universal Humanoid Pose Control [Paper Link] [Project Link] [2024]
Mobile-TeleVision: Predictive Motion Priors for Humanoid Whole-Body Control [Paper Link] [Project Link] [2024]
HumanPlus: Humanoid Shadowing and Imitation from Humans [Paper Link] [Project Link] [2024]
Humanoid-VLA: Towards Universal Humanoid Control with Visual Integration [Paper Link] [2025]
XBG: End-to-End Imitation Learning for Autonomous Behaviour in Human-Robot Interaction and Collaboration [Paper Link] [2024]
EMOTION: Expressive Motion Sequence Generation for Humanoid Robots with In-Context Learning [Paper Link] [Project Link] [2024]
HARMON: Whole-Body Motion Generation of Humanoid Robots from Language Descriptions [Paper Link] [Project Link] [2024]
ImitationNet: Unsupervised Human-to-Robot Motion Retargeting via Shared Latent Space [Paper Link] [Project Link] [2023]
FABG : End-to-end Imitation Learning for Embodied Affective Human-Robot Interaction [Paper Link] [Project Link] [2025]
HAPI: A Model for Learning Robot Facial Expressions from Human Preferences [Paper Link] [2025]
Human-robot facial coexpression [Paper Link] [2024]
Unlocking Human-Like Facial Expressions in Humanoid Robots: A Novel Approach for Action Unit Driven Facial Expression Disentangled Synthesis [Paper Link] [2024]
UGotMe: An Embodied System for Affective Human-Robot Interaction [Paper Link] [Project Link] [2024]
Knowing Where to Look: A Planning-based Architecture to Automate the Gaze Behavior of Social Robots* [Paper Link] [2022]
Naturalistic Head Motion Generation from Speech [Paper Link] [2022]
Transitioning to Human Interaction with AI Systems: New Challenges and Opportunities for HCI Professionals to Enable Human-Centered AI [Paper Link] [2023]
Roots and Requirements for Collaborative AI [Paper Link] [2023]
From Human-Computer Interaction to Human-AI Interaction:New Challenges and Opportunities for Enabling Human-Centered AI [Paper Link] [2021]
From explainable to interactive AI: A literature review on current trends in human-AI interaction [Paper Link] [2024]
Treat robots as humans? Perspective choice in human-human and human-robot spatial language interaction [Paper Link] [2023]
Advances in Large Language Models for Robotics [Paper Link] [2024]
Grounding Language to Natural Human-Robot Interaction in Robot Navigation Tasks [Paper Link] [2021]
Multi-modal interaction with transformers: bridging robots and human with natural language [Paper Link] [2024]
Robot Control Platform for Multimodal Interactions with Humans Based on ChatGPT [Paper Link] [2024]
Multi-Grained Multimodal Interaction Network for Sentiment Analysis [Paper Link] [2024]
Vision-Language Navigation with Embodied Intelligence: A Survey [Paper Link] [2024]
SweepMM: A High-Quality Multimodal Dataset for Sweeping Robots in Home Scenarios for Vision-Language Model [Paper Link] [2024]
Recent advancements in multimodal human–robot interaction [Paper Link] [2023]
Multi-Modal Data Fusion in Enhancing Human-Machine Interaction for Robotic Applications: A Survey [Paper Link] [2022]
LaMI: Large Language Models for Multi-Modal Human-Robot Interaction [Paper Link] [2024]
"Help Me Help the AI": Understanding How Explainability Can Support Human-AI Interaction [Paper Link] [2022]
Employing Co-Learning to Evaluate the Explainability of Multimodal Sentiment Analysis [Paper Link] [2024]
Towards Responsible AI: Developing Explanations to Increase Human-AI Collaboration [Paper Link] [2023]
Toward Affective XAI: Facial Affect Analysis for Understanding Explainable Human-AI Interactions [Paper Link] [2021]
As embodied AI systems are deployed in safety-critical environments (autonomous driving, healthcare, household robotics), ensuring their safety becomes technically challenging and socially indispensable. This section highlights representative works on attacks and defenses across five safety layers. We intentionally select ~80 representative papers rather than the full 500+ to avoid overwhelming this repo -- for the complete collection, see Awesome-Embodied-AI-Safety.
Visual Perception — adversarial attacks and backdoors on visual recognition, detection, and tracking:
Auditory Perception — voice command injection, audio adversarial examples, and defenses:
Spatial Perception — LiDAR spoofing, point cloud attacks, and 3D perception robustness:
Motion Perception — IMU/GPS/radar sensor spoofing and drone attacks:
Cross-Modal Perception — attacks exploiting multi-sensor fusion inconsistencies:
Instruction Understanding — attacks on embodied instruction following and VQA:
World Model — hallucination, robustness, and safety in learned world models:
Reasoning — jailbreaking chain-of-thought and embodied reasoning:
Task Planning — jailbreaking LLM planners and backdooring robotic task plans:
Trajectory Planning — adversarial scenarios for autonomous driving trajectory prediction:
Multi-Agent Planning — Byzantine resilience and adversarial communication in swarms:
Robot Control — adversarial RL, backdoors in policies, and safe VLA models:
Human-Agent Interaction — perceived safety and psychological risks:
Multi-Agent Collaboration — inter-agent infection and collusion:
Tool Use — prompt injection and skill poisoning in tool-using agents:
Memory — memory poisoning, privacy leakage, and prompt extraction:
Self-Evolving — risks from self-improving and hallucinating agents:
Cascading Risks — cross-layer failures, supply chain attacks, and system-level vulnerabilities:
OmniSim [Project Link] [2026]
SIMPLE: Simulation-Based Policy Learning and Evaluation for Humanoid Loco-manipulation [Paper Link] [Project Link] [2026]
An Embodied Simulation Platform, Benchmark, and Data-Efficient Augmentation Framework for Wet-Lab Robotics [Paper Link] [2026]
Embodied-BenchClaw: An Autonomous Multi-Agent System for Embodied Spatial Intelligence Benchmark Construction [Paper Link] [2026]
ORBIT: A Unified Simulation Framework for Interactive Robot Learning Environments [Paper Link] [Project Link] [2023]
Gazebo [Paper Link] [Project Link] [2004]
Pybullet, a python module for physics simulation for games, robotics and machine learning [Project Link] [2021]
Mujoco: A physics engine for model-based control [Paper Link] [Project Link] [2012]
V-REP: A versatile and scalable robot simulation framework [Project Link] [2013]
AI2-THOR: An Interactive 3D Environment for Visual AI [Paper Link] [Project Link] [2017]
CLIPORT: What and Where Pathways for Robotic Manipulation [Paper Link] [Project Link] [2021]
BEHAVIOR-1K: A Human-Centered, Embodied AI Benchmark with 1,000 Everyday Activities and Realistic Simulation [Paper Link] [Project Link] [2024]
RLBench: The Robot Learning Benchmark & Learning Environment [Paper Link] [Project Link] [2019]
MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations [Paper Link] [Project Link] [2023]
CALVIN: A Benchmark for Language-Conditioned Policy Learning for Long-Horizon Robot Manipulation Tasks [Paper Link] [Project Link] [2022]
Meta-World: A Benchmark and Evaluation for Multi-Task and Meta Reinforcement Learning [Paper Link] [Project Link] [2019]
ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI [Paper Link] [Project Link] [2024]
HomeRobot: Open-Vocabulary Mobile Manipulation [Paper Link] [Project Link] [2023]
ARNOLD: A Benchmark for Language-Grounded Task Learning With Continuous States in Realistic 3D Scenes [Paper Link] [Project Link] [2023]
Habitat 3.0: A Co-Habitat for Humans, Avatars and Robots [Paper Link] [Project Link] [2023]
InfiniteWorld: A Unified Scalable Simulation Framework for General Visual-Language Robot Interaction [Paper Link] [Project Link] [2024]
ProcTHOR: Large-Scale Embodied AI Using Procedural Generation [Paper Link] [Project Link] [2022]
Holodeck: Language Guided Generation of 3D Embodied AI Environments [Paper Link] [Project Link] [2023]
PhyScene: Physically Interactable 3D Scene Synthesis for Embodied AI [Paper Link] [Project Link] [2024]
RoboGen: Towards Unleashing Infinite Data for Automated Robot Learning via Generative Simulation [Paper Link] [Project Link] [2023]
Genesis: A Universal and Generative Physics Engine for Robotics and Beyond [Project Link] [2025]
Webots: open-source robot simulator [Paper Link] [Project Link] [2018]
Unity: A General Platform for Intelligent Agents [Paper Link] [Project Link] [2020]
ThreeDWorld: A Platform for Interactive Multi-Modal Physical Simulation [Paper Link] [Project Link] [2021]
iGibson 1.0: A Simulation Environment for Interactive Tasks in Large Realistic Scenes [Paper Link] [Project Link] [2021]
SAPIEN: A SimulAted Part-based Interactive ENvironment [Paper Link] [Project Link] [2020]
VirtualHome: Simulating Household Activities via Programs [Paper Link] [Project Link] [2018]
Modular Open Robots Simulation Engine: MORSE [Paper Link] [Project Link] [2011]
VRKitchen: an Interactive 3D Virtual Environment for Task-oriented Learning [Paper Link] [Project Link] [2019]
CHALET: Cornell House Agent Learning Environment [Paper Link] [Project Link] [2018]
Habitat: A Platform for Embodied AI Research [Paper Link] [Project Link] [2019]
MineDojo: Building Open-Ended Embodied Agents with Internet-Scale Knowledge [Paper Link] [Project Link] [2022]
ALFRED: A Benchmark for Interpreting Grounded Instructions for Everyday Tasks [Paper Link] [Project Link] [2019]
BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning [Paper Link] [Project Link] [2019]
Gibson Env: Real-World Perception for Embodied Agents [Paper Link] [Project Link] [2018]
iGibson 2.0: Object-Centric Simulation for Robot Learning of Everyday Household Tasks [Paper Link] [Project Link] [2021]
RoboTHOR: An Open Simulation-to-Real Embodied AI Platform [Paper Link] [Project Link] [2020]
LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning [Paper Link] [Project Link] [2023]
robosuite: A Modular Simulation Framework and Benchmark for Robot Learning [Paper Link] [Project Link] [2020]
Demonstrating HumanTHOR: A Simulation Platform and Benchmark for Human-Robot Collaboration in a Shared Workspace [Paper Link] [Project Link] [2024]
Robomimic: What Matters in Learning from Offline Human Demonstrations for Robot Manipulation [Paper Link] [Project Link] [2021]
Adroit: Manipulators and Manipulation in high dimensional spaces [Paper Link] [Project Link] [2016]
Gymnasium-Robotics [Paper Link] [Project Link] [2024]
RoboHive: A Unified Framework for Robot Learning [Paper Link] [Project Link] [2024]
If this repo helps your work, please use the metadata in CITATION.cff or cite it as:
@misc{yin2025awesomeembodiedai,
title = {Awesome-Embodied-AI},
author = {Cheng Yin and Chenyu Yang and Zhiwen Hu and Yunxiang Mi and Weichen Lin and Yimeng Wang},
year = {2025},
howpublished = {\url{https://github.com/wadeKeith/Awesome-Embodied-AI}},
note = {Curated repository of embodied AI resources}
}
This repo builds on and cross-links with several strong community collections:
Python
100.0%