WILLOSCAR/Awesome-HCI-LLM

Awesome-HCI (Ubiquitous, LLM, MLLM, Agent, RAG, Embodied-AI, RLHF)

Python

26

0 commits

updated Mar 8, 2026

See the code

README

Awesome HCI-LLM-Agent Papers

Awesome

A curated collection of research papers on HCI, LLM, MLLM, Agent, RAG, Agentic-RL, and Embodied AI (2021–present).

[Jan 2025] Added new sections: Agentic-RL and MLLM. Regular updates resumed.

Quick Start

python -m pip install -e .                # Install CLI
paper add 2312.00752 LLM -t "llm, mamba"  # Add paper
paper search transformer -t IMU           # Search
paper stats                               # Statistics

Documentation


HCI

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026An Expert Schema for Evaluating Large Language Model Errors in Scholarly Question-Answering SystemsAnna Martin-Boyle, et al.LLMcs.HC, cs.CL24 pages, 2 figures. Accepted at ACM CHI conference on Human Factors in Computing Systems, 20262026.02
arXiv(v1) 2026Codesigning Ripplet: an LLM-Assisted Assessment Authoring System Grounded in a Conceptual Model of Teachers' WorkflowsYuan Cui, et al.LLMcs.HCProceedings of the 2026 CHI Conference on Human Factors in Computing Systems2026.02
arXiv(v1) 2026Detecting UX smells in Visual Studio Code using LLMsAndrés Rodriguez, et al.LLMcs.SE, cs.HC4 pages, 2 figures, 1 table, 3rd International Workshop on Integrated Development Environments (IDE 2026)2026.02
arXiv(v1) 2026LLM Novice Uplift on Dual-Use, In Silico Biology TasksChen Bo Calvin Zhang, et al.LLMcs.AI, cs.CL, cs.CR, cs.CY, cs.HC59 pages, 33 figures2026.02
arXiv(v1) 2026PaperTrail: A Claim-Evidence Interface for Grounding Provenance in LLM-based Scholarly Q&AAnna Martin-Boyle, et al.LLMcs.HC, cs.CL25 pages, 3 figures. Accepted at the ACM CHI conference on Human Factors in Computing Systems 20262026.02
arXiv(v1) 2026Scaling Search Relevance: Augmenting App Store Ranking with LLM-Generated JudgmentsEvangelia Christakopoulou, et al.LLMcs.IR, cs.AI, cs.LG2026.02
arXiv(v1) 2026SparkMe: Adaptive Semi-Structured Interviewing for Qualitative Insight DiscoveryDavid Anugraha, et al.LLMcs.HC, cs.AI, cs.CY2026.02
arXiv(v1) 2026 (Proceedings of the 34th ACM International Conference on Information and Knowledge Management (CIKM '25), November 10--14, 2025, Seoul, Republic of Korea)UXSim: Towards a Hybrid User Search SimulationSaber Zerhoudi, et al.LLMcs.IR, cs.HC2026.02
arXiv(v1) 2026Understanding Usage and Engagement in AI-Powered Scientific Research Tools: The Asta Interaction DatasetDany Haddad, et al.LLMcs.HC, cs.AI, cs.IR2026.02
arXiv(v1) 2026When LLMs Help -- and Hurt -- Teaching Assistants in Proof-Based CoursesRomina Mahinpei, et al.LLMcs.HC2026.02
Ubicomp25 (IMWUT Vol 9 Issue 4)Ads that Talk Back: Implications and Perceptions of Injecting Personalized Advertising into LLM ChatbotsBrian Jay Tang, et al.LLM, personalization, advertising, chatbot2025.12
Ubicomp25 (IMWUT Vol 9 Issue 4)CHEF-VL: Detecting Cognitive Sequencing Errors in Cooking with Vision-language ModelsRuiqi Wang, et al.VLM, cooking, error detection2025.12
Ubicomp25 (IMWUT Vol 9 Issue 4)Design and Evaluation of Generative Agent-based Platform for Human-Assistant Interaction Research: A Tale of 10 User StudiesZiyi Xuan, et al.LLM, generative agent, platform, human-assistant interaction, user study2025.12
Ubicomp25 (IMWUT Vol 9 Issue 4)Gestura: A LVLM-Powered System Bridging Motion and Semantics for Real-Time Free-Form Gesture UnderstandingZhuoming Li, et al.LVLM, gesture, motion, semantics2025.12
Ubicomp25 (IMWUT Vol 9 Issue 4)IMUZero: Zero-Shot Human Activity Recognition by Language-Based Cross Modality FusionJie Su, et al.LLM, HAR, zero-shot, cross-modality, language2025.12
arXiv(v2) 2026LLM-Guided Exemplar Selection for Few-Shot Wearable-Sensor Human Activity RecognitionElsen Ronando, et al.LLM, HAR, wearable, few-shot, exemplar selectioncs.CL, cs.AI, cs.CV88.78% F1 on UCI-HAR2025.12
Ubicomp25 (IMWUT Vol 9 Issue 4)Large Language Model-guided Semantic Alignment for Human Activity RecognitionHua Yan, et al.LLM, HAR, semantic alignment2025.12
Ubicomp25 (IMWUT Vol 9 Issue 4)TourismMinds: A Geo-augmented LLM Framework for Semantic-aware Trajectory Analytics and GenerationZhuohan Ye, et al.LLM, trajectory, geo-augmented2025.12
CSCW25An Emergent Understanding of Human-AI Collaboration in DeliberationAuthors TBDLLM, human-AI collaboration, deliberation, citizen assembly2025.11
SenSys25Demo: An LLM-Powered Multimodal Mobile Sensing System for Personalized Health Behavior AnalysisAuthors TBDLLM, mobile sensing, multimodal, health, personalizedDemo paper2025.11
CSCW25Exploring Collaboration Patterns and Strategies in Human-AI Co-creation through the Lens of Agency: A Scoping ReviewAuthors TBDhuman-AI co-creation, agency, collaboration, scoping reviewPACM HCI2025.11
HAI25Human-Like Remembering and Forgetting in LLM Agents: An ACT-R-Inspired Memory ArchitectureYudai Honda, et al.LLM, agent, memory, ACT-R, forgetting2025.11
HAI25Robots with Attitudes: Influence of LLM-Driven Robot Personalities on Motivation and PerformanceDennis Becker, et al.LLM, robot, personality, motivation, performance2025.11
HAI25The Double-Edged Sword: Exploring Older Adults' Interaction and Imagination with an LLM-Enhanced Health AgentLeon Paul Mondrian Munz, et al.LLM, health agent, older adults, interaction2025.11
SenSys25Toward Sensor-In-the-Loop LLM Agent: Benchmarks and ImplicationsZechen Li, et al.LLM, agent, sensor, wearable, benchmark2025.11
UIST25ImaginationVellum: Generative-AI Ideation Canvas with Spatial PromptsAuthors TBDgenerative AI, ideation, canvas, spatial prompts, co-creation2025.10
Ubicomp25LLM Powered Memory Consolidation for Ubiquitous ComputingParampuneet Kaur Thind, et al.LLM, memory consolidation, ubiquitous computingUbiComp 2025 Companion2025.10
UIST25SketchGPT: A Sketch-based Multimodal Interface for Application-Agnostic LLM InteractionAuthors TBDLLM, sketch, multimodal, interface2025.10
UIST25"This is My Fault", Really? Understanding Blind and Low-Vision People’s Perception of Hallucination in Large Vision Language ModelsYilin Tang, et al.VLM2025.09
UIST25BloomIntent: Automating Search Evaluation with LLM-Generated Fine-Grained User IntentsYoonseo Choi, et al.LLM2025.09
UIST25Can You Move These Over There? Exploring an LLM-based VR Mover to Support Natural Multi-object ManipulationXiangzhi Eric Wang, et al.LLM2025.09
UIST25CoGrader: Transforming Instructors' Assessment of Project Reports through Collaborative LLM IntegrationZixin Chen, et al.LLM2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Contact-free Vital Signs Monitoring and Separating Distinct Vital SignsAuthors TBDvital signs, contactless, respiration, monitoring2025.09
UIST25DxHF: Providing High-Quality Human Feedback for LLM Alignment with Interactive DecompositionDanqing Shi, et al.LLM, alignment2025.09
UIST25GestureCoach: Rehearsing for Engaging Talks with LLM-Driven Gesture RecommendationsAshwin Ram, et al.LLM2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Hapt-Aids: Self-Powered, On-Body Haptics for Activity MonitoringAuthors TBDhaptic, self-powered, activity monitoring, wearable2025.09
UIST25InReAcTable: LLM-powered Interactive Visual Data Story Construction from Tabular DataGerile Aodeng, et al.LLM2025.09
arXiv(v4) 2025LLaSA: A Sensor-Aware LLM for Natural Language Reasoning of Human Activity from IMU DataSheikh Asif Imran, et al.IMU, LLM, wearable, sensor, HARcs.CL2025.09
UIST25LegisFlow: Enhancing Korean Legal Research with Temporal-Aware LLM InterfacesJunghwan Kim, et al.LLM2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)MASTER: A Multi-modal Foundation Model for Human Activity RecognitionGuanzhou Zhu, et al.foundation model, HAR, multimodal2025.09
UIST25MapStory: Prototyping Editable Map Animations with LLM AgentsAditya Gunturu, et al.LLM, agent2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Mindfulness Meditation and Respiration: Accelerometer-based Respiration Rate EstimationAuthors TBDmindfulness, respiration, accelerometer, meditation2025.09
UIST25NarraGuide: an LLM-based Narrative Mobile Robot for Remote Place ExplorationYaxin Hu, et al.LLM2025.09
UIST25NeuroSync: Intent-Aware Code-Based Problem Solving via Direct LLM Understanding ModificationWenshuo Zhang, et al.LLM2025.09
UIST25Oak Story: Improving Learner Outcomes with LLM-Mediated Interactive NarrativesAlan Y. Cheng, et al.LLM2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)One Model to Fit Them All: Universal IMU-based Human Activity Recognition with LLM-assisted Cross-dataset RepresentationWei Wei, et al.IMU, LLM, HAR, universal model, cross-dataset2025.09
UIST25Policy Maps: Tools for Guiding the Unbounded Space of LLM BehaviorsMichelle S. Lam, et al.LLM2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Pulse-PPG: An Open-Source Field-Trained PPG Foundation Model for Wearable Applications across Lab and Field SettingsMithun Saha, et al.foundation model, wearable, PPG2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)RouteLLM: A Large Language Model with Native Route Context Understanding to Enable Context-Aware ReasoningPhilipp Hallgarten, et al.LLM, navigation, route, context-aware2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)SELA: Smart Edge LLM Agent to Optimize Response Trade-offs of AI AssistantsShreshth Tuli, et al.LLM agent, edge, assistant, optimization2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Sleep Monitoring with Continuous Posture, Heart Rate, Respiratory Rate TrackingAuthors TBDsleep, posture, heart rate, respiration, monitoring2025.09
UIST25Streaming, Fast and Slow: Cognitive Load-Aware Streaming for Efficient LLM ServingChang Xiao, et al.LLM2025.09
arXiv(v1) 2025Through the Lens of Human-Human Collaboration: A Configurable Research Platform for Exploring Human-Agent CollaborationBingsheng Yao, et al.LLM, agent, human-agent collaboration, CSCW, platformcs.HC, cs.AI, cs.CL2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Towards Customizable Foundation Models for Human Activity Recognition with Wearable DevicesMinghui Qiu, et al.foundation model, HAR, customizable2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Vinci: A Real-time Smart Assistant Based on Egocentric Vision-language Model for Portable DevicesYifei Huang, et al.VLM, egocentric, assistant, wearable2025.09
UIST25ViseGPT: Towards Better Alignment of LLM-generated Data Wrangling Scripts and User PromptsJiajun Zhu, et al.LLM, alignment2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Vital Insight: Assisting Experts' Context-Driven Sensemaking of Multi-modal Personal Tracking Data Using Visualization and Human-in-the-Loop LLMJiachen Li, et al.LLM, sensemaking, visualization, personal data2025.09
UIST25agentAR: Creating Augmented Reality Applications with Tool-Augmented LLM-based Autonomous AgentsChenfei Zhu, et al.LLM, agent2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)mmPencil: Toward Writing-Style-Independent In-Air Handwriting Recognition via mmWave Radar and Large Vision-Language ModelYifan Guo, et al.mmWave, radar, handwriting, VLM2025.09
arXiv(v1) 2025BaroPoser: Real-time Human Motion Tracking from IMUs and Barometers in Everyday DevicesRiku Arakawa, et al.IMU, barometer, pose estimation, motion trackingcs.CV, cs.HC2025.08
arXiv(v1) 2025DiffCap: Diffusion-based Real-time Human Motion Capture using Sparse IMUs and a Monocular CameraZongmian Li, et al.IMU, diffusion, motion capture, real-time, cameracs.CV2025.08
arXiv(v4) 2025SensorLLM: Aligning Large Language Models with Motion Sensors for Human Activity RecognitionZechen Li, et al.IMU, LLM, motion sensor, HARcs.CLAccepted by EMNLP 2025 Main Conference2025.08
Ubicomp25 (IMWUT Vol 9 Issue 2)CataractBot: An LLM-powered Expert-in-the-Loop Chatbot for Cataract PatientsPragnya Ramjee, et al.LLM, chatbot, healthcare, expert-in-the-loop2025.06
arXiv(v1) 2025Garment Inertial Poser: Human Motion Capture from Loose and Sparse Inertial Sensors with Garment-aware Diffusion ModelsTao Wang, et al.IMU, pose estimation, loose sensor, diffusion, garmentcs.CV2025.06
Ubicomp25 (IMWUT Vol 9 Issue 2)LEGO: Synthesizing IoT Device Components Based on Static Analysis and Large Language ModelsLiwei Liu, et al.LLM, IoT, static analysis2025.06
arXiv(v1) 2025SensorLM: Learning the Language of Wearable SensorsTong Xia, et al.wearable, sensor, LLM, activity recognition, foundation modelcs.LG, cs.HC2025.06
arXiv(v2) 2025 (PRX Quantum 6, 020311 (April 2025))Device-Independent Quantum Key Distribution Based on Routed Bell TestsTristan Le Roy-Deloison, et al.IMU, RGB, human object interaction, dataset, 3D trackingquant-phVersion2: Slight improvements in the text. Close to published version2025.05
arXiv(v4) 2025LLM-Based Human-Agent Collaboration and Interaction Systems: A SurveyHengyi Peng, et al.LLM, agent, human-agent collaboration, HCI, surveycs.HC, cs.AI, cs.CL2025.05
CHI25"A Great Start, But...": Evaluating LLM-Generated Mind Maps for Information Mapping in Video-Based DesignTianhao He, et al.LLM2025.04
CHI25"Ask Sir Oliver Ingham": LLM-based Social Simulations for History EducationKieun Park, et al.LLM2025.04
CHI25"Create a Fear of Missing Out" - ChatGPT Implements Unsolicited Deceptive Designs in Generated Websites Without WarningVeronika Krauß, et al.LLM2025.04
CHI25"It Warned Me Just at the Right Moment": Exploring LLM-based Real-time Detection of Phone ScamsZitong Shen, et al.LLM2025.04
CHI25"Kya family planning after marriage hoti hai?": Integrating Cultural Sensitivity in an LLM Chatbot for Reproductive HealthRoshini Deva, et al.LLM, chatbot2025.04
CHI25"We do use it, but not how hearing people think": How the Deaf and Hard of Hearing Community Uses Large Language Model ToolsShuxu Huffman, et al.LLM2025.04
CHI25"When AI Writes Personas": Analyzing Lexical Diversity in LLM-Generated Persona DescriptionsSankalp Sethi, et al.LLM, persona2025.04
CHI25"You Don't Need a University Degree to Comprehend Data Protection This Way": LLM-Powered Interactive Privacy Policy AssessmentVincent Freiberger, et al.LLM, privacy2025.04
CHI25A Matter of Perspective(s): Contrasting Human and LLM Argumentation in Subjective Decision-Making on Subtle SexismPaula Akemi Aoyagui, et al.LLM2025.04
CHI25AI on My Shoulder: Supporting Emotional Labor in Front-Office Roles with an LLM-based Empathetic CoworkerVedant Das Swain, et al.LLM2025.04
CHI25AI-Instruments: Embodying Prompts as InstrumentsAuthors TBDLLM, prompt, instruments, direct manipulation2025.04
CHI25ASHABot: An LLM-Powered Chatbot to Support the Informational Needs of Community Health WorkersPragnya Ramjee, et al.LLM, chatbot2025.04
CHI25Adaptive Human-LLMs Interaction Collaboration: Reinforcement Learning driven Vision-Language Models for Medical Report GenerationYiming Cao, et al.VLM2025.04
CHI25Align with Me, Not TO Me: How People Perceive Concept Alignment with LLM-Powered Conversational AgentsShengchen Zhang, et al.LLM, agent, alignment2025.04
CHI25Applying the Gricean Maxims to a Human-LLM Interaction Cycle: Design Insights from a Participatory ApproachYoonsu Kim, et al.LLM2025.04
CHI25Artificial Intimacy: Exploring Normativity and Personalization Through Fine-tuning LLM ChatbotsMirabelle Jones, et al.LLM, chatbot, personalization, fine-tuning, normativity2025.04
CHI25Assessing Critical Thinking through a Multi-Agent LLM-Based Debate ChatbotBogyeom Park, et al.LLM, agent, chatbot2025.04
CHI25AutoPBL: An LLM-powered Platform to Guide and Support Individual Learners Through Self Project-based LearningYihao Zhu, et al.LLM2025.04
CHI25BallistoBud: Heart Rate Variability Monitoring using Earbud Accelerometry for Stress AssessmentAuthors TBDearbuds, accelerometer, BCG, heart rate, stress2025.04
CHI25Beyond Adaptation: an LLM-Supported Self-Presentation Ideation Tool for Cross-Cultural MinglingChien-Yin Wu, et al.LLM2025.04
CHI25Beyond Code Generation: LLM-supported Exploration of the Program Design SpaceJ.D. Zamfirescu-Pereira, et al.LLM2025.04
CHI25BioSpark: Beyond Analogical Inspiration to LLM-augmented TransferHyeonsu B Kang, et al.LLM2025.04
CHI25Boosting Diary Study Outcomes with a Fine-Tuned Large Language ModelSunggyeol Oh, et al.LLM2025.04
CHI25Breaking Barriers or Building Dependency? Exploring Team-LLM Collaboration in AI-infused Classroom DebateZihan Zhang, et al.LLM2025.04
CHI25Bridging the Treatment Gap: A Novel LLM-Driven System for Scalable Initial Patient Assessments in Mental HealthcareNiclas Rosteck, et al.LLM2025.04
CHI25BudsID: Mobile-Ready and Expressive Finger Identification Input for EarbudsAuthors TBDearbuds, finger identification, magnetometer, wearable96.9% accuracy2025.04
CHI25COMETIC: Enhancing Smartphone Eye Tracking with Cursor-Based Implicit CalibrationAuthors TBDeye tracking, smartphone, calibration, cursor27.2% improvement2025.04
CHI25Canvil: Designerly Adaptation for LLM-Powered User ExperiencesKenneth Li, et al.LLM, design, UX, user experience, adaptation2025.04
CHI25CaseMaster: Designing a Probe for Oral Case Presentation Training with LLM AssistanceYang Ouyang, et al.LLM2025.04
CHI25ChainBuddy: An AI-assisted Agent System for Generating LLM PipelinesXinyue Chen, et al.LLM, agent, pipeline, AI-assisted, prompt engineering2025.04
CHI25Characterizing LLM-Empowered Personalized Story Reading and Interaction for Children: Insights From Multi-Stakeholder PerspectivesJiaju Chen, et al.LLM, persona, personalization2025.04
CHI25Closing the Loop between User Stories and GUI Prototypes: An LLM-Based Assistant for Cross-Functional Integration in Software DevelopmentFelix Kretzer, et al.LLM, assistant2025.04
CHI25Co-designing Large Language Model Tools for Project-Based Learning with K12 EducatorsPrerna Ravi, et al.LLM2025.04
CHI25Context over Categories: Implementing the Theory of Constructed Emotion with LLM-Guided User AnalysisNils Klüwer, et al.LLM2025.04
CHI25ConversAR: Exploring Embodied LLM-Powered Group Conversations in Augmented Reality for Second Language LearnersJad Bendarkawi, et al.LLM2025.04
CHI25Cross, Dwell, or Pinch: Around-Device Selection Methods for Unmodified SmartwatchesAuthors TBDsmartwatch, sonar, around-device, inputFirst sonar-based around-device input on consumer smartwatch2025.04
CHI25Customizing Emotional Support: How Do Individuals Construct and Interact With LLM-Powered ChatbotsXi Zheng, et al.LLM, chatbot, personalization2025.04
CHI25DBox: Scaffolding Algorithmic Programming Learning through Learner-LLM Co-DecompositionShuai Ma, et al.LLM2025.04
CHI25Dango: A Mixed-Initiative Data Wrangling System using Large Language ModelWei-Hao Chen, et al.LLM2025.04
CHI25DanmuA11y: Making Time-Synced Video Comments Accessible to BLV UsersAuthors TBDaccessibility, BLV, video, Danmu, audio2025.04
CHI25Demonstration of GazeNoter: Enhancing AR Note-Taking Through Gaze-Based Selection of LLM SuggestionsShih-Kang Chiu, et al.LLM2025.04
CHI25Demystifying Mental Health Reports Through an LLM-based ApproachShyama Sastha Krishnamoorthy Srinivasan, et al.LLM2025.04
CHI25Design Principles and Guidelines for LLM Observability: Insights from DevelopersXin Chen, et al.LLM2025.04
CHI25Designing Accessible Audio Nudges for Voice InterfacesHira Jamshed, et al.voice interface, audio, nudging, older adults, accessibility2025.04
CHI25Designing LLM-Powered Multimodal Instructions to Support Rich Hands-on Skills Remote Learning: A Case Study with Massage Instructors and LearnersChutian Jiang, et al.LLM, multimodal2025.04
CHI25Development of an LLM-Based Chatbot to Support Learnability in Stardew Valley: A Diary Study ApproachJungmin Lee, et al.LLM, chatbot2025.04
CHI25Effects of Acoustic Transparency of Wearable Audio Devices on Audio ARYuki Watanabe, et al.audio AR, wearable, acoustic transparency, hearables2025.04
CHI25Effects of LLM-based Search on Decision Making: Speed, Accuracy, and OverrelianceSofia Eleni Spatharioti, et al.LLM2025.04
CHI25Efficient Management of LLM-Based Coaching Agents' Reasoning While Maintaining Interaction Quality and SpeedAndreas Göldi, et al.LLM, agent2025.04
arXiv(v1) 2025Ego4o: Egocentric Human Motion Capture and Understanding from Multi-Modal InputJian Wang, et al.egocentric, IMU, pose estimation, multimodal, VRcs.CV, cs.HC2025.04
CHI25End User Authoring of Personalized Content Classifiers: Comparing Example Labeling, Rule Writing, and LLM PromptingLeijie Wang, et al.LLM, personalization, content classifier, end-user, prompting2025.04
CHI25Enhancing AI Explainability for Non-technical Users with LLM-Driven Narrative GamificationYuzhe You, et al.LLM2025.04
CHI25EvAlignUX: Advancing UX Evaluation through LLM-Supported Metrics ExplorationQingxiao Zheng, et al.LLM2025.04
CHI25Explaining Complex ML Models to Domain Experts Using LLM & Visualization: An Exploration in the French Breadmaking IndustryBriggs Twitchell, et al.LLM2025.04
CHI25Exploring Culturally Informed AI Assistants: A Comparative Study of ChatBlackGPT and ChatGPTLisa Egede, et al.LLM, assistant2025.04
CHI25Exploring Gender Biases in LLM-based Voice Chatbots for Job InterviewsSumin Heo, et al.LLM, chatbot2025.04
CHI25Exploring LLM-Powered Role and Action-Switching Pedagogical Agents for History Education in Virtual RealityZihao Zhu, et al.LLM, agent2025.04
CHI25Exploring Mobile Touch Interaction with Large Language ModelsAuthors TBDLLM, touch, mobile, gesture, interaction2025.04
CHI25Exploring Older Adults Personality Preferences for LLM-powered Conversational CompanionsAjwa Shahid, et al.LLM, persona, personalization2025.04
CHI25Exploring Personalized Health Support through Data-Driven, Theory-Guided LLMs: A Case Study in Sleep HealthXin Tong, et al.LLM, wearable, health, sleep, personalized, chatbotHealthGuru multi-agent framework2025.04
CHI25Exploring the Design Space of Real-time LLM Knowledge Support Systems: A Case Study of Jargon ExplanationsYuhan Liu, et al.LLM2025.04
CHI25Exploring the Design of LLM-based Agent in Enhancing Self-disclosure Among the Older AdultsYijie Guo, et al.LLM, agent2025.04
CHI25Exploring the Impact of Explainability in Large Language Model (LLM) Applications on User ExperienceYanyun Wang, et al.LLM2025.04
CHI25Exploring the Impact of Intervention Methods on Developers’ Security Behavior in a Manipulated ChatGPT StudyRaphael Serafini, et al.LLM2025.04
CHI25FIP: Endowing Robust Motion Capture on Daily Garment by Fusing Flex and Inertial SensorsYiwei Zhao, et al.IMU, flex sensor, motion capture, pose estimation, garment19.5% improvement over SOTA2025.04
CHI25Fact or Fiction? Exploring Explanations to Identify Factual Confabulations in RAG-Based LLM SystemsPhilipp Reinhard, et al.LLM2025.04
CHI25FineType: Fine-grained Tapping Gesture Recognition for Text EntryAuthors TBDtapping, gesture, text entry, IMU, wristband2025.04
CHI25FingerGlass: Enhancing Smart Glasses Interaction via Fingerprint SensingAuthors TBDsmart glasses, fingerprint, gesture recognition, CNN, LSTM2025.04
CHI25Friction: Deciphering Writing Feedback into Writing Revisions through LLM-Assisted ReflectionChao Zhang, et al.LLM2025.04
CHI25From Text to Trust: Empowering AI-assisted Decision Making with Adaptive LLM-powered AnalysisZhuoyan Li, et al.LLM2025.04
CHI25FusAIn: Composing Generative AI Visual Prompts Using Pen-based InteractionXiaohan Peng, et al.generative AI, pen-based, prompt, visual design2025.04
CHI25GPTCoach: Towards LLM-Based Physical Activity CoachingMatthew Jörke, et al.LLM2025.04
CHI25GazeNoter: Co-Piloted AR Note-Taking via Gaze Selection of LLM Suggestions to Match Users' IntentionsZhiyi Rong, et al.LLM, AR, gaze, note-taking, eye tracking2025.04
CHI25Gesture and Audio-Haptic Guidance Techniques to Direct Conversations with Intelligent Voice InterfacesXiyuan Shen, et al.LLM, wearable, voice interface, gesture, haptic, smart glassesRay-Ban Meta Glasses, GPT-4o2025.04
CHI25HaptiCoil: Soft Programmable Buttons with Hydraulically Coupled Haptic Feedback and SensingAuthors TBDhaptic, soft button, sensing, feedback1-500Hz bandwidth2025.04
CHI25Human Robot Interaction for Blind and Low Vision People: A Systematic Literature ReviewAuthors TBDrobot, accessibility, BLV, HRI, survey2025.04
CHI25Human Subjects Research in the Age of Generative AI: Opportunities and Challenges of Applying LLM-Simulated Data to HCI StudiesAngel Hsing-Chi Hwang, et al.LLM2025.04
CHI25IdeationWeb: Tracking the Evolution of Design Ideas in Human-AI Co-CreationAuthors TBDLLM, human-AI co-creation, ideation, design2025.04
CHI25Improving User Engagement and Learning Outcomes in LLM-Based Python Tutor: A Study of PACEMuhtasim Ibteda Shochcho, et al.LLM2025.04
CHI25Inkspire: Supporting Design Exploration with Generative AI through Analogical SketchingAuthors TBDgenerative AI, T2I, sketching, design exploration2025.04
CHI25Interactive Debugging and Steering of Multi-Agent AI SystemsWill Epperson, et al.LLM, agent, multi-agent, debugging, HCI2025.04
CHI25Investigating LLM-Driven Curiosity in Human-Robot InteractionJan Leusmann, et al.LLM2025.04
CHI25LEGOLAS: Learning & Enhancing Golf Skills through LLM-Augmented SystemKangbeen Ko, et al.LLM2025.04
CHI25LIGS: Developing an LLM-infused Game System for Emergent NarrativeJin Jeong, et al.LLM2025.04
CHI25LLM Adoption in Data Curation Workflows: Industry Practices and InsightsCrystal Qian, et al.LLM2025.04
CHI25LLM Integration in Extended Reality: A Comprehensive Review of Current Trends, Challenges, and Future PerspectivesChengkun Wu, et al.LLM, VR, AR, XR, survey, extended realitySurvey paper2025.04
CHI25LLM Powered Text Entry Decoding and Flexible Typing on SmartphonesAuthors TBDLLM, text entry, typing, smartphone, gesture93.1% top-1 accuracy2025.04
CHI25LLM Whisperer: An Inconspicuous Attack to Bias LLM ResponsesWeiran Lin, et al.LLM2025.04
CHI25LearnMate: Enhancing Online Education with LLM-Powered Personalized Learning Plans and SupportXinyu Jessica Wang, et al.LLM, persona, personalization2025.04
CHI25Letters from Future Self: Augmenting the Letter-Exchange Exercise with LLM-based Agents to Enhance Young Adults' Career ExplorationHayeon Jeon, et al.LLM, agent2025.04
CHI25Leveraging Multimodal LLM for Inspirational User Interface SearchSeokhyeon Park, et al.LLM, multimodal2025.04
CHI25LifeInsight: Design and Evaluation of an AI-Powered Assistive Wearable for Blind and Low Vision PeopleAuthors TBDAI, wearable, accessibility, BLV, assistive2025.04
CHI25LittleToDo: Large Language Model Driven Intervention Tool for Adolescent Academic Procrastination with Affective ComputingXiaofan Hu, et al.LLM2025.04
CHI25Lookee: Gaze Tracking-based Infant Vocabulary Comprehension AssessmentAuthors TBDgaze, infant, vocabulary, assessment, AI2025.04
CHI25M2SILENT: Enabling Multi-user Silent Speech Interactions via Multi-directional SpeakersAuthors TBDsilent speech, multi-user, speaker, shared space2025.04
CHI25MAP: Multi-user Personalization with Collaborative LLM-powered AgentsChristine P. Lee, et al.LLM, agent, persona, personalization2025.04
CHI25Maintaining Long-Distance Relationships with (Mediocre) LLM-based Chatbots: A Collaborative Ethnographic StudyBernd Ploderer, et al.LLM, chatbot2025.04
arXiv(v1) 2025MobilePoser: Real-Time Full-Body Pose Estimation and 3D Human Translation from IMUs in Mobile Consumer DevicesVimal Mollyn, et al.IMU, pose estimation, mobile, consumer device, real-timecs.CV, cs.HC2025.04
CHI25MotionBlocks: Modular Geometric Motion Remapping for Accessible VRAuthors TBDVR, accessibility, motion, remapping, limited mobility2025.04
CHI25Objection Overruled! Lay People can Distinguish Large Language Models from Lawyers, but still Favour Advice from an LLMEike Schneiders, et al.LLM2025.04
CHI25Online-EYE: Multimodal Implicit Eye Tracking Calibration for XRAuthors TBDeye tracking, XR, VR, calibration, implicit2025.04
CHI25PPG Earring: Wireless Smart Earring for Heart Health MonitoringAuthors TBDPPG, earring, heart rate, wearable, health monitoring14mm, 2g, 21h battery2025.04
CHI25PaperWave: Listening to Research Papers as Conversational Podcasts Scripted by LLMYuchi Yahagi, et al.LLM2025.04
CHI25Parents, Children, and ChatGPT in Home Environments: The Conversation Content and the Interaction ModeShuang Quan, et al.LLM2025.04
CHI25Piecing Together Teamwork: A Responsible Approach to an LLM-based Educational Jigsaw AgentEmily Doherty, et al.LLM, agent2025.04
CHI25Plan-Then-Execute: An Empirical Study of User Trust and Team Performance When Using LLM Agents As A Daily AssistantGaole He, et al.LLM, agent, assistant2025.04
CHI25Playing Dumb to Get Smart: Creating and Evaluating an LLM-based Teachable Agent within University Computer Science ClassesNaiming Liu, et al.LLM, teachable agent, education, learning2025.04
CHI25PolicyPulse: LLM-Synthesis Tool for Policy ResearchersMaggie Wang, et al.LLM2025.04
CHI25Privacy Meets Explainability: Managing Confidential Data and Transparency Policies in LLM-Empowered ScienceYashothara Shanmugarasa, et al.LLM, privacy2025.04
CHI25Private Yet Social: How LLM Chatbots Support and Challenge Eating Disorder RecoveryRyuhaerang Choi, et al.LLM, chatbot2025.04
CHI25Promoting Cognitive Health in Elder Care with Large Language Model-Powered Socially Assistive RobotsMaria R. Lima, et al.LLM2025.04
CHI25PropType: Everyday Props as Typing Surfaces in Augmented RealityAuthors TBDAR, typing, props, text entry2025.04
CHI25Prototyping with Prompts: Emerging Approaches and Challenges in Generative AI DesignHari Subramonyam, et al.generative AI, prompt engineering, design, prototyping2025.04
CHI25Proxona: Supporting Creators' Sensemaking and Ideation with LLM-Powered Audience PersonasYoonseo Choi, et al.LLM, persona2025.04
CHI25RadEye: Tracking Eye Motion Using FMCW RadarAuthors TBDradar, eye tracking, FMCW, gaze2025.04
CHI25Rambler in the Wild: A Diary Study of LLM-Assisted Writing With SpeechXuyu Yang, et al.LLM2025.04
CHI25Redefining Research Crowdsourcing: Incorporating Human Feedback with LLM-Powered Digital TwinsAmanda Chan, et al.LLM2025.04
CHI25Rescriber: Smaller-LLM-Powered User-Led Data Minimization for LLM-Based ChatbotsJijie Zhou, et al.LLM, chatbot2025.04
CHI25SPECTRA: Personalizable Sound Recognition for DHH Users through Interactive MLSteven M. Goodman, et al.sound recognition, DHH, accessibility, interactive ML2025.04
CHI25Scaffolded Turns and Logical Conversations: Designing Humanized LLM-Powered Conversational Agents for Hospital Admission InterviewsDingdong Liu, et al.LLM, agent2025.04
CHI25Script&Shift: A Layered Interface Paradigm for Integrating Content Development and Rhetorical Strategy with LLM Writing AssistantsMomin N Siddiqui, et al.LLM, assistant2025.04
CHI25Seeing and Touching the Air: Eye-Hand Coordination in Mid-Air Gesture Typing for MRAuthors TBDgesture typing, MR, mid-air, eye-hand coordination2025.04
CHI25Seeking Inspiration through Human-LLM InteractionXinrui Lin, et al.LLM2025.04
CHI25SocialEyes: Scaling Mobile Eye-tracking to Multi-person Social SettingsAuthors TBDeye tracking, mobile, social, multi-person2025.04
CHI25Sonora: Human-AI Co-Creation of 3D Audio WorldsFernanda M De La Torre, et al.AI, audio, 3D, soundscape, LLM, co-creationUses LLM for voice commands2025.04
CHI25Spatial Hand Actions: Hand Actions for Spatial Thinking in 3D AssemblingAuthors TBDhand, spatial, 3D, assembling, gesture2025.04
CHI25Spatial Haptics: A Sensory Substitution Method for Distal Object Detection Using Tactile CuesAuthors TBDhaptic, tactile, sensory substitution, localization2025.04
CHI25Spatial Speech Translation: Translating Across Space With Binaural HearablesAuthors TBDspeech translation, binaural, hearables, spatial audio2025.04
CHI25SpellRing: Recognizing Continuous Fingerspelling in ASL using a RingAuthors TBDring, ASL, sign language, fingerspelling, wearable2025.04
CHI25TableNarrator: Making Image Tables Accessible to Blind and Low Vision PeopleAuthors TBDaccessibility, BLV, tables, image, data2025.04
CHI25Talk to the Hand: an LLM-powered Chatbot with Visual Pointer as Proactive Companion for On-Screen TasksZhepeng Wang, et al.LLM, chatbot, UI, on-screen tasks, visual pointer2025.04
CHI25Tap&Say: Touch Location-Informed LLM for Multimodal Text CorrectionAuthors TBDLLM, touch, voice, multimodal, text correction2025.04
CHI25The Interaction Layer: An Exploration for Co-Designing User-LLM Interactions in Parental Wellbeing Support SystemsSruthi Viswanathan, et al.LLM2025.04
CHI25The Voice of Endo: Leveraging Speech for Illness Flare-up ForecastingAuthors TBDspeech, health, voice, endometriosis, forecasting2025.04
CHI25Through the Lens of Privacy: Exploring Privacy Protection in Vision-Language Model Interactions on Smart GlassesZiyang Zhang, et al.VLM, privacy2025.04
CHI25Too Much Information? Investigating Information Disclosure in Auction Systems with LLM SimulationsYue YinLLM2025.04
CHI25Toward Enabling Natural Conversation with Older Adults via the Design of LLM-Powered Voice Agents that Support Interruptions and BackchannelsChao Liu, et al.LLM, agent2025.04
CHI25Towards Human-AI Deliberation: Design and Evaluation of LLM-Empowered Deliberative AI for AI-Assisted Decision-MakingShuai Ma, et al.LLM2025.04
CHI25UXAgent: An LLM Agent-Based Usability Testing Framework for Web DesignYuxuan Lu, et al.LLM, agent2025.04
CHI25Understanding the Effects of Large Language Model (LLM)-driven Adversarial Social Influences in Online Information SpreadZhuoran Lu, et al.LLM2025.04
CHI25Understanding the LLM-ification of CHI: Unpacking the Impact of LLMs at CHI through a Systematic Literature ReviewRock Yuren Pang, et al.LLM, HCI, CHI, survey, systematic review2025.04
CHI25Unlocking Scientific Concepts: How Effective Are LLM-Generated Analogies for Student Understanding and Classroom Practice?Zekai Shao, et al.LLM2025.04
CHI25Unpacking Trust Dynamics in the LLM Supply Chain: An Empirical Exploration to Foster Trustworthy LLM Production & UseAgathe Balayn, et al.LLM2025.04
CHI25User Experience with LLM-powered Conversational Recommendation Systems: A Case of Music RecommendationSojeong Yun, et al.LLM2025.04
CHI25Users' Expectations and Practices with Agent MemoryBrennan Jones, et al.LLM, agent, memory, user expectations, HCI2025.04
CHI25Utilizing ChatGPT in a Data Structures and Algorithms Course: A Teaching Assistant's PerspectivePooriya Jamie, et al.LLM, assistant2025.04
CHI25VibWalk: Mapping Lower-limb Haptic Experiences of Everyday WalkingAuthors TBDhaptic, walking, vibration, wearable, foot2025.04
CHI25Visiobo Demo: Augmenting Static Prints with Projection-based Visual Cueing and Concept Mapping via LLM ReasoningJiaqi Jiang, et al.LLM2025.04
CHI25Wearable Meets LLM for Stress Management: A Duoethnographic Study Integrating Wearable-Triggered Stressors and LLM Chatbots for Personalized InterventionsSameer Neupane, et al.LLM, wearable, stress, chatbot, personalized intervention2025.04
CHI25Weaving Sound Information to Support Real-Time Sensemaking for DHH UsersJeremy Zhengqi Huang, et al.sound, DHH, accessibility, AI, sensemaking2025.04
CHI25What If Smart Homes Could See Our Homes?: Exploring DIY Smart Home Building Experiences with VLM-Based Camera SensorsSojeong Yun, et al.VLM2025.04
CHI25What Social Media Use Do People Regret? An Analysis of 34K Smartphone Screenshots with Multimodal LLMLongjie Guo, et al.LLM, multimodal2025.04
CHI25WritingRing: Enabling Natural Handwriting Input with a Single IMU RingXiaoying Yang, et al.IMU, ring, handwriting, text entry, wearableSingle IMU ring2025.04
CHI25Your Hands Can Tell: Detecting Redirected Hand Movements in VRAuthors TBDVR, hand, redirection, detection2025.04
arXiv(v1) 2025Broadband shot-to-shot transient absorption anisotropyMaximilian Binzer, et al.VR, AR, egocentric, motion capture, FRAMEphysics.opticsThe following article has been submitted to The Journal of Physical Chemistry. After it is published, it will be found at https://pubs.aip.org/aip/jcp2025.03
IUI25CAIM: A Cognitive AI Memory Framework for Long-term Interaction with LLMsRebecca Westhäußer, et al.LLM, memory, long-term interaction, framework2025.03
Ubicomp25 (IMWUT Vol 9 Issue 1)HandSAW: Wearable Hand-based Event Recognition via On-Body Surface Acoustic WavesKaylee Yaxuan Li, et al.SAW, wrist, hand, object interaction, wearable2025.03
arXiv(v2) 2025Modeling Future Conversation Turns to Teach LLMs to Ask Clarifying QuestionsMichael J. Q. Zhang, et al.IMU, diffusion, pose estimation, loose sensorcs.CLPresented at ICLR 20252025.03
IUI25NoTeeline: Supporting Real-Time, Personalized Notetaking with LLM-Enhanced MicronotesFaria Huq, et al.LLM, note-taking, personalization, real-time, micronotes2025.03
Ubicomp25 (IMWUT Vol 9 Issue 1)Respiration Rate Estimation via Smartwatch-based PPG and Accelerometer DataAuthors TBDrespiration, smartwatch, PPG, accelerometer, transfer learning2025.03
Ubicomp25 (IMWUT Vol 9 Issue 1)SocialMind: LLM-based Proactive AR Social Assistive System with Human-like Perception for In-situ Live InteractionsBufang Yang, et al.LLM, AR, social assistive, proactive2025.03
TEI25Tangible LLMs: Tangible Sense-Making For Trustworthy Large Language ModelsAuthors TBDLLM, tangible, trustworthy AI, physical interface2025.02
CSCW24Is Human-AI Interaction CSCW?Meredith Ringel Morris, et al.human-AI collaboration, CSCW, LLM, panelPanel discussion2024.11
Ubicomp24 (IMWUT Vol 8 Issue 4)Ring-a-Pose: A Ring for Continuous Hand Pose TrackingTianhong Catherine Yu, et al.ring, hand pose, tracking, wearable2024.11
Ubicomp24Sensor2Text: Enabling Natural Language Interactions for Daily Activity Tracking Using Wearable SensorsWenqiang Chen, et al.LLM, wearable, natural language, activity tracking2024.11
Ubicomp24Leveraging LLMs to Predict Affective States via Smartphone Sensor FeaturesAuthors TBDLLM, smartphone, sensing, affective state, digital phenotypingFirst LLM work for affective state prediction2024.10
Ubicomp24Leveraging Large Language Models for Generating Mobile Sensing Strategies in Human Behavior ModelingAuthors TBDLLM, mobile sensing, behavior modeling, strategy generation2024.10
UIST24Patchview: LLM-powered Worldbuilding with Generative Dust and Magnet VisualizationJeongyeon Kim, et al.LLM, worldbuilding, writing, creative, visualization2024.10
UIST24SHAPE-IT: Exploring Text-to-Shape-Display for Generative Shape-Changing Behaviors with LLMsWanli Qian, et al.LLM, shape display, tangible, generative, text-to-shapeAI-chaining approach2024.10
UIST24SituationAdapt: Contextual UI Optimization in Mixed Reality with Situation Awareness via LLM ReasoningZhipeng Li, et al.LLM, MR, UI, adaptive, context-aware, mixed reality2024.10
UIST24VizAbility: Enhancing Chart Accessibility with LLM-based Conversational InteractionMandi Cai, et al.LLM, accessibility, chart, visualization, conversational2024.10
Ubicomp24 (IMWUT Vol 8 Issue 3)IMUGPT 2.0: Language-Based Cross Modality Transfer for Sensor-Based Human Activity RecognitionZikang Leng, et al.IMU, LLM, HAR, cross-modality, motion synthesis20 citations2024.09
arXiv(v1) 2024WheelPoser: Sparse-IMU Based Body Pose Estimation for Wheelchair UsersYunzhi Li, et al.IMU, pose estimation, wheelchair, accessibilitycs.GR, cs.CV, cs.HCAccepted by ASSETS 20242024.09
arXiv(v1) 2024EMHI: A Multimodal Egocentric Human Motion Dataset with HMD and Body-Worn IMUsZelin Ye, et al.IMU, VR, HMD, dataset, egocentric, pose estimationcs.CV, cs.HC885 sequences, 58 subjects, 28.5 hours2024.08
arXiv(v1) 2024Evaluating Text Classification Robustness to Part-of-Speech Adversarial ExamplesAnahita Samadi, et al.IMU, transformer, pose estimation, calibrationcs.CL, cs.LG2024.08
CHI24"My agent understands me better": Integrating Dynamic Human-like Memory Recall and Consolidation in LLM-Based AgentsYuki Hou, et al.LLM, agent, memory, recall, consolidation2024.05
CHI24As an AI language model I cannot: Investigating LLM Denials of User RequestsAuthors TBDLLM, denial, user request, perception2024.05
CHI24Bridging the Gulf of Envisioning: Cognitive Challenges in Prompt Based Interactions with LLMsAuthors TBDLLM, prompt, cognitive challenges, user study2024.05
CHI24ChaCha: Leveraging Large Language Models to Prompt Children to Share Their EmotionsAuthors TBDLLM, chatbot, children, emotion, conversation2024.05
CHI24CharacterMeet: Supporting Creative Writers' Character Construction Through LLM-Powered Chatbot AvatarsAuthors TBDLLM, chatbot, creative writing, character design2024.05
arXiv(v2) 2024Finding Candidate TeV Halos among Very-High Energy SourcesDong Zheng, et al.VR, AR, egocentric, pose estimationastro-ph.HE15 pages, 7 figures, 4 tables, referee's comments incorporated, accepted for publication in ApJ2024.05
CHI24How AI Processing Delays Foster Creativity: CoQuestAuthors TBDLLM, agent, research question, creativity, co-creation2024.05
CHI24Learning Agent-based Modeling with LLM Companions: ChatGPT and NetLogo ChatAuthors TBDLLM, agent-based modeling, NetLogo, learning2024.05
CHI24The HaLLMark Effect: Supporting Provenance and Transparent Use of LLMs in WritingAuthors TBDLLM, writing, provenance, visualization, transparency2024.05
CHI24Towards Robotic Companions: Understanding Handler-Guide Dog Interactions for Informed Guide Dog Robot DesignHochul Hwang, et al.LLM, HCI, CHI, interaction2024.05
CHI24Understanding the Impact of Long-Term Memory on Self-Disclosure with LLM-Driven ChatbotsAuthors TBDLLM, chatbot, long-term memory, self-disclosure, health2024.05
arXiv(v1) 2024Exploring Text-to-Motion Generation with Human PreferenceJenny Sheng, et al.IMU, human object interaction, datasetcs.LG, cs.AI, cs.CVAccepted to CVPR 2024 HuMoGen Workshop2024.04
arXiv(v2) 2024Health-LLM: Large Language Models for Health Prediction via Wearable Sensor DataYubin Kim, et al.LLM, wearable, health predictioncs.CL, cs.AI, cs.LG2024.04
arXiv(v2) 2024PRISM: Patient Records Interpretation for Semantic Clinical Trial Matching using Large Language ModelsShashi Kant Gupta, et al.IMU, RGB, HOI, dataset, trackingcs.CL, cs.AI30 Pages, 8 Figures, Supplementary Work Attached2024.04
arXiv(v1) 2024Bayesian Learned Models Can Detect Adversarial Malware For FreeBao Gia Doan, et al.VR, AR, simulated avatar, headsetcs.CRAccepted to the 29th European Symposium on Research in Computer Security (ESORICS) 2024 Conference2024.03
Ubicomp24 (IMWUT Vol 8 Issue 1)Capturing the College Experience: A Four-Year Mobile Sensing Study of Mental HealthAuthors TBDmobile sensing, mental health, college, longitudinal19 citations2024.03
arXiv(v1) (CVPR24)Dynamic Inertial Poser (DynaIP): Part-Based Motion Dynamics Learning for Enhanced Human Pose Estimation with Sparse Inertial SensorsYu Zhang, et al.IMU, sparse inertial sensorscs.CV2024.03
Ubicomp24HyperHARNafees Ahmad, et al.LLM, passive sensing, sensemaking2024.03
arXiv(v1) 2024LLMs in HCI Data Work: Bridging the Gap Between Information Retrieval and Responsible Research PracticesNeda Taghizadeh Serajeh, et al.VR, AR, avatar control, pose estimationcs.HC, cs.IR5 pages, CHI2024 Workshop on LLMs as Research Tools: Applications and Evaluations in HCI Data Work2024.03
arXiv(v1) 2024Modeling and optimization for arrays of water turbine OWC devicesM. Gambarini, et al.VR, AR, motion capture, egocentric, stereo cameramath.OC, physics.flu-dyn2024.03
arXiv(v1) 2024Modeling stock price dynamics on the Ghana Stock Exchange: A Geometric Brownian Motion approachDennis Lartey Quayesam, et al.VR, AR, motion capture, egocentricmath.OC, q-fin.ST2024.03
arXiv(v1) 2024On depth prediction for autonomous driving using self-supervised learningHoussem BoulahbalVR, AR, avatar, pose estimation, headsetcs.CVPhD thesis2024.03
Ubicomp24ViObjectWenqiang Chen, et al.mental health, LLM, text data2024.03
arXiv(v1) 2024IMUGPT 2.0: Language-Based Cross Modality Transfer for Sensor-Based Human Activity RecognitionZikang Leng, et al.IMU, LLM, cross-modality, HARcs.CV2024.02
arXiv(v2) 2024IMUOptimize: A Data-Driven Approach to Optimal IMU Placement for Human Pose Estimation with Transformer ArchitectureVarun Ramani, et al.IMU, transformer, interpretability, data driven, time seriescs.LG2024.02
arXiv(v1) 2024IMUSIC: IMU-based Facial Expression CaptureYoujia Wang, et al.IMU, generation, simulate, transformer diffusioncs.CVcode coming soon (link)2024.02
Ubicomp23CAvatarWenqiang Chen, et al.human activity, 3D mesh, tactile, pressure2023.12
Ubicomp23A Data-Driven Context-Aware Health Inference System for Children during School Closuresdata analysis, school closures, health inference, risk factor analysis
Ubicomp23Abacus Gestures: A Large Set of Math-Based Usable Finger-Counting Gestures for Mid-Air Interactionsvision, mid-air, gesture interaction, math, finger counting, abacus
ISWC 2023C-Auth: Exploring the Feasibility of Using Egocentric View of Face Contour for User Authentication on Glassessmart glasses, authentication, ecocentric view
Ubicomp24CAvatar: Real-time Human Activity Mesh Reconstruction via Tactile Carpetshuman activity reconstruction, 3D human mesh, pressure and vibrations, tactile sensor
Ubicomp23Contact Tracing for Healthcare Workers in an Intensive Care Unitcontact Tracing, Internet of things (IoT), bluetooth low energy, Covid-19
Ubicomp23DRG-Keyboard: Enabling Subtle Gesture Typing on the Fingertip with Dual IMU Ringstext entry, gesture keyboard, fingertip interaction, smart ring
arXiv(v5) 2024Evaluating Human-Language Model InteractionMina Lee, et al.LM, human-centered, evaluationcs.CL
Ubicomp23Exploring the Opportunities of AR for Enriching Storytelling with Family Photos between Grandparents and GrandchildrenAR, storytelling, intergenerational communication
Ubicomp23Fingerprinting IoT Devices Using Latent Physical Side-Channelsphysical side-channels, fingerprinting, internet-of-things
Ubicomp23From 2D to 3D: Facilitating Single-Finger Mid-Air Typing on QWERTY Keyboards with Probabilistic Touch Modelingmid air, text entry, VR
Ubicomp23GC-Loc: A Graph Attention Based Framework for Collaborative Indoor Localization Using Infrastructure-free Signalscollaborative indoor localization, graph neural network, geomagnetism
Ubicomp23GLOBEM: Cross-Dataset Generalization of Longitudinal Human Behavior Modelinggeneralizability, behavior modeling, passive sensing
Ubicomp23HIPPO: Pervasive Hand-Grip Estimation from Everyday Interactions
arXiv(v1) (CHI23)HOOV: Hand Out-Of-View Tracking for Proprioceptive Interaction using Inertial SensingPaul Streli, et al.IMU, VR, transformercs.HC, cs.CV, I.2; I.5; H.5
Ubicomp23Headar: Sensing Head Gestures for Confirmation Dialogs on Smartwatches with Wearable Millimeter-Wave Radarwearable interaction, gestural input, millimeter-wave radar, head gestures, smartwatch
Ubicomp23HyWay: Enabling Mingling in the Hybrid World∗hybrid mingling, unstructured and semi-structured conversations, awareness, agency, porosity, reciprocity
Ubicomp23I Know Your Intent: Graph-enhanced Intent-aware User Device Interaction Prediction via Contrastive Learninguser device interaction, graph, attention, contrastive learning
Ubicomp22IF-ConvTransformer: A Framework for Human Activity Recognition Using IMU Fusion and ConvTransformerIMU, fusion, multimodal, transformer, attention
arXiv(v1) (CHI23)IMUPoser: Full-Body Pose Estimation using IMUs in Phones, Watches, and EarbudsVimal Mollyn, et al.IMU, pose estimation, BiLSTMcs.HC, cs.CV
Ubicomp23LT-Fall: The Design and Implementation of a Life-threatening Fall Detection and Alarming System
Ubicomp23LapTouch: Using the Lap for Seated Touch Interaction with HMDsVR, seated, touch, on-body
Ubicomp23MI-Poser: Human Body Pose Tracking Using Magnetic and Inertial Sensor Fusion with Metal Interference MitigationEMF, body pose tracking, inverse kinematics, sensor fusion
Ubicomp24Mental-LLM: Leveraging Large Language Models for Mental Health Prediction via Online Text Data
Ubicomp23MoCaPose: Motion Capturing with Textile-integrated Capacitive Sensors in Loose-fitting Smart Garmentsmotion capture, wearable sensing, capacitive sensing, deep learning, motion tracking, smart textile
Ubicomp23N-euro Predictor: A Neural Network Approach for Smoothing and Predicting Motion Trajectoryvision-based interactions, motion-to-photon latency, motion prediction, neural network, perceived jitter and lag
Ubicomp23NF-Heart: A Near-field Non-contact Continuous User Authentication System via Ballistocardiogramcontinuous authentication, ballistocardiogram (BCG), biometrics, non-contact sensing, smart chair
Ubicomp23Naturalistic E-Scooter Maneuver Recognition with Federated Contrastive Rider Interaction LearningIMU, DCT, contrastive learning, asynchronous federated learning, ehavior analysis
ISWC 2023On the Utility of Virtual On-body Acceleration Data for Fine-grained Human Activity RecognitionHAR, virtual, IMU
Ubicomp23PoseSonic: 3D Upper Body Pose Estimation Through Egocentric Acoustic Sensing on Smartglasseshuman pose estimation, acoustic sensing, smart/AR glasses, deep learning, cross-modal supervision
Ubicomp23PrintShear: Shear Input Based on Fingerprint Deformationtouch, finger input
Ubicomp23Privacy-Enhancing Technology and Everyday Augmented Reality: Understanding Bystanders’ Varying Needs for Awareness and ConsentAR, privacy, bystanders, altered reality, extended perception, biometrics
Ubicomp23Radio2Text: Streaming Speech Recognition Using mmWave Radio Signals
Ubicomp23SkinLink: On-body Construction and Prototyping of Reconfigurable Epidermal Interfaces
Ubicomp23Spectral-Loc: Indoor Localization Using Light Spectral Informationindoor localization, spectral information, ambient light
Ubicomp23StructureSense: Inferring Constructive Assembly Structures from User Behaviorstangible user interfaces, TUI, RFID, user modeling, bayesian inference
Ubicomp23Synthetic Smartwatch IMU Data Generation from In-the-wild ASL VideosIMU, synthetic, ASL recognition
Ubicomp23TAO: Context Detection from Daily Activity Patterns Using Temporal Analysis and Ontologybehavioral context recognition, activity recognition, ontology, deep learning
Ubicomp23ThumbAir: In-Air Typing for Head Mounted Displaysmid air, text entry, VR, HMD, user study
ISWC 2023Towards a Haptic Taxonomy of Emotions: Exploring Vibrotactile Stimulation in the Dorsal Region
Ubicomp23TwinkleTwinkle: Interacting with Your Smart Devices by Eye Blinkacoustic sensing, eye blink, signal process
Ubicomp23ViSig: Automatic Interpretation of Visual Body Signals Using On-Body Sensorsvisual signalling, on-body sensors, UWB, IMU, body signals, fallback communication, sports automation, postures, gestures
Ubicomp23VibPath: Two-Factor Authentication with Your Hand's Vibration Response to Unlock Your Phoneuser authentication, vibration, IMU, smartphone, wearables, smartwatch
Ubicomp23Voicify Your UI: Towards Android App Control with Voice Commandsdesign, smartphones, sound-based input, dl parser, UI
Ubicomp23WristAcoustic: Through-Wrist Acoustic Response Based Authentication for Smartwatchessmartwatch authentication, bone conduction, acoustic response
Ubicomp23sUrban: Stable Prediction for Unseen Urban Data from Location-based Sensorsurban computing, location-based data, spatial-temporal prediction, out-of-distribution data

LLM

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026Evaluating the Usage of African-American Vernacular English in Large Language ModelsDeja Dunlap, et al.LLMcs.CL, cs.HC2026.02
arXiv(v1) 2026GUI-Eyes: Tool-Augmented Perception for Visual Grounding in GUI AgentsShuofei Qiao, et al.GUI agent, visual grounding, tool use, perceptioncs.CV, cs.AI2026.01
arXiv(v1) 2025 (WACV 2026)AFRAgent: An Adaptive Feature Renormalization Based High Resolution Aware GUI agentNeeraj Anand, et al.GUI agent, VLM, smartphone automation, multimodalcs.CVWACV 20262025.12
arXiv(v2) 2025Mobile-Agent-v3: Fundamental Agents for GUI AutomationJunyang Wang, et al.GUI agent, mobile, smartphone automation, VLMcs.CV, cs.AI2025.08
arXiv(v6) 2025LLaVA-CoT: Let Vision Language Models Reason Step-by-StepGuowei Xu, et al.LLM, GUI agent, survey, computer usecs.CV17 pages, ICCV 20252025.07
arXiv(v3) 2025Beyond 2:4: exploring V:N:M sparsity for efficient transformer inference on GPUsKang Zhao, et al.LLM, GUI agent, interfacecs.LG, cs.AI2025.06
arXiv(v3) 2025Lai Loss: A Novel Loss for Gradient ControlYuFei LaiLLM, agent, interface, UIcs.LGThe experiment in this article is not very rigorous and may require further testing for its effectiveness2025.05
arXiv(v3) 2025SignLLM: Sign Language Production Large Language ModelsSen Fang, et al.LLM, agent, user interface, HCI, interactioncs.CV, cs.CLwebsite at https://signllm.github.io/2025.04
arXiv(v1) 2024ShowUI: One Vision-Language-Action Model for GUI Visual AgentKevin Qinghong Lin, et al.LLM, UI, vision-language, GUIcs.CV, cs.AI, cs.CL, cs.HCTechnical Report. Github: https://github.com/showlab/ShowUI2024.11
arXiv(v1) 2024A Scalable Communication Protocol for Networks of Large Language ModelsSamuele Marro, et al.LLM, agent, GUI, computer use, interfacecs.AI, cs.LG2024.10
arXiv(v1) 2024In-Band Full-Duplex MIMO Systems for Simultaneous Communications and Sensing: Challenges, Methods, and Future PerspectivesBesma Smida, et al.LLM, GUI agent, computer usecs.IT, cs.ET, eess.SP12 pages, 5 figures, White Paper to appear at IEEE SPM2024.10
arXiv(v2) 2024Massively parallel CMA-ES with increasing populationDavid Redon, et al.LLM, agent, HCI, user interfacecs.DC2024.10
arXiv(v1) 2024OS-ATLAS: A Foundation Action Model for Generalist GUI AgentsZhiyong Wu, et al.LLM, agent, computer use, GUI, foundation modelcs.CL, cs.CV, cs.HC2024.10
arXiv(v1) 2024OrientedFormer: An End-to-End Transformer-Based Oriented Object Detector in Remote Sensing ImagesJiaqi Zhao, et al.LLM, agent, GUI, interfacecs.CVThe paper is accepted by IEEE Transactions on Geoscience and Remote Sensing (TGRS)2024.09
arXiv(v1) 2024Unlocking the Power of Environment Assumptions for Unit ProofsSiddharth Priya, et al.LLM, GUI, agent, surveycs.SE, cs.PLSEFM 20242024.09
arXiv(v2) 2024 (ACL 2025)GUICourse: From General Vision Language Models to Versatile GUI AgentsWentong Chen, et al.GUI agent, VLM, training data, OCR, groundingcs.CV, cs.CL, cs.HCACL 20252024.06
arXiv(v1) 2024Ferret-UI: Grounded Mobile UI Understanding with Multimodal LLMsKeen You, et al.ui, mllm, benchmark, any-resolutioncs.CV, cs.CL, cs.HC2024.04
arXiv(v1) 2024Are You Being Tracked? Discover the Power of Zero-Shot Trajectory Tracing with LLMs!Huanqi Yang, et al.iot, imu, cot, promptcs.CL, cs.AI, cs.HC, cs.LG2024.03
arXiv(v1) 2024Design2Code: How Far Are We From Automating Front-End Engineering?Chenglei Si, et al.llm, auto, googlecs.CL, cs.CV, cs.CY2024.03
arXiv(v2) 2023The Good, The Bad, and Why: Unveiling Emotions in Generative AICheng Li, et al.emotion, prompt, attack, decodecs.AI, cs.CL, cs.HCextension of Large language models understand and can be enhanced by emotional stimuli2023.12
arXiv(v1) (NIPS23)Large Language Model as Attributed Training Data Generator: A Tale of Diversity and BiasYue Yu, et al.synthetic data generationcs.CL, cs.AI, cs.LGarXiv(v2) 2023.10
arXiv(v1) 2023Multimodal Foundation Models: From Specialists to General-Purpose AssistantsChunyuan Li, et al.surveycs.CV, cs.CL2023.09
arXiv(v7) 2023Attention Is All You NeedAshish Vaswani, et al.arxivcs.CL, cs.LG15 pages, 5 figures2023.08

RAG

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026AgenticOCR: Parsing Only What You Need for Efficient Retrieval-Augmented GenerationZhengren Wang, et al.RAGcs.CV, cs.CL2026.02
arXiv(v1) 2026CLFEC: A New Task for Unified Linguistic and Factual Error Correction in paragraph-level Chinese Professional WritingJian Kai, et al.RAGcs.CL2026.02
arXiv(v1) 2026CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM EraZhengqing Yuan, et al.RAGcs.CL, cs.DL2026.02
arXiv(v1) 2026CiteLLM: An Agentic Platform for Trustworthy Scientific Reference DiscoveryMengze Hong, et al.RAGcs.CL, cs.IRAccepted by TheWebConf 2026 Demo Track2026.02
arXiv(v2) 2026MoDora: Tree-Based Semi-Structured Document Analysis SystemBangrui Xu, et al.RAGcs.IR, cs.AI, cs.CL, cs.DB, cs.LGExtension of our SIGMOD 2026 paper. Please refer to source code available at https://github.com/weAIDB/MoDora2026.02
arXiv(v1) 2026Search-P1: Path-Centric Reward Shaping for Stable and Efficient Agentic RAG TrainingTianle Xia, et al.RAGcs.CL, cs.IR, cs.LG2026.02
arXiv(v1) 2026TCM-DiffRAG: Personalized Syndrome Differentiation Reasoning Method for Traditional Chinese Medicine based on Knowledge Graph and Chain of ThoughtJianmin Li, et al.RAGcs.CL, cs.AI2026.02
arXiv(v1) 2026TRIZ-RAGNER: A Retrieval-Augmented Large Language Model for TRIZ-Aware Named Entity Recognition in Patent-Based Contradiction MiningZitong Xu, et al.RAGcs.CL, cs.AI2026.02
arXiv(v1) 2026Truncated Step-Level Sampling with Process Rewards for Retrieval-Augmented ReasoningChris Samarinas, et al.RAGcs.CL, cs.IR2026.02
arXiv(v1) 2026Vectorizing the Trie: Efficient Constrained Decoding for LLM-based Generative Retrieval on AcceleratorsZhengyang Su, et al.RAGcs.IR, cs.CL, cs.LG14 pages, 4 figures2026.02
arXiv(v1) 2026L-RAG: Balancing Context and Retrieval with Entropy-Based Lazy LoadingAuthors TBDRAG, lazy loading, entropy, contextcs.CL, cs.IR2026.01
arXiv(v1) 2025RAGLens: Toward Faithful RAG with Sparse AutoencodersAuthors TBDRAG, hallucination, faithfulness, detectioncs.CL2025.12
arXiv(v1) 2025CDTA: Cross-Document Topic-Aligned Chunking for RAGAuthors TBDRAG, chunking, cross-document, topic alignmentcs.CL, cs.IR0.93 faithfulness on HotpotQA2025.11
arXiv(v1) 2025Agentic RAG for Fintech: Design and EvaluationAuthors TBDRAG, agentic, fintech, query reformulationcs.CL, cs.IR2025.10
arXiv(v1) 2025Practical Code RAG at Scale: Task-Aware Retrieval DesignAuthors TBDRAG, code, retrieval, hybrid, densecs.CL, cs.IRBM25 + dense hybrid2025.10
arXiv(v1) 2025A Systematic Review of Key RAG Systems: Progress, Gaps, and Future DirectionsAuthors TBDRAG, survey, systematic review, knowledge basecs.CL, cs.IR2025.07
arXiv(v1) 2025Late Chunking: Contextual Chunk Embeddings for RAGAuthors TBDRAG, chunking, embedding, contextualcs.CL, cs.IRUpdated July 20252025.07
arXiv(v1) 2025GraphRAG-Bench: When to Use Graphs in RAGAuthors TBDRAG, graph, benchmark, evaluationcs.CL, cs.IR2025.06
arXiv(v1) 2025RAG Survey: Architectures, Enhancements, and Robustness FrontiersAuthors TBDRAG, survey, architecture, robustnesscs.CL, cs.IR2025.06
arXiv(v1) 2025Rethinking Chunk Size for Long-Document Retrieval: Multi-Dataset AnalysisAuthors TBDRAG, chunking, chunk size, retrievalcs.CL, cs.IR64-1024 tokens optimal2025.05
arXiv(v1) 2025A Survey of Multimodal Retrieval-Augmented GenerationZihan Zhao, et al.multimodal, RAG, retrieval, survey, vision-languagecs.CV, cs.CL2025.04
arXiv(v1) 2025RAG Evaluation in the Era of LLMs: A Comprehensive SurveyAuthors TBDRAG, evaluation, benchmark, LLMcs.CL, cs.IR2025.04
arXiv(v1) 2025HiRAG: Retrieval-Augmented Generation with Hierarchical KnowledgeAuthors TBDRAG, hierarchical, knowledge, indexingcs.CL, cs.IR2025.03
arXiv(v1) 2025Ask in Any Modality: A Comprehensive Survey on Multimodal Retrieval-Augmented GenerationZihan Wang, et al.multimodal, RAG, retrieval, survey, LLMcs.CL, cs.IR2025.02
arXiv(v1) 2025GFM-RAG: Graph Foundation Model for Retrieval Augmented GenerationAuthors TBDRAG, graph, foundation model, knowledge graphcs.CL, cs.IR8M params, 60 KGs, 14M triples2025.02
arXiv(v1) 2025KG2RAG: Knowledge Graph-Guided Retrieval Augmented GenerationAuthors TBDRAG, knowledge graph, retrieval, fact-levelcs.CL, cs.IR2025.02
arXiv(v1) 2025RAG-Fusion: Query Expansion and Multi-Source RetrievalAuthors TBDRAG, query expansion, multi-source, fusioncs.CL, cs.IR2025.02
arXiv(v1) 2025Vendi-RAG: Adaptively Trading Off Diversity and Quality in RAGAuthors TBDRAG, diversity, quality, adaptivecs.CL, cs.IR2025.02
arXiv(v1) 2025Agentic RAG: A SurveyAuthors TBDRAG, agentic, survey, LLM agentcs.CL, cs.IR2025.01
arXiv(v1) 2025CG-RAG: Citation Graph RAG for Research Question AnsweringAuthors TBDRAG, citation graph, research QAcs.CL, cs.IR2025.01
arXiv(v1) 2025ChunkRAG: Novel Context-Aware Chunking for RAG SystemsAuthors TBDRAG, chunking, context-aware, retrievalcs.CL, cs.IR2025.01
arXiv(v1) 2024LLM-Augmented Retrieval: Enhancing Retrieval Models Through Language Models and Doc-Level Embeddingrelevant query, doc-Level embedding, embedding-based retrieval, dense retrieval2024.04
arXiv(v6) 2024Health-LLM: Personalized Retrieval-Augmented Disease Prediction SystemQinkai Yu, et al.RAG, XGBoost, AutoMLcs.CL2024.03
arXiv(v1) 2024CRUD-RAG: A Comprehensive Chinese Benchmark for Retrieval-Augmented Generation of Large Language Modelsretrieval-augmented generation, large language models, evaluation2024.02

Agent

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026"Are You Sure?": An Empirical Study of Human Perception Vulnerability in LLM-Driven Agentic SystemsXinfeng Li, et al.LLM, agentcs.HC, cs.AI, cs.CR, cs.SI2026.02
arXiv(v1) 2026AgentDropoutV2: Optimizing Information Flow in Multi-Agent Systems via Test-Time Rectify-or-Reject PruningYutong Wang, et al.LLM, agentcs.AI, cs.CL2026.02
arXiv(v1) 2026E3VA: Enhancing Emotional Expressiveness in Virtual Conversational AgentsAbhishek Kulkarni, et al.LLM, agentcs.HC5 pages2026.02
arXiv(v1) 2026ESAA: Event Sourcing for Autonomous Agents in LLM-Based Software EngineeringElzo Brito dos Santos FilhoLLM, agentcs.AI13 pages, 1 figure, 4 tables. Includes 5 technical appendices2026.02
arXiv(v1) 2026From Flat Logs to Causal Graphs: Hierarchical Failure Attribution for LLM-based Multi-Agent SystemsYawen Wang, et al.LLM, agentcs.AI, cs.SE2026.02
arXiv(v1) 2026HotelQuEST: Balancing Quality and Efficiency in Agentic SearchGuy Hadad, et al.LLM, agentcs.IR, cs.AITo be published in EACL 20262026.02
arXiv(v1) 2026Hybrid LLM-Embedded Dialogue Agents for Learner Reflection: Designing Responsive and Theory-Driven InteractionsParas Sharma, et al.LLM, agentcs.HC, cs.AI2026.02
arXiv(v1) 2026ProductResearch: Training E-Commerce Deep Research Agents via Multi-Agent Synthetic Trajectory DistillationJiangyuan Wang, et al.LLM, agentcs.AI2026.02
arXiv(v1) 2026PseudoAct: Leveraging Pseudocode Synthesis for Flexible Planning and Action Control in Large Language Model AgentsYihan, et al.LLM, agentcs.AI, eess.SY2026.02
arXiv(v1) 2026SafeGen-LLM: Enhancing Safety Generalization in Task Planning for Robotic SystemsJialiang Fan, et al.LLM, agentcs.RO, cs.AI12 pages, 6 figures2026.02
arXiv(v1) 2026The Auton Agentic AI FrameworkSheng Cao, et al.LLM, agentcs.AI2026.02
arXiv(v1) 2026InfiniteWeb: Scalable Web Environment Synthesis for GUI Agent TrainingAuthors TBDagent, web, GUI, training, environmentcs.CV, cs.AI600 tasks across websites2026.01
arXiv(v1) 2026LLM-Based Agentic Systems for Software Engineering: Challenges and OpportunitiesAuthors TBDcode agent, software engineering, surveycs.SE, cs.AI2026.01
arXiv(v1) 2025Beyond Task Completion: Assessment Framework for Agentic AI SystemsAuthors TBDLLM agent, evaluation, framework, benchmarkcs.AI, cs.CL2025.12
arXiv(v1) 2025DeepCode: Open Agentic CodingAuthors TBDcode agent, agentic coding, open sourcecs.SE, cs.AI2025.12
arXiv(v1) 2025LongVideoAgent: Multi-Agent Reasoning with Long VideosJingyi Zhang, et al.multi-agent, video understanding, LLM, reasoningcs.CV, cs.AI2025.12
arXiv(v1) 2025MAR: Multi-Agent Reflexion Improves Reasoning Abilities in LLMsXinyuan Lu, et al.multi-agent, LLM, reflexion, reasoning, debatecs.CL, cs.AI2025.12
arXiv(v1) 2025SWE-RL: Training Superintelligent Software Agents through Self-PlayAuthors TBDcode agent, self-play, RL, software engineeringcs.SE, cs.LG2025.12
arXiv(v1) 2025Agent0: Self-Evolving Agents from Zero Data via Tool-Integrated ReasoningAuthors TBDLLM agent, self-evolving, tool use, reasoningcs.AI, cs.CL18% improvement on math, 24% on general2025.11
arXiv(v1) 2025Building Browser Agents: Architecture, Security, and Practical SolutionsAuthors TBDagent, browser, security, architecturecs.AI, cs.CL2025.11
arXiv(v1) 2025ReMA: Unlocking the Power of Multi-Agent LLM for Reasoning: From Lazy Agents to DeliberationZhiwei Zhang, et al.multi-agent, LLM, reasoning, lazy agent, deliberationcs.AI, cs.CL2025.11
arXiv(v1) 2025BrowserAgent: Web Agents with Human-Inspired Browsing ActionsAuthors TBDagent, browser, web, automationcs.AI, cs.CL2025.10
arXiv(v1) 2025Dark Patterns Impact on LLM-Based Web AgentsAuthors TBDagent, web, dark patterns, decision makingcs.HC, cs.AI2025.10
arXiv(v1) 2025Architecting Resilient LLM AgentsAuthors TBDLLM agent, resilience, architecturecs.AI, cs.CL2025.09
arXiv(v1) 2025A Survey on Code Generation with LLM-based AgentsAuthors TBDcode agent, LLM, code generation, surveycs.SE, cs.AI2025.08
arXiv(v1) 2025LLM-based Agentic Reasoning Frameworks: A Survey from Methods to ScenariosBingxi Zhao, et al.LLM, agent, reasoning, survey, frameworkcs.AI, cs.CL2025.08
arXiv(v1) 2025MCP-Bench: Benchmarking Tool-Using LLM AgentsAuthors TBDagent, benchmark, tool use, MCPcs.AI, cs.CL2025.08
arXiv(v4) 2025Towards Embodied Agentic AI: Review and Classification of LLM- and VLM-Driven Robot Autonomy and InteractionHarshitha Manoj, et al.embodied AI, robot, LLM, VLM, autonomy, surveycs.RO, cs.AI2025.08
arXiv(v1) 2025Understanding Tool-Integrated ReasoningHeng Lin, et al.LLM, agent, tool use, reasoning, TIRcs.LG, cs.AI, stat.ML2025.08
arXiv(v2) 2025Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic ToolsJunde Wu, et al.LLM, agent, agentic reasoning, tool use, frameworkcs.AI, cs.CLACL 20252025.07
arXiv(v3) 2025Embodied AI Agents: Modeling the WorldPascale Fung, et al.embodied AI, world model, VLM, robot, avatarcs.AIMeta AI2025.07
arXiv(v1) 2025Routine: A Structural Planning Framework for LLM Agent System in EnterpriseAuthors TBDLLM agent, planning, enterprise, frameworkcs.AI, cs.CL2025.07
arXiv(v1) 2025Toward a Theory of Agents as Tool-Use Decision-MakersAuthors TBDLLM agent, tool use, theory, decision makingcs.AI, cs.CL2025.06
arXiv(v1) 2025WebRL: Training LLM Web Agents via Self-Evolving Curriculum RLAuthors TBDagent, web, RL, curriculum, self-evolvingcs.LG, cs.AI2025.05
arXiv(v1) 2025AgentRewardBench: Benchmark for LLM Judges in Web Agent EvaluationAuthors TBDagent, benchmark, web, evaluation, rewardcs.AI, cs.CL1,302 trajectories2025.04
arXiv(v1) 2025From LLM Reasoning to Autonomous AI Agents: A Comprehensive ReviewMohamed Amine Ferrag, et al.LLM, agent, survey, autonomous, comprehensive reviewcs.AI, cs.LG2025.04
arXiv(v1) 2024Simultaneous identification of the parameters in the plasticity function for power hardening materials : A Bayesian approachSalih Tatar, et al.LLM, agent, computer use, evaluationmath.NA, math.AP2024.12
arXiv(v2) 2024Feasibility Consistent Representation Learning for Safe Reinforcement LearningZhepeng Cen, et al.LLM, agent, UI, LAUI, interfacecs.LGICML 20242024.06
arXiv(v1) 2024DataDreamer: A Tool for Synthetic Data Generation and Reproducible LLM WorkflowsAjay Patel, et al.pipeline, liarbry, generationcs.CL, cs.LG2024.02
arXiv(v1) 2024More Agents Is All You NeedJunyou Li, et al.multi agent, vote, taskcs.CL, cs.AI, cs.LG2024.02
arXiv(v1) (NIPS23)HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging FaceYongliang Shen, et al.hugging face, APIcs.CL, cs.AI, cs.CV, cs.LG2023.12
arXiv(v1) (NIPS23)CAMEL: Communicative Agents for "Mind" Exploration of Large Language Model SocietyGuohao Li, et al.role play, autonomous, user&assistantcs.AI, cs.CL, cs.CY, cs.LG, cs.MA2023.11(v2)
arXiv(v2) 2023MusicAgent: An AI Agent for Music Understanding and Generation with Large Language ModelsDingyao Yu, et al.pipelinecs.CL, cs.MM, eess.AS2023.10
arXiv(v2) 2023VOYAGER: An Open-Ended Embodied Agent with Large Language ModelsGuanzhi Wang, et al.multi, autonomous, microcraft, gamecs.AI, cs.LG2023.10
arXiv(v3) 2023The Rise and Potential of Large Language Model Based Agents: A SurveyZhiheng Xi, et al.survey, github paper listcs.AI, cs.CL2023.09
arXiv(v3) 2023A Survey on Large Language Model based Autonomous AgentsLei Wang, et al.survey, autonomouscs.AI, cs.CLLatest version is v4(2024.03), double columns. But v3(2023.09) single columns is easy to read.
arXiv(v1) (ICLR24)MetaGPT: Meta Programming for A Multi-Agent Collaborative FrameworkSirui Hong, et al.autonomous system, SOP, multi-agent, frameworkcs.AI, cs.MA

Agentic-RL

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026CUDA Agent: Large-Scale Agentic RL for High-Performance CUDA Kernel GenerationWeinan Dai, et al.LLM, agentic-RLcs.LG, cs.AI2026.02
arXiv(v1) 2026Exploratory Memory-Augmented LLM Agent via Hybrid On- and Off-Policy OptimizationZeyuan Liu, et al.LLM, agentic-RLcs.LG, cs.AIAccepted to ICLR 20262026.02
arXiv(v1) 2026FactGuard: Agentic Video Misinformation Detection via Reinforcement LearningZehao Li, et al.LLM, agentic-RLcs.AI2026.02
arXiv(v1) 2026RF-Agent: Automated Reward Function Design via Language Agent Tree SearchNing Gao, et al.LLM, agentic-RLcs.AI, cs.LG39 pages, 9 tables, 11 figures, Project page see https://github.com/deng-ai-lab/RF-Agent2026.02
arXiv(v1) 2026RUMAD: Reinforcement-Unifying Multi-Agent DebateChao Wang, et al.LLM, agentic-RLcs.AI13 pages, 3 figures2026.02
arXiv(v1) 2026Recycling Failures: Salvaging Exploration in RLVR via Fine-Grained Off-Policy GuidanceYanwei Ren, et al.LLM, agentic-RLcs.AI, cs.CL2026.02
arXiv(v2) 2026Regularized Online RLHF with Generalized Bilinear PreferencesJunghyun Lee, et al.LLM, agentic-RLcs.LG, stat.ML43 pages, 1 table (ver2: more colorful boxes, fixed some typos)2026.02
arXiv(v1) 2026RewardUQ: A Unified Framework for Uncertainty-Aware Reward ModelsDaniel Yang, et al.LLM, agentic-RLcs.LG, cs.AI, cs.CL2026.02
arXiv(v1) 2026Agent Drift: Quantifying Behavioral Degradation in Multi-Agent LLM Systems Over Extended InteractionsAbhishek Rathmulti-agent, behavioral degradation, LLM, agentcs.AI2026.01
arXiv(v5) 2026AgentOrchestra: Orchestrating Multi-Agent Intelligence with the Tool-Environment-Agent(TEA) ProtocolWentao Zhang, et al.hierarchical multi-agent, task solving, LLM, agentcs.AI2026.01
arXiv(v1) 2026Beyond Perfect APIs: A Comprehensive Evaluation of LLM Agents Under Real-World API ComplexityDoyoung Kim, et al.evaluation, benchmark, API complexity, LLM, agentcs.CL, cs.AI26 pages2026.01
arXiv(v1) 2026Beyond Rule-Based Workflows: An Information-Flow-Orchestrated Multi-Agents Paradigm via Agent-to-Agent Communication from CORALXinxing Ren, et al.information flow, multi-agent communication, LLM, agentcs.AI2026.01
arXiv(v2) 2026Cochain: Balancing Insufficient and Excessive Collaboration in LLM Agent WorkflowsJiaxing Zhao, et al.chain-of-collaboration, multi-agent, LLM, agentcs.CL35 pages, 23 figures2026.01
arXiv(v2) 2026Collaborate, Deliberate, Evaluate: How LLM Alignment Affects Coordinated Multi-Agent OutcomesAbhijnan Nath, et al.alignment, multi-agent coordination, LLM, agentcs.CL, cs.AI, cs.LGThis submission is a new version of arXiv:2509.05882v1. with a substantially revised experimental pipeline and new metrics. In particular, collaborator agents are now instantiated independently via separate API calls, rather than generated autoregressively by a single agent. All experimental results are new. Accepted as an extended abstract at AAMAS 20262026.01
arXiv(v1) 2026DemMA: Dementia Multi-Turn Dialogue Agent with Expert-Guided Reasoning and Action SimulationYutong Song, et al.dementia dialogue, multi-turn, medical, LLM, agentcs.MA2026.01
arXiv(v1) 2026EduSim-LLM: An Educational Platform Integrating Large Language Models and Robotic Simulation for BeginnersShenqi Lu, et al.educational platform, robotics, simulation, LLMcs.RO2026.01
arXiv(v1) 2026Game-Theoretic Lens on LLM-based Multi-Agent SystemsJianing Hao, et al.game theory, multi-agent, LLM, agentcs.MA, cs.GT9 pages, 5 figures2026.01
arXiv(v2) 2026 (Spotlight paper of NeurIPS 2025)KARMA: Leveraging Multi-Agent LLMs for Automated Knowledge Graph EnrichmentYuxing Lu, et al.knowledge graph enrichment, multi-agent, LLM, agentcs.CL, cs.AI, cs.CE, cs.DL24 pages, 3 figures, 2 tables2026.01
arXiv(v1) 2026LLM-in-Sandbox Elicits General Agentic IntelligenceDaixuan Cheng, et al.sandbox, agentic intelligence, LLMcs.CL, cs.AIProject Page: https://llm-in-sandbox.github.io2026.01
arXiv(v2) 2026Lifelong Learning of Large Language Model based Agents: A RoadmapJunhao Zheng, et al.lifelong learning, continual learning, LLM, agentcs.AIAccepted to IEEE TPAMI2026.01
arXiv(v1) 2026Nalar: An agent serving frameworkMarco Laju, et al.workflow serving, agent workflows, LLM, agentcs.DC, cs.MA2026.01
arXiv(v1) 2026Orchestral AI: A Framework for Agent OrchestrationAlexander Roman, et al.agent orchestration, framework, LLM, agentcs.AI, astro-ph.IM, hep-ph17 pages, 3 figures. For more information visit https://orchestral-ai.com2026.01
arXiv(v1) 2025PRL: Process Reward Learning Improves LLM ReasoningAuthors TBDagentic RL, PRM, process reward, reasoningcs.LG, cs.AI2026.01
arXiv(v1) 2026ReliabilityBench: Evaluating LLM Agent Reliability Under Production-Like Stress ConditionsAayush Guptareliability evaluation, stress testing, LLM, agentcs.AI18 pages, 5 figures, 8 tables. Evaluates ReAct vs Reflexion across four tool-using domains with perturbation (epsilon) and fault-injection (lambda) stress testing; 1,280 total episodes2026.01
arXiv(v2) 2026SimWorld: An Open-ended Realistic Simulator for Autonomous Agents in Physical and Social WorldsJiawei Ren, et al.simulation, autonomous agent, world model, LLMcs.AI2026.01
arXiv(v2) 2026The Agentic Leash: Extracting Causal Feedback Fuzzy Cognitive Maps with LLMsAkash Kumar Panda, et al.causal extraction, fuzzy cognitive maps, LLM, agentcs.AI, cs.CL, cs.HC, cs.IR15 figures2026.01
arXiv(v2) 2026The Molecular Structure of Thought: Mapping the Topology of Long Chain-of-Thought ReasoningQiguang Chen, et al.chain-of-thought, reasoning topology, LLMcs.CL, cs.AIPreprint2026.01
arXiv(v1) 2026The Two-Stage Decision-Sampling Hypothesis: Understanding the Emergence of Self-Reflection in RL-Trained LLMsZibo Zhao, et al.self-reflection, reinforcement learning, LLM, agentcs.LG, cs.AI2026.01
arXiv(v1) 2026When Numbers Start Talking: Implicit Numerical Coordination Among LLM-Based AgentsAlessio Buscemi, et al.numerical coordination, multi-agent, LLM, agentcs.MA, cs.AI2026.01
arXiv(v2) 2025DoVer: Intervention-Driven Auto Debugging for LLM Multi-Agent SystemsMing Ma, et al.auto debugging, multi-agent, intervention, LLM, agentcs.AI, cs.SE2025.12
arXiv(v1) 2025Enhancing Agentic RL with Progressive Reward Shaping and VSPOAuthors TBDagentic RL, reward shaping, GRPO, tool usecs.LG, cs.AI2025.12
arXiv(v1) 2025From Word to World: Can Large Language Models be Implicit Text-based World Models?Yixia Li, et al.world model, text-based, LLM, agentcs.CL2025.12
arXiv(v2) 2025ToTRL: Unlock LLM Tree-of-Thoughts Reasoning Potential through Puzzles SolvingHaoyuan Wu, et al.tree-of-thought, reasoning, puzzle solving, LLMcs.CL2025.12
arXiv(v2) 2025Understanding LLM Agent Behaviours via Game Theory: Strategy Recognition, Biases and Multi-Agent DynamicsTrung-Kiet Huynh, et al.game theory, agent behavior, LLM, agentcs.MA, cs.AI, cs.GT, cs.LG, math.DS2025.12
arXiv(v2) 2025Large Language Model-based Data Science Agent: A SurveyKe Chen, et al.data science, agent, LLMcs.AI2025.11
arXiv(v1) 2025MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline ParallelismShulin Liu, et al.multi-agent, agentic RL, reasoning, LLM, pipeline parallelismcs.AI10 pages2025.11
arXiv(v1) 2025AGENTRL: Scaling Agentic Reinforcement LearningAuthors TBDagentic RL, RLHF, LLM agent, scalingcs.LG, cs.AIOutperforms GPT-5 and Claude-Sonnet-42025.10
arXiv(v2) 2025BrowserArena: Evaluating LLM Agents on Real-World Web Navigation TasksSagnik Anupam, et al.web navigation, evaluation, real-world, LLM, agentcs.AI, cs.LG2025.10
arXiv(v1) 2025Consistently Simulating Human Personas with Multi-Turn Reinforcement LearningMarwa Abdulhai, et al.persona simulation, multi-turn RL, LLM, agentcs.CL, cs.AI2025.10
arXiv(v3) 2025DS-STAR: Data Science Agent via Iterative Planning and VerificationJaehyun Nam, et al.data science, iterative planning, verification, LLM, agentcs.AI2025.10
arXiv(v1) 2025Deliberate Lab: A Platform for Real-Time Human-AI Social ExperimentsCrystal Qian, et al.social experiments, human-AI interaction, LLM, agentcs.HC, cs.AI2025.10
arXiv(v1) 2025Demystifying Reinforcement Learning in Agentic ReasoningZhaochen Yu, et al.agentic RL, reasoning, LLM, agentcs.CLCode and models: https://github.com/Gen-Verse/Open-AgentRL2025.10
arXiv(v3) 2025DoctorAgent-RL: A Multi-Agent Collaborative Reinforcement Learning System for Multi-Turn Clinical DialogueYichun Feng, et al.clinical dialogue, multi-agent RL, medical, LLM, agentcs.CL2025.10
arXiv(v1) 2025GEM: A Gym for Agentic LLMsZichen Liu, et al.training environment, agentic LLM, gym, agentcs.LG, cs.AI, cs.CL2025.10
arXiv(v3) 2025MLR-Bench: Evaluating AI Agents on Open-Ended Machine Learning ResearchHui Chen, et al.ML research, evaluation, benchmark, LLM, agentcs.LG, cs.AI, cs.CL49 pages, 9 figures. Accepted by NeurIPS 2025 D&B Track2025.10
arXiv(v1) 2025Natural Language Tools: A Natural Language Approach to Tool Calling In Large Language AgentsReid T. Johnson, et al.natural language tool calling, LLM, agentcs.CL31 pages, 7 figures2025.10
arXiv(v1) 2025On Designing Effective RL Reward at Training Time for LLM ReasoningAuthors TBDagentic RL, reward design, reasoning, trainingcs.LG, cs.AI2025.10
arXiv(v1) 2025RLSR: Reinforcement Learning with Supervised RewardAuthors TBDagentic RL, supervised reward, instruction followingcs.LG, cs.AI2025.10
arXiv(v2) 2025SimuRA: A World-Model-Driven Simulative Reasoning Architecture for General Goal-Oriented AgentsMingkai Deng, et al.world model, simulative reasoning, goal-oriented, LLM, agentcs.AI, cs.CL, cs.LG, cs.ROThis submission has been updated to adjust the scope and presentation of the work2025.10
arXiv(v2) 2025 (Am. Statist. (2025) 1-14)A Survey on Large Language Model-based Agents for Statistics and Data ScienceMaojun Sun, et al.data science, statistics, survey, LLM, agentcs.AI, cs.CL, cs.LG, stat.OT2025.09
arXiv(v1) 2025OPPO: Accelerating PPO-based RLHF via Pipeline OverlapAuthors TBDagentic RL, PPO, RLHF, efficiency, overlapcs.LG, cs.AI2025.09
arXiv(v1) 2025RL Foundations for Deep Research Systems: A SurveyAuthors TBDagentic RL, deep research, survey, post-DeepSeekcs.LG, cs.AIPost Feb 2025 papers2025.09
arXiv(v1) 2025Reward Hacking Mitigation using Verifiable Composite RewardsAuthors TBDRLHF, reward hacking, RLVR, verificationcs.LG, cs.AI2025.09
arXiv(v1) 2025SAMULE: Self-Learning Agents Enhanced by Multi-level ReflectionYubin Ge, et al.self-learning, multi-level reflection, LLM, agentcs.AIAccepted at EMNLP 2025 Main Conference2025.09
arXiv(v1) 2025Teaching LLMs to Plan: Logical Chain-of-Thought Instruction Tuning for Symbolic PlanningPulkit Verma, et al.logical chain-of-thought, symbolic planning, LLMcs.AI, cs.CL2025.09
arXiv(v1) 2025The Landscape of Agentic Reinforcement Learning for LLMs: A SurveyAuthors TBDagentic RL, survey, LLM, POMDPcs.LG, cs.AI2025.09
arXiv(v1) 2025Where LLM Agents Fail and How They can Learn From FailuresKunlun Zhu, et al.failure detection, learning from failures, LLM, agentcs.AI2025.09
arXiv(v1) 2025iStar: Agentic Reinforcement Learning with Implicit Step RewardsAuthors TBDagentic RL, credit assignment, implicit PRMcs.LG, cs.AI2025.09
arXiv(v3) 2025MLE-STAR: Machine Learning Engineering Agent via Search and Targeted RefinementJaehyun Nam, et al.ML engineering, code generation, search, LLM, agentcs.LG2025.08
arXiv(v1) 2025Agent Safety Alignment via Reinforcement LearningZeyang Sha, et al.safety alignment, reinforcement learning, LLM, agentcs.AI, cs.CR2025.07
arXiv(v1) 2025AgentMesh: A Cooperative Multi-Agent Generative AI Framework for Software Development AutomationSourena Khanzadehmulti-agent, software development, code generation, LLM, agentcs.SE, cs.AI2025.07
arXiv(v1) 2025Technical Survey of RL Techniques for Large Language ModelsAuthors TBDagentic RL, survey, PPO, DPO, GRPOcs.LG, cs.AI2025.07
arXiv(v2) 2025ToolACE: Winning the Points of LLM Function CallingWeiwen Liu, et al.function calling, tool use, LLM, agentcs.LG, cs.AI, cs.CL21 pages, 22 figures2025.07
arXiv(v1) 2025A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI AutonomyHenry Peng Zou, et al.human-agent systems, collaboration, LLM, agentcs.AI, cs.CL, cs.HC, cs.LG, cs.MA2025.06
arXiv(v2) 2025Beyond Self-Talk: A Communication-Centric Survey of LLM-Based Multi-Agent SystemsBingyu Yan, et al.multi-agent communication, survey, LLM, agentcs.MA, cs.CL2025.06
arXiv(v1) 2025Enhancing Decision-Making of Large Language Models via Actor-CriticHeng Dong, et al.decision-making, actor-critic, LLM, agentcs.CL, cs.AIForty-second International Conference on Machine Learning (ICML 2025)2025.06
arXiv(v1) 2025GENMANIP: LLM-driven Simulation for Generalizable Instruction-Following ManipulationNing Gao, et al.simulation, robotic manipulation, LLM, agentcs.RO2025.06
arXiv(v1) 2025OPT-BENCH: Evaluating LLM Agent on Large-Scale Search Spaces Optimization ProblemsXiaozhe Li, et al.optimization, benchmark, evaluation, LLM, agentcs.AI, cs.LG2025.06
arXiv(v1) 2025RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World EnvironmentsYuchuan Fu, et al.security evaluation, benchmark, LLM, agentcs.CR, cs.AI12 pages, 8 figures2025.06
arXiv(v1) 2025Sailing by the Stars: Survey on Reward Models and Learning StrategiesAuthors TBDagentic RL, reward model, survey, learningcs.LG, cs.AI2025.06
arXiv(v1) 2025ML-Agent: Reinforcing LLM Agents for Autonomous Machine Learning EngineeringZexi Liu, et al.agentic RL, ML engineering, LLM, agentcs.CL, cs.AI, cs.LG2025.05
arXiv(v1) 2025Multi-Agent Systems for Robotic Autonomy with LLMsJunhong Chen, et al.multi-agent, robotics, LLM, agentcs.RO, cs.AI11 pages, 2 figures, 5 tables, submitted for publication2025.05
arXiv(v1) 2025Training LLM-Based Agents with Synthetic Self-Reflected Trajectories and Partial MaskingYihan Chen, et al.synthetic trajectories, self-reflection, training, LLM, agentcs.CL2025.05
arXiv(v1) 2025Agentic Reasoning and Tool Integration for LLMs via Reinforcement LearningJoykirat Singh, et al.agentic RL, tool integration, reasoning, LLMcs.AI2025.04
arXiv(v1) 2025Comprehensive Survey of Reward Models: Taxonomy and ApplicationsAuthors TBDagentic RL, reward model, survey, taxonomycs.LG, cs.AI2025.04
arXiv(v1) 2025DPO Meets PPO: Reinforced Token Optimization for RLHFHan Zhong, et al.agentic RL, DPO, PPO, RLHF, token-levelcs.LG, cs.AIRTO framework2025.04
arXiv(v1) 2025Hierarchical Multi-Step Reward Models for Enhanced ReasoningAuthors TBDagentic RL, reward model, hierarchical, reasoningcs.LG, cs.AI2025.03
arXiv(v1) 2025Look Before You Leap: Using Serialized State Machine for Language Conditioned Robotic ManipulationTong Mu, et al.finite state machine, robotic manipulation, LLMcs.RO, cs.AI7 pages, 4 figures2025.03
arXiv(v1) 2025SafePlan: Leveraging Formal Logic and Chain-of-Thought Reasoning for Enhanced Safety in LLM-based Robotic Task PlanningIke Obi, et al.formal logic, chain-of-thought, safety, robotic planning, LLMcs.RO2025.03
arXiv(v2) 2025Web Agents with World Models: Learning and Leveraging Environment Dynamics in Web NavigationHyungjoo Chae, et al.web navigation, environment dynamics, LLM, agentcs.CLICLR 20252025.03
arXiv(v1) 2025Every Software as an Agent: Blueprint and Case StudyMengwei Xusoftware agent, autonomous, LLMcs.SE, cs.AI2025.02
arXiv(v2) 2025Flow: Modularized Agentic Workflow AutomationBoye Niu, et al.workflow generation, modular, LLM, agentcs.AI, cs.LG, cs.MA2025.02
arXiv(v1) 2025Policy Learning with a Natural Language Action Space: A Causal ApproachBohan Zhang, et al.policy learning, natural language action, LLM, agentcs.CL2025.02
arXiv(v1) 2025Process Reward Models for LLM Agents: Practical FrameworkAuthors TBDPRM, reward model, LLM agent, RLHFcs.LG, cs.AIInversePRM2025.02
arXiv(v1) 2025Provably Efficient Online RLHF with One-Pass Reward ModelingAuthors TBDonline RLHF, reward modeling, efficiencycs.LG, cs.AI2025.02
arXiv(v1) 2025 (Proceedings of the 2024 IEEE International Japan-Africa Conference on Electronics communications and Computations (JAC ECC))Guided Code Generation with LLMs: A Multi-Agent Framework for Complex Code TasksAmr Almorsi, et al.multi-agent, code generation, LLM, agentcs.AI4 pages, 3 figures2025.01
arXiv(v1) 2025REINFORCE++: Critic-Free Policy Optimization with Global Advantage NormalizationAuthors TBDagentic RL, REINFORCE, critic-free, GRPOcs.LG, cs.AIOutperforms PPO2025.01
arXiv(v4) 2024Planning with Multi-Constraints via Collaborative Language AgentsCong Zhang, et al.meta-task planning, multi-agent, LLM, agentcs.AI, cs.CL, cs.LG2024.12
arXiv(v2) 2024AutoWebGLM: A Large Language Model-based Web Navigating AgentHanyu Lai, et al.web navigation, browsing, LLM, agentcs.CLAccepted to KDD 20242024.10

MLLM

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v2) 2026A Taxonomy of Human--MLLM Interaction in Early-Stage Sketch-Based Design IdeationWeiayn Shi, et al.VLMcs.HCAccepted at CHI 2026 Posters2026.02
arXiv(v1) 2026Beyond Dominant Patches: Spatial Credit Redistribution For Grounded Vision-Language ModelsNiamul Hassan Samin, et al.VLMcs.CV, cs.AI2026.02
arXiv(v1) 2026Beyond Static Artifacts: A Forensic Benchmark for Video Deepfake Reasoning in Vision Language ModelsZheyuan Gu, et al.VLMcs.CV, cs.AI16 pages, 9 figures. Submitted to CVPR 20262026.02
arXiv(v1) 2026Causal Decoding for Hallucination-Resistant Multimodal Large Language ModelsShiwei Tan, et al.VLMcs.LG, cs.AI, cs.CVPublished in Transactions on Machine Learning Research (TMLR), 20262026.02
arXiv(v1) 2026Dynamic Multimodal Activation Steering for Hallucination Mitigation in Large Vision-Language ModelsJianghao Yin, et al.VLMcs.CV, cs.AIAccepted by ICLR 20262026.02
arXiv(v1) 2026GuardAlign: Test-time Safety Alignment in Multimodal Large Language ModelsXingyu Zhu, et al.VLMcs.CV, cs.MMICLR 20262026.02
arXiv(v1) 2026HulluEdit: Single-Pass Evidence-Consistent Subspace Editing for Mitigating Hallucinations in Large Vision-Language ModelsYangguang Lin, et al.VLMcs.CVaccepted at CVPR 20262026.02
arXiv(v1) 2026Large Multimodal Models as General In-Context ClassifiersMarco Garosi, et al.VLMcs.CVCVPR Findings 2026. Project website at https://circle-lmm.github.io/2026.02
arXiv(v1) 2026Look Carefully: Adaptive Visual Reinforcements in Multimodal Large Language Models for Hallucination MitigationXingyu Zhu, et al.VLMcs.CVICLR 20262026.02
arXiv(v1) 2026MediX-R1: Open Ended Medical Reinforcement LearningSahal Shaji Mullappilly, et al.VLMcs.CV2026.02
arXiv(v1) 2026NoLan: Mitigating Object Hallucinations in Large Vision-Language Models via Dynamic Suppression of Language PriorsLingfeng Ren, et al.VLMcs.CV, cs.AI, cs.CLCode: https://github.com/lingfengren/NoLan2026.02
arXiv(v1) 2026See It, Say It, Sorted: An Iterative Training-Free Framework for Visually-Grounded Multimodal Reasoning in LVLMsYongchang Zhang, et al.VLMcs.CVCVPR2026 Accepted2026.02
arXiv(v1) 2026Seeing Graphs Like Humans: Benchmarking Computational Measures and MLLMs for Similarity AssessmentSeokweon Jung, et al.VLMcs.HC21 pages including 1 page of appendix, 9 figures, 4 tables2026.02
arXiv(v1) 2026Suppressing Prior-Comparison Hallucinations in Radiology Report Generation via Semantically Decoupled Latent SteeringAo Li, et al.VLMcs.CV15 pages, 5 figures2026.02
arXiv(v1) 2026SurGo-R1: Benchmarking and Modeling Contextual Reasoning for Operative Zone in Surgical VideoGuanyi Qin, et al.VLMcs.CV, cs.AI2026.02
arXiv(v1) 2026Toward Guarantees for Clinical Reasoning in Vision Language Models via Formal VerificationVikash Singh, et al.VLMcs.CV, cs.AI, cs.CL, cs.LO2026.02
arXiv(v1) 2026VAUQ: Vision-Aware Uncertainty Quantification for LVLM Self-EvaluationSeongheon Park, et al.VLMcs.CV, cs.AI, cs.CL2026.02
arXiv(v1) 2026Ground What You See: Hallucination-Resistant MLLMs via Caption FeedbackAuthors TBDMLLM, hallucination, caption feedback, groundingcs.CV, cs.CL2026.01
arXiv(v1) 2026Innovator-VL: A Multimodal Large Language Model for Scientific DiscoveryZichen Wen, et al.scientific discovery, MLLM, vision-languagecs.CV, cs.AIInnovator-VL tech report2026.01
arXiv(v2) 2026MGPC: Multimodal Network for Generalizable Point Cloud Completion With Modality Dropout and Progressive DecodingJiangyuan Liu, et al.point cloud completion, multimodal, MLLMcs.CVCode and dataset are available at https://github.com/L-J-Yuan/MGPC2026.01
arXiv(v1) 2026Multimodal In-context Learning for ASR of Low-resource LanguagesZhaolin Li, et al.multimodal in-context learning, ASR, low-resource, MLLMcs.CL, cs.AIUnder review2026.01
arXiv(v3) 2026Omni-AVSR: Towards Unified Multimodal Speech Recognition with Large Language ModelsUmberto Cappellazzo, et al.unified speech recognition, multimodal, MLLMeess.AS, cs.CV, cs.SDAccepted to IEEE ICASSP 2026 (camera-ready version). Project website (code and model weights): https://umbertocappellazzo.github.io/Omni-AVSR/2026.01
arXiv(v2) 2026Table as a Modality for Large Language ModelsLiyao Li, et al.table modality, structured data, MLLMcs.CL, cs.AIAccepted to NeurIPS 20252026.01
arXiv(v1) 2026The Paradigm Shift: A Comprehensive Survey on Large Vision Language Models for Multimodal Fake News DetectionWei Ai, et al.fake news detection, vision-language, MLLM, surveycs.AI, cs.CV2026.01
arXiv(v3) 2026UniVideo: Unified Understanding, Generation, and Editing for VideosCong Wei, et al.video understanding, generation, editing, unified, MLLMcs.CVProject Website https://congwei1230.github.io/UniVideo/2026.01
arXiv(v1) 2026VLingNav: Embodied Navigation with Adaptive Reasoning and Visual-Assisted Linguistic MemoryShaoan Wang, et al.embodied navigation, adaptive reasoning, MLLM, vision-languagecs.RO, cs.CVProject page: https://wsakobe.github.io/VLingNav-web/2026.01
arXiv(v1) 2026VideoLoom: A Video Large Language Model for Joint Spatial-Temporal UnderstandingJiapeng Shi, et al.video understanding, spatial-temporal, MLLM, vision-languagecs.CV2026.01
arXiv(v1) 2025A Medical Multimodal Diagnostic Framework Integrating Vision-Language Models and Logic Tree ReasoningZelin Zang, et al.medical diagnosis, logic tree reasoning, vision-language, MLLMcs.AI2025.12
arXiv(v1) 2025DiffThinker: Towards Generative Multimodal Reasoning with Diffusion ModelsZefeng He, et al.generative reasoning, diffusion model, MLLMcs.CVProject page: https://diffthinker-project.github.io2025.12
arXiv(v1) 2025From Indoor to Open World: Revealing the Spatial Reasoning Gap in MLLMsAuthors TBDMLLM, spatial reasoning, benchmark, open worldcs.CV, cs.CL2025.12
arXiv(v1) 2025Kling-Omni Technical ReportKling Team, et al.video generation, multimodal synthesis, MLLMcs.CVKling-Omni Technical Report2025.12
arXiv(v1) 2025Lemon: A Unified and Scalable 3D Multimodal Model for Universal Spatial UnderstandingYongyuan Liang, et al.3D understanding, point cloud, spatial understanding, MLLMcs.CV, cs.AI2025.12
arXiv(v3) 2025 (Proc. 2025 IEEE 8th International Conference on Multimedia Information Processing and Retrieval (MIPR), pp. 456-462, 2025)MedChat: A Multi-Agent Framework for Multimodal Diagnosis with Large Language ModelsPhilip R. Liu, et al.multi-agent, medical diagnosis, MLLMcs.MA, cs.AI, cs.CV, cs.LG2025.12
arXiv(v2) 2025TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement LearningTao Wu, et al.temporal understanding, reinforcement learning, MLLM, videocs.CV2025.12
arXiv(v1) 2025MVU-Eval: Multi-Video Understanding Evaluation for MLLMsAuthors TBDMLLM, multi-video, evaluation, benchmarkcs.CV, cs.CL2025.11
arXiv(v1) 2025 (Proceedings of the Conference on Language Modeling (COLM 2025))REM: Evaluating LLM Embodied Spatial Reasoning through Multi-Frame TrajectoriesJacob Thompson, et al.embodied spatial reasoning, trajectory, MLLMcs.LG, cs.AI, cs.CV2025.11
arXiv(v1) 2025Seeing is Believing: Rich-Context Hallucination Detection via Backward Visual GroundingAuthors TBDMLLM, hallucination, detection, visual groundingcs.CV, cs.CLOutperforms GPT-4o2025.11
arXiv(v1) 2025SpatialThinker: Reinforcing 3D Reasoning in Multimodal LLMs via Spatial RewardsAuthors TBDMLLM, 3D reasoning, spatial, RLcs.CV, cs.CLOutperforms GPT-4o2025.11
arXiv(v1) 2025MT-Video-Bench: Video Understanding Benchmark for MLLMs in Multi-Turn DialoguesAuthors TBDMLLM, video understanding, benchmark, multi-turncs.CV, cs.CL2025.10
arXiv(v1) 2025MemVR: Memory-Space Visual Retracing for Hallucination Mitigation in MLLMsAuthors TBDMLLM, hallucination, mitigation, memorycs.CV, cs.CLPlug-and-play2025.10
arXiv(v2) 2025Revealing Multimodal Causality with Large Language ModelsJin Li, et al.causal discovery, MLLM, multimodal causalitycs.LG, cs.AIAccepted at NeurIPS 20252025.10
arXiv(v1) 2025Video-STR: Reinforcing MLLMs in Video Spatio-Temporal Reasoning with Relation GraphWentao Wang, et al.spatio-temporal reasoning, relation graph, MLLM, videocs.AI2025.10
arXiv(v1) 2025Two Causes, Not One: Rethinking Omission and Fabrication Hallucinations in MLLMsAuthors TBDMLLM, hallucination, omission, fabricationcs.CV, cs.CL2025.09
arXiv(v1) 2025VIRAL: Visual Representation Alignment for MLLMsAuthors TBDMLLM, visual alignment, fine-grained understandingcs.CV, cs.CL2025.09
arXiv(v1) 2025Bifrost-1: Bridging Multimodal LLMs and Diffusion Models with Patch-level CLIP LatentsHan Lin, et al.diffusion model, patch-level CLIP, image generation, MLLMcs.CV, cs.AI, cs.CLProject Page: https://bifrost-1.github.io2025.08
arXiv(v1) 2025Grounding the Ungrounded: Spectral-Graph Framework for Quantifying Hallucinations in MLLMsAuthors TBDMLLM, hallucination, grounding, detectioncs.CV, cs.CL2025.08
arXiv(v1) 2025Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A SurveyAuthors TBDMLLM, VLA, robotics, manipulation, surveycs.RO, cs.CV2025.08
arXiv(v1) 2025Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge PromptingMiaosen Luo, et al.affective computing, emotion recognition, MLLMcs.AI, cs.LG2025.08
arXiv(v1) 2025RynnEC: Bringing MLLMs into Embodied WorldAuthors TBDMLLM, embodied, video, spatial reasoningcs.CV, cs.RO2025.08
arXiv(v3) 2025SimVecVis: A Dataset for Enhancing MLLMs in Visualization UnderstandingCan Liu, et al.visualization understanding, dataset, MLLMcs.HC, cs.CV2025.07
arXiv(v2) 2025UniCode$^2$: Cascaded Large-scale Codebooks for Unified Multimodal Understanding and GenerationYanzhe Chen, et al.codebook, multimodal generation, MLLMcs.CV, cs.MM19 pages, 5 figures2025.07
arXiv(v1) 2025CLiViS: Unleashing Cognitive Map through Linguistic-Visual Synergy for Embodied Visual ReasoningKailing Li, et al.embodied visual reasoning, cognitive map, MLLMcs.CV, cs.AI, cs.CL2025.06
arXiv(v1) 2025Insight-V: Exploring Long-Chain Visual Reasoning with MLLMsAuthors TBDMLLM, visual reasoning, long-chain, CVPRcs.CVCVPR 20252025.06
arXiv(v2) 2025LLaDA-V: Large Language Diffusion Models with Visual Instruction TuningZebin You, et al.diffusion model, visual instruction tuning, MLLMcs.LG, cs.CL, cs.CVProject page and codes: \url{https://ml-gsai.github.io/LLaDA-V-demo/}2025.06
arXiv(v2) 2025LLaVA-ST: A Multimodal Large Language Model for Fine-Grained Spatial-Temporal UnderstandingHongyu Li, et al.spatial-temporal understanding, MLLM, vision-languagecs.CVAccepted by CVPR20252025.06
arXiv(v1) 2025 (CVPR 2025)LLaVA-ST: Multimodal LLM for Fine-Grained Spatial-Temporal UnderstandingAuthors TBDMLLM, spatial-temporal, video, CVPRcs.CVCVPR 20252025.06
arXiv(v1) 2025Manager: Aggregating Insights from Unimodal Experts in VLMs and MLLMsAuthors TBDMLLM, VLM, unimodal experts, fusioncs.CV, cs.CL2025.06
arXiv(v1) 2025MedTVT-R1: A Multimodal LLM Empowering Medical Reasoning and DiagnosisYuting Zhang, et al.medical reasoning, diagnosis, MLLMeess.IV, cs.CL, cs.CV, q-bio.QM2025.06
arXiv(v1) 2025Multimodal Tabular Reasoning with Privileged Structured InformationJun-Peng Jiang, et al.tabular reasoning, structured information, MLLMcs.LG, cs.AI, cs.CL, cs.CV2025.06
arXiv(v1) 2025Pts3D-LLM: Studying the Impact of Token Structure for 3D Scene Understanding With Large Language ModelsHugues Thomas, et al.3D scene understanding, point cloud, token structure, MLLMcs.CVMain paper and appendix2025.06
CVPR25 (CVPR 2025)Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention ReweightingAuthors TBDMLLM, hallucination, attention, CVPRcs.CVCVPR 20252025.06
arXiv(v3) 2025SpatialLLM: A Compound 3D-Informed Design towards Spatially-Intelligent Large Multimodal ModelsWufei Ma, et al.3D-informed, spatial intelligence, MLLMcs.CVCVPR 2025 highlight2025.06
arXiv(v1) 2025Structured Attention Matters to Multimodal LLMs in Document UnderstandingChang Liu, et al.document understanding, structured attention, MLLMcs.CL, cs.AI, cs.IR2025.06
arXiv(v2) 2025Watch and Listen: Understanding Audio-Visual-Speech Moments with Multimodal LLMZinuo Li, et al.audio-visual-speech, multimodal, MLLM, video understandingcs.CL2025.06
arXiv(v3) 2025Cosmos-Reason1: From Physical Common Sense To Embodied ReasoningNVIDIA, et al.physical common sense, embodied reasoning, MLLMcs.AI, cs.CV, cs.LG, cs.RO2025.05
arXiv(v1) 2025HoloLLM: Multisensory Foundation Model for Language-Grounded Human Sensing and ReasoningChuhao Zhou, et al.multisensory, human sensing, reasoning, MLLMcs.CV, cs.AI, cs.CL, cs.LG, cs.MM18 pages, 13 figures, 6 tables2025.05
arXiv(v2) 2025Kubrick: Multimodal Agent Collaborations for Synthetic Video GenerationLiu He, et al.video generation, multi-agent collaboration, MLLMcs.CV, cs.GR, cs.MMAccepted by CVPR 2025 AI4CC Workshop2025.05
arXiv(v1) 2025MedBridge: Bridging Foundation Vision-Language Models to Medical Image DiagnosisAuthors TBDMLLM, medical, VLM, diagnosiscs.CV2025.05
arXiv(v1) 2025VideoLLM Benchmarks and Evaluation: A SurveyAuthors TBDMLLM, video, benchmark, evaluation, surveycs.CV, cs.CL2025.05
arXiv(v2) 2025Can Large Language Models Help Multimodal Language Analysis? MMLA: A Comprehensive BenchmarkHanlei Zhang, et al.multimodal language analysis, MLLM, semanticscs.CL, cs.AI, cs.MM23 pages, 5 figures2025.04
arXiv(v2) 2025Dual Diffusion for Unified Image Generation and UnderstandingZijie Li, et al.dual diffusion, unified generation, understanding, MLLMcs.CV, cs.AI, cs.LG2025.04
arXiv(v1) 2025Multimodal LLMs for OCR, OCR Post-Correction, and Named Entity Recognition in Historical DocumentsGavin Greif, et al.OCR, historical documents, named entity recognition, MLLMcs.CL, cs.AI, cs.DL2025.04
arXiv(v1) 2025Socratic Chart: Cooperating Multiple Agents for Robust SVG Chart UnderstandingYuyang Ji, et al.chart understanding, SVG, multi-agent, MLLMcs.CV2025.04
arXiv(v1) 2025VLM-R1: A Stable and Generalizable R1-Style Large VLMAuthors TBDMLLM, VLM, reasoning, R1-stylecs.CV, cs.CL2025.04
arXiv(v1) 2025MLLM-For3D: Adapting Multimodal Large Language Model for 3D Reasoning SegmentationAuthors TBDMLLM, 3D, segmentation, reasoningcs.CV2025.03
arXiv(v1) 2025MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMsErik Daxberger, et al.MLLM, 3D, spatial, understanding, benchmarkcs.CV, cs.CLICCV 20252025.03
arXiv(v1) 2025Med3DVLM: An Efficient Vision-Language Model for 3D Medical Image AnalysisAuthors TBDMLLM, medical, 3D, VLM, CTcs.CV2025.03
arXiv(v1) 2025Mobile-VideoGPT: Fast and Accurate Video Understanding Language ModelAuthors TBDMLLM, video, mobile, efficientcs.CV, cs.CL2025.03
arXiv(v1) 2025R1-Zero's Aha Moment in Visual Reasoning on a 2B Non-SFT ModelAuthors TBDMLLM, visual reasoning, R1-Zero, emergentcs.CV, cs.CL2025.03
arXiv(v1) 2025SpaceVLLM: Endowing MLLM with Spatio-Temporal Video GroundingAuthors TBDMLLM, video, spatio-temporal, groundingcs.CV, cs.CL2025.03
arXiv(v1) 2025Vision-R1: Incentivizing Reasoning Capability in MLLMsAuthors TBDMLLM, reasoning, visual reasoningcs.CV, cs.CL2025.03
arXiv(v2) 2025Does Table Source Matter? Benchmarking and Improving Multimodal Scientific Table Understanding and ReasoningBohao Yang, et al.table understanding, scientific data, MLLMcs.CL2025.02
arXiv(v4) 2025LMFusion: Adapting Pretrained Language Models for Multimodal GenerationWeijia Shi, et al.multimodal generation, LLM adaptation, MLLMcs.CL, cs.AI, cs.CV, cs.LGName change: LlamaFusion to LMFusion2025.02
arXiv(v1) 2025Multimodal Large Language Models for Text-rich Image Understanding: A Comprehensive ReviewPei Fu, et al.text-rich image understanding, MLLM, vision-languagecs.CV2025.02
arXiv(v1) 2025Visual Perception Token for Multimodal Large Language ModelsAuthors TBDMLLM, visual perception, token, autonomous controlcs.CV, cs.CL829k training samples2025.02
arXiv(v1) 2025Weak Supervision Dynamic KL-Weighted Diffusion Models Guided by Large Language ModelsJulian Perry, et al.diffusion model, LLM guidance, weak supervision, image generationcs.CL2025.02
arXiv(v1) 2025Bridging Visualization and Optimization: Multimodal Large Language Models on Graph-Structured Combinatorial OptimizationJie Zhao, et al.graph-structured optimization, combinatorial, MLLMcs.AI, cs.LG2025.01
arXiv(v1) 2025Exploring the Role of Explicit Temporal Modeling in Multimodal Large Language Models for Video UnderstandingYun Li, et al.temporal modeling, video understanding, MLLMcs.CV, cs.CL2025.01
arXiv(v5) 2025Harnessing Multimodal Large Language Models for Multimodal Sequential RecommendationYuyang Ye, et al.sequential recommendation, MLLM, multimodalcs.IR, cs.AI2025.01

Memory

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026Contextual Memory Virtualisation: DAG-Based State Management and Structurally Lossless Trimming for LLM AgentsCosmo SantoniLLM, memorycs.SE, cs.AI, cs.HC, cs.OS11 pages. 6 figures. Introduces a DAG-based state management system for LLM agents. Evaluation on 76 coding sessions shows up to 86% token reduction (mean 20%) while remaining economically viable under prompt caching. Includes reference implementation for Claude Code2026.02
arXiv(v1) 2026A Dynamic Retrieval-Augmented Generation System with Selective Memory and RemembranceOkan Bursadynamic RAG, selective memory, retrieval, LLMcs.IR, cs.AI6 Pages, 2 figures2026.01
arXiv(v1) 2026Active Context Compression: Autonomous Memory Management in LLM AgentsAuthors TBDmemory, compression, context, autonomous, Focuscs.CL, cs.AI22.7% token savings2026.01
arXiv(v1) 2025Agentic Memory: Learning Unified Long-Term and Short-Term Memory ManagementAuthors TBDmemory, unified, long-term, short-term, managementcs.AI, cs.CL2026.01
arXiv(v1) 2026Beyond Dialogue Time: Temporal Semantic Memory for Personalized LLM AgentsAuthors TBDmemory, temporal, semantic, personalizedcs.CL, cs.AIDurative memory, Zep architecture2026.01
arXiv(v1) 2026Beyond Static Summarization: Proactive Memory Extraction for LLM AgentsChengyuan Yang, et al.memory, extraction, summarization, agentcs.CL, cs.AI2026.01
arXiv(v1) 2026Continuum Memory Architectures for Long-Horizon LLM AgentsAuthors TBDmemory, continuum, long-horizon, consolidationcs.AI, cs.CLEpisodic-to-semantic conversion2026.01
arXiv(v2) 2026Cost and accuracy of long-term memory in Distributed Multi-Agent Systems based on Large Language ModelsBenedict Wolff, et al.graph memory, distributed multi-agent, LLMcs.IR23 pages, 4 figures, 7 tables2026.01
arXiv(v1) 2026Explore with Long-term Memory: A Benchmark and Multimodal LLM-based Reinforcement Learning Framework for Embodied ExplorationSen Wang, et al.memory, benchmark, multimodal, RL, MLLMcs.AI, cs.CVOur dataset and code will be released at our \href{https://wangsen99.github.io/papers/lmee/}{website}2026.01
arXiv(v1) 2026Fine-Mem: Fine-Grained Feedback Alignment for Long-Horizon Memory ManagementWeitao Ma, et al.memory, management, feedback, alignment, agentcs.CL18 pages, 5 figures2026.01
arXiv(v3) 2026HaluMem: Evaluating Hallucinations in Memory Systems of AgentsDing Chen, et al.memory, hallucination, evaluation, agentcs.CL2026.01
arXiv(v1) 2026HiMeS: Hippocampus-inspired Memory System for Personalized AI AssistantsHailong Li, et al.memory, personalized assistant, hippocampus, long-termcs.AI2026.01
arXiv(v1) 2026HiMem: Hierarchical Long-Term Memory for LLM Long-Horizon AgentsNingning Zhang, et al.memory, long-term, hierarchical, agent, LLMcs.AI2026.01
arXiv(v2) 2026Intrinsic Memory Agents: Heterogeneous Multi-Agent LLM Systems through Structured Contextual MemorySizhe Yuen, et al.multi-agent, structured memory, LLM, agentcs.AI2026.01
arXiv(v1) 2026LLMs Can't Play Hangman: On the Necessity of a Private Working Memory for Language AgentsDavide Baldelli, et al.working memory, evaluation, agent, LLMcs.CL2026.01
arXiv(v2) 2026LiCoMemory: Lightweight and Cognitive Agentic Memory for Efficient Long-Term ReasoningZhengjun Huang, et al.cognitive memory, long-term reasoning, LLM, agentcs.IR2026.01
arXiv(v1) 2026MAGMA: A Multi-Graph based Agentic Memory Architecture for AI AgentsDongming Jiang, et al.multi-graph, agentic memory, LLM, agentcs.AI2026.01
arXiv(v1) 2026Mem-Gallery: Benchmarking Multimodal Long-Term Conversational Memory for MLLM AgentsYuanchen Bei, et al.memory, multimodal, conversational, MLLM benchmarkcs.CL, cs.AI34 pages, 18 figures2026.01
arXiv(v1) 2026MemBuilder: Reinforcing LLMs for Long-Term Memory Construction via Attributed Dense RewardsAuthors TBDmemory, long-term, construction, dense rewardcs.CL, cs.AI2026.01
arXiv(v1) 2026RealMem: Benchmarking LLMs in Real-World Memory-Driven InteractionHaonan Bian, et al.memory, benchmark, interaction, agentcs.CL, cs.AI2026.01
arXiv(v3) 2026SimpleMem: Efficient Lifelong Memory for LLM AgentsJiaqi Liu, et al.lifelong memory, efficient, LLM, agentcs.AI2026.01
arXiv(v1) 2026SwiftMem: Fast Agentic Memory via Query-aware IndexingAnxin Tian, et al.memory, retrieval, indexing, agent, LLMcs.CL, cs.AI2026.01
arXiv(v3) 2026TeleMem: Building Long-Term and Multimodal Memory for Agentic AIChunliang Chen, et al.memory, multimodal, long-term, agent, LLMcs.CL, cs.AI, cs.CV2026.01
arXiv(v1) 2025Tool-Memory Conflicts in Tool-Augmented LLMsAuthors TBDmemory, tool use, conflict, LLMcs.CL, cs.AI2026.01
arXiv(v1) (NeurIPS25)A-MEM: Agentic Memory for LLM Agents (NeurIPS)Authors TBDmemory, agentic, Zettelkasten, NeurIPScs.AI, cs.CLNeurIPS 2025 publication2025.12
arXiv(v1) 2025AI Meets Brain: Memory Systems from Cognitive Neuroscience to Autonomous AgentsJiafeng Liang, et al.memory, cognitive, survey, agentcs.CL, cs.AI, cs.CV57 pages, 5 figures2025.12
arXiv(v1) 2025Audited Skill-Graph Self-Improvement for Agentic LLMs via Verifiable Rewards, Experience Synthesis, and Continual MemoryKen Huang, et al.memory, continual, self-improvement, agentcs.CR, cs.AI11 pages, 4 figures. Includes a complete runnable reference implementation and audit logging framework2025.12
arXiv(v1) 2025Beyond Heuristics: A Decision-Theoretic Framework for Agent Memory ManagementChangzhi Sun, et al.memory, agent, decision-theoretic, managementcs.CL2025.12
arXiv(v1) 2025Cache What Lasts: Token Retention for Memory-Bounded KV Cache in LLMsNgoc Bui, et al.memory, KV cache, retention, long-contextcs.LG, cs.AI2025.12
arXiv(v1) 2025Context as a Tool: Context Management for Long-Horizon SWE-AgentsShukai Liu, et al.memory, context management, long-horizon, SWE-agentcs.CL2025.12
arXiv(v2) 2025Evaluating Long-Term Memory for Long-Context Question AnsweringAlessandra Terranova, et al.memory, long-context, evaluation, QAcs.CLAccepted as a poster at Metacognition in Generative AI EurIPS workshop2025.12
arXiv(v1) 2025Hindsight is 20/20: Building Agent Memory that Retains, Recalls, and ReflectsAuthors TBDmemory, hindsight, reflection, retentioncs.AI, cs.CLRetain-Recall-Reflect framework2025.12
arXiv(v1) 2025Learning Hierarchical Procedural Memory for LLM Agents through Bayesian Selection and Contrastive RefinementSaman Forouzandeh, et al.memory, procedural, hierarchical, agentcs.LG, cs.AIAccepted at The 25th International Conference on Autonomous Agents and Multi-Agent Systems (AAMAS 2026). 21 pages including references, with 7 figures and 8 tables. Code is publicly available at the authors GitHub repository: https://github.com/S-Forouzandeh/MACLA-LLM-Agents-AAMAS-Conference2025.12
arXiv(v2) 2025MMAG: Mixed Memory-Augmented Generation for Large Language Models ApplicationsStefano Zeppierimemory, memory-augmented generation, RAG, LLMcs.CL, cs.IR2025.12
arXiv(v1) 2025MemEvolve: Meta-Evolution of Agent Memory SystemsGuibin Zhang, et al.memory, evolution, meta-learning, agentcs.CL, cs.MA2025.12
arXiv(v1) 2025MemR3: Memory Retrieval via Reflective Reasoning for LLM AgentsAuthors TBDmemory, retrieval, reflection, reasoningcs.AI, cs.CLRouter + evidence-gap tracker2025.12
arXiv(v2) 2025Memento 2: Learning by Stateful Reflective MemoryJun Wangmemory, agent, reflection, statefulcs.AI, cs.CV, cs.LG35 pages, four figures2025.12
arXiv(v1) 2025Memory in the Age of AI AgentsYuyang Hu, et al.memory, survey, LLM, agentcs.CL, cs.AIComprehensive survey on agent memory2025.12
arXiv(v1) 2025MemoryGraft: Persistent Compromise of LLM Agents via Poisoned Experience RetrievalSaksham Sahai Srivastava, et al.memory, security, agent, attackcs.CR, cs.AI, cs.LG14 pages, 1 figure, includes appendix2025.12
arXiv(v3) 2025O-Mem: Omni Memory System for Personalized, Long Horizon, Self-Evolving AgentsPiaohong Wang, et al.memory, long-horizon, self-evolving, agentcs.CL2025.12
arXiv(v1) 2025R-Debater: Retrieval-Augmented Debate Generation through Argumentative MemoryMaoyuan Li, et al.memory, retrieval, debate, RAGcs.CL, cs.AIAccepteed by AAMAS 2026 full paper2025.12
arXiv(v2) 2025Significant Other AI: Identity, Memory, and Emotional Regulation as Long-Term Relational IntelligenceSung Parkmemory, identity, relational, long-termcs.HC, cs.AI2025.12
arXiv(v1) 2025Adaptive Focus Memory for Language ModelsAuthors TBDmemory, adaptive, focus, compression, AFMcs.CL, cs.AI2/3 token reduction2025.11
arXiv(v1) 2025BudgetMem: Learning Selective Memory Policies for Cost-Efficient Long-Context ProcessingAuthors TBDmemory, selective, budget, efficientcs.CL, cs.AILearned gating + BM252025.11
arXiv(v1) 2025CoEdge-RAG: Optimizing Hierarchical Scheduling for Retrieval-Augmented LLMs in Collaborative Edge ComputingGuihang Hong, et al.RAG, edge computing, optimization, LLMcs.DCAccepted by RTSS 2025 (Real-Time Systems Symposium, 2025)2025.11
arXiv(v1) 2025EMem: Event-Centric Memory for Long-Term Conversational AgentsAuthors TBDmemory, event-centric, conversation, neo-Davidsoniancs.CL, cs.AIEvent-like propositions2025.11
arXiv(v1) 2025Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving MemoryTianxin Wei, et al.memory, benchmark, test-time learning, agentcs.CL, cs.AI2025.11
arXiv(v1) 2025G-KV: Decoding-Time KV Cache Eviction with Global AttentionMengqi Liao, et al.memory, KV cache, eviction, efficiencycs.CL, cs.AI2025.11
arXiv(v1) 2025GCAgent: Long-Video Understanding via Schematic and Narrative Episodic MemoryJeong Hun Yeo, et al.memory, episodic, video, MLLMcs.CV, cs.AI2025.11
arXiv(v1) 2025Goal-Directed Search Outperforms Goal-Agnostic Memory Compression in Long-Context Memory TasksYicong Zheng, et al.memory, compression, long-context, searchcs.CL, cs.AI, cs.LG2025.11
arXiv(v1) 2025KVzip: Memory Compression for LLM Chatbots via KV Cache OptimizationAuthors TBDmemory, compression, KV cache, chatbotcs.CL, cs.AI3-4x compression, 170K tokens2025.11
MobiCom25Poster: MemAura: Persistent Personalized Context Memory for LLM Services in Smart EnvironmentsSiyuan Liu, et al.LLM, memory, personalization, smart environment, context2025.11
arXiv(v1) 2025Trainable Graph Memory for LLM Agents: From Experience to StrategyAuthors TBDmemory, graph, trainable, strategycs.AI, cs.CLUtility assessment mechanism2025.11
arXiv(v1) 2025WebCoach: Self-Evolving Web Agents with Cross-Session Memory GuidanceGenglin Liu, et al.memory, web agent, cross-session, self-evolvingcs.AI, cs.CL18 pages; work in progress2025.11
arXiv(v1) 2025A Memory-Efficient Retrieval Architecture for RAG-Enabled Wearable Medical LLMs-AgentsZhipeng Liao, et al.memory-efficient, RAG, wearable, medical, LLM, agentcs.ARAccepted by BioCAS20252025.10
arXiv(v1) 2025Acon: Optimizing Context Compression for Long-horizon LLM AgentsAuthors TBDmemory, compression, context, optimizationcs.CL, cs.AI26-54% memory reduction2025.10
arXiv(v1) 2025Beyond a Million Tokens: Benchmarking and Enhancing Long-Term Memory in LLMsMohammad Tavakoli, et al.memory, long-context, benchmark, long-termcs.CL, cs.AI, cs.IR2025.10
arXiv(v1) 2025CAM: Contextual Augmentation Memory for LLM AgentsAuthors TBDmemory, contextual, augmentation, agentcs.AI, cs.CL2025.10
arXiv(v1) 2025Dynamic Affective Memory Management for Personalized LLM AgentsJunfeng Lu, et al.memory, affective, personalization, agentcs.CL12 pasges, 8 figures2025.10
arXiv(v1) 2025Enabling Personalized Long-term Interactions in LLM-based Agents through Persistent MemoryAuthors TBDmemory, personalized, long-term, user profilecs.CL, cs.AI2025.10
arXiv(v1) 2025LightMem: Lightweight Memory for Efficient LLM AgentsAuthors TBDmemory, lightweight, efficient, agentcs.AI, cs.CL2025.10
arXiv(v1) 2025MEMTRACK: Evaluating Long-Term Memory and State Tracking in Multi-Platform Dynamic Agent EnvironmentsDarshan Deshpande, et al.memory, benchmark, state tracking, agentcs.AI, cs.CLAccepted to NeurIPS 2025 SEA Workshop2025.10
arXiv(v1) 2025Memory-Augmented State Machine Prompting: A Novel LLM Agent Framework for Real-Time Strategy GamesRunnan Qi, et al.memory, prompting, state machine, agentcs.AI10 pages, 4 figures, 1 table, 1 algorithm. Submitted to conference2025.10
arXiv(v1) 2025Pre-Storage Reasoning for Episodic Memory in LLM AgentsAuthors TBDmemory, episodic, pre-storage, reasoningcs.AI, cs.CL2025.10
arXiv(v2) 2025Evaluating Memory in LLM Agents via Incremental Multi-Turn InteractionsYuanzhe Hu, et al.memory evaluation, multi-turn, LLM, agentcs.CL, cs.AIY. Hu and Y. Wang contribute equally2025.09
arXiv(v1) 2025HopRAG: Multi-Hop Reasoning with Graph Memory for LLM AgentsAuthors TBDmemory, multi-hop, graph, reasoning, RAGcs.AI, cs.CL2025.09
arXiv(v1) 2025Mem-α: Learning Memory Construction via Reinforcement LearningYu Wang, et al.memory, RL, construction, agentcs.CL2025.09
arXiv(v1) 2025Mem-α: Memory with Adaptive Forgetting for LLM AgentsAuthors TBDmemory, forgetting, adaptive, agentcs.AI, cs.CL2025.09
arXiv(v1) 2025Memory in LLM-based Multi-agent Systems: Mechanisms, Challenges, and CollectiveAuthors TBDmemory, multi-agent, collective, surveycs.AI, cs.MA2025.09
arXiv(v1) 2025Multiple Memory Systems for Enhancing Long-term Memory of LLM AgentsAuthors TBDmemory, multiple systems, long-term, enhancementcs.AI, cs.CL2025.09
arXiv(v1) 2025Nemori: Neural Memory Organization for LLM AgentsAuthors TBDmemory, neural, organization, agentcs.AI, cs.CL2025.09
arXiv(v1) 2025SGMem: Sentence Graph Memory for Long-Term Conversational AgentsYaxiong Wu, et al.memory, graph, conversation, retrievalcs.CL, cs.IR19 pages, 6 figures, 1 table2025.09
arXiv(v1) 2025SGMem: Structured Graph Memory for LLM AgentsAuthors TBDmemory, graph, structured, agentcs.AI, cs.CL2025.09
IJCAI25 (IJCAI 2025)AriGraph: Learning Knowledge Graph World Models with Episodic MemoryAuthors TBDmemory, knowledge graph, episodic, world modelcs.AIIJCAI 20252025.08
arXiv(v1) 2025Cognitive Workspace: Active Memory Management for LLMs - Functional Infinite ContextAuthors TBDmemory, cognitive, workspace, active managementcs.CL, cs.AIMetacognitive control2025.08
arXiv(v1) 2025Learn to Memorize: Optimizing LLM-based Agents with Adaptive Memory FrameworkZeyu Zhang, et al.adaptive memory, optimization, LLM, agentcs.LG, cs.AI, cs.CL, cs.IR17 pages, 4 figures, 5 tables2025.08
arXiv(v1) 2025Memory-Augmented Transformers: A Systematic ReviewAuthors TBDmemory, transformer, survey, augmentedcs.CL, cs.AISystematic review2025.08
arXiv(v1) 2025Memory-R1: Enhancing LLM Agents to Manage and Utilize Memories via RLAuthors TBDmemory, RL, memory manager, ADD/UPDATE/DELETEcs.AI, cs.LGMemory Manager + Answer Agent2025.08
arXiv(v1) 2025Recursive Summarization for Long-Term Dialogue Memory in LLMsAuthors TBDmemory, summarization, dialogue, recursivecs.CL, cs.AIUpdated 20252025.08
arXiv(v2) 2025In Prospect and Retrospect: Reflective Memory Management for Long-term Personalized Dialogue AgentsZhen Tan, et al.reflective memory, dialogue agent, personalization, LLMcs.CL, cs.AIAccepted to ACL 20252025.07
ACL25 (ACL 2025)Pretraining Context Compressor for LLMs with Embedding-Based MemoryAuthors TBDmemory, compression, embedding, contextcs.CLPCC framework2025.07
arXiv(v1) 2025Cross-Attention Networks for Memory Retrieval in Generative AgentsAuthors TBDmemory, retrieval, cross-attention, generativecs.AI, cs.CLFrontiers in Psychology2025.04
arXiv(v1) 2025From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMsAuthors TBDmemory, survey, episodic, semantic, working memorycs.CL, cs.AIPersonal/system, parametric/non-parametric2025.04
arXiv(v3) 2025LIFT: Improving Long Context Understanding of Large Language Models through Long Input Fine-TuningYansheng Mao, et al.long context, fine-tuning, LLMcs.CL2025.04
arXiv(v1) 2025Mem0: Building Production-Ready AI Agents with Scalable Long-Term MemoryAuthors TBDmemory, long-term, scalable, productioncs.AI, cs.CL26% improvement over OpenAI2025.04
arXiv(v1) 2025In Prospect and Retrospect: Reflective Memory Management for Long-term Dialogue AgentsAuthors TBDmemory, reflective, dialogue, long-termcs.CL, cs.AI2025.03
arXiv(v1) 2025Tuning LLMs by RAG Principles: Towards LLM-native MemoryJiale Wei, et al.RAG, fine-tuning, optimization, LLMcs.CL, cs.AI, cs.IR2025.03
arXiv(v1) 2025A-MEM: Agentic Memory for LLM AgentsAuthors TBDmemory, agentic, self-organizing, Zettelkastencs.AI, cs.CLDynamic memory organization2025.02
arXiv(v1) 2025Position: Episodic Memory is the Missing Piece for Long-Term LLM AgentsAuthors TBDmemory, episodic, long-term, position papercs.AI, cs.CLEncoding and retrieval2025.02
arXiv(v1) 2025Zep: Temporal Knowledge Graph Architecture for Agent MemoryAuthors TBDmemory, temporal, knowledge graph, agentcs.AI, cs.CLEpisodic + semantic + community2025.02
arXiv(v1) 2024Memory-Augmented Agent Training for Business Document UnderstandingJiale Liu, et al.memory, agent, training, documentcs.CL, cs.AI11 pages, 8 figures2024.12
arXiv(v1) 2024On the Structural Memory of LLM AgentsRuihong Zeng, et al.memory, agent, analysis, structuralcs.CL, cs.AI2024.12
arXiv(v1) 2024XKV: Personalized KV Cache Memory Reduction for Long-Context LLM InferenceWeizhuo Li, et al.memory, KV cache, long-context, personalizationcs.LG, cs.CL2024.12
arXiv(v1) 2024MELODI: Exploring Memory Compression for Long ContextsYinpeng Chen, et al.memory compression, long context, LLMcs.LG, cs.AI2024.10
arXiv(v1) 2024A Survey on the Memory Mechanism of Large Language Model based AgentsZeyu Zhang, et al.memory, survey, LLM, agentcs.AIACM TOIS, 39 pages2024.04

Personalization

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026DeepInterestGR: Mining Deep Multi-Interest Using Multi-Modal LLMs for Generative RecommendationYangchen ZengLLM, personalizationcs.LG, cs.CV, cs.CY2026.02
arXiv(v1) 2026Dynamic Personality Adaptation in Large Language Models via State MachinesLeon Pielage, et al.LLM, personalizationcs.CL, cs.HC, cs.LG22 pages, 5 figures, submitted to ICPR 20262026.02
arXiv(v1) 2026Facet-Level Persona Control by Trait-Activated Routing with Contrastive SAE for Role-Playing LLMsWenqiu Tang, et al.LLM, personalizationcs.CLAccepted in PAKDD 2026 special session on Data Science :Foundation and Applications2026.02
arXiv(v1) 2026InterviewSim: A Scalable Framework for Interview-Grounded Personality SimulationYu Li, et al.LLM, personalizationcs.CL, cs.AI, cs.CY2026.02
arXiv(v1) 2026Learning to Reason for Multi-Step Retrieval of Personal Context in Personalized Question AnsweringMaryam Amirizaniani, et al.LLM, personalizationcs.CL, cs.AI, cs.IR2026.02
arXiv(v1) 2026Long Context, Less Focus: A Scaling Gap in LLMs Revealed through Privacy and PersonalizationShangding GuLLM, personalizationcs.LG, cs.AI2026.02
arXiv(v1) 2026Multi-Agent Large Language Model Based Emotional Detoxification Through Personalized Intensity Control for Consumer ProtectionKeito InoshitaLLM, personalizationcs.AI2026.02
arXiv(v1) 2026Offline Reasoning for Efficient Recommendation: LLM-Empowered Persona-Profiled Item IndexingDeogyong Kim, et al.LLM, personalizationcs.IR, cs.LGUnder review2026.02
arXiv(v1) 2026PERSONA: Dynamic and Compositional Inference-Time Personality Control via Activation Vector AlgebraXiachong Feng, et al.LLM, personalizationcs.AIICLR 20262026.02
arXiv(v1) 2026PRECTR-V2:Unified Relevance-CTR Framework with Cross-User Preference Mining, Exposure Bias Correction, and LLM-Distilled Encoder OptimizationShuzhi Cao, et al.LLM, personalizationcs.IR, cs.AIarXiv admin note: text overlap with arXiv:2503.183952026.02
arXiv(v1) 2026Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User HistorySerin Kim, et al.LLM, personalizationcs.CL, cs.AI2026.02
arXiv(v1) 2026Personalized Graph-Empowered Large Language Model for Proactive Information AccessChia Cheng Chang, et al.LLM, personalizationcs.CL2026.02
arXiv(v1) 2026Personalized Prediction of Perceived Message Effectiveness Using Large Language Model Based Digital TwinsJasmin Han, et al.LLM, personalizationcs.CL, stat.AP31 pages, 5 figures, submitted to Journal of the American Medical Informatics Association (JAMIA). Drs. Chen and Thrul share last authorship2026.02
arXiv(v1) 2026Sydney Telling Fables on AI and Humans: A Corpus Tracing Memetic Transfer of Persona between LLMsJiří Milička, et al.LLM, personalizationcs.CL, cs.AI2026.02
arXiv(v1) 2026Toward Personalized LLM-Powered Agents: Foundations, Evaluation, and Future DirectionsYue Xu, et al.LLM, personalizationcs.AI2026.02
arXiv(v1) 2026CURP: Codebook-based Continuous User Representation for Personalized Generation with LLMsLiang Wang, et al.arxivcs.CL2026.01
arXiv(v1) 2026Enriching Semantic Profiles into Knowledge Graph for Recommender Systems Using Large Language ModelsSeokho Ahn, et al.knowledge graph, semantic profiles, recommendation, LLMcs.IR, cs.AI, cs.LGAccepted at KDD 20262026.01
ICLR 2026FingerTip 20K: A Benchmark for Proactive and Personalized Mobile LLM AgentsICLR,OpenReviewMobile Agent, LLM Agent, GUI, Proactive Agent, PersonalizationOpenReview ID: n3iFV0gLMc2026.01
arXiv(v1) 2026HumanLLM: Towards Personalized Understanding and Simulation of Human NatureYuxuan Lei, et al.arxivcs.CL12 pages, 5 figures, 7 tables, to be published in KDD 20262026.01
arXiv(v1) 2026Improving User Privacy in Personalized Generation: Client-Side Retrieval-Augmented Modification of Server-Side Generated SpeculationsAlireza Salemi, et al.arxivcs.CL, cs.AI, cs.CR, cs.IR2026.01
arXiv(v2) 2026Linear Personality Probing and Steering in LLMs: A Big Five StudyMichel Frising, et al.personalization, personality, probing, steeringcs.CL29 pages, 6 figures2026.01
arXiv(v1) 2026Me-Agent: A Personalized Mobile Agent with Two-Level User Habit Learning for Enhanced InteractionShuoxin Wang, et al.arxivcs.CL2026.01
ICLR 2026Meta-Router: Bridging Gold-standard and Preference-based Evaluations in LLM RoutingICLR,OpenReviewCausal learning, Meta-learner, Large Language Model, query routingOpenReview ID: r0BFucF2dH2026.01
ICLR 2026NextQuill: Causal Preference Modeling for Enhancing LLM PersonalizationICLR,OpenReviewPersonalized text generation, Large Language Models, LLM PersonalizationOpenReview ID: xYpVlKMFqv2026.01
arXiv(v1) 2026One Adapts to Any: Meta Reward Modeling for Personalized LLM AlignmentHongru Cai, et al.meta reward modeling, alignment, personalization, LLMcs.CL, cs.AI2026.01
arXiv(v1) 2026Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM PersonalizationLinfeng Du, et al.arxivcs.CL, cs.IR2026.01
arXiv(v1) 2026PRISP: Privacy-Safe Few-Shot Personalization via Lightweight AdaptationJunho Park, et al.few-shot, privacy-safe, lightweight adaptation, personalization, LLMcs.CL, cs.AI, cs.LG16 pages, 9 figures2026.01
arXiv(v1) 2026PersonaDual: Balancing Personalization and Objectivity via Adaptive ReasoningXiaoyou Liu, et al.personalization, persona, reasoning, LLMcs.AI2026.01
ICLR 2026Preference Leakage: A Contamination Problem in LLM-as-a-judgeICLR,OpenReviewLLM-as-a-judge, Preference Leakage, Data ContaminationOpenReview ID: grIvSXVJ652026.01
ICLR 2026ProPerSim: Developing Proactive and Personalized AI Assistants through User-Assistant SimulationICLR,OpenReviewBenchmark, Agent Simulation, Personalization, ProactivityOpenReview ID: RV2aeCgxdB2026.01
arXiv(v1) 2026SPRInG: Continual LLM Personalization via Selective Parametric Adaptation and Retrieval-Interpolated GenerationSeoyeon Kim, et al.personalization, continual, retrieval, parametric adaptation, LLMcs.AI, cs.CLunder review, 23 pages2026.01
ICLR 2026Solving the Granularity Mismatch: Hierarchical Preference Learning for Long-Horizon LLM AgentsICLR,OpenReviewLLM-based Agents, Process Supervision, Curriculum LearningOpenReview ID: s8usvGHYlk2026.01
arXiv(v1) 2026Structured Personality Control and Adaptation for LLM AgentsJinpeng Wang, et al.personalization, personality, control, agent, LLMcs.AI2026.01
arXiv(v1) 2026Styles + Persona-plug = Customized LLMsYutong Song, et al.style customization, persona-plug, LLMcs.AI2026.01
arXiv(v1) 2026The Assistant Axis: Situating and Stabilizing the Default Persona of Language ModelsChristina Lu, et al.persona, default persona, alignment, LLMcs.CL2026.01
arXiv(v2) 2026The Reward Model Selection Crisis in Personalized AlignmentFady Rezk, et al.personalization, alignment, reward model, RLHFcs.AI, cs.LG2026.01
ICLR 2026Towards Understanding Valuable Preference Data for Large Language Model AlignmentICLR,OpenReviewLarge language model alignment, preference data, influence functionOpenReview ID: FUp0KeEEBs2026.01
ICLR 2026Verification and Co-Alignment via Heterogeneous Consistency for Preference-Aligned LLM AnnotationsICLR,OpenReviewVerification, Co-Alignment, Preference-Aligned LLM Annotations, Reference-Free MetricOpenReview ID: jugY302BAh2026.01
arXiv(v1) 2026When Personalization Misleads: Understanding and Mitigating Hallucinations in Personalized LLMsZhongxiang Sun, et al.personalization, hallucination, safety, evaluationcs.CL, cs.AI20 pages, 15 figures2026.01
arXiv(v1) 2025Agentic Multi-Persona Framework for Evidence-Aware Fake News DetectionRoopa Bukke, et al.persona, multi-agent, fake news, LLMcs.IR, cs.LG12 pages, 8 tables, 2 figures2025.12
arXiv(v1) 2025Interpolative Decoding: Exploring the Spectrum of Personality Traits in LLMsEric Yeh, et al.personalization, personality, decoding, controlcs.AI20 pages, 5 figures2025.12
arXiv(v1) 2025LLM Personas as a Substitute for Field Experiments in Method BenchmarkingEnoch Hyunwook Kangpersona, evaluation, methodology, LLMcs.AI, cs.LG, econ.EM2025.12
arXiv(v1) 2025Memoria: A Scalable Agentic Memory Framework for Personalized Conversational AISamarth Sarin, et al.personalization, memory, conversational AI, agentcs.AI, cs.CLPaper accepted at 5th International Conference of AIML Systems 2025, Bangalore, India2025.12
arXiv(v1) 2025PILAR: Personalizing Augmented Reality Interactions with LLM-based Human-Centric and Trustworthy Explanations for Daily Use CasesRipan Kumar Kundu, et al.personalization, AR, explanations, LLMcs.HC, cs.AIPublished in the 2025 IEEE International Symposium on Mixed and Augmented Reality Adjunct (ISMAR-Adjunct)2025.12
arXiv(v1) 2025PRISM: A Personality-Driven Multi-Agent Framework for Social Media SimulationZhixiang Lu, et al.personality, multi-agent, social simulation, LLMcs.CL2025.12
arXiv(v1) 2025PersonaMem-v2: Towards Personalized Intelligence via Learning Implicit User PersonasAuthors TBDpersonalization, persona, implicit, user modelingcs.CL, cs.AI1000 personas, 20k preferences, 128k context2025.12
arXiv(v1) 2025Personalized Multimodal Large Language Models: A SurveyAuthors TBDpersonalization, MLLM, survey, multimodalcs.CV, cs.CLComprehensive MLLM personalization survey2025.12
arXiv(v1) 2025PrefGen: Multimodal Preference Learning for Image GenerationAuthors TBDpersonalization, MLLM, image generation, preferencecs.CV, cs.CLUser-specific conditioning2025.12
arXiv(v2) 2025ProEx: A Unified Framework Leveraging Large Language Model with Profile Extrapolation for RecommendationYi Zhang, et al.profile extrapolation, recommendation, LLMcs.IRAccepted by KDD 2026 (First Cycle)2025.12
arXiv(v1) 2025SPARK: Search Personalization via Agent-Driven Retrieval and Knowledge-sharingGaurab Chhetri, et al.personalization, search, agent, retrievalcs.AIAccepted to WEB&GRAPH 2026 (WSDM 2026 workshop)2025.12
arXiv(v1) 2025TAME: Long-Context MLLM Personalization with Double MemoriesAuthors TBDpersonalization, MLLM, memory, training-freecs.CV, cs.CLRA2G recipe2025.12
arXiv(v1) 2025The Mental World of Large Language Models in Recommendation: A Benchmark on Association, Personalization, and KnowledgeabilityGuangneng Hupersonalization, recommendation, benchmark, LLMcs.IR21 pages, 13 figures, 27 tables, submission to KDD 20252025.12
arXiv(v1) 2025Towards Proactive Personalization through Profile Customization for Individual Users in DialoguesXiaotian Zhang, et al.proactive personalization, profile customization, dialogue, LLMcs.CL2025.12
TiiS25User Perceptions of Personalized and Generic Explanations in LLM-Driven Recommender SystemsÍtallo De Sousa Silva, et al.LLM, recommender system, personalization, explanations, user study2025.12
arXiv(v1) 2025Fixed-Persona SLMs with Modular Memory: Scalable NPC Dialogue on Consumer HardwareMartin Braas, et al.personalization, persona, modular memory, NPCcs.AI, cs.IR2025.11
arXiv(v2) 2025Mem-PAL: Towards Memory-based Personalized Dialogue Assistants for Long-term User-Agent InteractionZhaopei Huang, et al.personalization, dialogue, memory, long-termcs.CLAccepted by AAAI 2026 (Oral)2025.11
arXiv(v1) 2025PLUM: Learning to Remember User Conversations for PersonalizationAuthors TBDpersonalization, memory, conversation, LoRAcs.CL, cs.AIParameter-efficient2025.11
arXiv(v1) 2025PersonaAgent with GraphRAG: Community-Aware KG for Personalized LLMAuthors TBDpersonalization, GraphRAG, knowledge graph, agentcs.CL, cs.AI11.1% F1 improvement on LaMP2025.11
arXiv(v1) 2025PersonalizedRouter: Personalized LLM Routing via Graph-based User Preference ModelingAuthors TBDpersonalization, routing, GNN, user preferencecs.CL, cs.AI2025.11
arXiv(v1) 2025Profile-LLM: Dynamic Profile Optimization for Realistic Personality ExpressionAuthors TBDpersonalization, profile, personality, dynamiccs.CL, cs.AIEducation, therapy, entertainment2025.11
arXiv(v1) 2025LLMDiRec: LLM-Enhanced Intent Diffusion for Sequential RecommendationBo-Chian Chen, et al.intent diffusion, sequential recommendation, LLMcs.IRUnder review2025.10
arXiv(v1) 2025MemWeaver: A Hierarchical Memory from Textual Interactive Behaviors for Personalized GenerationShuo Yu, et al.personalization, memory, user behavior, generationcs.CL12 pages, 8 figures2025.10
arXiv(v1) 2025P2P: Instant Personalized LLM Adaptation via HypernetworkAuthors TBDpersonalization, hypernetwork, instant adaptationcs.CL, cs.AISingle-pass generation2025.10
arXiv(v1) 2025Preference-Aware Memory Update for Long-Term LLM AgentsHaoran Sun, et al.personalization, preference, memory update, long-termcs.CL, cs.AI2025.10
arXiv(v1) 2025RGMem: Renormalization Group-based Memory Evolution for Language Agent User ProfileAo Tian, et al.personalization, user profile, memory, agentcs.AI11 pages,3 figures2025.10
arXiv(v1) 2025Real-Time Personalization for LLM-based Recommendation with Customized ICLAuthors TBDpersonalization, recommendation, ICL, real-time, onlinecs.IR, cs.AINo model update needed2025.10
arXiv(v2) 2025CoPL: Collaborative Preference Learning for Personalizing LLMsYoungbin Choi, et al.collaborative preference learning, personalization, LLMcs.LG, cs.AI, cs.IR19pages, 13 figures, 11 tables2025.09
arXiv(v1) 2025DP-FedLoRA: Privacy-Enhanced Federated Fine-Tuning for On-Device Large Language ModelsHonghui Xu, et al.federated learning, privacy-enhanced, on-device LLM, fine-tuningcs.CR, cs.AI2025.09
arXiv(v1) 2025HumAIne-Chatbot: Real-Time Personalized Conversational AI via RLAuthors TBDpersonalization, chatbot, RL, real-time, industrycs.CL, cs.AIProduction deployment2025.09
arXiv(v1) 2025MMPB: Multi-Modal Personalization Benchmark for VLMsAuthors TBDpersonalization, MLLM, benchmark, VLMcs.CV, cs.CLFirst MLLM personalization benchmark2025.09
arXiv(v1) 2025Personalized Reasoning: Just-In-Time Personalization and Why LLMs Fail At ItShuyue Stella Li, et al.just-in-time personalization, LLM, user preferencecs.CL, cs.AI57 pages, 6 figures2025.09
RecSys25Revisiting Prompt Engineering: A Comprehensive Evaluation for LLM-based Personalized RecommendationGenki Kusano, et al.LLM, recommendation, personalization, prompt engineering2025.09
arXiv(v1) 2025T-POP: Test-Time Personalization with Online Preference FeedbackZikun Qu, et al.test-time personalization, online preference, LLMcs.LG, cs.AIPreprint2025.09
arXiv(v1) 2025DGDPO: Diagnostic-Guided Dynamic Profile Optimization for User SimulatorsAuthors TBDpersonalization, user simulation, profile, dynamiccs.IR, cs.AIBidirectional evolution2025.08
arXiv(v1) 2025End-to-End Personalization: Unifying Recommender Systems with Large Language ModelsDanial Ebrat, et al.recommendation system, unifying, LLMcs.IR, cs.LGSecond Workshop on Generative AI for Recommender Systems and Personalization at the ACM Conference on Knowledge Discovery and Data Mining (GenAIRecP@KDD 2025)2025.08
arXiv(v1) 2025MLLMRec: MLLMs in Recommender SystemsAuthors TBDpersonalization, MLLM, recommendation, visualcs.CV, cs.IRVisual attribute extraction2025.08
arXiv(v1) 2025MM-R1: Unified MLLMs for Personalized Image GenerationAuthors TBDpersonalization, MLLM, image generation, GRPOcs.CV, cs.CLX-CoT reasoning2025.08
arXiv(v1) 2025MSPA: Multimodal Self-Corrective Preference Alignment for RecommendationAuthors TBDpersonalization, MLLM, recommendation, self-correctivecs.CV, cs.IR4D multimodal signals2025.08
arXiv(v2) 2025Personalized LLM for Generating Customized Responses to the Same Query from Different UsersHang Zeng, et al.personalization, response generation, user-specific, LLMcs.CLAccepted by CIKM'252025.08
arXiv(v1) 2025RLHF Fine-Tuning of LLMs for Alignment with Implicit User Feedback in Conversational RecommendersAuthors TBDpersonalization, RLHF, implicit feedback, recommendercs.IR, cs.AIDwell time, sentiment signals2025.08
arXiv(v1) (SIGIR25)CoT-Rec: Enhancing LLM-Based Recommendations Through Personalized ReasoningAuthors TBDpersonalization, recommendation, CoT, reasoningcs.IR, cs.AISIGIR 20252025.07
arXiv(v1) 2025Comprehensive Review on LLMs for Recommender SystemsAuthors TBDpersonalization, recommendation, survey, LLMcs.IR, cs.AIHybrid RAG approaches2025.07
arXiv(v1) 2025DEP: Latent Inter-User Difference Modeling for LLM PersonalizationAuthors TBDpersonalization, latent, embedding, difference-awarecs.CL, cs.AISparse autoencoder2025.07
arXiv(v1) 2025PITA: Preference-Guided Inference-Time Alignment for LLM Post-TrainingAuthors TBDpersonalization, inference-time, alignment, preferencecs.CL, cs.AINo reward model needed2025.07
arXiv(v1) 2025PLUS: Learning to Summarize User Information for Personalized RLHFAuthors TBDpersonalization, RLHF, user summary, preferencecs.LG, cs.AI11-77% reward model improvement2025.07
arXiv(v1) 2025PRIME: LLM Personalization with Cognitive Memory and Thought ProcessesAuthors TBDpersonalization, memory, episodic, semantic, cognitivecs.CL, cs.AIDual-memory model2025.07
arXiv(v1) 2025PURE: LLM-based User Profile Management for Recommender SystemAuthors TBDpersonalization, user profile, recommendation, managementcs.IR, cs.AIProfile extraction and updating2025.07
arXiv(v1) 2025Personalization of Large Language Models: A SurveyZhehao Zhang, et al.personalization, survey, LLM, user profilecs.CL, cs.AIComprehensive taxonomy2025.07
arXiv(v2) 2025Comparison-based Active Preference Learning for Multi-dimensional PersonalizationMinhyeon Oh, et al.active preference learning, multi-dimensional, personalization, LLMcs.LG2025.06
arXiv(v1) 2025PersonalAI: KG Storage and Retrieval for Personalized LLM AgentsAuthors TBDpersonalization, knowledge graph, memory, agentcs.AI, cs.CLHybrid graph with hyperedges2025.06
UMAP25Personalizing LLM Responses to Combat Political MisinformationAdiba Proma, et al.LLM, personalization, misinformation, user modeling2025.06
arXiv(v1) 2025ProfiLLM: LLM-Based Framework for Implicit User ProfilingAuthors TBDpersonalization, profiling, implicit, chatbotcs.CL, cs.AIIT/cybersecurity domain2025.06
arXiv(v1) 2025SEAL: Self-Adapting Language ModelsAuthors TBDpersonalization, self-adaptation, online learningcs.CL, cs.AISelf-generated finetuning data2025.06
arXiv(v3) 2025Drift: Decoding-time Personalized Alignments with Implicit User PreferencesMinbeom Kim, et al.decoding-time alignment, implicit preferences, personalization, LLMcs.CL19 pages, 6 figures2025.05
arXiv(v2) 2025HyPerAlign: Interpretable Personalized LLM Alignment via Hypothesis GenerationCristina Garbacea, et al.alignment, hypothesis generation, personalization, LLMcs.CL2025.05
arXiv(v1) 2025MAP: Memory Assisted LLM for Personalized Recommendation SystemAuthors TBDpersonalization, recommendation, memory, historycs.IR, cs.AI2025.05
arXiv(v1) 2025PROSE: Aligning LLMs by Predicting Preferences from User Writing SamplesAuthors TBDpersonalization, preference prediction, writing samplescs.CL, cs.AIIterative refinement2025.05
arXiv(v1) 2025Privacy-preserving Prompt Personalization in Federated Learning for Multimodal Large Language ModelsSizai Hou, et al.federated learning, privacy-preserving, prompt personalization, MLLMcs.CRUnder Review2025.05
arXiv(v1) 2025RLPA: Teaching LLMs to Evolve with Users via Dynamic Profile ModelingAuthors TBDpersonalization, RLHF, dynamic profile, user evolutioncs.CL, cs.AIOutperforms Claude-3.5, DeepSeek-V32025.05
arXiv(v1) 2025Steerable Chatbots: Personalizing LLMs with Preference-Based Activation SteeringAuthors TBDpersonalization, chatbot, activation steering, inferencecs.CL, cs.AITraining-free2025.05
arXiv(v1) 2025Towards Explainable Temporal User Profiling with LLMsMilad Sabouri, et al.temporal user profiling, explainable, LLMcs.IR, cs.AI2025.05
arXiv(v1) 2025Towards a unified user modeling language for engineering human centered AI systemsAaron Conrardy, et al.user modeling, LLM, personalizationcs.SEAccepted at the Third Workshop on Engineering Interactive Systems Embedding AI Technologies (EISEAIT workshop at EICS 2025)2025.05
arXiv(v1) 2025A Survey on Personalized and Pluralistic Preference Alignment in Large Language ModelsAuthors TBDpersonalization, preference alignment, survey, pluralisticcs.CL, cs.AITraining and inference-time methods2025.04
arXiv(v2) 2025Differential Privacy Personalized Federated Learning Based on Dynamically Sparsified Client UpdatesChuanyin Wang, et al.differential privacy, federated learning, personalization, LLMcs.LG, cs.CR10 pages,2 figures2025.04
arXiv(v1) 2025Know Me, Respond to Me: Benchmarking LLMs for Dynamic User ProfilingAuthors TBDpersonalization, benchmark, user profiling, dynamiccs.CL, cs.AI2025.04
arXiv(v1) 2025LoRe: Personalizing LLMs via Low-Rank Reward ModelingAvinandan Bose, et al.low-rank reward modeling, personalization, LLM, RLHFcs.LG, cs.AI, cs.CL2025.04
arXiv(v1) 2025PaRT: Enhancing Proactive Social Chatbots with Personalized Real-Time RetrievalAuthors TBDpersonalization, chatbot, retrieval, real-timecs.CL, cs.AI21.77% dialogue improvement, production deployed2025.04
arXiv(v3) 2025Tuning-Free Personalized Alignment via Trial-Error-Explain In-Context LearningHyundong Cho, et al.in-context learning, trial-error-explain, personalization, LLMcs.CL, cs.AINAACL 2025 Findings2025.04
arXiv(v1) 2025User Feedback Alignment for LLM-powered Exploration in Large-scale RecommendationAuthors TBDpersonalization, recommendation, feedback, explorationcs.IR, cs.AIClick and dwell time signals2025.04
arXiv(v1) 2025A Shared Low-Rank Adaptation Approach to Personalized RLHFRenpu Liu, et al.RLHF, low-rank adaptation, personalization, LLMcs.LG, cs.AIPublished as a conference paper at AISTATS 20252025.03
arXiv(v1) 2025Agentic Recommender Systems in the Era of Multimodal LLMs: SurveyAuthors TBDpersonalization, recommendation, agent, MLLM, surveycs.IR, cs.AIUser agent simulation2025.03
arXiv(v1) 2025BAHE: LLM-Enhanced CTR Prediction in Long Textual User BehaviorsAuthors TBDpersonalization, CTR, user behavior, industrycs.IR, cs.AIDeployed on 50M daily data2025.03
arXiv(v1) 2025Can LLM Agents Simulate Multi-Turn Human Behavior? Evidence from Online ShoppingAuthors TBDpersonalization, agent, user simulation, behaviorcs.AI, cs.HC31,865 shopping sessions2025.03
arXiv(v1) 2025Language Model Personalization via Reward FactorizationIdan Shenfeld, et al.reward factorization, RLHF, personalization, LLMcs.LG2025.03
arXiv(v1) 2025Measuring What Makes You Unique: Difference-Aware User Modeling for LLM PersonalizationAuthors TBDpersonalization, user modeling, difference-awarecs.CL, cs.AI2025.03
arXiv(v7) 2025PAD: Personalized Alignment of LLMs at Decoding-TimeRuizhe Chen, et al.decoding-time alignment, personalization, LLMcs.CL, cs.AIICLR 20252025.03
arXiv(v1) 2025PersonaX: A Recommendation Agent Oriented User Modeling FrameworkAuthors TBDpersonalization, recommendation, user modeling, agentcs.IR, cs.AI3-11% improvement on AgentCF2025.03
arXiv(v1) 2025A Survey of Personalized Large Language Models: Progress and Future DirectionsAuthors TBDpersonalization, survey, LLM, prompting, finetuningcs.CL, cs.AIInput/model/objective level2025.02
arXiv(v1) 2025FSPO: Few-Shot Preference Optimization of Synthetic Preference Data in LLMs Elicits Effective Personalization to Real UsersAnikait Singh, et al.few-shot, preference optimization, personalization, LLMcs.LG, cs.AI, cs.CL, cs.HC, stat.MLWebsite: https://fewshot-preference-optimization.github.io/2025.02
arXiv(v1) 2025LoCoMo: Evaluating Very Long-Term Conversational Memory of LLM AgentsAuthors TBDpersonalization, memory, conversation, benchmarkcs.CL, cs.AI300 turns, 9K tokens, 35 sessions2025.02
arXiv(v1) 2025PrefEval: Do LLMs Recognize Your Preferences? Evaluating Personalized Preference FollowingAuthors TBDpersonalization, benchmark, preference, evaluationcs.CL3000 preference-query pairs, 20 topics2025.02
arXiv(v3) 2025Privacy-Preserving Personalized Federated Prompt Learning for Multimodal Large Language ModelsLinh Tran, et al.federated learning, privacy-preserving, personalized prompt, MLLMcs.LG2025.02
arXiv(v1) 2025RLTHF: Targeted Human Feedback for LLM AlignmentAuthors TBDpersonalization, RLHF, targeted feedback, efficientcs.CL, cs.AI6-7% human annotation effort2025.02
arXiv(v3) 2025SmartAgent: Chain-of-User-Thought for Embodied Personalized Agent in Cyber WorldJiaqi Zhang, et al.personalization, agent, user modeling, embodiedcs.AI2025.02
arXiv(v1) 2025User Profile Construction and Updating with LLMs: BenchmarkAuthors TBDpersonalization, user profile, construction, updatingcs.CL, cs.AIStatic and dynamic profiling2025.02
arXiv(v1) 2025When Personalization Meets Reality: Multi-Faceted Analysis of Personalized Preference LearningAuthors TBDpersonalization, preference learning, fairness, evaluationcs.CL, cs.AI2025.02
arXiv(v1) 2025Advancing Personalized Federated Learning: Integrative Approaches with AI for Enhanced Privacy and CustomizationKevin Cooper, et al.federated learning, privacy, personalization, LLMcs.LG, eess.SParXiv admin note: substantial text overlap with arXiv:2501.167582025.01
arXiv(v2) 2025Identifying and Manipulating Personality Traits in LLMs Through Activation EngineeringRumi Allbert, et al.personalization, personality, activation engineering, steeringcs.CL, cs.AI2025.01
arXiv(v1) 2025PerRecBench: Can LLMs Understand Preferences in Personalized Recommendation?Authors TBDpersonalization, benchmark, recommendation, preferencecs.IR, cs.CL19 LLMs evaluated2025.01
arXiv(v2) 2025PsychAdapter: Adapting LLM Transformers to Reflect Traits, Personality and Mental HealthHuy Vu, et al.personalization, personality, traits, mental healthcs.AI, cs.CL2025.01
arXiv(v1) 2024AI PERSONA: Towards Life-long Personalization of LLMsTiannan Wang, et al.personalization, persona, lifelong, LLMcs.CL, cs.AIWork in progress2024.12
arXiv(v1) 2024Beyond Discrete Personas: Personality Modeling Through Journal Intensive ConversationsSayantan Pal, et al.personalization, persona, personality modeling, long-termcs.CL, cs.AIAccepted in COLING 20252024.12
arXiv(v1) 2024Can Large Language Models Understand You Better? An MBTI Personality Detection Dataset Aligned with Population TraitsBohan Li, et al.personalization, personality, dataset, MBTIcs.CL, cs.CYAccepted by COLING 2025. 28 papges, 20 figures, 10 tables2024.12
arXiv(v1) 2024Disentangling Preference Representation and Text Generation for Efficient Individual Preference AlignmentJianfei Zhang, et al.personalization, preference alignment, efficiency, LLMcs.CL, cs.AIColing 20252024.12
arXiv(v2) 2024Molar: Multimodal LLMs with Collaborative Filtering Alignment for Enhanced Sequential RecommendationYucong Luo, et al.collaborative filtering, sequential recommendation, MLLMcs.IR, cs.AI2024.12
arXiv(v1) 2024Semantic Convergence: Harmonizing Recommender Systems via Two-Stage Alignment and Behavioral Semantic TokenizationGuanghan Li, et al.recommendation system, alignment, behavioral semantic, LLMcs.IR, cs.AI, cs.CL7 pages, 3 figures, AAAI 20252024.12
arXiv(v1) 2024ULMRec: User-centric Large Language Model for Sequential RecommendationMinglai Shao, et al.personalization, recommendation, sequential, LLMcs.IR2024.12
arXiv(v1) 2024Personalizing Reinforcement Learning from Human Feedback with Variational Preference LearningSriyash Poddar, et al.RLHF, variational preference learning, personalization, LLMcs.LG, cs.AI, cs.CL, cs.ROweirdlabuw.github.io/vpl2024.08
arXiv(v2) 2024RLHF from Heterogeneous Feedback via Personalization and Preference AggregationChanwoo Park, et al.RLHF, heterogeneous feedback, preference aggregation, personalization, LLMcs.AI, cs.LGAdded experiments2024.05
agent
awesome
awesome-list
embodied-ai
hci
human-computer-interaction
imu
llm
mllm
multimodal
rag
rl
rlhf
sensor
sensors
survey
ubiquitous
vision
vllm
vlm

WILLOSCAR/Awesome-HCI-LLM

Awesome-HCI (Ubiquitous, LLM, MLLM, Agent, RAG, Embodied-AI, RLHF)

Python

26

0 commits

updated Mar 8, 2026

See the code

README

Awesome HCI-LLM-Agent Papers

Awesome

A curated collection of research papers on HCI, LLM, MLLM, Agent, RAG, Agentic-RL, and Embodied AI (2021–present).

[Jan 2025] Added new sections: Agentic-RL and MLLM. Regular updates resumed.

Quick Start

python -m pip install -e .                # Install CLI
paper add 2312.00752 LLM -t "llm, mamba"  # Add paper
paper search transformer -t IMU           # Search
paper stats                               # Statistics

Documentation


HCI

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026An Expert Schema for Evaluating Large Language Model Errors in Scholarly Question-Answering SystemsAnna Martin-Boyle, et al.LLMcs.HC, cs.CL24 pages, 2 figures. Accepted at ACM CHI conference on Human Factors in Computing Systems, 20262026.02
arXiv(v1) 2026Codesigning Ripplet: an LLM-Assisted Assessment Authoring System Grounded in a Conceptual Model of Teachers' WorkflowsYuan Cui, et al.LLMcs.HCProceedings of the 2026 CHI Conference on Human Factors in Computing Systems2026.02
arXiv(v1) 2026Detecting UX smells in Visual Studio Code using LLMsAndrés Rodriguez, et al.LLMcs.SE, cs.HC4 pages, 2 figures, 1 table, 3rd International Workshop on Integrated Development Environments (IDE 2026)2026.02
arXiv(v1) 2026LLM Novice Uplift on Dual-Use, In Silico Biology TasksChen Bo Calvin Zhang, et al.LLMcs.AI, cs.CL, cs.CR, cs.CY, cs.HC59 pages, 33 figures2026.02
arXiv(v1) 2026PaperTrail: A Claim-Evidence Interface for Grounding Provenance in LLM-based Scholarly Q&AAnna Martin-Boyle, et al.LLMcs.HC, cs.CL25 pages, 3 figures. Accepted at the ACM CHI conference on Human Factors in Computing Systems 20262026.02
arXiv(v1) 2026Scaling Search Relevance: Augmenting App Store Ranking with LLM-Generated JudgmentsEvangelia Christakopoulou, et al.LLMcs.IR, cs.AI, cs.LG2026.02
arXiv(v1) 2026SparkMe: Adaptive Semi-Structured Interviewing for Qualitative Insight DiscoveryDavid Anugraha, et al.LLMcs.HC, cs.AI, cs.CY2026.02
arXiv(v1) 2026 (Proceedings of the 34th ACM International Conference on Information and Knowledge Management (CIKM '25), November 10--14, 2025, Seoul, Republic of Korea)UXSim: Towards a Hybrid User Search SimulationSaber Zerhoudi, et al.LLMcs.IR, cs.HC2026.02
arXiv(v1) 2026Understanding Usage and Engagement in AI-Powered Scientific Research Tools: The Asta Interaction DatasetDany Haddad, et al.LLMcs.HC, cs.AI, cs.IR2026.02
arXiv(v1) 2026When LLMs Help -- and Hurt -- Teaching Assistants in Proof-Based CoursesRomina Mahinpei, et al.LLMcs.HC2026.02
Ubicomp25 (IMWUT Vol 9 Issue 4)Ads that Talk Back: Implications and Perceptions of Injecting Personalized Advertising into LLM ChatbotsBrian Jay Tang, et al.LLM, personalization, advertising, chatbot2025.12
Ubicomp25 (IMWUT Vol 9 Issue 4)CHEF-VL: Detecting Cognitive Sequencing Errors in Cooking with Vision-language ModelsRuiqi Wang, et al.VLM, cooking, error detection2025.12
Ubicomp25 (IMWUT Vol 9 Issue 4)Design and Evaluation of Generative Agent-based Platform for Human-Assistant Interaction Research: A Tale of 10 User StudiesZiyi Xuan, et al.LLM, generative agent, platform, human-assistant interaction, user study2025.12
Ubicomp25 (IMWUT Vol 9 Issue 4)Gestura: A LVLM-Powered System Bridging Motion and Semantics for Real-Time Free-Form Gesture UnderstandingZhuoming Li, et al.LVLM, gesture, motion, semantics2025.12
Ubicomp25 (IMWUT Vol 9 Issue 4)IMUZero: Zero-Shot Human Activity Recognition by Language-Based Cross Modality FusionJie Su, et al.LLM, HAR, zero-shot, cross-modality, language2025.12
arXiv(v2) 2026LLM-Guided Exemplar Selection for Few-Shot Wearable-Sensor Human Activity RecognitionElsen Ronando, et al.LLM, HAR, wearable, few-shot, exemplar selectioncs.CL, cs.AI, cs.CV88.78% F1 on UCI-HAR2025.12
Ubicomp25 (IMWUT Vol 9 Issue 4)Large Language Model-guided Semantic Alignment for Human Activity RecognitionHua Yan, et al.LLM, HAR, semantic alignment2025.12
Ubicomp25 (IMWUT Vol 9 Issue 4)TourismMinds: A Geo-augmented LLM Framework for Semantic-aware Trajectory Analytics and GenerationZhuohan Ye, et al.LLM, trajectory, geo-augmented2025.12
CSCW25An Emergent Understanding of Human-AI Collaboration in DeliberationAuthors TBDLLM, human-AI collaboration, deliberation, citizen assembly2025.11
SenSys25Demo: An LLM-Powered Multimodal Mobile Sensing System for Personalized Health Behavior AnalysisAuthors TBDLLM, mobile sensing, multimodal, health, personalizedDemo paper2025.11
CSCW25Exploring Collaboration Patterns and Strategies in Human-AI Co-creation through the Lens of Agency: A Scoping ReviewAuthors TBDhuman-AI co-creation, agency, collaboration, scoping reviewPACM HCI2025.11
HAI25Human-Like Remembering and Forgetting in LLM Agents: An ACT-R-Inspired Memory ArchitectureYudai Honda, et al.LLM, agent, memory, ACT-R, forgetting2025.11
HAI25Robots with Attitudes: Influence of LLM-Driven Robot Personalities on Motivation and PerformanceDennis Becker, et al.LLM, robot, personality, motivation, performance2025.11
HAI25The Double-Edged Sword: Exploring Older Adults' Interaction and Imagination with an LLM-Enhanced Health AgentLeon Paul Mondrian Munz, et al.LLM, health agent, older adults, interaction2025.11
SenSys25Toward Sensor-In-the-Loop LLM Agent: Benchmarks and ImplicationsZechen Li, et al.LLM, agent, sensor, wearable, benchmark2025.11
UIST25ImaginationVellum: Generative-AI Ideation Canvas with Spatial PromptsAuthors TBDgenerative AI, ideation, canvas, spatial prompts, co-creation2025.10
Ubicomp25LLM Powered Memory Consolidation for Ubiquitous ComputingParampuneet Kaur Thind, et al.LLM, memory consolidation, ubiquitous computingUbiComp 2025 Companion2025.10
UIST25SketchGPT: A Sketch-based Multimodal Interface for Application-Agnostic LLM InteractionAuthors TBDLLM, sketch, multimodal, interface2025.10
UIST25"This is My Fault", Really? Understanding Blind and Low-Vision People’s Perception of Hallucination in Large Vision Language ModelsYilin Tang, et al.VLM2025.09
UIST25BloomIntent: Automating Search Evaluation with LLM-Generated Fine-Grained User IntentsYoonseo Choi, et al.LLM2025.09
UIST25Can You Move These Over There? Exploring an LLM-based VR Mover to Support Natural Multi-object ManipulationXiangzhi Eric Wang, et al.LLM2025.09
UIST25CoGrader: Transforming Instructors' Assessment of Project Reports through Collaborative LLM IntegrationZixin Chen, et al.LLM2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Contact-free Vital Signs Monitoring and Separating Distinct Vital SignsAuthors TBDvital signs, contactless, respiration, monitoring2025.09
UIST25DxHF: Providing High-Quality Human Feedback for LLM Alignment with Interactive DecompositionDanqing Shi, et al.LLM, alignment2025.09
UIST25GestureCoach: Rehearsing for Engaging Talks with LLM-Driven Gesture RecommendationsAshwin Ram, et al.LLM2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Hapt-Aids: Self-Powered, On-Body Haptics for Activity MonitoringAuthors TBDhaptic, self-powered, activity monitoring, wearable2025.09
UIST25InReAcTable: LLM-powered Interactive Visual Data Story Construction from Tabular DataGerile Aodeng, et al.LLM2025.09
arXiv(v4) 2025LLaSA: A Sensor-Aware LLM for Natural Language Reasoning of Human Activity from IMU DataSheikh Asif Imran, et al.IMU, LLM, wearable, sensor, HARcs.CL2025.09
UIST25LegisFlow: Enhancing Korean Legal Research with Temporal-Aware LLM InterfacesJunghwan Kim, et al.LLM2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)MASTER: A Multi-modal Foundation Model for Human Activity RecognitionGuanzhou Zhu, et al.foundation model, HAR, multimodal2025.09
UIST25MapStory: Prototyping Editable Map Animations with LLM AgentsAditya Gunturu, et al.LLM, agent2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Mindfulness Meditation and Respiration: Accelerometer-based Respiration Rate EstimationAuthors TBDmindfulness, respiration, accelerometer, meditation2025.09
UIST25NarraGuide: an LLM-based Narrative Mobile Robot for Remote Place ExplorationYaxin Hu, et al.LLM2025.09
UIST25NeuroSync: Intent-Aware Code-Based Problem Solving via Direct LLM Understanding ModificationWenshuo Zhang, et al.LLM2025.09
UIST25Oak Story: Improving Learner Outcomes with LLM-Mediated Interactive NarrativesAlan Y. Cheng, et al.LLM2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)One Model to Fit Them All: Universal IMU-based Human Activity Recognition with LLM-assisted Cross-dataset RepresentationWei Wei, et al.IMU, LLM, HAR, universal model, cross-dataset2025.09
UIST25Policy Maps: Tools for Guiding the Unbounded Space of LLM BehaviorsMichelle S. Lam, et al.LLM2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Pulse-PPG: An Open-Source Field-Trained PPG Foundation Model for Wearable Applications across Lab and Field SettingsMithun Saha, et al.foundation model, wearable, PPG2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)RouteLLM: A Large Language Model with Native Route Context Understanding to Enable Context-Aware ReasoningPhilipp Hallgarten, et al.LLM, navigation, route, context-aware2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)SELA: Smart Edge LLM Agent to Optimize Response Trade-offs of AI AssistantsShreshth Tuli, et al.LLM agent, edge, assistant, optimization2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Sleep Monitoring with Continuous Posture, Heart Rate, Respiratory Rate TrackingAuthors TBDsleep, posture, heart rate, respiration, monitoring2025.09
UIST25Streaming, Fast and Slow: Cognitive Load-Aware Streaming for Efficient LLM ServingChang Xiao, et al.LLM2025.09
arXiv(v1) 2025Through the Lens of Human-Human Collaboration: A Configurable Research Platform for Exploring Human-Agent CollaborationBingsheng Yao, et al.LLM, agent, human-agent collaboration, CSCW, platformcs.HC, cs.AI, cs.CL2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Towards Customizable Foundation Models for Human Activity Recognition with Wearable DevicesMinghui Qiu, et al.foundation model, HAR, customizable2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Vinci: A Real-time Smart Assistant Based on Egocentric Vision-language Model for Portable DevicesYifei Huang, et al.VLM, egocentric, assistant, wearable2025.09
UIST25ViseGPT: Towards Better Alignment of LLM-generated Data Wrangling Scripts and User PromptsJiajun Zhu, et al.LLM, alignment2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)Vital Insight: Assisting Experts' Context-Driven Sensemaking of Multi-modal Personal Tracking Data Using Visualization and Human-in-the-Loop LLMJiachen Li, et al.LLM, sensemaking, visualization, personal data2025.09
UIST25agentAR: Creating Augmented Reality Applications with Tool-Augmented LLM-based Autonomous AgentsChenfei Zhu, et al.LLM, agent2025.09
Ubicomp25 (IMWUT Vol 9 Issue 3)mmPencil: Toward Writing-Style-Independent In-Air Handwriting Recognition via mmWave Radar and Large Vision-Language ModelYifan Guo, et al.mmWave, radar, handwriting, VLM2025.09
arXiv(v1) 2025BaroPoser: Real-time Human Motion Tracking from IMUs and Barometers in Everyday DevicesRiku Arakawa, et al.IMU, barometer, pose estimation, motion trackingcs.CV, cs.HC2025.08
arXiv(v1) 2025DiffCap: Diffusion-based Real-time Human Motion Capture using Sparse IMUs and a Monocular CameraZongmian Li, et al.IMU, diffusion, motion capture, real-time, cameracs.CV2025.08
arXiv(v4) 2025SensorLLM: Aligning Large Language Models with Motion Sensors for Human Activity RecognitionZechen Li, et al.IMU, LLM, motion sensor, HARcs.CLAccepted by EMNLP 2025 Main Conference2025.08
Ubicomp25 (IMWUT Vol 9 Issue 2)CataractBot: An LLM-powered Expert-in-the-Loop Chatbot for Cataract PatientsPragnya Ramjee, et al.LLM, chatbot, healthcare, expert-in-the-loop2025.06
arXiv(v1) 2025Garment Inertial Poser: Human Motion Capture from Loose and Sparse Inertial Sensors with Garment-aware Diffusion ModelsTao Wang, et al.IMU, pose estimation, loose sensor, diffusion, garmentcs.CV2025.06
Ubicomp25 (IMWUT Vol 9 Issue 2)LEGO: Synthesizing IoT Device Components Based on Static Analysis and Large Language ModelsLiwei Liu, et al.LLM, IoT, static analysis2025.06
arXiv(v1) 2025SensorLM: Learning the Language of Wearable SensorsTong Xia, et al.wearable, sensor, LLM, activity recognition, foundation modelcs.LG, cs.HC2025.06
arXiv(v2) 2025 (PRX Quantum 6, 020311 (April 2025))Device-Independent Quantum Key Distribution Based on Routed Bell TestsTristan Le Roy-Deloison, et al.IMU, RGB, human object interaction, dataset, 3D trackingquant-phVersion2: Slight improvements in the text. Close to published version2025.05
arXiv(v4) 2025LLM-Based Human-Agent Collaboration and Interaction Systems: A SurveyHengyi Peng, et al.LLM, agent, human-agent collaboration, HCI, surveycs.HC, cs.AI, cs.CL2025.05
CHI25"A Great Start, But...": Evaluating LLM-Generated Mind Maps for Information Mapping in Video-Based DesignTianhao He, et al.LLM2025.04
CHI25"Ask Sir Oliver Ingham": LLM-based Social Simulations for History EducationKieun Park, et al.LLM2025.04
CHI25"Create a Fear of Missing Out" - ChatGPT Implements Unsolicited Deceptive Designs in Generated Websites Without WarningVeronika Krauß, et al.LLM2025.04
CHI25"It Warned Me Just at the Right Moment": Exploring LLM-based Real-time Detection of Phone ScamsZitong Shen, et al.LLM2025.04
CHI25"Kya family planning after marriage hoti hai?": Integrating Cultural Sensitivity in an LLM Chatbot for Reproductive HealthRoshini Deva, et al.LLM, chatbot2025.04
CHI25"We do use it, but not how hearing people think": How the Deaf and Hard of Hearing Community Uses Large Language Model ToolsShuxu Huffman, et al.LLM2025.04
CHI25"When AI Writes Personas": Analyzing Lexical Diversity in LLM-Generated Persona DescriptionsSankalp Sethi, et al.LLM, persona2025.04
CHI25"You Don't Need a University Degree to Comprehend Data Protection This Way": LLM-Powered Interactive Privacy Policy AssessmentVincent Freiberger, et al.LLM, privacy2025.04
CHI25A Matter of Perspective(s): Contrasting Human and LLM Argumentation in Subjective Decision-Making on Subtle SexismPaula Akemi Aoyagui, et al.LLM2025.04
CHI25AI on My Shoulder: Supporting Emotional Labor in Front-Office Roles with an LLM-based Empathetic CoworkerVedant Das Swain, et al.LLM2025.04
CHI25AI-Instruments: Embodying Prompts as InstrumentsAuthors TBDLLM, prompt, instruments, direct manipulation2025.04
CHI25ASHABot: An LLM-Powered Chatbot to Support the Informational Needs of Community Health WorkersPragnya Ramjee, et al.LLM, chatbot2025.04
CHI25Adaptive Human-LLMs Interaction Collaboration: Reinforcement Learning driven Vision-Language Models for Medical Report GenerationYiming Cao, et al.VLM2025.04
CHI25Align with Me, Not TO Me: How People Perceive Concept Alignment with LLM-Powered Conversational AgentsShengchen Zhang, et al.LLM, agent, alignment2025.04
CHI25Applying the Gricean Maxims to a Human-LLM Interaction Cycle: Design Insights from a Participatory ApproachYoonsu Kim, et al.LLM2025.04
CHI25Artificial Intimacy: Exploring Normativity and Personalization Through Fine-tuning LLM ChatbotsMirabelle Jones, et al.LLM, chatbot, personalization, fine-tuning, normativity2025.04
CHI25Assessing Critical Thinking through a Multi-Agent LLM-Based Debate ChatbotBogyeom Park, et al.LLM, agent, chatbot2025.04
CHI25AutoPBL: An LLM-powered Platform to Guide and Support Individual Learners Through Self Project-based LearningYihao Zhu, et al.LLM2025.04
CHI25BallistoBud: Heart Rate Variability Monitoring using Earbud Accelerometry for Stress AssessmentAuthors TBDearbuds, accelerometer, BCG, heart rate, stress2025.04
CHI25Beyond Adaptation: an LLM-Supported Self-Presentation Ideation Tool for Cross-Cultural MinglingChien-Yin Wu, et al.LLM2025.04
CHI25Beyond Code Generation: LLM-supported Exploration of the Program Design SpaceJ.D. Zamfirescu-Pereira, et al.LLM2025.04
CHI25BioSpark: Beyond Analogical Inspiration to LLM-augmented TransferHyeonsu B Kang, et al.LLM2025.04
CHI25Boosting Diary Study Outcomes with a Fine-Tuned Large Language ModelSunggyeol Oh, et al.LLM2025.04
CHI25Breaking Barriers or Building Dependency? Exploring Team-LLM Collaboration in AI-infused Classroom DebateZihan Zhang, et al.LLM2025.04
CHI25Bridging the Treatment Gap: A Novel LLM-Driven System for Scalable Initial Patient Assessments in Mental HealthcareNiclas Rosteck, et al.LLM2025.04
CHI25BudsID: Mobile-Ready and Expressive Finger Identification Input for EarbudsAuthors TBDearbuds, finger identification, magnetometer, wearable96.9% accuracy2025.04
CHI25COMETIC: Enhancing Smartphone Eye Tracking with Cursor-Based Implicit CalibrationAuthors TBDeye tracking, smartphone, calibration, cursor27.2% improvement2025.04
CHI25Canvil: Designerly Adaptation for LLM-Powered User ExperiencesKenneth Li, et al.LLM, design, UX, user experience, adaptation2025.04
CHI25CaseMaster: Designing a Probe for Oral Case Presentation Training with LLM AssistanceYang Ouyang, et al.LLM2025.04
CHI25ChainBuddy: An AI-assisted Agent System for Generating LLM PipelinesXinyue Chen, et al.LLM, agent, pipeline, AI-assisted, prompt engineering2025.04
CHI25Characterizing LLM-Empowered Personalized Story Reading and Interaction for Children: Insights From Multi-Stakeholder PerspectivesJiaju Chen, et al.LLM, persona, personalization2025.04
CHI25Closing the Loop between User Stories and GUI Prototypes: An LLM-Based Assistant for Cross-Functional Integration in Software DevelopmentFelix Kretzer, et al.LLM, assistant2025.04
CHI25Co-designing Large Language Model Tools for Project-Based Learning with K12 EducatorsPrerna Ravi, et al.LLM2025.04
CHI25Context over Categories: Implementing the Theory of Constructed Emotion with LLM-Guided User AnalysisNils Klüwer, et al.LLM2025.04
CHI25ConversAR: Exploring Embodied LLM-Powered Group Conversations in Augmented Reality for Second Language LearnersJad Bendarkawi, et al.LLM2025.04
CHI25Cross, Dwell, or Pinch: Around-Device Selection Methods for Unmodified SmartwatchesAuthors TBDsmartwatch, sonar, around-device, inputFirst sonar-based around-device input on consumer smartwatch2025.04
CHI25Customizing Emotional Support: How Do Individuals Construct and Interact With LLM-Powered ChatbotsXi Zheng, et al.LLM, chatbot, personalization2025.04
CHI25DBox: Scaffolding Algorithmic Programming Learning through Learner-LLM Co-DecompositionShuai Ma, et al.LLM2025.04
CHI25Dango: A Mixed-Initiative Data Wrangling System using Large Language ModelWei-Hao Chen, et al.LLM2025.04
CHI25DanmuA11y: Making Time-Synced Video Comments Accessible to BLV UsersAuthors TBDaccessibility, BLV, video, Danmu, audio2025.04
CHI25Demonstration of GazeNoter: Enhancing AR Note-Taking Through Gaze-Based Selection of LLM SuggestionsShih-Kang Chiu, et al.LLM2025.04
CHI25Demystifying Mental Health Reports Through an LLM-based ApproachShyama Sastha Krishnamoorthy Srinivasan, et al.LLM2025.04
CHI25Design Principles and Guidelines for LLM Observability: Insights from DevelopersXin Chen, et al.LLM2025.04
CHI25Designing Accessible Audio Nudges for Voice InterfacesHira Jamshed, et al.voice interface, audio, nudging, older adults, accessibility2025.04
CHI25Designing LLM-Powered Multimodal Instructions to Support Rich Hands-on Skills Remote Learning: A Case Study with Massage Instructors and LearnersChutian Jiang, et al.LLM, multimodal2025.04
CHI25Development of an LLM-Based Chatbot to Support Learnability in Stardew Valley: A Diary Study ApproachJungmin Lee, et al.LLM, chatbot2025.04
CHI25Effects of Acoustic Transparency of Wearable Audio Devices on Audio ARYuki Watanabe, et al.audio AR, wearable, acoustic transparency, hearables2025.04
CHI25Effects of LLM-based Search on Decision Making: Speed, Accuracy, and OverrelianceSofia Eleni Spatharioti, et al.LLM2025.04
CHI25Efficient Management of LLM-Based Coaching Agents' Reasoning While Maintaining Interaction Quality and SpeedAndreas Göldi, et al.LLM, agent2025.04
arXiv(v1) 2025Ego4o: Egocentric Human Motion Capture and Understanding from Multi-Modal InputJian Wang, et al.egocentric, IMU, pose estimation, multimodal, VRcs.CV, cs.HC2025.04
CHI25End User Authoring of Personalized Content Classifiers: Comparing Example Labeling, Rule Writing, and LLM PromptingLeijie Wang, et al.LLM, personalization, content classifier, end-user, prompting2025.04
CHI25Enhancing AI Explainability for Non-technical Users with LLM-Driven Narrative GamificationYuzhe You, et al.LLM2025.04
CHI25EvAlignUX: Advancing UX Evaluation through LLM-Supported Metrics ExplorationQingxiao Zheng, et al.LLM2025.04
CHI25Explaining Complex ML Models to Domain Experts Using LLM & Visualization: An Exploration in the French Breadmaking IndustryBriggs Twitchell, et al.LLM2025.04
CHI25Exploring Culturally Informed AI Assistants: A Comparative Study of ChatBlackGPT and ChatGPTLisa Egede, et al.LLM, assistant2025.04
CHI25Exploring Gender Biases in LLM-based Voice Chatbots for Job InterviewsSumin Heo, et al.LLM, chatbot2025.04
CHI25Exploring LLM-Powered Role and Action-Switching Pedagogical Agents for History Education in Virtual RealityZihao Zhu, et al.LLM, agent2025.04
CHI25Exploring Mobile Touch Interaction with Large Language ModelsAuthors TBDLLM, touch, mobile, gesture, interaction2025.04
CHI25Exploring Older Adults Personality Preferences for LLM-powered Conversational CompanionsAjwa Shahid, et al.LLM, persona, personalization2025.04
CHI25Exploring Personalized Health Support through Data-Driven, Theory-Guided LLMs: A Case Study in Sleep HealthXin Tong, et al.LLM, wearable, health, sleep, personalized, chatbotHealthGuru multi-agent framework2025.04
CHI25Exploring the Design Space of Real-time LLM Knowledge Support Systems: A Case Study of Jargon ExplanationsYuhan Liu, et al.LLM2025.04
CHI25Exploring the Design of LLM-based Agent in Enhancing Self-disclosure Among the Older AdultsYijie Guo, et al.LLM, agent2025.04
CHI25Exploring the Impact of Explainability in Large Language Model (LLM) Applications on User ExperienceYanyun Wang, et al.LLM2025.04
CHI25Exploring the Impact of Intervention Methods on Developers’ Security Behavior in a Manipulated ChatGPT StudyRaphael Serafini, et al.LLM2025.04
CHI25FIP: Endowing Robust Motion Capture on Daily Garment by Fusing Flex and Inertial SensorsYiwei Zhao, et al.IMU, flex sensor, motion capture, pose estimation, garment19.5% improvement over SOTA2025.04
CHI25Fact or Fiction? Exploring Explanations to Identify Factual Confabulations in RAG-Based LLM SystemsPhilipp Reinhard, et al.LLM2025.04
CHI25FineType: Fine-grained Tapping Gesture Recognition for Text EntryAuthors TBDtapping, gesture, text entry, IMU, wristband2025.04
CHI25FingerGlass: Enhancing Smart Glasses Interaction via Fingerprint SensingAuthors TBDsmart glasses, fingerprint, gesture recognition, CNN, LSTM2025.04
CHI25Friction: Deciphering Writing Feedback into Writing Revisions through LLM-Assisted ReflectionChao Zhang, et al.LLM2025.04
CHI25From Text to Trust: Empowering AI-assisted Decision Making with Adaptive LLM-powered AnalysisZhuoyan Li, et al.LLM2025.04
CHI25FusAIn: Composing Generative AI Visual Prompts Using Pen-based InteractionXiaohan Peng, et al.generative AI, pen-based, prompt, visual design2025.04
CHI25GPTCoach: Towards LLM-Based Physical Activity CoachingMatthew Jörke, et al.LLM2025.04
CHI25GazeNoter: Co-Piloted AR Note-Taking via Gaze Selection of LLM Suggestions to Match Users' IntentionsZhiyi Rong, et al.LLM, AR, gaze, note-taking, eye tracking2025.04
CHI25Gesture and Audio-Haptic Guidance Techniques to Direct Conversations with Intelligent Voice InterfacesXiyuan Shen, et al.LLM, wearable, voice interface, gesture, haptic, smart glassesRay-Ban Meta Glasses, GPT-4o2025.04
CHI25HaptiCoil: Soft Programmable Buttons with Hydraulically Coupled Haptic Feedback and SensingAuthors TBDhaptic, soft button, sensing, feedback1-500Hz bandwidth2025.04
CHI25Human Robot Interaction for Blind and Low Vision People: A Systematic Literature ReviewAuthors TBDrobot, accessibility, BLV, HRI, survey2025.04
CHI25Human Subjects Research in the Age of Generative AI: Opportunities and Challenges of Applying LLM-Simulated Data to HCI StudiesAngel Hsing-Chi Hwang, et al.LLM2025.04
CHI25IdeationWeb: Tracking the Evolution of Design Ideas in Human-AI Co-CreationAuthors TBDLLM, human-AI co-creation, ideation, design2025.04
CHI25Improving User Engagement and Learning Outcomes in LLM-Based Python Tutor: A Study of PACEMuhtasim Ibteda Shochcho, et al.LLM2025.04
CHI25Inkspire: Supporting Design Exploration with Generative AI through Analogical SketchingAuthors TBDgenerative AI, T2I, sketching, design exploration2025.04
CHI25Interactive Debugging and Steering of Multi-Agent AI SystemsWill Epperson, et al.LLM, agent, multi-agent, debugging, HCI2025.04
CHI25Investigating LLM-Driven Curiosity in Human-Robot InteractionJan Leusmann, et al.LLM2025.04
CHI25LEGOLAS: Learning & Enhancing Golf Skills through LLM-Augmented SystemKangbeen Ko, et al.LLM2025.04
CHI25LIGS: Developing an LLM-infused Game System for Emergent NarrativeJin Jeong, et al.LLM2025.04
CHI25LLM Adoption in Data Curation Workflows: Industry Practices and InsightsCrystal Qian, et al.LLM2025.04
CHI25LLM Integration in Extended Reality: A Comprehensive Review of Current Trends, Challenges, and Future PerspectivesChengkun Wu, et al.LLM, VR, AR, XR, survey, extended realitySurvey paper2025.04
CHI25LLM Powered Text Entry Decoding and Flexible Typing on SmartphonesAuthors TBDLLM, text entry, typing, smartphone, gesture93.1% top-1 accuracy2025.04
CHI25LLM Whisperer: An Inconspicuous Attack to Bias LLM ResponsesWeiran Lin, et al.LLM2025.04
CHI25LearnMate: Enhancing Online Education with LLM-Powered Personalized Learning Plans and SupportXinyu Jessica Wang, et al.LLM, persona, personalization2025.04
CHI25Letters from Future Self: Augmenting the Letter-Exchange Exercise with LLM-based Agents to Enhance Young Adults' Career ExplorationHayeon Jeon, et al.LLM, agent2025.04
CHI25Leveraging Multimodal LLM for Inspirational User Interface SearchSeokhyeon Park, et al.LLM, multimodal2025.04
CHI25LifeInsight: Design and Evaluation of an AI-Powered Assistive Wearable for Blind and Low Vision PeopleAuthors TBDAI, wearable, accessibility, BLV, assistive2025.04
CHI25LittleToDo: Large Language Model Driven Intervention Tool for Adolescent Academic Procrastination with Affective ComputingXiaofan Hu, et al.LLM2025.04
CHI25Lookee: Gaze Tracking-based Infant Vocabulary Comprehension AssessmentAuthors TBDgaze, infant, vocabulary, assessment, AI2025.04
CHI25M2SILENT: Enabling Multi-user Silent Speech Interactions via Multi-directional SpeakersAuthors TBDsilent speech, multi-user, speaker, shared space2025.04
CHI25MAP: Multi-user Personalization with Collaborative LLM-powered AgentsChristine P. Lee, et al.LLM, agent, persona, personalization2025.04
CHI25Maintaining Long-Distance Relationships with (Mediocre) LLM-based Chatbots: A Collaborative Ethnographic StudyBernd Ploderer, et al.LLM, chatbot2025.04
arXiv(v1) 2025MobilePoser: Real-Time Full-Body Pose Estimation and 3D Human Translation from IMUs in Mobile Consumer DevicesVimal Mollyn, et al.IMU, pose estimation, mobile, consumer device, real-timecs.CV, cs.HC2025.04
CHI25MotionBlocks: Modular Geometric Motion Remapping for Accessible VRAuthors TBDVR, accessibility, motion, remapping, limited mobility2025.04
CHI25Objection Overruled! Lay People can Distinguish Large Language Models from Lawyers, but still Favour Advice from an LLMEike Schneiders, et al.LLM2025.04
CHI25Online-EYE: Multimodal Implicit Eye Tracking Calibration for XRAuthors TBDeye tracking, XR, VR, calibration, implicit2025.04
CHI25PPG Earring: Wireless Smart Earring for Heart Health MonitoringAuthors TBDPPG, earring, heart rate, wearable, health monitoring14mm, 2g, 21h battery2025.04
CHI25PaperWave: Listening to Research Papers as Conversational Podcasts Scripted by LLMYuchi Yahagi, et al.LLM2025.04
CHI25Parents, Children, and ChatGPT in Home Environments: The Conversation Content and the Interaction ModeShuang Quan, et al.LLM2025.04
CHI25Piecing Together Teamwork: A Responsible Approach to an LLM-based Educational Jigsaw AgentEmily Doherty, et al.LLM, agent2025.04
CHI25Plan-Then-Execute: An Empirical Study of User Trust and Team Performance When Using LLM Agents As A Daily AssistantGaole He, et al.LLM, agent, assistant2025.04
CHI25Playing Dumb to Get Smart: Creating and Evaluating an LLM-based Teachable Agent within University Computer Science ClassesNaiming Liu, et al.LLM, teachable agent, education, learning2025.04
CHI25PolicyPulse: LLM-Synthesis Tool for Policy ResearchersMaggie Wang, et al.LLM2025.04
CHI25Privacy Meets Explainability: Managing Confidential Data and Transparency Policies in LLM-Empowered ScienceYashothara Shanmugarasa, et al.LLM, privacy2025.04
CHI25Private Yet Social: How LLM Chatbots Support and Challenge Eating Disorder RecoveryRyuhaerang Choi, et al.LLM, chatbot2025.04
CHI25Promoting Cognitive Health in Elder Care with Large Language Model-Powered Socially Assistive RobotsMaria R. Lima, et al.LLM2025.04
CHI25PropType: Everyday Props as Typing Surfaces in Augmented RealityAuthors TBDAR, typing, props, text entry2025.04
CHI25Prototyping with Prompts: Emerging Approaches and Challenges in Generative AI DesignHari Subramonyam, et al.generative AI, prompt engineering, design, prototyping2025.04
CHI25Proxona: Supporting Creators' Sensemaking and Ideation with LLM-Powered Audience PersonasYoonseo Choi, et al.LLM, persona2025.04
CHI25RadEye: Tracking Eye Motion Using FMCW RadarAuthors TBDradar, eye tracking, FMCW, gaze2025.04
CHI25Rambler in the Wild: A Diary Study of LLM-Assisted Writing With SpeechXuyu Yang, et al.LLM2025.04
CHI25Redefining Research Crowdsourcing: Incorporating Human Feedback with LLM-Powered Digital TwinsAmanda Chan, et al.LLM2025.04
CHI25Rescriber: Smaller-LLM-Powered User-Led Data Minimization for LLM-Based ChatbotsJijie Zhou, et al.LLM, chatbot2025.04
CHI25SPECTRA: Personalizable Sound Recognition for DHH Users through Interactive MLSteven M. Goodman, et al.sound recognition, DHH, accessibility, interactive ML2025.04
CHI25Scaffolded Turns and Logical Conversations: Designing Humanized LLM-Powered Conversational Agents for Hospital Admission InterviewsDingdong Liu, et al.LLM, agent2025.04
CHI25Script&Shift: A Layered Interface Paradigm for Integrating Content Development and Rhetorical Strategy with LLM Writing AssistantsMomin N Siddiqui, et al.LLM, assistant2025.04
CHI25Seeing and Touching the Air: Eye-Hand Coordination in Mid-Air Gesture Typing for MRAuthors TBDgesture typing, MR, mid-air, eye-hand coordination2025.04
CHI25Seeking Inspiration through Human-LLM InteractionXinrui Lin, et al.LLM2025.04
CHI25SocialEyes: Scaling Mobile Eye-tracking to Multi-person Social SettingsAuthors TBDeye tracking, mobile, social, multi-person2025.04
CHI25Sonora: Human-AI Co-Creation of 3D Audio WorldsFernanda M De La Torre, et al.AI, audio, 3D, soundscape, LLM, co-creationUses LLM for voice commands2025.04
CHI25Spatial Hand Actions: Hand Actions for Spatial Thinking in 3D AssemblingAuthors TBDhand, spatial, 3D, assembling, gesture2025.04
CHI25Spatial Haptics: A Sensory Substitution Method for Distal Object Detection Using Tactile CuesAuthors TBDhaptic, tactile, sensory substitution, localization2025.04
CHI25Spatial Speech Translation: Translating Across Space With Binaural HearablesAuthors TBDspeech translation, binaural, hearables, spatial audio2025.04
CHI25SpellRing: Recognizing Continuous Fingerspelling in ASL using a RingAuthors TBDring, ASL, sign language, fingerspelling, wearable2025.04
CHI25TableNarrator: Making Image Tables Accessible to Blind and Low Vision PeopleAuthors TBDaccessibility, BLV, tables, image, data2025.04
CHI25Talk to the Hand: an LLM-powered Chatbot with Visual Pointer as Proactive Companion for On-Screen TasksZhepeng Wang, et al.LLM, chatbot, UI, on-screen tasks, visual pointer2025.04
CHI25Tap&Say: Touch Location-Informed LLM for Multimodal Text CorrectionAuthors TBDLLM, touch, voice, multimodal, text correction2025.04
CHI25The Interaction Layer: An Exploration for Co-Designing User-LLM Interactions in Parental Wellbeing Support SystemsSruthi Viswanathan, et al.LLM2025.04
CHI25The Voice of Endo: Leveraging Speech for Illness Flare-up ForecastingAuthors TBDspeech, health, voice, endometriosis, forecasting2025.04
CHI25Through the Lens of Privacy: Exploring Privacy Protection in Vision-Language Model Interactions on Smart GlassesZiyang Zhang, et al.VLM, privacy2025.04
CHI25Too Much Information? Investigating Information Disclosure in Auction Systems with LLM SimulationsYue YinLLM2025.04
CHI25Toward Enabling Natural Conversation with Older Adults via the Design of LLM-Powered Voice Agents that Support Interruptions and BackchannelsChao Liu, et al.LLM, agent2025.04
CHI25Towards Human-AI Deliberation: Design and Evaluation of LLM-Empowered Deliberative AI for AI-Assisted Decision-MakingShuai Ma, et al.LLM2025.04
CHI25UXAgent: An LLM Agent-Based Usability Testing Framework for Web DesignYuxuan Lu, et al.LLM, agent2025.04
CHI25Understanding the Effects of Large Language Model (LLM)-driven Adversarial Social Influences in Online Information SpreadZhuoran Lu, et al.LLM2025.04
CHI25Understanding the LLM-ification of CHI: Unpacking the Impact of LLMs at CHI through a Systematic Literature ReviewRock Yuren Pang, et al.LLM, HCI, CHI, survey, systematic review2025.04
CHI25Unlocking Scientific Concepts: How Effective Are LLM-Generated Analogies for Student Understanding and Classroom Practice?Zekai Shao, et al.LLM2025.04
CHI25Unpacking Trust Dynamics in the LLM Supply Chain: An Empirical Exploration to Foster Trustworthy LLM Production & UseAgathe Balayn, et al.LLM2025.04
CHI25User Experience with LLM-powered Conversational Recommendation Systems: A Case of Music RecommendationSojeong Yun, et al.LLM2025.04
CHI25Users' Expectations and Practices with Agent MemoryBrennan Jones, et al.LLM, agent, memory, user expectations, HCI2025.04
CHI25Utilizing ChatGPT in a Data Structures and Algorithms Course: A Teaching Assistant's PerspectivePooriya Jamie, et al.LLM, assistant2025.04
CHI25VibWalk: Mapping Lower-limb Haptic Experiences of Everyday WalkingAuthors TBDhaptic, walking, vibration, wearable, foot2025.04
CHI25Visiobo Demo: Augmenting Static Prints with Projection-based Visual Cueing and Concept Mapping via LLM ReasoningJiaqi Jiang, et al.LLM2025.04
CHI25Wearable Meets LLM for Stress Management: A Duoethnographic Study Integrating Wearable-Triggered Stressors and LLM Chatbots for Personalized InterventionsSameer Neupane, et al.LLM, wearable, stress, chatbot, personalized intervention2025.04
CHI25Weaving Sound Information to Support Real-Time Sensemaking for DHH UsersJeremy Zhengqi Huang, et al.sound, DHH, accessibility, AI, sensemaking2025.04
CHI25What If Smart Homes Could See Our Homes?: Exploring DIY Smart Home Building Experiences with VLM-Based Camera SensorsSojeong Yun, et al.VLM2025.04
CHI25What Social Media Use Do People Regret? An Analysis of 34K Smartphone Screenshots with Multimodal LLMLongjie Guo, et al.LLM, multimodal2025.04
CHI25WritingRing: Enabling Natural Handwriting Input with a Single IMU RingXiaoying Yang, et al.IMU, ring, handwriting, text entry, wearableSingle IMU ring2025.04
CHI25Your Hands Can Tell: Detecting Redirected Hand Movements in VRAuthors TBDVR, hand, redirection, detection2025.04
arXiv(v1) 2025Broadband shot-to-shot transient absorption anisotropyMaximilian Binzer, et al.VR, AR, egocentric, motion capture, FRAMEphysics.opticsThe following article has been submitted to The Journal of Physical Chemistry. After it is published, it will be found at https://pubs.aip.org/aip/jcp2025.03
IUI25CAIM: A Cognitive AI Memory Framework for Long-term Interaction with LLMsRebecca Westhäußer, et al.LLM, memory, long-term interaction, framework2025.03
Ubicomp25 (IMWUT Vol 9 Issue 1)HandSAW: Wearable Hand-based Event Recognition via On-Body Surface Acoustic WavesKaylee Yaxuan Li, et al.SAW, wrist, hand, object interaction, wearable2025.03
arXiv(v2) 2025Modeling Future Conversation Turns to Teach LLMs to Ask Clarifying QuestionsMichael J. Q. Zhang, et al.IMU, diffusion, pose estimation, loose sensorcs.CLPresented at ICLR 20252025.03
IUI25NoTeeline: Supporting Real-Time, Personalized Notetaking with LLM-Enhanced MicronotesFaria Huq, et al.LLM, note-taking, personalization, real-time, micronotes2025.03
Ubicomp25 (IMWUT Vol 9 Issue 1)Respiration Rate Estimation via Smartwatch-based PPG and Accelerometer DataAuthors TBDrespiration, smartwatch, PPG, accelerometer, transfer learning2025.03
Ubicomp25 (IMWUT Vol 9 Issue 1)SocialMind: LLM-based Proactive AR Social Assistive System with Human-like Perception for In-situ Live InteractionsBufang Yang, et al.LLM, AR, social assistive, proactive2025.03
TEI25Tangible LLMs: Tangible Sense-Making For Trustworthy Large Language ModelsAuthors TBDLLM, tangible, trustworthy AI, physical interface2025.02
CSCW24Is Human-AI Interaction CSCW?Meredith Ringel Morris, et al.human-AI collaboration, CSCW, LLM, panelPanel discussion2024.11
Ubicomp24 (IMWUT Vol 8 Issue 4)Ring-a-Pose: A Ring for Continuous Hand Pose TrackingTianhong Catherine Yu, et al.ring, hand pose, tracking, wearable2024.11
Ubicomp24Sensor2Text: Enabling Natural Language Interactions for Daily Activity Tracking Using Wearable SensorsWenqiang Chen, et al.LLM, wearable, natural language, activity tracking2024.11
Ubicomp24Leveraging LLMs to Predict Affective States via Smartphone Sensor FeaturesAuthors TBDLLM, smartphone, sensing, affective state, digital phenotypingFirst LLM work for affective state prediction2024.10
Ubicomp24Leveraging Large Language Models for Generating Mobile Sensing Strategies in Human Behavior ModelingAuthors TBDLLM, mobile sensing, behavior modeling, strategy generation2024.10
UIST24Patchview: LLM-powered Worldbuilding with Generative Dust and Magnet VisualizationJeongyeon Kim, et al.LLM, worldbuilding, writing, creative, visualization2024.10
UIST24SHAPE-IT: Exploring Text-to-Shape-Display for Generative Shape-Changing Behaviors with LLMsWanli Qian, et al.LLM, shape display, tangible, generative, text-to-shapeAI-chaining approach2024.10
UIST24SituationAdapt: Contextual UI Optimization in Mixed Reality with Situation Awareness via LLM ReasoningZhipeng Li, et al.LLM, MR, UI, adaptive, context-aware, mixed reality2024.10
UIST24VizAbility: Enhancing Chart Accessibility with LLM-based Conversational InteractionMandi Cai, et al.LLM, accessibility, chart, visualization, conversational2024.10
Ubicomp24 (IMWUT Vol 8 Issue 3)IMUGPT 2.0: Language-Based Cross Modality Transfer for Sensor-Based Human Activity RecognitionZikang Leng, et al.IMU, LLM, HAR, cross-modality, motion synthesis20 citations2024.09
arXiv(v1) 2024WheelPoser: Sparse-IMU Based Body Pose Estimation for Wheelchair UsersYunzhi Li, et al.IMU, pose estimation, wheelchair, accessibilitycs.GR, cs.CV, cs.HCAccepted by ASSETS 20242024.09
arXiv(v1) 2024EMHI: A Multimodal Egocentric Human Motion Dataset with HMD and Body-Worn IMUsZelin Ye, et al.IMU, VR, HMD, dataset, egocentric, pose estimationcs.CV, cs.HC885 sequences, 58 subjects, 28.5 hours2024.08
arXiv(v1) 2024Evaluating Text Classification Robustness to Part-of-Speech Adversarial ExamplesAnahita Samadi, et al.IMU, transformer, pose estimation, calibrationcs.CL, cs.LG2024.08
CHI24"My agent understands me better": Integrating Dynamic Human-like Memory Recall and Consolidation in LLM-Based AgentsYuki Hou, et al.LLM, agent, memory, recall, consolidation2024.05
CHI24As an AI language model I cannot: Investigating LLM Denials of User RequestsAuthors TBDLLM, denial, user request, perception2024.05
CHI24Bridging the Gulf of Envisioning: Cognitive Challenges in Prompt Based Interactions with LLMsAuthors TBDLLM, prompt, cognitive challenges, user study2024.05
CHI24ChaCha: Leveraging Large Language Models to Prompt Children to Share Their EmotionsAuthors TBDLLM, chatbot, children, emotion, conversation2024.05
CHI24CharacterMeet: Supporting Creative Writers' Character Construction Through LLM-Powered Chatbot AvatarsAuthors TBDLLM, chatbot, creative writing, character design2024.05
arXiv(v2) 2024Finding Candidate TeV Halos among Very-High Energy SourcesDong Zheng, et al.VR, AR, egocentric, pose estimationastro-ph.HE15 pages, 7 figures, 4 tables, referee's comments incorporated, accepted for publication in ApJ2024.05
CHI24How AI Processing Delays Foster Creativity: CoQuestAuthors TBDLLM, agent, research question, creativity, co-creation2024.05
CHI24Learning Agent-based Modeling with LLM Companions: ChatGPT and NetLogo ChatAuthors TBDLLM, agent-based modeling, NetLogo, learning2024.05
CHI24The HaLLMark Effect: Supporting Provenance and Transparent Use of LLMs in WritingAuthors TBDLLM, writing, provenance, visualization, transparency2024.05
CHI24Towards Robotic Companions: Understanding Handler-Guide Dog Interactions for Informed Guide Dog Robot DesignHochul Hwang, et al.LLM, HCI, CHI, interaction2024.05
CHI24Understanding the Impact of Long-Term Memory on Self-Disclosure with LLM-Driven ChatbotsAuthors TBDLLM, chatbot, long-term memory, self-disclosure, health2024.05
arXiv(v1) 2024Exploring Text-to-Motion Generation with Human PreferenceJenny Sheng, et al.IMU, human object interaction, datasetcs.LG, cs.AI, cs.CVAccepted to CVPR 2024 HuMoGen Workshop2024.04
arXiv(v2) 2024Health-LLM: Large Language Models for Health Prediction via Wearable Sensor DataYubin Kim, et al.LLM, wearable, health predictioncs.CL, cs.AI, cs.LG2024.04
arXiv(v2) 2024PRISM: Patient Records Interpretation for Semantic Clinical Trial Matching using Large Language ModelsShashi Kant Gupta, et al.IMU, RGB, HOI, dataset, trackingcs.CL, cs.AI30 Pages, 8 Figures, Supplementary Work Attached2024.04
arXiv(v1) 2024Bayesian Learned Models Can Detect Adversarial Malware For FreeBao Gia Doan, et al.VR, AR, simulated avatar, headsetcs.CRAccepted to the 29th European Symposium on Research in Computer Security (ESORICS) 2024 Conference2024.03
Ubicomp24 (IMWUT Vol 8 Issue 1)Capturing the College Experience: A Four-Year Mobile Sensing Study of Mental HealthAuthors TBDmobile sensing, mental health, college, longitudinal19 citations2024.03
arXiv(v1) (CVPR24)Dynamic Inertial Poser (DynaIP): Part-Based Motion Dynamics Learning for Enhanced Human Pose Estimation with Sparse Inertial SensorsYu Zhang, et al.IMU, sparse inertial sensorscs.CV2024.03
Ubicomp24HyperHARNafees Ahmad, et al.LLM, passive sensing, sensemaking2024.03
arXiv(v1) 2024LLMs in HCI Data Work: Bridging the Gap Between Information Retrieval and Responsible Research PracticesNeda Taghizadeh Serajeh, et al.VR, AR, avatar control, pose estimationcs.HC, cs.IR5 pages, CHI2024 Workshop on LLMs as Research Tools: Applications and Evaluations in HCI Data Work2024.03
arXiv(v1) 2024Modeling and optimization for arrays of water turbine OWC devicesM. Gambarini, et al.VR, AR, motion capture, egocentric, stereo cameramath.OC, physics.flu-dyn2024.03
arXiv(v1) 2024Modeling stock price dynamics on the Ghana Stock Exchange: A Geometric Brownian Motion approachDennis Lartey Quayesam, et al.VR, AR, motion capture, egocentricmath.OC, q-fin.ST2024.03
arXiv(v1) 2024On depth prediction for autonomous driving using self-supervised learningHoussem BoulahbalVR, AR, avatar, pose estimation, headsetcs.CVPhD thesis2024.03
Ubicomp24ViObjectWenqiang Chen, et al.mental health, LLM, text data2024.03
arXiv(v1) 2024IMUGPT 2.0: Language-Based Cross Modality Transfer for Sensor-Based Human Activity RecognitionZikang Leng, et al.IMU, LLM, cross-modality, HARcs.CV2024.02
arXiv(v2) 2024IMUOptimize: A Data-Driven Approach to Optimal IMU Placement for Human Pose Estimation with Transformer ArchitectureVarun Ramani, et al.IMU, transformer, interpretability, data driven, time seriescs.LG2024.02
arXiv(v1) 2024IMUSIC: IMU-based Facial Expression CaptureYoujia Wang, et al.IMU, generation, simulate, transformer diffusioncs.CVcode coming soon (link)2024.02
Ubicomp23CAvatarWenqiang Chen, et al.human activity, 3D mesh, tactile, pressure2023.12
Ubicomp23A Data-Driven Context-Aware Health Inference System for Children during School Closuresdata analysis, school closures, health inference, risk factor analysis
Ubicomp23Abacus Gestures: A Large Set of Math-Based Usable Finger-Counting Gestures for Mid-Air Interactionsvision, mid-air, gesture interaction, math, finger counting, abacus
ISWC 2023C-Auth: Exploring the Feasibility of Using Egocentric View of Face Contour for User Authentication on Glassessmart glasses, authentication, ecocentric view
Ubicomp24CAvatar: Real-time Human Activity Mesh Reconstruction via Tactile Carpetshuman activity reconstruction, 3D human mesh, pressure and vibrations, tactile sensor
Ubicomp23Contact Tracing for Healthcare Workers in an Intensive Care Unitcontact Tracing, Internet of things (IoT), bluetooth low energy, Covid-19
Ubicomp23DRG-Keyboard: Enabling Subtle Gesture Typing on the Fingertip with Dual IMU Ringstext entry, gesture keyboard, fingertip interaction, smart ring
arXiv(v5) 2024Evaluating Human-Language Model InteractionMina Lee, et al.LM, human-centered, evaluationcs.CL
Ubicomp23Exploring the Opportunities of AR for Enriching Storytelling with Family Photos between Grandparents and GrandchildrenAR, storytelling, intergenerational communication
Ubicomp23Fingerprinting IoT Devices Using Latent Physical Side-Channelsphysical side-channels, fingerprinting, internet-of-things
Ubicomp23From 2D to 3D: Facilitating Single-Finger Mid-Air Typing on QWERTY Keyboards with Probabilistic Touch Modelingmid air, text entry, VR
Ubicomp23GC-Loc: A Graph Attention Based Framework for Collaborative Indoor Localization Using Infrastructure-free Signalscollaborative indoor localization, graph neural network, geomagnetism
Ubicomp23GLOBEM: Cross-Dataset Generalization of Longitudinal Human Behavior Modelinggeneralizability, behavior modeling, passive sensing
Ubicomp23HIPPO: Pervasive Hand-Grip Estimation from Everyday Interactions
arXiv(v1) (CHI23)HOOV: Hand Out-Of-View Tracking for Proprioceptive Interaction using Inertial SensingPaul Streli, et al.IMU, VR, transformercs.HC, cs.CV, I.2; I.5; H.5
Ubicomp23Headar: Sensing Head Gestures for Confirmation Dialogs on Smartwatches with Wearable Millimeter-Wave Radarwearable interaction, gestural input, millimeter-wave radar, head gestures, smartwatch
Ubicomp23HyWay: Enabling Mingling in the Hybrid World∗hybrid mingling, unstructured and semi-structured conversations, awareness, agency, porosity, reciprocity
Ubicomp23I Know Your Intent: Graph-enhanced Intent-aware User Device Interaction Prediction via Contrastive Learninguser device interaction, graph, attention, contrastive learning
Ubicomp22IF-ConvTransformer: A Framework for Human Activity Recognition Using IMU Fusion and ConvTransformerIMU, fusion, multimodal, transformer, attention
arXiv(v1) (CHI23)IMUPoser: Full-Body Pose Estimation using IMUs in Phones, Watches, and EarbudsVimal Mollyn, et al.IMU, pose estimation, BiLSTMcs.HC, cs.CV
Ubicomp23LT-Fall: The Design and Implementation of a Life-threatening Fall Detection and Alarming System
Ubicomp23LapTouch: Using the Lap for Seated Touch Interaction with HMDsVR, seated, touch, on-body
Ubicomp23MI-Poser: Human Body Pose Tracking Using Magnetic and Inertial Sensor Fusion with Metal Interference MitigationEMF, body pose tracking, inverse kinematics, sensor fusion
Ubicomp24Mental-LLM: Leveraging Large Language Models for Mental Health Prediction via Online Text Data
Ubicomp23MoCaPose: Motion Capturing with Textile-integrated Capacitive Sensors in Loose-fitting Smart Garmentsmotion capture, wearable sensing, capacitive sensing, deep learning, motion tracking, smart textile
Ubicomp23N-euro Predictor: A Neural Network Approach for Smoothing and Predicting Motion Trajectoryvision-based interactions, motion-to-photon latency, motion prediction, neural network, perceived jitter and lag
Ubicomp23NF-Heart: A Near-field Non-contact Continuous User Authentication System via Ballistocardiogramcontinuous authentication, ballistocardiogram (BCG), biometrics, non-contact sensing, smart chair
Ubicomp23Naturalistic E-Scooter Maneuver Recognition with Federated Contrastive Rider Interaction LearningIMU, DCT, contrastive learning, asynchronous federated learning, ehavior analysis
ISWC 2023On the Utility of Virtual On-body Acceleration Data for Fine-grained Human Activity RecognitionHAR, virtual, IMU
Ubicomp23PoseSonic: 3D Upper Body Pose Estimation Through Egocentric Acoustic Sensing on Smartglasseshuman pose estimation, acoustic sensing, smart/AR glasses, deep learning, cross-modal supervision
Ubicomp23PrintShear: Shear Input Based on Fingerprint Deformationtouch, finger input
Ubicomp23Privacy-Enhancing Technology and Everyday Augmented Reality: Understanding Bystanders’ Varying Needs for Awareness and ConsentAR, privacy, bystanders, altered reality, extended perception, biometrics
Ubicomp23Radio2Text: Streaming Speech Recognition Using mmWave Radio Signals
Ubicomp23SkinLink: On-body Construction and Prototyping of Reconfigurable Epidermal Interfaces
Ubicomp23Spectral-Loc: Indoor Localization Using Light Spectral Informationindoor localization, spectral information, ambient light
Ubicomp23StructureSense: Inferring Constructive Assembly Structures from User Behaviorstangible user interfaces, TUI, RFID, user modeling, bayesian inference
Ubicomp23Synthetic Smartwatch IMU Data Generation from In-the-wild ASL VideosIMU, synthetic, ASL recognition
Ubicomp23TAO: Context Detection from Daily Activity Patterns Using Temporal Analysis and Ontologybehavioral context recognition, activity recognition, ontology, deep learning
Ubicomp23ThumbAir: In-Air Typing for Head Mounted Displaysmid air, text entry, VR, HMD, user study
ISWC 2023Towards a Haptic Taxonomy of Emotions: Exploring Vibrotactile Stimulation in the Dorsal Region
Ubicomp23TwinkleTwinkle: Interacting with Your Smart Devices by Eye Blinkacoustic sensing, eye blink, signal process
Ubicomp23ViSig: Automatic Interpretation of Visual Body Signals Using On-Body Sensorsvisual signalling, on-body sensors, UWB, IMU, body signals, fallback communication, sports automation, postures, gestures
Ubicomp23VibPath: Two-Factor Authentication with Your Hand's Vibration Response to Unlock Your Phoneuser authentication, vibration, IMU, smartphone, wearables, smartwatch
Ubicomp23Voicify Your UI: Towards Android App Control with Voice Commandsdesign, smartphones, sound-based input, dl parser, UI
Ubicomp23WristAcoustic: Through-Wrist Acoustic Response Based Authentication for Smartwatchessmartwatch authentication, bone conduction, acoustic response
Ubicomp23sUrban: Stable Prediction for Unseen Urban Data from Location-based Sensorsurban computing, location-based data, spatial-temporal prediction, out-of-distribution data

LLM

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026Evaluating the Usage of African-American Vernacular English in Large Language ModelsDeja Dunlap, et al.LLMcs.CL, cs.HC2026.02
arXiv(v1) 2026GUI-Eyes: Tool-Augmented Perception for Visual Grounding in GUI AgentsShuofei Qiao, et al.GUI agent, visual grounding, tool use, perceptioncs.CV, cs.AI2026.01
arXiv(v1) 2025 (WACV 2026)AFRAgent: An Adaptive Feature Renormalization Based High Resolution Aware GUI agentNeeraj Anand, et al.GUI agent, VLM, smartphone automation, multimodalcs.CVWACV 20262025.12
arXiv(v2) 2025Mobile-Agent-v3: Fundamental Agents for GUI AutomationJunyang Wang, et al.GUI agent, mobile, smartphone automation, VLMcs.CV, cs.AI2025.08
arXiv(v6) 2025LLaVA-CoT: Let Vision Language Models Reason Step-by-StepGuowei Xu, et al.LLM, GUI agent, survey, computer usecs.CV17 pages, ICCV 20252025.07
arXiv(v3) 2025Beyond 2:4: exploring V:N:M sparsity for efficient transformer inference on GPUsKang Zhao, et al.LLM, GUI agent, interfacecs.LG, cs.AI2025.06
arXiv(v3) 2025Lai Loss: A Novel Loss for Gradient ControlYuFei LaiLLM, agent, interface, UIcs.LGThe experiment in this article is not very rigorous and may require further testing for its effectiveness2025.05
arXiv(v3) 2025SignLLM: Sign Language Production Large Language ModelsSen Fang, et al.LLM, agent, user interface, HCI, interactioncs.CV, cs.CLwebsite at https://signllm.github.io/2025.04
arXiv(v1) 2024ShowUI: One Vision-Language-Action Model for GUI Visual AgentKevin Qinghong Lin, et al.LLM, UI, vision-language, GUIcs.CV, cs.AI, cs.CL, cs.HCTechnical Report. Github: https://github.com/showlab/ShowUI2024.11
arXiv(v1) 2024A Scalable Communication Protocol for Networks of Large Language ModelsSamuele Marro, et al.LLM, agent, GUI, computer use, interfacecs.AI, cs.LG2024.10
arXiv(v1) 2024In-Band Full-Duplex MIMO Systems for Simultaneous Communications and Sensing: Challenges, Methods, and Future PerspectivesBesma Smida, et al.LLM, GUI agent, computer usecs.IT, cs.ET, eess.SP12 pages, 5 figures, White Paper to appear at IEEE SPM2024.10
arXiv(v2) 2024Massively parallel CMA-ES with increasing populationDavid Redon, et al.LLM, agent, HCI, user interfacecs.DC2024.10
arXiv(v1) 2024OS-ATLAS: A Foundation Action Model for Generalist GUI AgentsZhiyong Wu, et al.LLM, agent, computer use, GUI, foundation modelcs.CL, cs.CV, cs.HC2024.10
arXiv(v1) 2024OrientedFormer: An End-to-End Transformer-Based Oriented Object Detector in Remote Sensing ImagesJiaqi Zhao, et al.LLM, agent, GUI, interfacecs.CVThe paper is accepted by IEEE Transactions on Geoscience and Remote Sensing (TGRS)2024.09
arXiv(v1) 2024Unlocking the Power of Environment Assumptions for Unit ProofsSiddharth Priya, et al.LLM, GUI, agent, surveycs.SE, cs.PLSEFM 20242024.09
arXiv(v2) 2024 (ACL 2025)GUICourse: From General Vision Language Models to Versatile GUI AgentsWentong Chen, et al.GUI agent, VLM, training data, OCR, groundingcs.CV, cs.CL, cs.HCACL 20252024.06
arXiv(v1) 2024Ferret-UI: Grounded Mobile UI Understanding with Multimodal LLMsKeen You, et al.ui, mllm, benchmark, any-resolutioncs.CV, cs.CL, cs.HC2024.04
arXiv(v1) 2024Are You Being Tracked? Discover the Power of Zero-Shot Trajectory Tracing with LLMs!Huanqi Yang, et al.iot, imu, cot, promptcs.CL, cs.AI, cs.HC, cs.LG2024.03
arXiv(v1) 2024Design2Code: How Far Are We From Automating Front-End Engineering?Chenglei Si, et al.llm, auto, googlecs.CL, cs.CV, cs.CY2024.03
arXiv(v2) 2023The Good, The Bad, and Why: Unveiling Emotions in Generative AICheng Li, et al.emotion, prompt, attack, decodecs.AI, cs.CL, cs.HCextension of Large language models understand and can be enhanced by emotional stimuli2023.12
arXiv(v1) (NIPS23)Large Language Model as Attributed Training Data Generator: A Tale of Diversity and BiasYue Yu, et al.synthetic data generationcs.CL, cs.AI, cs.LGarXiv(v2) 2023.10
arXiv(v1) 2023Multimodal Foundation Models: From Specialists to General-Purpose AssistantsChunyuan Li, et al.surveycs.CV, cs.CL2023.09
arXiv(v7) 2023Attention Is All You NeedAshish Vaswani, et al.arxivcs.CL, cs.LG15 pages, 5 figures2023.08

RAG

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026AgenticOCR: Parsing Only What You Need for Efficient Retrieval-Augmented GenerationZhengren Wang, et al.RAGcs.CV, cs.CL2026.02
arXiv(v1) 2026CLFEC: A New Task for Unified Linguistic and Factual Error Correction in paragraph-level Chinese Professional WritingJian Kai, et al.RAGcs.CL2026.02
arXiv(v1) 2026CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM EraZhengqing Yuan, et al.RAGcs.CL, cs.DL2026.02
arXiv(v1) 2026CiteLLM: An Agentic Platform for Trustworthy Scientific Reference DiscoveryMengze Hong, et al.RAGcs.CL, cs.IRAccepted by TheWebConf 2026 Demo Track2026.02
arXiv(v2) 2026MoDora: Tree-Based Semi-Structured Document Analysis SystemBangrui Xu, et al.RAGcs.IR, cs.AI, cs.CL, cs.DB, cs.LGExtension of our SIGMOD 2026 paper. Please refer to source code available at https://github.com/weAIDB/MoDora2026.02
arXiv(v1) 2026Search-P1: Path-Centric Reward Shaping for Stable and Efficient Agentic RAG TrainingTianle Xia, et al.RAGcs.CL, cs.IR, cs.LG2026.02
arXiv(v1) 2026TCM-DiffRAG: Personalized Syndrome Differentiation Reasoning Method for Traditional Chinese Medicine based on Knowledge Graph and Chain of ThoughtJianmin Li, et al.RAGcs.CL, cs.AI2026.02
arXiv(v1) 2026TRIZ-RAGNER: A Retrieval-Augmented Large Language Model for TRIZ-Aware Named Entity Recognition in Patent-Based Contradiction MiningZitong Xu, et al.RAGcs.CL, cs.AI2026.02
arXiv(v1) 2026Truncated Step-Level Sampling with Process Rewards for Retrieval-Augmented ReasoningChris Samarinas, et al.RAGcs.CL, cs.IR2026.02
arXiv(v1) 2026Vectorizing the Trie: Efficient Constrained Decoding for LLM-based Generative Retrieval on AcceleratorsZhengyang Su, et al.RAGcs.IR, cs.CL, cs.LG14 pages, 4 figures2026.02
arXiv(v1) 2026L-RAG: Balancing Context and Retrieval with Entropy-Based Lazy LoadingAuthors TBDRAG, lazy loading, entropy, contextcs.CL, cs.IR2026.01
arXiv(v1) 2025RAGLens: Toward Faithful RAG with Sparse AutoencodersAuthors TBDRAG, hallucination, faithfulness, detectioncs.CL2025.12
arXiv(v1) 2025CDTA: Cross-Document Topic-Aligned Chunking for RAGAuthors TBDRAG, chunking, cross-document, topic alignmentcs.CL, cs.IR0.93 faithfulness on HotpotQA2025.11
arXiv(v1) 2025Agentic RAG for Fintech: Design and EvaluationAuthors TBDRAG, agentic, fintech, query reformulationcs.CL, cs.IR2025.10
arXiv(v1) 2025Practical Code RAG at Scale: Task-Aware Retrieval DesignAuthors TBDRAG, code, retrieval, hybrid, densecs.CL, cs.IRBM25 + dense hybrid2025.10
arXiv(v1) 2025A Systematic Review of Key RAG Systems: Progress, Gaps, and Future DirectionsAuthors TBDRAG, survey, systematic review, knowledge basecs.CL, cs.IR2025.07
arXiv(v1) 2025Late Chunking: Contextual Chunk Embeddings for RAGAuthors TBDRAG, chunking, embedding, contextualcs.CL, cs.IRUpdated July 20252025.07
arXiv(v1) 2025GraphRAG-Bench: When to Use Graphs in RAGAuthors TBDRAG, graph, benchmark, evaluationcs.CL, cs.IR2025.06
arXiv(v1) 2025RAG Survey: Architectures, Enhancements, and Robustness FrontiersAuthors TBDRAG, survey, architecture, robustnesscs.CL, cs.IR2025.06
arXiv(v1) 2025Rethinking Chunk Size for Long-Document Retrieval: Multi-Dataset AnalysisAuthors TBDRAG, chunking, chunk size, retrievalcs.CL, cs.IR64-1024 tokens optimal2025.05
arXiv(v1) 2025A Survey of Multimodal Retrieval-Augmented GenerationZihan Zhao, et al.multimodal, RAG, retrieval, survey, vision-languagecs.CV, cs.CL2025.04
arXiv(v1) 2025RAG Evaluation in the Era of LLMs: A Comprehensive SurveyAuthors TBDRAG, evaluation, benchmark, LLMcs.CL, cs.IR2025.04
arXiv(v1) 2025HiRAG: Retrieval-Augmented Generation with Hierarchical KnowledgeAuthors TBDRAG, hierarchical, knowledge, indexingcs.CL, cs.IR2025.03
arXiv(v1) 2025Ask in Any Modality: A Comprehensive Survey on Multimodal Retrieval-Augmented GenerationZihan Wang, et al.multimodal, RAG, retrieval, survey, LLMcs.CL, cs.IR2025.02
arXiv(v1) 2025GFM-RAG: Graph Foundation Model for Retrieval Augmented GenerationAuthors TBDRAG, graph, foundation model, knowledge graphcs.CL, cs.IR8M params, 60 KGs, 14M triples2025.02
arXiv(v1) 2025KG2RAG: Knowledge Graph-Guided Retrieval Augmented GenerationAuthors TBDRAG, knowledge graph, retrieval, fact-levelcs.CL, cs.IR2025.02
arXiv(v1) 2025RAG-Fusion: Query Expansion and Multi-Source RetrievalAuthors TBDRAG, query expansion, multi-source, fusioncs.CL, cs.IR2025.02
arXiv(v1) 2025Vendi-RAG: Adaptively Trading Off Diversity and Quality in RAGAuthors TBDRAG, diversity, quality, adaptivecs.CL, cs.IR2025.02
arXiv(v1) 2025Agentic RAG: A SurveyAuthors TBDRAG, agentic, survey, LLM agentcs.CL, cs.IR2025.01
arXiv(v1) 2025CG-RAG: Citation Graph RAG for Research Question AnsweringAuthors TBDRAG, citation graph, research QAcs.CL, cs.IR2025.01
arXiv(v1) 2025ChunkRAG: Novel Context-Aware Chunking for RAG SystemsAuthors TBDRAG, chunking, context-aware, retrievalcs.CL, cs.IR2025.01
arXiv(v1) 2024LLM-Augmented Retrieval: Enhancing Retrieval Models Through Language Models and Doc-Level Embeddingrelevant query, doc-Level embedding, embedding-based retrieval, dense retrieval2024.04
arXiv(v6) 2024Health-LLM: Personalized Retrieval-Augmented Disease Prediction SystemQinkai Yu, et al.RAG, XGBoost, AutoMLcs.CL2024.03
arXiv(v1) 2024CRUD-RAG: A Comprehensive Chinese Benchmark for Retrieval-Augmented Generation of Large Language Modelsretrieval-augmented generation, large language models, evaluation2024.02

Agent

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026"Are You Sure?": An Empirical Study of Human Perception Vulnerability in LLM-Driven Agentic SystemsXinfeng Li, et al.LLM, agentcs.HC, cs.AI, cs.CR, cs.SI2026.02
arXiv(v1) 2026AgentDropoutV2: Optimizing Information Flow in Multi-Agent Systems via Test-Time Rectify-or-Reject PruningYutong Wang, et al.LLM, agentcs.AI, cs.CL2026.02
arXiv(v1) 2026E3VA: Enhancing Emotional Expressiveness in Virtual Conversational AgentsAbhishek Kulkarni, et al.LLM, agentcs.HC5 pages2026.02
arXiv(v1) 2026ESAA: Event Sourcing for Autonomous Agents in LLM-Based Software EngineeringElzo Brito dos Santos FilhoLLM, agentcs.AI13 pages, 1 figure, 4 tables. Includes 5 technical appendices2026.02
arXiv(v1) 2026From Flat Logs to Causal Graphs: Hierarchical Failure Attribution for LLM-based Multi-Agent SystemsYawen Wang, et al.LLM, agentcs.AI, cs.SE2026.02
arXiv(v1) 2026HotelQuEST: Balancing Quality and Efficiency in Agentic SearchGuy Hadad, et al.LLM, agentcs.IR, cs.AITo be published in EACL 20262026.02
arXiv(v1) 2026Hybrid LLM-Embedded Dialogue Agents for Learner Reflection: Designing Responsive and Theory-Driven InteractionsParas Sharma, et al.LLM, agentcs.HC, cs.AI2026.02
arXiv(v1) 2026ProductResearch: Training E-Commerce Deep Research Agents via Multi-Agent Synthetic Trajectory DistillationJiangyuan Wang, et al.LLM, agentcs.AI2026.02
arXiv(v1) 2026PseudoAct: Leveraging Pseudocode Synthesis for Flexible Planning and Action Control in Large Language Model AgentsYihan, et al.LLM, agentcs.AI, eess.SY2026.02
arXiv(v1) 2026SafeGen-LLM: Enhancing Safety Generalization in Task Planning for Robotic SystemsJialiang Fan, et al.LLM, agentcs.RO, cs.AI12 pages, 6 figures2026.02
arXiv(v1) 2026The Auton Agentic AI FrameworkSheng Cao, et al.LLM, agentcs.AI2026.02
arXiv(v1) 2026InfiniteWeb: Scalable Web Environment Synthesis for GUI Agent TrainingAuthors TBDagent, web, GUI, training, environmentcs.CV, cs.AI600 tasks across websites2026.01
arXiv(v1) 2026LLM-Based Agentic Systems for Software Engineering: Challenges and OpportunitiesAuthors TBDcode agent, software engineering, surveycs.SE, cs.AI2026.01
arXiv(v1) 2025Beyond Task Completion: Assessment Framework for Agentic AI SystemsAuthors TBDLLM agent, evaluation, framework, benchmarkcs.AI, cs.CL2025.12
arXiv(v1) 2025DeepCode: Open Agentic CodingAuthors TBDcode agent, agentic coding, open sourcecs.SE, cs.AI2025.12
arXiv(v1) 2025LongVideoAgent: Multi-Agent Reasoning with Long VideosJingyi Zhang, et al.multi-agent, video understanding, LLM, reasoningcs.CV, cs.AI2025.12
arXiv(v1) 2025MAR: Multi-Agent Reflexion Improves Reasoning Abilities in LLMsXinyuan Lu, et al.multi-agent, LLM, reflexion, reasoning, debatecs.CL, cs.AI2025.12
arXiv(v1) 2025SWE-RL: Training Superintelligent Software Agents through Self-PlayAuthors TBDcode agent, self-play, RL, software engineeringcs.SE, cs.LG2025.12
arXiv(v1) 2025Agent0: Self-Evolving Agents from Zero Data via Tool-Integrated ReasoningAuthors TBDLLM agent, self-evolving, tool use, reasoningcs.AI, cs.CL18% improvement on math, 24% on general2025.11
arXiv(v1) 2025Building Browser Agents: Architecture, Security, and Practical SolutionsAuthors TBDagent, browser, security, architecturecs.AI, cs.CL2025.11
arXiv(v1) 2025ReMA: Unlocking the Power of Multi-Agent LLM for Reasoning: From Lazy Agents to DeliberationZhiwei Zhang, et al.multi-agent, LLM, reasoning, lazy agent, deliberationcs.AI, cs.CL2025.11
arXiv(v1) 2025BrowserAgent: Web Agents with Human-Inspired Browsing ActionsAuthors TBDagent, browser, web, automationcs.AI, cs.CL2025.10
arXiv(v1) 2025Dark Patterns Impact on LLM-Based Web AgentsAuthors TBDagent, web, dark patterns, decision makingcs.HC, cs.AI2025.10
arXiv(v1) 2025Architecting Resilient LLM AgentsAuthors TBDLLM agent, resilience, architecturecs.AI, cs.CL2025.09
arXiv(v1) 2025A Survey on Code Generation with LLM-based AgentsAuthors TBDcode agent, LLM, code generation, surveycs.SE, cs.AI2025.08
arXiv(v1) 2025LLM-based Agentic Reasoning Frameworks: A Survey from Methods to ScenariosBingxi Zhao, et al.LLM, agent, reasoning, survey, frameworkcs.AI, cs.CL2025.08
arXiv(v1) 2025MCP-Bench: Benchmarking Tool-Using LLM AgentsAuthors TBDagent, benchmark, tool use, MCPcs.AI, cs.CL2025.08
arXiv(v4) 2025Towards Embodied Agentic AI: Review and Classification of LLM- and VLM-Driven Robot Autonomy and InteractionHarshitha Manoj, et al.embodied AI, robot, LLM, VLM, autonomy, surveycs.RO, cs.AI2025.08
arXiv(v1) 2025Understanding Tool-Integrated ReasoningHeng Lin, et al.LLM, agent, tool use, reasoning, TIRcs.LG, cs.AI, stat.ML2025.08
arXiv(v2) 2025Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic ToolsJunde Wu, et al.LLM, agent, agentic reasoning, tool use, frameworkcs.AI, cs.CLACL 20252025.07
arXiv(v3) 2025Embodied AI Agents: Modeling the WorldPascale Fung, et al.embodied AI, world model, VLM, robot, avatarcs.AIMeta AI2025.07
arXiv(v1) 2025Routine: A Structural Planning Framework for LLM Agent System in EnterpriseAuthors TBDLLM agent, planning, enterprise, frameworkcs.AI, cs.CL2025.07
arXiv(v1) 2025Toward a Theory of Agents as Tool-Use Decision-MakersAuthors TBDLLM agent, tool use, theory, decision makingcs.AI, cs.CL2025.06
arXiv(v1) 2025WebRL: Training LLM Web Agents via Self-Evolving Curriculum RLAuthors TBDagent, web, RL, curriculum, self-evolvingcs.LG, cs.AI2025.05
arXiv(v1) 2025AgentRewardBench: Benchmark for LLM Judges in Web Agent EvaluationAuthors TBDagent, benchmark, web, evaluation, rewardcs.AI, cs.CL1,302 trajectories2025.04
arXiv(v1) 2025From LLM Reasoning to Autonomous AI Agents: A Comprehensive ReviewMohamed Amine Ferrag, et al.LLM, agent, survey, autonomous, comprehensive reviewcs.AI, cs.LG2025.04
arXiv(v1) 2024Simultaneous identification of the parameters in the plasticity function for power hardening materials : A Bayesian approachSalih Tatar, et al.LLM, agent, computer use, evaluationmath.NA, math.AP2024.12
arXiv(v2) 2024Feasibility Consistent Representation Learning for Safe Reinforcement LearningZhepeng Cen, et al.LLM, agent, UI, LAUI, interfacecs.LGICML 20242024.06
arXiv(v1) 2024DataDreamer: A Tool for Synthetic Data Generation and Reproducible LLM WorkflowsAjay Patel, et al.pipeline, liarbry, generationcs.CL, cs.LG2024.02
arXiv(v1) 2024More Agents Is All You NeedJunyou Li, et al.multi agent, vote, taskcs.CL, cs.AI, cs.LG2024.02
arXiv(v1) (NIPS23)HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging FaceYongliang Shen, et al.hugging face, APIcs.CL, cs.AI, cs.CV, cs.LG2023.12
arXiv(v1) (NIPS23)CAMEL: Communicative Agents for "Mind" Exploration of Large Language Model SocietyGuohao Li, et al.role play, autonomous, user&assistantcs.AI, cs.CL, cs.CY, cs.LG, cs.MA2023.11(v2)
arXiv(v2) 2023MusicAgent: An AI Agent for Music Understanding and Generation with Large Language ModelsDingyao Yu, et al.pipelinecs.CL, cs.MM, eess.AS2023.10
arXiv(v2) 2023VOYAGER: An Open-Ended Embodied Agent with Large Language ModelsGuanzhi Wang, et al.multi, autonomous, microcraft, gamecs.AI, cs.LG2023.10
arXiv(v3) 2023The Rise and Potential of Large Language Model Based Agents: A SurveyZhiheng Xi, et al.survey, github paper listcs.AI, cs.CL2023.09
arXiv(v3) 2023A Survey on Large Language Model based Autonomous AgentsLei Wang, et al.survey, autonomouscs.AI, cs.CLLatest version is v4(2024.03), double columns. But v3(2023.09) single columns is easy to read.
arXiv(v1) (ICLR24)MetaGPT: Meta Programming for A Multi-Agent Collaborative FrameworkSirui Hong, et al.autonomous system, SOP, multi-agent, frameworkcs.AI, cs.MA

Agentic-RL

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026CUDA Agent: Large-Scale Agentic RL for High-Performance CUDA Kernel GenerationWeinan Dai, et al.LLM, agentic-RLcs.LG, cs.AI2026.02
arXiv(v1) 2026Exploratory Memory-Augmented LLM Agent via Hybrid On- and Off-Policy OptimizationZeyuan Liu, et al.LLM, agentic-RLcs.LG, cs.AIAccepted to ICLR 20262026.02
arXiv(v1) 2026FactGuard: Agentic Video Misinformation Detection via Reinforcement LearningZehao Li, et al.LLM, agentic-RLcs.AI2026.02
arXiv(v1) 2026RF-Agent: Automated Reward Function Design via Language Agent Tree SearchNing Gao, et al.LLM, agentic-RLcs.AI, cs.LG39 pages, 9 tables, 11 figures, Project page see https://github.com/deng-ai-lab/RF-Agent2026.02
arXiv(v1) 2026RUMAD: Reinforcement-Unifying Multi-Agent DebateChao Wang, et al.LLM, agentic-RLcs.AI13 pages, 3 figures2026.02
arXiv(v1) 2026Recycling Failures: Salvaging Exploration in RLVR via Fine-Grained Off-Policy GuidanceYanwei Ren, et al.LLM, agentic-RLcs.AI, cs.CL2026.02
arXiv(v2) 2026Regularized Online RLHF with Generalized Bilinear PreferencesJunghyun Lee, et al.LLM, agentic-RLcs.LG, stat.ML43 pages, 1 table (ver2: more colorful boxes, fixed some typos)2026.02
arXiv(v1) 2026RewardUQ: A Unified Framework for Uncertainty-Aware Reward ModelsDaniel Yang, et al.LLM, agentic-RLcs.LG, cs.AI, cs.CL2026.02
arXiv(v1) 2026Agent Drift: Quantifying Behavioral Degradation in Multi-Agent LLM Systems Over Extended InteractionsAbhishek Rathmulti-agent, behavioral degradation, LLM, agentcs.AI2026.01
arXiv(v5) 2026AgentOrchestra: Orchestrating Multi-Agent Intelligence with the Tool-Environment-Agent(TEA) ProtocolWentao Zhang, et al.hierarchical multi-agent, task solving, LLM, agentcs.AI2026.01
arXiv(v1) 2026Beyond Perfect APIs: A Comprehensive Evaluation of LLM Agents Under Real-World API ComplexityDoyoung Kim, et al.evaluation, benchmark, API complexity, LLM, agentcs.CL, cs.AI26 pages2026.01
arXiv(v1) 2026Beyond Rule-Based Workflows: An Information-Flow-Orchestrated Multi-Agents Paradigm via Agent-to-Agent Communication from CORALXinxing Ren, et al.information flow, multi-agent communication, LLM, agentcs.AI2026.01
arXiv(v2) 2026Cochain: Balancing Insufficient and Excessive Collaboration in LLM Agent WorkflowsJiaxing Zhao, et al.chain-of-collaboration, multi-agent, LLM, agentcs.CL35 pages, 23 figures2026.01
arXiv(v2) 2026Collaborate, Deliberate, Evaluate: How LLM Alignment Affects Coordinated Multi-Agent OutcomesAbhijnan Nath, et al.alignment, multi-agent coordination, LLM, agentcs.CL, cs.AI, cs.LGThis submission is a new version of arXiv:2509.05882v1. with a substantially revised experimental pipeline and new metrics. In particular, collaborator agents are now instantiated independently via separate API calls, rather than generated autoregressively by a single agent. All experimental results are new. Accepted as an extended abstract at AAMAS 20262026.01
arXiv(v1) 2026DemMA: Dementia Multi-Turn Dialogue Agent with Expert-Guided Reasoning and Action SimulationYutong Song, et al.dementia dialogue, multi-turn, medical, LLM, agentcs.MA2026.01
arXiv(v1) 2026EduSim-LLM: An Educational Platform Integrating Large Language Models and Robotic Simulation for BeginnersShenqi Lu, et al.educational platform, robotics, simulation, LLMcs.RO2026.01
arXiv(v1) 2026Game-Theoretic Lens on LLM-based Multi-Agent SystemsJianing Hao, et al.game theory, multi-agent, LLM, agentcs.MA, cs.GT9 pages, 5 figures2026.01
arXiv(v2) 2026 (Spotlight paper of NeurIPS 2025)KARMA: Leveraging Multi-Agent LLMs for Automated Knowledge Graph EnrichmentYuxing Lu, et al.knowledge graph enrichment, multi-agent, LLM, agentcs.CL, cs.AI, cs.CE, cs.DL24 pages, 3 figures, 2 tables2026.01
arXiv(v1) 2026LLM-in-Sandbox Elicits General Agentic IntelligenceDaixuan Cheng, et al.sandbox, agentic intelligence, LLMcs.CL, cs.AIProject Page: https://llm-in-sandbox.github.io2026.01
arXiv(v2) 2026Lifelong Learning of Large Language Model based Agents: A RoadmapJunhao Zheng, et al.lifelong learning, continual learning, LLM, agentcs.AIAccepted to IEEE TPAMI2026.01
arXiv(v1) 2026Nalar: An agent serving frameworkMarco Laju, et al.workflow serving, agent workflows, LLM, agentcs.DC, cs.MA2026.01
arXiv(v1) 2026Orchestral AI: A Framework for Agent OrchestrationAlexander Roman, et al.agent orchestration, framework, LLM, agentcs.AI, astro-ph.IM, hep-ph17 pages, 3 figures. For more information visit https://orchestral-ai.com2026.01
arXiv(v1) 2025PRL: Process Reward Learning Improves LLM ReasoningAuthors TBDagentic RL, PRM, process reward, reasoningcs.LG, cs.AI2026.01
arXiv(v1) 2026ReliabilityBench: Evaluating LLM Agent Reliability Under Production-Like Stress ConditionsAayush Guptareliability evaluation, stress testing, LLM, agentcs.AI18 pages, 5 figures, 8 tables. Evaluates ReAct vs Reflexion across four tool-using domains with perturbation (epsilon) and fault-injection (lambda) stress testing; 1,280 total episodes2026.01
arXiv(v2) 2026SimWorld: An Open-ended Realistic Simulator for Autonomous Agents in Physical and Social WorldsJiawei Ren, et al.simulation, autonomous agent, world model, LLMcs.AI2026.01
arXiv(v2) 2026The Agentic Leash: Extracting Causal Feedback Fuzzy Cognitive Maps with LLMsAkash Kumar Panda, et al.causal extraction, fuzzy cognitive maps, LLM, agentcs.AI, cs.CL, cs.HC, cs.IR15 figures2026.01
arXiv(v2) 2026The Molecular Structure of Thought: Mapping the Topology of Long Chain-of-Thought ReasoningQiguang Chen, et al.chain-of-thought, reasoning topology, LLMcs.CL, cs.AIPreprint2026.01
arXiv(v1) 2026The Two-Stage Decision-Sampling Hypothesis: Understanding the Emergence of Self-Reflection in RL-Trained LLMsZibo Zhao, et al.self-reflection, reinforcement learning, LLM, agentcs.LG, cs.AI2026.01
arXiv(v1) 2026When Numbers Start Talking: Implicit Numerical Coordination Among LLM-Based AgentsAlessio Buscemi, et al.numerical coordination, multi-agent, LLM, agentcs.MA, cs.AI2026.01
arXiv(v2) 2025DoVer: Intervention-Driven Auto Debugging for LLM Multi-Agent SystemsMing Ma, et al.auto debugging, multi-agent, intervention, LLM, agentcs.AI, cs.SE2025.12
arXiv(v1) 2025Enhancing Agentic RL with Progressive Reward Shaping and VSPOAuthors TBDagentic RL, reward shaping, GRPO, tool usecs.LG, cs.AI2025.12
arXiv(v1) 2025From Word to World: Can Large Language Models be Implicit Text-based World Models?Yixia Li, et al.world model, text-based, LLM, agentcs.CL2025.12
arXiv(v2) 2025ToTRL: Unlock LLM Tree-of-Thoughts Reasoning Potential through Puzzles SolvingHaoyuan Wu, et al.tree-of-thought, reasoning, puzzle solving, LLMcs.CL2025.12
arXiv(v2) 2025Understanding LLM Agent Behaviours via Game Theory: Strategy Recognition, Biases and Multi-Agent DynamicsTrung-Kiet Huynh, et al.game theory, agent behavior, LLM, agentcs.MA, cs.AI, cs.GT, cs.LG, math.DS2025.12
arXiv(v2) 2025Large Language Model-based Data Science Agent: A SurveyKe Chen, et al.data science, agent, LLMcs.AI2025.11
arXiv(v1) 2025MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline ParallelismShulin Liu, et al.multi-agent, agentic RL, reasoning, LLM, pipeline parallelismcs.AI10 pages2025.11
arXiv(v1) 2025AGENTRL: Scaling Agentic Reinforcement LearningAuthors TBDagentic RL, RLHF, LLM agent, scalingcs.LG, cs.AIOutperforms GPT-5 and Claude-Sonnet-42025.10
arXiv(v2) 2025BrowserArena: Evaluating LLM Agents on Real-World Web Navigation TasksSagnik Anupam, et al.web navigation, evaluation, real-world, LLM, agentcs.AI, cs.LG2025.10
arXiv(v1) 2025Consistently Simulating Human Personas with Multi-Turn Reinforcement LearningMarwa Abdulhai, et al.persona simulation, multi-turn RL, LLM, agentcs.CL, cs.AI2025.10
arXiv(v3) 2025DS-STAR: Data Science Agent via Iterative Planning and VerificationJaehyun Nam, et al.data science, iterative planning, verification, LLM, agentcs.AI2025.10
arXiv(v1) 2025Deliberate Lab: A Platform for Real-Time Human-AI Social ExperimentsCrystal Qian, et al.social experiments, human-AI interaction, LLM, agentcs.HC, cs.AI2025.10
arXiv(v1) 2025Demystifying Reinforcement Learning in Agentic ReasoningZhaochen Yu, et al.agentic RL, reasoning, LLM, agentcs.CLCode and models: https://github.com/Gen-Verse/Open-AgentRL2025.10
arXiv(v3) 2025DoctorAgent-RL: A Multi-Agent Collaborative Reinforcement Learning System for Multi-Turn Clinical DialogueYichun Feng, et al.clinical dialogue, multi-agent RL, medical, LLM, agentcs.CL2025.10
arXiv(v1) 2025GEM: A Gym for Agentic LLMsZichen Liu, et al.training environment, agentic LLM, gym, agentcs.LG, cs.AI, cs.CL2025.10
arXiv(v3) 2025MLR-Bench: Evaluating AI Agents on Open-Ended Machine Learning ResearchHui Chen, et al.ML research, evaluation, benchmark, LLM, agentcs.LG, cs.AI, cs.CL49 pages, 9 figures. Accepted by NeurIPS 2025 D&B Track2025.10
arXiv(v1) 2025Natural Language Tools: A Natural Language Approach to Tool Calling In Large Language AgentsReid T. Johnson, et al.natural language tool calling, LLM, agentcs.CL31 pages, 7 figures2025.10
arXiv(v1) 2025On Designing Effective RL Reward at Training Time for LLM ReasoningAuthors TBDagentic RL, reward design, reasoning, trainingcs.LG, cs.AI2025.10
arXiv(v1) 2025RLSR: Reinforcement Learning with Supervised RewardAuthors TBDagentic RL, supervised reward, instruction followingcs.LG, cs.AI2025.10
arXiv(v2) 2025SimuRA: A World-Model-Driven Simulative Reasoning Architecture for General Goal-Oriented AgentsMingkai Deng, et al.world model, simulative reasoning, goal-oriented, LLM, agentcs.AI, cs.CL, cs.LG, cs.ROThis submission has been updated to adjust the scope and presentation of the work2025.10
arXiv(v2) 2025 (Am. Statist. (2025) 1-14)A Survey on Large Language Model-based Agents for Statistics and Data ScienceMaojun Sun, et al.data science, statistics, survey, LLM, agentcs.AI, cs.CL, cs.LG, stat.OT2025.09
arXiv(v1) 2025OPPO: Accelerating PPO-based RLHF via Pipeline OverlapAuthors TBDagentic RL, PPO, RLHF, efficiency, overlapcs.LG, cs.AI2025.09
arXiv(v1) 2025RL Foundations for Deep Research Systems: A SurveyAuthors TBDagentic RL, deep research, survey, post-DeepSeekcs.LG, cs.AIPost Feb 2025 papers2025.09
arXiv(v1) 2025Reward Hacking Mitigation using Verifiable Composite RewardsAuthors TBDRLHF, reward hacking, RLVR, verificationcs.LG, cs.AI2025.09
arXiv(v1) 2025SAMULE: Self-Learning Agents Enhanced by Multi-level ReflectionYubin Ge, et al.self-learning, multi-level reflection, LLM, agentcs.AIAccepted at EMNLP 2025 Main Conference2025.09
arXiv(v1) 2025Teaching LLMs to Plan: Logical Chain-of-Thought Instruction Tuning for Symbolic PlanningPulkit Verma, et al.logical chain-of-thought, symbolic planning, LLMcs.AI, cs.CL2025.09
arXiv(v1) 2025The Landscape of Agentic Reinforcement Learning for LLMs: A SurveyAuthors TBDagentic RL, survey, LLM, POMDPcs.LG, cs.AI2025.09
arXiv(v1) 2025Where LLM Agents Fail and How They can Learn From FailuresKunlun Zhu, et al.failure detection, learning from failures, LLM, agentcs.AI2025.09
arXiv(v1) 2025iStar: Agentic Reinforcement Learning with Implicit Step RewardsAuthors TBDagentic RL, credit assignment, implicit PRMcs.LG, cs.AI2025.09
arXiv(v3) 2025MLE-STAR: Machine Learning Engineering Agent via Search and Targeted RefinementJaehyun Nam, et al.ML engineering, code generation, search, LLM, agentcs.LG2025.08
arXiv(v1) 2025Agent Safety Alignment via Reinforcement LearningZeyang Sha, et al.safety alignment, reinforcement learning, LLM, agentcs.AI, cs.CR2025.07
arXiv(v1) 2025AgentMesh: A Cooperative Multi-Agent Generative AI Framework for Software Development AutomationSourena Khanzadehmulti-agent, software development, code generation, LLM, agentcs.SE, cs.AI2025.07
arXiv(v1) 2025Technical Survey of RL Techniques for Large Language ModelsAuthors TBDagentic RL, survey, PPO, DPO, GRPOcs.LG, cs.AI2025.07
arXiv(v2) 2025ToolACE: Winning the Points of LLM Function CallingWeiwen Liu, et al.function calling, tool use, LLM, agentcs.LG, cs.AI, cs.CL21 pages, 22 figures2025.07
arXiv(v1) 2025A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI AutonomyHenry Peng Zou, et al.human-agent systems, collaboration, LLM, agentcs.AI, cs.CL, cs.HC, cs.LG, cs.MA2025.06
arXiv(v2) 2025Beyond Self-Talk: A Communication-Centric Survey of LLM-Based Multi-Agent SystemsBingyu Yan, et al.multi-agent communication, survey, LLM, agentcs.MA, cs.CL2025.06
arXiv(v1) 2025Enhancing Decision-Making of Large Language Models via Actor-CriticHeng Dong, et al.decision-making, actor-critic, LLM, agentcs.CL, cs.AIForty-second International Conference on Machine Learning (ICML 2025)2025.06
arXiv(v1) 2025GENMANIP: LLM-driven Simulation for Generalizable Instruction-Following ManipulationNing Gao, et al.simulation, robotic manipulation, LLM, agentcs.RO2025.06
arXiv(v1) 2025OPT-BENCH: Evaluating LLM Agent on Large-Scale Search Spaces Optimization ProblemsXiaozhe Li, et al.optimization, benchmark, evaluation, LLM, agentcs.AI, cs.LG2025.06
arXiv(v1) 2025RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World EnvironmentsYuchuan Fu, et al.security evaluation, benchmark, LLM, agentcs.CR, cs.AI12 pages, 8 figures2025.06
arXiv(v1) 2025Sailing by the Stars: Survey on Reward Models and Learning StrategiesAuthors TBDagentic RL, reward model, survey, learningcs.LG, cs.AI2025.06
arXiv(v1) 2025ML-Agent: Reinforcing LLM Agents for Autonomous Machine Learning EngineeringZexi Liu, et al.agentic RL, ML engineering, LLM, agentcs.CL, cs.AI, cs.LG2025.05
arXiv(v1) 2025Multi-Agent Systems for Robotic Autonomy with LLMsJunhong Chen, et al.multi-agent, robotics, LLM, agentcs.RO, cs.AI11 pages, 2 figures, 5 tables, submitted for publication2025.05
arXiv(v1) 2025Training LLM-Based Agents with Synthetic Self-Reflected Trajectories and Partial MaskingYihan Chen, et al.synthetic trajectories, self-reflection, training, LLM, agentcs.CL2025.05
arXiv(v1) 2025Agentic Reasoning and Tool Integration for LLMs via Reinforcement LearningJoykirat Singh, et al.agentic RL, tool integration, reasoning, LLMcs.AI2025.04
arXiv(v1) 2025Comprehensive Survey of Reward Models: Taxonomy and ApplicationsAuthors TBDagentic RL, reward model, survey, taxonomycs.LG, cs.AI2025.04
arXiv(v1) 2025DPO Meets PPO: Reinforced Token Optimization for RLHFHan Zhong, et al.agentic RL, DPO, PPO, RLHF, token-levelcs.LG, cs.AIRTO framework2025.04
arXiv(v1) 2025Hierarchical Multi-Step Reward Models for Enhanced ReasoningAuthors TBDagentic RL, reward model, hierarchical, reasoningcs.LG, cs.AI2025.03
arXiv(v1) 2025Look Before You Leap: Using Serialized State Machine for Language Conditioned Robotic ManipulationTong Mu, et al.finite state machine, robotic manipulation, LLMcs.RO, cs.AI7 pages, 4 figures2025.03
arXiv(v1) 2025SafePlan: Leveraging Formal Logic and Chain-of-Thought Reasoning for Enhanced Safety in LLM-based Robotic Task PlanningIke Obi, et al.formal logic, chain-of-thought, safety, robotic planning, LLMcs.RO2025.03
arXiv(v2) 2025Web Agents with World Models: Learning and Leveraging Environment Dynamics in Web NavigationHyungjoo Chae, et al.web navigation, environment dynamics, LLM, agentcs.CLICLR 20252025.03
arXiv(v1) 2025Every Software as an Agent: Blueprint and Case StudyMengwei Xusoftware agent, autonomous, LLMcs.SE, cs.AI2025.02
arXiv(v2) 2025Flow: Modularized Agentic Workflow AutomationBoye Niu, et al.workflow generation, modular, LLM, agentcs.AI, cs.LG, cs.MA2025.02
arXiv(v1) 2025Policy Learning with a Natural Language Action Space: A Causal ApproachBohan Zhang, et al.policy learning, natural language action, LLM, agentcs.CL2025.02
arXiv(v1) 2025Process Reward Models for LLM Agents: Practical FrameworkAuthors TBDPRM, reward model, LLM agent, RLHFcs.LG, cs.AIInversePRM2025.02
arXiv(v1) 2025Provably Efficient Online RLHF with One-Pass Reward ModelingAuthors TBDonline RLHF, reward modeling, efficiencycs.LG, cs.AI2025.02
arXiv(v1) 2025 (Proceedings of the 2024 IEEE International Japan-Africa Conference on Electronics communications and Computations (JAC ECC))Guided Code Generation with LLMs: A Multi-Agent Framework for Complex Code TasksAmr Almorsi, et al.multi-agent, code generation, LLM, agentcs.AI4 pages, 3 figures2025.01
arXiv(v1) 2025REINFORCE++: Critic-Free Policy Optimization with Global Advantage NormalizationAuthors TBDagentic RL, REINFORCE, critic-free, GRPOcs.LG, cs.AIOutperforms PPO2025.01
arXiv(v4) 2024Planning with Multi-Constraints via Collaborative Language AgentsCong Zhang, et al.meta-task planning, multi-agent, LLM, agentcs.AI, cs.CL, cs.LG2024.12
arXiv(v2) 2024AutoWebGLM: A Large Language Model-based Web Navigating AgentHanyu Lai, et al.web navigation, browsing, LLM, agentcs.CLAccepted to KDD 20242024.10

MLLM

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v2) 2026A Taxonomy of Human--MLLM Interaction in Early-Stage Sketch-Based Design IdeationWeiayn Shi, et al.VLMcs.HCAccepted at CHI 2026 Posters2026.02
arXiv(v1) 2026Beyond Dominant Patches: Spatial Credit Redistribution For Grounded Vision-Language ModelsNiamul Hassan Samin, et al.VLMcs.CV, cs.AI2026.02
arXiv(v1) 2026Beyond Static Artifacts: A Forensic Benchmark for Video Deepfake Reasoning in Vision Language ModelsZheyuan Gu, et al.VLMcs.CV, cs.AI16 pages, 9 figures. Submitted to CVPR 20262026.02
arXiv(v1) 2026Causal Decoding for Hallucination-Resistant Multimodal Large Language ModelsShiwei Tan, et al.VLMcs.LG, cs.AI, cs.CVPublished in Transactions on Machine Learning Research (TMLR), 20262026.02
arXiv(v1) 2026Dynamic Multimodal Activation Steering for Hallucination Mitigation in Large Vision-Language ModelsJianghao Yin, et al.VLMcs.CV, cs.AIAccepted by ICLR 20262026.02
arXiv(v1) 2026GuardAlign: Test-time Safety Alignment in Multimodal Large Language ModelsXingyu Zhu, et al.VLMcs.CV, cs.MMICLR 20262026.02
arXiv(v1) 2026HulluEdit: Single-Pass Evidence-Consistent Subspace Editing for Mitigating Hallucinations in Large Vision-Language ModelsYangguang Lin, et al.VLMcs.CVaccepted at CVPR 20262026.02
arXiv(v1) 2026Large Multimodal Models as General In-Context ClassifiersMarco Garosi, et al.VLMcs.CVCVPR Findings 2026. Project website at https://circle-lmm.github.io/2026.02
arXiv(v1) 2026Look Carefully: Adaptive Visual Reinforcements in Multimodal Large Language Models for Hallucination MitigationXingyu Zhu, et al.VLMcs.CVICLR 20262026.02
arXiv(v1) 2026MediX-R1: Open Ended Medical Reinforcement LearningSahal Shaji Mullappilly, et al.VLMcs.CV2026.02
arXiv(v1) 2026NoLan: Mitigating Object Hallucinations in Large Vision-Language Models via Dynamic Suppression of Language PriorsLingfeng Ren, et al.VLMcs.CV, cs.AI, cs.CLCode: https://github.com/lingfengren/NoLan2026.02
arXiv(v1) 2026See It, Say It, Sorted: An Iterative Training-Free Framework for Visually-Grounded Multimodal Reasoning in LVLMsYongchang Zhang, et al.VLMcs.CVCVPR2026 Accepted2026.02
arXiv(v1) 2026Seeing Graphs Like Humans: Benchmarking Computational Measures and MLLMs for Similarity AssessmentSeokweon Jung, et al.VLMcs.HC21 pages including 1 page of appendix, 9 figures, 4 tables2026.02
arXiv(v1) 2026Suppressing Prior-Comparison Hallucinations in Radiology Report Generation via Semantically Decoupled Latent SteeringAo Li, et al.VLMcs.CV15 pages, 5 figures2026.02
arXiv(v1) 2026SurGo-R1: Benchmarking and Modeling Contextual Reasoning for Operative Zone in Surgical VideoGuanyi Qin, et al.VLMcs.CV, cs.AI2026.02
arXiv(v1) 2026Toward Guarantees for Clinical Reasoning in Vision Language Models via Formal VerificationVikash Singh, et al.VLMcs.CV, cs.AI, cs.CL, cs.LO2026.02
arXiv(v1) 2026VAUQ: Vision-Aware Uncertainty Quantification for LVLM Self-EvaluationSeongheon Park, et al.VLMcs.CV, cs.AI, cs.CL2026.02
arXiv(v1) 2026Ground What You See: Hallucination-Resistant MLLMs via Caption FeedbackAuthors TBDMLLM, hallucination, caption feedback, groundingcs.CV, cs.CL2026.01
arXiv(v1) 2026Innovator-VL: A Multimodal Large Language Model for Scientific DiscoveryZichen Wen, et al.scientific discovery, MLLM, vision-languagecs.CV, cs.AIInnovator-VL tech report2026.01
arXiv(v2) 2026MGPC: Multimodal Network for Generalizable Point Cloud Completion With Modality Dropout and Progressive DecodingJiangyuan Liu, et al.point cloud completion, multimodal, MLLMcs.CVCode and dataset are available at https://github.com/L-J-Yuan/MGPC2026.01
arXiv(v1) 2026Multimodal In-context Learning for ASR of Low-resource LanguagesZhaolin Li, et al.multimodal in-context learning, ASR, low-resource, MLLMcs.CL, cs.AIUnder review2026.01
arXiv(v3) 2026Omni-AVSR: Towards Unified Multimodal Speech Recognition with Large Language ModelsUmberto Cappellazzo, et al.unified speech recognition, multimodal, MLLMeess.AS, cs.CV, cs.SDAccepted to IEEE ICASSP 2026 (camera-ready version). Project website (code and model weights): https://umbertocappellazzo.github.io/Omni-AVSR/2026.01
arXiv(v2) 2026Table as a Modality for Large Language ModelsLiyao Li, et al.table modality, structured data, MLLMcs.CL, cs.AIAccepted to NeurIPS 20252026.01
arXiv(v1) 2026The Paradigm Shift: A Comprehensive Survey on Large Vision Language Models for Multimodal Fake News DetectionWei Ai, et al.fake news detection, vision-language, MLLM, surveycs.AI, cs.CV2026.01
arXiv(v3) 2026UniVideo: Unified Understanding, Generation, and Editing for VideosCong Wei, et al.video understanding, generation, editing, unified, MLLMcs.CVProject Website https://congwei1230.github.io/UniVideo/2026.01
arXiv(v1) 2026VLingNav: Embodied Navigation with Adaptive Reasoning and Visual-Assisted Linguistic MemoryShaoan Wang, et al.embodied navigation, adaptive reasoning, MLLM, vision-languagecs.RO, cs.CVProject page: https://wsakobe.github.io/VLingNav-web/2026.01
arXiv(v1) 2026VideoLoom: A Video Large Language Model for Joint Spatial-Temporal UnderstandingJiapeng Shi, et al.video understanding, spatial-temporal, MLLM, vision-languagecs.CV2026.01
arXiv(v1) 2025A Medical Multimodal Diagnostic Framework Integrating Vision-Language Models and Logic Tree ReasoningZelin Zang, et al.medical diagnosis, logic tree reasoning, vision-language, MLLMcs.AI2025.12
arXiv(v1) 2025DiffThinker: Towards Generative Multimodal Reasoning with Diffusion ModelsZefeng He, et al.generative reasoning, diffusion model, MLLMcs.CVProject page: https://diffthinker-project.github.io2025.12
arXiv(v1) 2025From Indoor to Open World: Revealing the Spatial Reasoning Gap in MLLMsAuthors TBDMLLM, spatial reasoning, benchmark, open worldcs.CV, cs.CL2025.12
arXiv(v1) 2025Kling-Omni Technical ReportKling Team, et al.video generation, multimodal synthesis, MLLMcs.CVKling-Omni Technical Report2025.12
arXiv(v1) 2025Lemon: A Unified and Scalable 3D Multimodal Model for Universal Spatial UnderstandingYongyuan Liang, et al.3D understanding, point cloud, spatial understanding, MLLMcs.CV, cs.AI2025.12
arXiv(v3) 2025 (Proc. 2025 IEEE 8th International Conference on Multimedia Information Processing and Retrieval (MIPR), pp. 456-462, 2025)MedChat: A Multi-Agent Framework for Multimodal Diagnosis with Large Language ModelsPhilip R. Liu, et al.multi-agent, medical diagnosis, MLLMcs.MA, cs.AI, cs.CV, cs.LG2025.12
arXiv(v2) 2025TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement LearningTao Wu, et al.temporal understanding, reinforcement learning, MLLM, videocs.CV2025.12
arXiv(v1) 2025MVU-Eval: Multi-Video Understanding Evaluation for MLLMsAuthors TBDMLLM, multi-video, evaluation, benchmarkcs.CV, cs.CL2025.11
arXiv(v1) 2025 (Proceedings of the Conference on Language Modeling (COLM 2025))REM: Evaluating LLM Embodied Spatial Reasoning through Multi-Frame TrajectoriesJacob Thompson, et al.embodied spatial reasoning, trajectory, MLLMcs.LG, cs.AI, cs.CV2025.11
arXiv(v1) 2025Seeing is Believing: Rich-Context Hallucination Detection via Backward Visual GroundingAuthors TBDMLLM, hallucination, detection, visual groundingcs.CV, cs.CLOutperforms GPT-4o2025.11
arXiv(v1) 2025SpatialThinker: Reinforcing 3D Reasoning in Multimodal LLMs via Spatial RewardsAuthors TBDMLLM, 3D reasoning, spatial, RLcs.CV, cs.CLOutperforms GPT-4o2025.11
arXiv(v1) 2025MT-Video-Bench: Video Understanding Benchmark for MLLMs in Multi-Turn DialoguesAuthors TBDMLLM, video understanding, benchmark, multi-turncs.CV, cs.CL2025.10
arXiv(v1) 2025MemVR: Memory-Space Visual Retracing for Hallucination Mitigation in MLLMsAuthors TBDMLLM, hallucination, mitigation, memorycs.CV, cs.CLPlug-and-play2025.10
arXiv(v2) 2025Revealing Multimodal Causality with Large Language ModelsJin Li, et al.causal discovery, MLLM, multimodal causalitycs.LG, cs.AIAccepted at NeurIPS 20252025.10
arXiv(v1) 2025Video-STR: Reinforcing MLLMs in Video Spatio-Temporal Reasoning with Relation GraphWentao Wang, et al.spatio-temporal reasoning, relation graph, MLLM, videocs.AI2025.10
arXiv(v1) 2025Two Causes, Not One: Rethinking Omission and Fabrication Hallucinations in MLLMsAuthors TBDMLLM, hallucination, omission, fabricationcs.CV, cs.CL2025.09
arXiv(v1) 2025VIRAL: Visual Representation Alignment for MLLMsAuthors TBDMLLM, visual alignment, fine-grained understandingcs.CV, cs.CL2025.09
arXiv(v1) 2025Bifrost-1: Bridging Multimodal LLMs and Diffusion Models with Patch-level CLIP LatentsHan Lin, et al.diffusion model, patch-level CLIP, image generation, MLLMcs.CV, cs.AI, cs.CLProject Page: https://bifrost-1.github.io2025.08
arXiv(v1) 2025Grounding the Ungrounded: Spectral-Graph Framework for Quantifying Hallucinations in MLLMsAuthors TBDMLLM, hallucination, grounding, detectioncs.CV, cs.CL2025.08
arXiv(v1) 2025Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A SurveyAuthors TBDMLLM, VLA, robotics, manipulation, surveycs.RO, cs.CV2025.08
arXiv(v1) 2025Multimodal Large Language Models for End-to-End Affective Computing: Benchmarking and Boosting with Generative Knowledge PromptingMiaosen Luo, et al.affective computing, emotion recognition, MLLMcs.AI, cs.LG2025.08
arXiv(v1) 2025RynnEC: Bringing MLLMs into Embodied WorldAuthors TBDMLLM, embodied, video, spatial reasoningcs.CV, cs.RO2025.08
arXiv(v3) 2025SimVecVis: A Dataset for Enhancing MLLMs in Visualization UnderstandingCan Liu, et al.visualization understanding, dataset, MLLMcs.HC, cs.CV2025.07
arXiv(v2) 2025UniCode$^2$: Cascaded Large-scale Codebooks for Unified Multimodal Understanding and GenerationYanzhe Chen, et al.codebook, multimodal generation, MLLMcs.CV, cs.MM19 pages, 5 figures2025.07
arXiv(v1) 2025CLiViS: Unleashing Cognitive Map through Linguistic-Visual Synergy for Embodied Visual ReasoningKailing Li, et al.embodied visual reasoning, cognitive map, MLLMcs.CV, cs.AI, cs.CL2025.06
arXiv(v1) 2025Insight-V: Exploring Long-Chain Visual Reasoning with MLLMsAuthors TBDMLLM, visual reasoning, long-chain, CVPRcs.CVCVPR 20252025.06
arXiv(v2) 2025LLaDA-V: Large Language Diffusion Models with Visual Instruction TuningZebin You, et al.diffusion model, visual instruction tuning, MLLMcs.LG, cs.CL, cs.CVProject page and codes: \url{https://ml-gsai.github.io/LLaDA-V-demo/}2025.06
arXiv(v2) 2025LLaVA-ST: A Multimodal Large Language Model for Fine-Grained Spatial-Temporal UnderstandingHongyu Li, et al.spatial-temporal understanding, MLLM, vision-languagecs.CVAccepted by CVPR20252025.06
arXiv(v1) 2025 (CVPR 2025)LLaVA-ST: Multimodal LLM for Fine-Grained Spatial-Temporal UnderstandingAuthors TBDMLLM, spatial-temporal, video, CVPRcs.CVCVPR 20252025.06
arXiv(v1) 2025Manager: Aggregating Insights from Unimodal Experts in VLMs and MLLMsAuthors TBDMLLM, VLM, unimodal experts, fusioncs.CV, cs.CL2025.06
arXiv(v1) 2025MedTVT-R1: A Multimodal LLM Empowering Medical Reasoning and DiagnosisYuting Zhang, et al.medical reasoning, diagnosis, MLLMeess.IV, cs.CL, cs.CV, q-bio.QM2025.06
arXiv(v1) 2025Multimodal Tabular Reasoning with Privileged Structured InformationJun-Peng Jiang, et al.tabular reasoning, structured information, MLLMcs.LG, cs.AI, cs.CL, cs.CV2025.06
arXiv(v1) 2025Pts3D-LLM: Studying the Impact of Token Structure for 3D Scene Understanding With Large Language ModelsHugues Thomas, et al.3D scene understanding, point cloud, token structure, MLLMcs.CVMain paper and appendix2025.06
CVPR25 (CVPR 2025)Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention ReweightingAuthors TBDMLLM, hallucination, attention, CVPRcs.CVCVPR 20252025.06
arXiv(v3) 2025SpatialLLM: A Compound 3D-Informed Design towards Spatially-Intelligent Large Multimodal ModelsWufei Ma, et al.3D-informed, spatial intelligence, MLLMcs.CVCVPR 2025 highlight2025.06
arXiv(v1) 2025Structured Attention Matters to Multimodal LLMs in Document UnderstandingChang Liu, et al.document understanding, structured attention, MLLMcs.CL, cs.AI, cs.IR2025.06
arXiv(v2) 2025Watch and Listen: Understanding Audio-Visual-Speech Moments with Multimodal LLMZinuo Li, et al.audio-visual-speech, multimodal, MLLM, video understandingcs.CL2025.06
arXiv(v3) 2025Cosmos-Reason1: From Physical Common Sense To Embodied ReasoningNVIDIA, et al.physical common sense, embodied reasoning, MLLMcs.AI, cs.CV, cs.LG, cs.RO2025.05
arXiv(v1) 2025HoloLLM: Multisensory Foundation Model for Language-Grounded Human Sensing and ReasoningChuhao Zhou, et al.multisensory, human sensing, reasoning, MLLMcs.CV, cs.AI, cs.CL, cs.LG, cs.MM18 pages, 13 figures, 6 tables2025.05
arXiv(v2) 2025Kubrick: Multimodal Agent Collaborations for Synthetic Video GenerationLiu He, et al.video generation, multi-agent collaboration, MLLMcs.CV, cs.GR, cs.MMAccepted by CVPR 2025 AI4CC Workshop2025.05
arXiv(v1) 2025MedBridge: Bridging Foundation Vision-Language Models to Medical Image DiagnosisAuthors TBDMLLM, medical, VLM, diagnosiscs.CV2025.05
arXiv(v1) 2025VideoLLM Benchmarks and Evaluation: A SurveyAuthors TBDMLLM, video, benchmark, evaluation, surveycs.CV, cs.CL2025.05
arXiv(v2) 2025Can Large Language Models Help Multimodal Language Analysis? MMLA: A Comprehensive BenchmarkHanlei Zhang, et al.multimodal language analysis, MLLM, semanticscs.CL, cs.AI, cs.MM23 pages, 5 figures2025.04
arXiv(v2) 2025Dual Diffusion for Unified Image Generation and UnderstandingZijie Li, et al.dual diffusion, unified generation, understanding, MLLMcs.CV, cs.AI, cs.LG2025.04
arXiv(v1) 2025Multimodal LLMs for OCR, OCR Post-Correction, and Named Entity Recognition in Historical DocumentsGavin Greif, et al.OCR, historical documents, named entity recognition, MLLMcs.CL, cs.AI, cs.DL2025.04
arXiv(v1) 2025Socratic Chart: Cooperating Multiple Agents for Robust SVG Chart UnderstandingYuyang Ji, et al.chart understanding, SVG, multi-agent, MLLMcs.CV2025.04
arXiv(v1) 2025VLM-R1: A Stable and Generalizable R1-Style Large VLMAuthors TBDMLLM, VLM, reasoning, R1-stylecs.CV, cs.CL2025.04
arXiv(v1) 2025MLLM-For3D: Adapting Multimodal Large Language Model for 3D Reasoning SegmentationAuthors TBDMLLM, 3D, segmentation, reasoningcs.CV2025.03
arXiv(v1) 2025MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMsErik Daxberger, et al.MLLM, 3D, spatial, understanding, benchmarkcs.CV, cs.CLICCV 20252025.03
arXiv(v1) 2025Med3DVLM: An Efficient Vision-Language Model for 3D Medical Image AnalysisAuthors TBDMLLM, medical, 3D, VLM, CTcs.CV2025.03
arXiv(v1) 2025Mobile-VideoGPT: Fast and Accurate Video Understanding Language ModelAuthors TBDMLLM, video, mobile, efficientcs.CV, cs.CL2025.03
arXiv(v1) 2025R1-Zero's Aha Moment in Visual Reasoning on a 2B Non-SFT ModelAuthors TBDMLLM, visual reasoning, R1-Zero, emergentcs.CV, cs.CL2025.03
arXiv(v1) 2025SpaceVLLM: Endowing MLLM with Spatio-Temporal Video GroundingAuthors TBDMLLM, video, spatio-temporal, groundingcs.CV, cs.CL2025.03
arXiv(v1) 2025Vision-R1: Incentivizing Reasoning Capability in MLLMsAuthors TBDMLLM, reasoning, visual reasoningcs.CV, cs.CL2025.03
arXiv(v2) 2025Does Table Source Matter? Benchmarking and Improving Multimodal Scientific Table Understanding and ReasoningBohao Yang, et al.table understanding, scientific data, MLLMcs.CL2025.02
arXiv(v4) 2025LMFusion: Adapting Pretrained Language Models for Multimodal GenerationWeijia Shi, et al.multimodal generation, LLM adaptation, MLLMcs.CL, cs.AI, cs.CV, cs.LGName change: LlamaFusion to LMFusion2025.02
arXiv(v1) 2025Multimodal Large Language Models for Text-rich Image Understanding: A Comprehensive ReviewPei Fu, et al.text-rich image understanding, MLLM, vision-languagecs.CV2025.02
arXiv(v1) 2025Visual Perception Token for Multimodal Large Language ModelsAuthors TBDMLLM, visual perception, token, autonomous controlcs.CV, cs.CL829k training samples2025.02
arXiv(v1) 2025Weak Supervision Dynamic KL-Weighted Diffusion Models Guided by Large Language ModelsJulian Perry, et al.diffusion model, LLM guidance, weak supervision, image generationcs.CL2025.02
arXiv(v1) 2025Bridging Visualization and Optimization: Multimodal Large Language Models on Graph-Structured Combinatorial OptimizationJie Zhao, et al.graph-structured optimization, combinatorial, MLLMcs.AI, cs.LG2025.01
arXiv(v1) 2025Exploring the Role of Explicit Temporal Modeling in Multimodal Large Language Models for Video UnderstandingYun Li, et al.temporal modeling, video understanding, MLLMcs.CV, cs.CL2025.01
arXiv(v5) 2025Harnessing Multimodal Large Language Models for Multimodal Sequential RecommendationYuyang Ye, et al.sequential recommendation, MLLM, multimodalcs.IR, cs.AI2025.01

Memory

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026Contextual Memory Virtualisation: DAG-Based State Management and Structurally Lossless Trimming for LLM AgentsCosmo SantoniLLM, memorycs.SE, cs.AI, cs.HC, cs.OS11 pages. 6 figures. Introduces a DAG-based state management system for LLM agents. Evaluation on 76 coding sessions shows up to 86% token reduction (mean 20%) while remaining economically viable under prompt caching. Includes reference implementation for Claude Code2026.02
arXiv(v1) 2026A Dynamic Retrieval-Augmented Generation System with Selective Memory and RemembranceOkan Bursadynamic RAG, selective memory, retrieval, LLMcs.IR, cs.AI6 Pages, 2 figures2026.01
arXiv(v1) 2026Active Context Compression: Autonomous Memory Management in LLM AgentsAuthors TBDmemory, compression, context, autonomous, Focuscs.CL, cs.AI22.7% token savings2026.01
arXiv(v1) 2025Agentic Memory: Learning Unified Long-Term and Short-Term Memory ManagementAuthors TBDmemory, unified, long-term, short-term, managementcs.AI, cs.CL2026.01
arXiv(v1) 2026Beyond Dialogue Time: Temporal Semantic Memory for Personalized LLM AgentsAuthors TBDmemory, temporal, semantic, personalizedcs.CL, cs.AIDurative memory, Zep architecture2026.01
arXiv(v1) 2026Beyond Static Summarization: Proactive Memory Extraction for LLM AgentsChengyuan Yang, et al.memory, extraction, summarization, agentcs.CL, cs.AI2026.01
arXiv(v1) 2026Continuum Memory Architectures for Long-Horizon LLM AgentsAuthors TBDmemory, continuum, long-horizon, consolidationcs.AI, cs.CLEpisodic-to-semantic conversion2026.01
arXiv(v2) 2026Cost and accuracy of long-term memory in Distributed Multi-Agent Systems based on Large Language ModelsBenedict Wolff, et al.graph memory, distributed multi-agent, LLMcs.IR23 pages, 4 figures, 7 tables2026.01
arXiv(v1) 2026Explore with Long-term Memory: A Benchmark and Multimodal LLM-based Reinforcement Learning Framework for Embodied ExplorationSen Wang, et al.memory, benchmark, multimodal, RL, MLLMcs.AI, cs.CVOur dataset and code will be released at our \href{https://wangsen99.github.io/papers/lmee/}{website}2026.01
arXiv(v1) 2026Fine-Mem: Fine-Grained Feedback Alignment for Long-Horizon Memory ManagementWeitao Ma, et al.memory, management, feedback, alignment, agentcs.CL18 pages, 5 figures2026.01
arXiv(v3) 2026HaluMem: Evaluating Hallucinations in Memory Systems of AgentsDing Chen, et al.memory, hallucination, evaluation, agentcs.CL2026.01
arXiv(v1) 2026HiMeS: Hippocampus-inspired Memory System for Personalized AI AssistantsHailong Li, et al.memory, personalized assistant, hippocampus, long-termcs.AI2026.01
arXiv(v1) 2026HiMem: Hierarchical Long-Term Memory for LLM Long-Horizon AgentsNingning Zhang, et al.memory, long-term, hierarchical, agent, LLMcs.AI2026.01
arXiv(v2) 2026Intrinsic Memory Agents: Heterogeneous Multi-Agent LLM Systems through Structured Contextual MemorySizhe Yuen, et al.multi-agent, structured memory, LLM, agentcs.AI2026.01
arXiv(v1) 2026LLMs Can't Play Hangman: On the Necessity of a Private Working Memory for Language AgentsDavide Baldelli, et al.working memory, evaluation, agent, LLMcs.CL2026.01
arXiv(v2) 2026LiCoMemory: Lightweight and Cognitive Agentic Memory for Efficient Long-Term ReasoningZhengjun Huang, et al.cognitive memory, long-term reasoning, LLM, agentcs.IR2026.01
arXiv(v1) 2026MAGMA: A Multi-Graph based Agentic Memory Architecture for AI AgentsDongming Jiang, et al.multi-graph, agentic memory, LLM, agentcs.AI2026.01
arXiv(v1) 2026Mem-Gallery: Benchmarking Multimodal Long-Term Conversational Memory for MLLM AgentsYuanchen Bei, et al.memory, multimodal, conversational, MLLM benchmarkcs.CL, cs.AI34 pages, 18 figures2026.01
arXiv(v1) 2026MemBuilder: Reinforcing LLMs for Long-Term Memory Construction via Attributed Dense RewardsAuthors TBDmemory, long-term, construction, dense rewardcs.CL, cs.AI2026.01
arXiv(v1) 2026RealMem: Benchmarking LLMs in Real-World Memory-Driven InteractionHaonan Bian, et al.memory, benchmark, interaction, agentcs.CL, cs.AI2026.01
arXiv(v3) 2026SimpleMem: Efficient Lifelong Memory for LLM AgentsJiaqi Liu, et al.lifelong memory, efficient, LLM, agentcs.AI2026.01
arXiv(v1) 2026SwiftMem: Fast Agentic Memory via Query-aware IndexingAnxin Tian, et al.memory, retrieval, indexing, agent, LLMcs.CL, cs.AI2026.01
arXiv(v3) 2026TeleMem: Building Long-Term and Multimodal Memory for Agentic AIChunliang Chen, et al.memory, multimodal, long-term, agent, LLMcs.CL, cs.AI, cs.CV2026.01
arXiv(v1) 2025Tool-Memory Conflicts in Tool-Augmented LLMsAuthors TBDmemory, tool use, conflict, LLMcs.CL, cs.AI2026.01
arXiv(v1) (NeurIPS25)A-MEM: Agentic Memory for LLM Agents (NeurIPS)Authors TBDmemory, agentic, Zettelkasten, NeurIPScs.AI, cs.CLNeurIPS 2025 publication2025.12
arXiv(v1) 2025AI Meets Brain: Memory Systems from Cognitive Neuroscience to Autonomous AgentsJiafeng Liang, et al.memory, cognitive, survey, agentcs.CL, cs.AI, cs.CV57 pages, 5 figures2025.12
arXiv(v1) 2025Audited Skill-Graph Self-Improvement for Agentic LLMs via Verifiable Rewards, Experience Synthesis, and Continual MemoryKen Huang, et al.memory, continual, self-improvement, agentcs.CR, cs.AI11 pages, 4 figures. Includes a complete runnable reference implementation and audit logging framework2025.12
arXiv(v1) 2025Beyond Heuristics: A Decision-Theoretic Framework for Agent Memory ManagementChangzhi Sun, et al.memory, agent, decision-theoretic, managementcs.CL2025.12
arXiv(v1) 2025Cache What Lasts: Token Retention for Memory-Bounded KV Cache in LLMsNgoc Bui, et al.memory, KV cache, retention, long-contextcs.LG, cs.AI2025.12
arXiv(v1) 2025Context as a Tool: Context Management for Long-Horizon SWE-AgentsShukai Liu, et al.memory, context management, long-horizon, SWE-agentcs.CL2025.12
arXiv(v2) 2025Evaluating Long-Term Memory for Long-Context Question AnsweringAlessandra Terranova, et al.memory, long-context, evaluation, QAcs.CLAccepted as a poster at Metacognition in Generative AI EurIPS workshop2025.12
arXiv(v1) 2025Hindsight is 20/20: Building Agent Memory that Retains, Recalls, and ReflectsAuthors TBDmemory, hindsight, reflection, retentioncs.AI, cs.CLRetain-Recall-Reflect framework2025.12
arXiv(v1) 2025Learning Hierarchical Procedural Memory for LLM Agents through Bayesian Selection and Contrastive RefinementSaman Forouzandeh, et al.memory, procedural, hierarchical, agentcs.LG, cs.AIAccepted at The 25th International Conference on Autonomous Agents and Multi-Agent Systems (AAMAS 2026). 21 pages including references, with 7 figures and 8 tables. Code is publicly available at the authors GitHub repository: https://github.com/S-Forouzandeh/MACLA-LLM-Agents-AAMAS-Conference2025.12
arXiv(v2) 2025MMAG: Mixed Memory-Augmented Generation for Large Language Models ApplicationsStefano Zeppierimemory, memory-augmented generation, RAG, LLMcs.CL, cs.IR2025.12
arXiv(v1) 2025MemEvolve: Meta-Evolution of Agent Memory SystemsGuibin Zhang, et al.memory, evolution, meta-learning, agentcs.CL, cs.MA2025.12
arXiv(v1) 2025MemR3: Memory Retrieval via Reflective Reasoning for LLM AgentsAuthors TBDmemory, retrieval, reflection, reasoningcs.AI, cs.CLRouter + evidence-gap tracker2025.12
arXiv(v2) 2025Memento 2: Learning by Stateful Reflective MemoryJun Wangmemory, agent, reflection, statefulcs.AI, cs.CV, cs.LG35 pages, four figures2025.12
arXiv(v1) 2025Memory in the Age of AI AgentsYuyang Hu, et al.memory, survey, LLM, agentcs.CL, cs.AIComprehensive survey on agent memory2025.12
arXiv(v1) 2025MemoryGraft: Persistent Compromise of LLM Agents via Poisoned Experience RetrievalSaksham Sahai Srivastava, et al.memory, security, agent, attackcs.CR, cs.AI, cs.LG14 pages, 1 figure, includes appendix2025.12
arXiv(v3) 2025O-Mem: Omni Memory System for Personalized, Long Horizon, Self-Evolving AgentsPiaohong Wang, et al.memory, long-horizon, self-evolving, agentcs.CL2025.12
arXiv(v1) 2025R-Debater: Retrieval-Augmented Debate Generation through Argumentative MemoryMaoyuan Li, et al.memory, retrieval, debate, RAGcs.CL, cs.AIAccepteed by AAMAS 2026 full paper2025.12
arXiv(v2) 2025Significant Other AI: Identity, Memory, and Emotional Regulation as Long-Term Relational IntelligenceSung Parkmemory, identity, relational, long-termcs.HC, cs.AI2025.12
arXiv(v1) 2025Adaptive Focus Memory for Language ModelsAuthors TBDmemory, adaptive, focus, compression, AFMcs.CL, cs.AI2/3 token reduction2025.11
arXiv(v1) 2025BudgetMem: Learning Selective Memory Policies for Cost-Efficient Long-Context ProcessingAuthors TBDmemory, selective, budget, efficientcs.CL, cs.AILearned gating + BM252025.11
arXiv(v1) 2025CoEdge-RAG: Optimizing Hierarchical Scheduling for Retrieval-Augmented LLMs in Collaborative Edge ComputingGuihang Hong, et al.RAG, edge computing, optimization, LLMcs.DCAccepted by RTSS 2025 (Real-Time Systems Symposium, 2025)2025.11
arXiv(v1) 2025EMem: Event-Centric Memory for Long-Term Conversational AgentsAuthors TBDmemory, event-centric, conversation, neo-Davidsoniancs.CL, cs.AIEvent-like propositions2025.11
arXiv(v1) 2025Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving MemoryTianxin Wei, et al.memory, benchmark, test-time learning, agentcs.CL, cs.AI2025.11
arXiv(v1) 2025G-KV: Decoding-Time KV Cache Eviction with Global AttentionMengqi Liao, et al.memory, KV cache, eviction, efficiencycs.CL, cs.AI2025.11
arXiv(v1) 2025GCAgent: Long-Video Understanding via Schematic and Narrative Episodic MemoryJeong Hun Yeo, et al.memory, episodic, video, MLLMcs.CV, cs.AI2025.11
arXiv(v1) 2025Goal-Directed Search Outperforms Goal-Agnostic Memory Compression in Long-Context Memory TasksYicong Zheng, et al.memory, compression, long-context, searchcs.CL, cs.AI, cs.LG2025.11
arXiv(v1) 2025KVzip: Memory Compression for LLM Chatbots via KV Cache OptimizationAuthors TBDmemory, compression, KV cache, chatbotcs.CL, cs.AI3-4x compression, 170K tokens2025.11
MobiCom25Poster: MemAura: Persistent Personalized Context Memory for LLM Services in Smart EnvironmentsSiyuan Liu, et al.LLM, memory, personalization, smart environment, context2025.11
arXiv(v1) 2025Trainable Graph Memory for LLM Agents: From Experience to StrategyAuthors TBDmemory, graph, trainable, strategycs.AI, cs.CLUtility assessment mechanism2025.11
arXiv(v1) 2025WebCoach: Self-Evolving Web Agents with Cross-Session Memory GuidanceGenglin Liu, et al.memory, web agent, cross-session, self-evolvingcs.AI, cs.CL18 pages; work in progress2025.11
arXiv(v1) 2025A Memory-Efficient Retrieval Architecture for RAG-Enabled Wearable Medical LLMs-AgentsZhipeng Liao, et al.memory-efficient, RAG, wearable, medical, LLM, agentcs.ARAccepted by BioCAS20252025.10
arXiv(v1) 2025Acon: Optimizing Context Compression for Long-horizon LLM AgentsAuthors TBDmemory, compression, context, optimizationcs.CL, cs.AI26-54% memory reduction2025.10
arXiv(v1) 2025Beyond a Million Tokens: Benchmarking and Enhancing Long-Term Memory in LLMsMohammad Tavakoli, et al.memory, long-context, benchmark, long-termcs.CL, cs.AI, cs.IR2025.10
arXiv(v1) 2025CAM: Contextual Augmentation Memory for LLM AgentsAuthors TBDmemory, contextual, augmentation, agentcs.AI, cs.CL2025.10
arXiv(v1) 2025Dynamic Affective Memory Management for Personalized LLM AgentsJunfeng Lu, et al.memory, affective, personalization, agentcs.CL12 pasges, 8 figures2025.10
arXiv(v1) 2025Enabling Personalized Long-term Interactions in LLM-based Agents through Persistent MemoryAuthors TBDmemory, personalized, long-term, user profilecs.CL, cs.AI2025.10
arXiv(v1) 2025LightMem: Lightweight Memory for Efficient LLM AgentsAuthors TBDmemory, lightweight, efficient, agentcs.AI, cs.CL2025.10
arXiv(v1) 2025MEMTRACK: Evaluating Long-Term Memory and State Tracking in Multi-Platform Dynamic Agent EnvironmentsDarshan Deshpande, et al.memory, benchmark, state tracking, agentcs.AI, cs.CLAccepted to NeurIPS 2025 SEA Workshop2025.10
arXiv(v1) 2025Memory-Augmented State Machine Prompting: A Novel LLM Agent Framework for Real-Time Strategy GamesRunnan Qi, et al.memory, prompting, state machine, agentcs.AI10 pages, 4 figures, 1 table, 1 algorithm. Submitted to conference2025.10
arXiv(v1) 2025Pre-Storage Reasoning for Episodic Memory in LLM AgentsAuthors TBDmemory, episodic, pre-storage, reasoningcs.AI, cs.CL2025.10
arXiv(v2) 2025Evaluating Memory in LLM Agents via Incremental Multi-Turn InteractionsYuanzhe Hu, et al.memory evaluation, multi-turn, LLM, agentcs.CL, cs.AIY. Hu and Y. Wang contribute equally2025.09
arXiv(v1) 2025HopRAG: Multi-Hop Reasoning with Graph Memory for LLM AgentsAuthors TBDmemory, multi-hop, graph, reasoning, RAGcs.AI, cs.CL2025.09
arXiv(v1) 2025Mem-α: Learning Memory Construction via Reinforcement LearningYu Wang, et al.memory, RL, construction, agentcs.CL2025.09
arXiv(v1) 2025Mem-α: Memory with Adaptive Forgetting for LLM AgentsAuthors TBDmemory, forgetting, adaptive, agentcs.AI, cs.CL2025.09
arXiv(v1) 2025Memory in LLM-based Multi-agent Systems: Mechanisms, Challenges, and CollectiveAuthors TBDmemory, multi-agent, collective, surveycs.AI, cs.MA2025.09
arXiv(v1) 2025Multiple Memory Systems for Enhancing Long-term Memory of LLM AgentsAuthors TBDmemory, multiple systems, long-term, enhancementcs.AI, cs.CL2025.09
arXiv(v1) 2025Nemori: Neural Memory Organization for LLM AgentsAuthors TBDmemory, neural, organization, agentcs.AI, cs.CL2025.09
arXiv(v1) 2025SGMem: Sentence Graph Memory for Long-Term Conversational AgentsYaxiong Wu, et al.memory, graph, conversation, retrievalcs.CL, cs.IR19 pages, 6 figures, 1 table2025.09
arXiv(v1) 2025SGMem: Structured Graph Memory for LLM AgentsAuthors TBDmemory, graph, structured, agentcs.AI, cs.CL2025.09
IJCAI25 (IJCAI 2025)AriGraph: Learning Knowledge Graph World Models with Episodic MemoryAuthors TBDmemory, knowledge graph, episodic, world modelcs.AIIJCAI 20252025.08
arXiv(v1) 2025Cognitive Workspace: Active Memory Management for LLMs - Functional Infinite ContextAuthors TBDmemory, cognitive, workspace, active managementcs.CL, cs.AIMetacognitive control2025.08
arXiv(v1) 2025Learn to Memorize: Optimizing LLM-based Agents with Adaptive Memory FrameworkZeyu Zhang, et al.adaptive memory, optimization, LLM, agentcs.LG, cs.AI, cs.CL, cs.IR17 pages, 4 figures, 5 tables2025.08
arXiv(v1) 2025Memory-Augmented Transformers: A Systematic ReviewAuthors TBDmemory, transformer, survey, augmentedcs.CL, cs.AISystematic review2025.08
arXiv(v1) 2025Memory-R1: Enhancing LLM Agents to Manage and Utilize Memories via RLAuthors TBDmemory, RL, memory manager, ADD/UPDATE/DELETEcs.AI, cs.LGMemory Manager + Answer Agent2025.08
arXiv(v1) 2025Recursive Summarization for Long-Term Dialogue Memory in LLMsAuthors TBDmemory, summarization, dialogue, recursivecs.CL, cs.AIUpdated 20252025.08
arXiv(v2) 2025In Prospect and Retrospect: Reflective Memory Management for Long-term Personalized Dialogue AgentsZhen Tan, et al.reflective memory, dialogue agent, personalization, LLMcs.CL, cs.AIAccepted to ACL 20252025.07
ACL25 (ACL 2025)Pretraining Context Compressor for LLMs with Embedding-Based MemoryAuthors TBDmemory, compression, embedding, contextcs.CLPCC framework2025.07
arXiv(v1) 2025Cross-Attention Networks for Memory Retrieval in Generative AgentsAuthors TBDmemory, retrieval, cross-attention, generativecs.AI, cs.CLFrontiers in Psychology2025.04
arXiv(v1) 2025From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMsAuthors TBDmemory, survey, episodic, semantic, working memorycs.CL, cs.AIPersonal/system, parametric/non-parametric2025.04
arXiv(v3) 2025LIFT: Improving Long Context Understanding of Large Language Models through Long Input Fine-TuningYansheng Mao, et al.long context, fine-tuning, LLMcs.CL2025.04
arXiv(v1) 2025Mem0: Building Production-Ready AI Agents with Scalable Long-Term MemoryAuthors TBDmemory, long-term, scalable, productioncs.AI, cs.CL26% improvement over OpenAI2025.04
arXiv(v1) 2025In Prospect and Retrospect: Reflective Memory Management for Long-term Dialogue AgentsAuthors TBDmemory, reflective, dialogue, long-termcs.CL, cs.AI2025.03
arXiv(v1) 2025Tuning LLMs by RAG Principles: Towards LLM-native MemoryJiale Wei, et al.RAG, fine-tuning, optimization, LLMcs.CL, cs.AI, cs.IR2025.03
arXiv(v1) 2025A-MEM: Agentic Memory for LLM AgentsAuthors TBDmemory, agentic, self-organizing, Zettelkastencs.AI, cs.CLDynamic memory organization2025.02
arXiv(v1) 2025Position: Episodic Memory is the Missing Piece for Long-Term LLM AgentsAuthors TBDmemory, episodic, long-term, position papercs.AI, cs.CLEncoding and retrieval2025.02
arXiv(v1) 2025Zep: Temporal Knowledge Graph Architecture for Agent MemoryAuthors TBDmemory, temporal, knowledge graph, agentcs.AI, cs.CLEpisodic + semantic + community2025.02
arXiv(v1) 2024Memory-Augmented Agent Training for Business Document UnderstandingJiale Liu, et al.memory, agent, training, documentcs.CL, cs.AI11 pages, 8 figures2024.12
arXiv(v1) 2024On the Structural Memory of LLM AgentsRuihong Zeng, et al.memory, agent, analysis, structuralcs.CL, cs.AI2024.12
arXiv(v1) 2024XKV: Personalized KV Cache Memory Reduction for Long-Context LLM InferenceWeizhuo Li, et al.memory, KV cache, long-context, personalizationcs.LG, cs.CL2024.12
arXiv(v1) 2024MELODI: Exploring Memory Compression for Long ContextsYinpeng Chen, et al.memory compression, long context, LLMcs.LG, cs.AI2024.10
arXiv(v1) 2024A Survey on the Memory Mechanism of Large Language Model based AgentsZeyu Zhang, et al.memory, survey, LLM, agentcs.AIACM TOIS, 39 pages2024.04

Personalization

SourceTitle (Link)AuthorsTagSubjectsAdditional infoDate
arXiv(v1) 2026DeepInterestGR: Mining Deep Multi-Interest Using Multi-Modal LLMs for Generative RecommendationYangchen ZengLLM, personalizationcs.LG, cs.CV, cs.CY2026.02
arXiv(v1) 2026Dynamic Personality Adaptation in Large Language Models via State MachinesLeon Pielage, et al.LLM, personalizationcs.CL, cs.HC, cs.LG22 pages, 5 figures, submitted to ICPR 20262026.02
arXiv(v1) 2026Facet-Level Persona Control by Trait-Activated Routing with Contrastive SAE for Role-Playing LLMsWenqiu Tang, et al.LLM, personalizationcs.CLAccepted in PAKDD 2026 special session on Data Science :Foundation and Applications2026.02
arXiv(v1) 2026InterviewSim: A Scalable Framework for Interview-Grounded Personality SimulationYu Li, et al.LLM, personalizationcs.CL, cs.AI, cs.CY2026.02
arXiv(v1) 2026Learning to Reason for Multi-Step Retrieval of Personal Context in Personalized Question AnsweringMaryam Amirizaniani, et al.LLM, personalizationcs.CL, cs.AI, cs.IR2026.02
arXiv(v1) 2026Long Context, Less Focus: A Scaling Gap in LLMs Revealed through Privacy and PersonalizationShangding GuLLM, personalizationcs.LG, cs.AI2026.02
arXiv(v1) 2026Multi-Agent Large Language Model Based Emotional Detoxification Through Personalized Intensity Control for Consumer ProtectionKeito InoshitaLLM, personalizationcs.AI2026.02
arXiv(v1) 2026Offline Reasoning for Efficient Recommendation: LLM-Empowered Persona-Profiled Item IndexingDeogyong Kim, et al.LLM, personalizationcs.IR, cs.LGUnder review2026.02
arXiv(v1) 2026PERSONA: Dynamic and Compositional Inference-Time Personality Control via Activation Vector AlgebraXiachong Feng, et al.LLM, personalizationcs.AIICLR 20262026.02
arXiv(v1) 2026PRECTR-V2:Unified Relevance-CTR Framework with Cross-User Preference Mining, Exposure Bias Correction, and LLM-Distilled Encoder OptimizationShuzhi Cao, et al.LLM, personalizationcs.IR, cs.AIarXiv admin note: text overlap with arXiv:2503.183952026.02
arXiv(v1) 2026Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User HistorySerin Kim, et al.LLM, personalizationcs.CL, cs.AI2026.02
arXiv(v1) 2026Personalized Graph-Empowered Large Language Model for Proactive Information AccessChia Cheng Chang, et al.LLM, personalizationcs.CL2026.02
arXiv(v1) 2026Personalized Prediction of Perceived Message Effectiveness Using Large Language Model Based Digital TwinsJasmin Han, et al.LLM, personalizationcs.CL, stat.AP31 pages, 5 figures, submitted to Journal of the American Medical Informatics Association (JAMIA). Drs. Chen and Thrul share last authorship2026.02
arXiv(v1) 2026Sydney Telling Fables on AI and Humans: A Corpus Tracing Memetic Transfer of Persona between LLMsJiří Milička, et al.LLM, personalizationcs.CL, cs.AI2026.02
arXiv(v1) 2026Toward Personalized LLM-Powered Agents: Foundations, Evaluation, and Future DirectionsYue Xu, et al.LLM, personalizationcs.AI2026.02
arXiv(v1) 2026CURP: Codebook-based Continuous User Representation for Personalized Generation with LLMsLiang Wang, et al.arxivcs.CL2026.01
arXiv(v1) 2026Enriching Semantic Profiles into Knowledge Graph for Recommender Systems Using Large Language ModelsSeokho Ahn, et al.knowledge graph, semantic profiles, recommendation, LLMcs.IR, cs.AI, cs.LGAccepted at KDD 20262026.01
ICLR 2026FingerTip 20K: A Benchmark for Proactive and Personalized Mobile LLM AgentsICLR,OpenReviewMobile Agent, LLM Agent, GUI, Proactive Agent, PersonalizationOpenReview ID: n3iFV0gLMc2026.01
arXiv(v1) 2026HumanLLM: Towards Personalized Understanding and Simulation of Human NatureYuxuan Lei, et al.arxivcs.CL12 pages, 5 figures, 7 tables, to be published in KDD 20262026.01
arXiv(v1) 2026Improving User Privacy in Personalized Generation: Client-Side Retrieval-Augmented Modification of Server-Side Generated SpeculationsAlireza Salemi, et al.arxivcs.CL, cs.AI, cs.CR, cs.IR2026.01
arXiv(v2) 2026Linear Personality Probing and Steering in LLMs: A Big Five StudyMichel Frising, et al.personalization, personality, probing, steeringcs.CL29 pages, 6 figures2026.01
arXiv(v1) 2026Me-Agent: A Personalized Mobile Agent with Two-Level User Habit Learning for Enhanced InteractionShuoxin Wang, et al.arxivcs.CL2026.01
ICLR 2026Meta-Router: Bridging Gold-standard and Preference-based Evaluations in LLM RoutingICLR,OpenReviewCausal learning, Meta-learner, Large Language Model, query routingOpenReview ID: r0BFucF2dH2026.01
ICLR 2026NextQuill: Causal Preference Modeling for Enhancing LLM PersonalizationICLR,OpenReviewPersonalized text generation, Large Language Models, LLM PersonalizationOpenReview ID: xYpVlKMFqv2026.01
arXiv(v1) 2026One Adapts to Any: Meta Reward Modeling for Personalized LLM AlignmentHongru Cai, et al.meta reward modeling, alignment, personalization, LLMcs.CL, cs.AI2026.01
arXiv(v1) 2026Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM PersonalizationLinfeng Du, et al.arxivcs.CL, cs.IR2026.01
arXiv(v1) 2026PRISP: Privacy-Safe Few-Shot Personalization via Lightweight AdaptationJunho Park, et al.few-shot, privacy-safe, lightweight adaptation, personalization, LLMcs.CL, cs.AI, cs.LG16 pages, 9 figures2026.01
arXiv(v1) 2026PersonaDual: Balancing Personalization and Objectivity via Adaptive ReasoningXiaoyou Liu, et al.personalization, persona, reasoning, LLMcs.AI2026.01
ICLR 2026Preference Leakage: A Contamination Problem in LLM-as-a-judgeICLR,OpenReviewLLM-as-a-judge, Preference Leakage, Data ContaminationOpenReview ID: grIvSXVJ652026.01
ICLR 2026ProPerSim: Developing Proactive and Personalized AI Assistants through User-Assistant SimulationICLR,OpenReviewBenchmark, Agent Simulation, Personalization, ProactivityOpenReview ID: RV2aeCgxdB2026.01
arXiv(v1) 2026SPRInG: Continual LLM Personalization via Selective Parametric Adaptation and Retrieval-Interpolated GenerationSeoyeon Kim, et al.personalization, continual, retrieval, parametric adaptation, LLMcs.AI, cs.CLunder review, 23 pages2026.01
ICLR 2026Solving the Granularity Mismatch: Hierarchical Preference Learning for Long-Horizon LLM AgentsICLR,OpenReviewLLM-based Agents, Process Supervision, Curriculum LearningOpenReview ID: s8usvGHYlk2026.01
arXiv(v1) 2026Structured Personality Control and Adaptation for LLM AgentsJinpeng Wang, et al.personalization, personality, control, agent, LLMcs.AI2026.01
arXiv(v1) 2026Styles + Persona-plug = Customized LLMsYutong Song, et al.style customization, persona-plug, LLMcs.AI2026.01
arXiv(v1) 2026The Assistant Axis: Situating and Stabilizing the Default Persona of Language ModelsChristina Lu, et al.persona, default persona, alignment, LLMcs.CL2026.01
arXiv(v2) 2026The Reward Model Selection Crisis in Personalized AlignmentFady Rezk, et al.personalization, alignment, reward model, RLHFcs.AI, cs.LG2026.01
ICLR 2026Towards Understanding Valuable Preference Data for Large Language Model AlignmentICLR,OpenReviewLarge language model alignment, preference data, influence functionOpenReview ID: FUp0KeEEBs2026.01
ICLR 2026Verification and Co-Alignment via Heterogeneous Consistency for Preference-Aligned LLM AnnotationsICLR,OpenReviewVerification, Co-Alignment, Preference-Aligned LLM Annotations, Reference-Free MetricOpenReview ID: jugY302BAh2026.01
arXiv(v1) 2026When Personalization Misleads: Understanding and Mitigating Hallucinations in Personalized LLMsZhongxiang Sun, et al.personalization, hallucination, safety, evaluationcs.CL, cs.AI20 pages, 15 figures2026.01
arXiv(v1) 2025Agentic Multi-Persona Framework for Evidence-Aware Fake News DetectionRoopa Bukke, et al.persona, multi-agent, fake news, LLMcs.IR, cs.LG12 pages, 8 tables, 2 figures2025.12
arXiv(v1) 2025Interpolative Decoding: Exploring the Spectrum of Personality Traits in LLMsEric Yeh, et al.personalization, personality, decoding, controlcs.AI20 pages, 5 figures2025.12
arXiv(v1) 2025LLM Personas as a Substitute for Field Experiments in Method BenchmarkingEnoch Hyunwook Kangpersona, evaluation, methodology, LLMcs.AI, cs.LG, econ.EM2025.12
arXiv(v1) 2025Memoria: A Scalable Agentic Memory Framework for Personalized Conversational AISamarth Sarin, et al.personalization, memory, conversational AI, agentcs.AI, cs.CLPaper accepted at 5th International Conference of AIML Systems 2025, Bangalore, India2025.12
arXiv(v1) 2025PILAR: Personalizing Augmented Reality Interactions with LLM-based Human-Centric and Trustworthy Explanations for Daily Use CasesRipan Kumar Kundu, et al.personalization, AR, explanations, LLMcs.HC, cs.AIPublished in the 2025 IEEE International Symposium on Mixed and Augmented Reality Adjunct (ISMAR-Adjunct)2025.12
arXiv(v1) 2025PRISM: A Personality-Driven Multi-Agent Framework for Social Media SimulationZhixiang Lu, et al.personality, multi-agent, social simulation, LLMcs.CL2025.12
arXiv(v1) 2025PersonaMem-v2: Towards Personalized Intelligence via Learning Implicit User PersonasAuthors TBDpersonalization, persona, implicit, user modelingcs.CL, cs.AI1000 personas, 20k preferences, 128k context2025.12
arXiv(v1) 2025Personalized Multimodal Large Language Models: A SurveyAuthors TBDpersonalization, MLLM, survey, multimodalcs.CV, cs.CLComprehensive MLLM personalization survey2025.12
arXiv(v1) 2025PrefGen: Multimodal Preference Learning for Image GenerationAuthors TBDpersonalization, MLLM, image generation, preferencecs.CV, cs.CLUser-specific conditioning2025.12
arXiv(v2) 2025ProEx: A Unified Framework Leveraging Large Language Model with Profile Extrapolation for RecommendationYi Zhang, et al.profile extrapolation, recommendation, LLMcs.IRAccepted by KDD 2026 (First Cycle)2025.12
arXiv(v1) 2025SPARK: Search Personalization via Agent-Driven Retrieval and Knowledge-sharingGaurab Chhetri, et al.personalization, search, agent, retrievalcs.AIAccepted to WEB&GRAPH 2026 (WSDM 2026 workshop)2025.12
arXiv(v1) 2025TAME: Long-Context MLLM Personalization with Double MemoriesAuthors TBDpersonalization, MLLM, memory, training-freecs.CV, cs.CLRA2G recipe2025.12
arXiv(v1) 2025The Mental World of Large Language Models in Recommendation: A Benchmark on Association, Personalization, and KnowledgeabilityGuangneng Hupersonalization, recommendation, benchmark, LLMcs.IR21 pages, 13 figures, 27 tables, submission to KDD 20252025.12
arXiv(v1) 2025Towards Proactive Personalization through Profile Customization for Individual Users in DialoguesXiaotian Zhang, et al.proactive personalization, profile customization, dialogue, LLMcs.CL2025.12
TiiS25User Perceptions of Personalized and Generic Explanations in LLM-Driven Recommender SystemsÍtallo De Sousa Silva, et al.LLM, recommender system, personalization, explanations, user study2025.12
arXiv(v1) 2025Fixed-Persona SLMs with Modular Memory: Scalable NPC Dialogue on Consumer HardwareMartin Braas, et al.personalization, persona, modular memory, NPCcs.AI, cs.IR2025.11
arXiv(v2) 2025Mem-PAL: Towards Memory-based Personalized Dialogue Assistants for Long-term User-Agent InteractionZhaopei Huang, et al.personalization, dialogue, memory, long-termcs.CLAccepted by AAAI 2026 (Oral)2025.11
arXiv(v1) 2025PLUM: Learning to Remember User Conversations for PersonalizationAuthors TBDpersonalization, memory, conversation, LoRAcs.CL, cs.AIParameter-efficient2025.11
arXiv(v1) 2025PersonaAgent with GraphRAG: Community-Aware KG for Personalized LLMAuthors TBDpersonalization, GraphRAG, knowledge graph, agentcs.CL, cs.AI11.1% F1 improvement on LaMP2025.11
arXiv(v1) 2025PersonalizedRouter: Personalized LLM Routing via Graph-based User Preference ModelingAuthors TBDpersonalization, routing, GNN, user preferencecs.CL, cs.AI2025.11
arXiv(v1) 2025Profile-LLM: Dynamic Profile Optimization for Realistic Personality ExpressionAuthors TBDpersonalization, profile, personality, dynamiccs.CL, cs.AIEducation, therapy, entertainment2025.11
arXiv(v1) 2025LLMDiRec: LLM-Enhanced Intent Diffusion for Sequential RecommendationBo-Chian Chen, et al.intent diffusion, sequential recommendation, LLMcs.IRUnder review2025.10
arXiv(v1) 2025MemWeaver: A Hierarchical Memory from Textual Interactive Behaviors for Personalized GenerationShuo Yu, et al.personalization, memory, user behavior, generationcs.CL12 pages, 8 figures2025.10
arXiv(v1) 2025P2P: Instant Personalized LLM Adaptation via HypernetworkAuthors TBDpersonalization, hypernetwork, instant adaptationcs.CL, cs.AISingle-pass generation2025.10
arXiv(v1) 2025Preference-Aware Memory Update for Long-Term LLM AgentsHaoran Sun, et al.personalization, preference, memory update, long-termcs.CL, cs.AI2025.10
arXiv(v1) 2025RGMem: Renormalization Group-based Memory Evolution for Language Agent User ProfileAo Tian, et al.personalization, user profile, memory, agentcs.AI11 pages,3 figures2025.10
arXiv(v1) 2025Real-Time Personalization for LLM-based Recommendation with Customized ICLAuthors TBDpersonalization, recommendation, ICL, real-time, onlinecs.IR, cs.AINo model update needed2025.10
arXiv(v2) 2025CoPL: Collaborative Preference Learning for Personalizing LLMsYoungbin Choi, et al.collaborative preference learning, personalization, LLMcs.LG, cs.AI, cs.IR19pages, 13 figures, 11 tables2025.09
arXiv(v1) 2025DP-FedLoRA: Privacy-Enhanced Federated Fine-Tuning for On-Device Large Language ModelsHonghui Xu, et al.federated learning, privacy-enhanced, on-device LLM, fine-tuningcs.CR, cs.AI2025.09
arXiv(v1) 2025HumAIne-Chatbot: Real-Time Personalized Conversational AI via RLAuthors TBDpersonalization, chatbot, RL, real-time, industrycs.CL, cs.AIProduction deployment2025.09
arXiv(v1) 2025MMPB: Multi-Modal Personalization Benchmark for VLMsAuthors TBDpersonalization, MLLM, benchmark, VLMcs.CV, cs.CLFirst MLLM personalization benchmark2025.09
arXiv(v1) 2025Personalized Reasoning: Just-In-Time Personalization and Why LLMs Fail At ItShuyue Stella Li, et al.just-in-time personalization, LLM, user preferencecs.CL, cs.AI57 pages, 6 figures2025.09
RecSys25Revisiting Prompt Engineering: A Comprehensive Evaluation for LLM-based Personalized RecommendationGenki Kusano, et al.LLM, recommendation, personalization, prompt engineering2025.09
arXiv(v1) 2025T-POP: Test-Time Personalization with Online Preference FeedbackZikun Qu, et al.test-time personalization, online preference, LLMcs.LG, cs.AIPreprint2025.09
arXiv(v1) 2025DGDPO: Diagnostic-Guided Dynamic Profile Optimization for User SimulatorsAuthors TBDpersonalization, user simulation, profile, dynamiccs.IR, cs.AIBidirectional evolution2025.08
arXiv(v1) 2025End-to-End Personalization: Unifying Recommender Systems with Large Language ModelsDanial Ebrat, et al.recommendation system, unifying, LLMcs.IR, cs.LGSecond Workshop on Generative AI for Recommender Systems and Personalization at the ACM Conference on Knowledge Discovery and Data Mining (GenAIRecP@KDD 2025)2025.08
arXiv(v1) 2025MLLMRec: MLLMs in Recommender SystemsAuthors TBDpersonalization, MLLM, recommendation, visualcs.CV, cs.IRVisual attribute extraction2025.08
arXiv(v1) 2025MM-R1: Unified MLLMs for Personalized Image GenerationAuthors TBDpersonalization, MLLM, image generation, GRPOcs.CV, cs.CLX-CoT reasoning2025.08
arXiv(v1) 2025MSPA: Multimodal Self-Corrective Preference Alignment for RecommendationAuthors TBDpersonalization, MLLM, recommendation, self-correctivecs.CV, cs.IR4D multimodal signals2025.08
arXiv(v2) 2025Personalized LLM for Generating Customized Responses to the Same Query from Different UsersHang Zeng, et al.personalization, response generation, user-specific, LLMcs.CLAccepted by CIKM'252025.08
arXiv(v1) 2025RLHF Fine-Tuning of LLMs for Alignment with Implicit User Feedback in Conversational RecommendersAuthors TBDpersonalization, RLHF, implicit feedback, recommendercs.IR, cs.AIDwell time, sentiment signals2025.08
arXiv(v1) (SIGIR25)CoT-Rec: Enhancing LLM-Based Recommendations Through Personalized ReasoningAuthors TBDpersonalization, recommendation, CoT, reasoningcs.IR, cs.AISIGIR 20252025.07
arXiv(v1) 2025Comprehensive Review on LLMs for Recommender SystemsAuthors TBDpersonalization, recommendation, survey, LLMcs.IR, cs.AIHybrid RAG approaches2025.07
arXiv(v1) 2025DEP: Latent Inter-User Difference Modeling for LLM PersonalizationAuthors TBDpersonalization, latent, embedding, difference-awarecs.CL, cs.AISparse autoencoder2025.07
arXiv(v1) 2025PITA: Preference-Guided Inference-Time Alignment for LLM Post-TrainingAuthors TBDpersonalization, inference-time, alignment, preferencecs.CL, cs.AINo reward model needed2025.07
arXiv(v1) 2025PLUS: Learning to Summarize User Information for Personalized RLHFAuthors TBDpersonalization, RLHF, user summary, preferencecs.LG, cs.AI11-77% reward model improvement2025.07
arXiv(v1) 2025PRIME: LLM Personalization with Cognitive Memory and Thought ProcessesAuthors TBDpersonalization, memory, episodic, semantic, cognitivecs.CL, cs.AIDual-memory model2025.07
arXiv(v1) 2025PURE: LLM-based User Profile Management for Recommender SystemAuthors TBDpersonalization, user profile, recommendation, managementcs.IR, cs.AIProfile extraction and updating2025.07
arXiv(v1) 2025Personalization of Large Language Models: A SurveyZhehao Zhang, et al.personalization, survey, LLM, user profilecs.CL, cs.AIComprehensive taxonomy2025.07
arXiv(v2) 2025Comparison-based Active Preference Learning for Multi-dimensional PersonalizationMinhyeon Oh, et al.active preference learning, multi-dimensional, personalization, LLMcs.LG2025.06
arXiv(v1) 2025PersonalAI: KG Storage and Retrieval for Personalized LLM AgentsAuthors TBDpersonalization, knowledge graph, memory, agentcs.AI, cs.CLHybrid graph with hyperedges2025.06
UMAP25Personalizing LLM Responses to Combat Political MisinformationAdiba Proma, et al.LLM, personalization, misinformation, user modeling2025.06
arXiv(v1) 2025ProfiLLM: LLM-Based Framework for Implicit User ProfilingAuthors TBDpersonalization, profiling, implicit, chatbotcs.CL, cs.AIIT/cybersecurity domain2025.06
arXiv(v1) 2025SEAL: Self-Adapting Language ModelsAuthors TBDpersonalization, self-adaptation, online learningcs.CL, cs.AISelf-generated finetuning data2025.06
arXiv(v3) 2025Drift: Decoding-time Personalized Alignments with Implicit User PreferencesMinbeom Kim, et al.decoding-time alignment, implicit preferences, personalization, LLMcs.CL19 pages, 6 figures2025.05
arXiv(v2) 2025HyPerAlign: Interpretable Personalized LLM Alignment via Hypothesis GenerationCristina Garbacea, et al.alignment, hypothesis generation, personalization, LLMcs.CL2025.05
arXiv(v1) 2025MAP: Memory Assisted LLM for Personalized Recommendation SystemAuthors TBDpersonalization, recommendation, memory, historycs.IR, cs.AI2025.05
arXiv(v1) 2025PROSE: Aligning LLMs by Predicting Preferences from User Writing SamplesAuthors TBDpersonalization, preference prediction, writing samplescs.CL, cs.AIIterative refinement2025.05
arXiv(v1) 2025Privacy-preserving Prompt Personalization in Federated Learning for Multimodal Large Language ModelsSizai Hou, et al.federated learning, privacy-preserving, prompt personalization, MLLMcs.CRUnder Review2025.05
arXiv(v1) 2025RLPA: Teaching LLMs to Evolve with Users via Dynamic Profile ModelingAuthors TBDpersonalization, RLHF, dynamic profile, user evolutioncs.CL, cs.AIOutperforms Claude-3.5, DeepSeek-V32025.05
arXiv(v1) 2025Steerable Chatbots: Personalizing LLMs with Preference-Based Activation SteeringAuthors TBDpersonalization, chatbot, activation steering, inferencecs.CL, cs.AITraining-free2025.05
arXiv(v1) 2025Towards Explainable Temporal User Profiling with LLMsMilad Sabouri, et al.temporal user profiling, explainable, LLMcs.IR, cs.AI2025.05
arXiv(v1) 2025Towards a unified user modeling language for engineering human centered AI systemsAaron Conrardy, et al.user modeling, LLM, personalizationcs.SEAccepted at the Third Workshop on Engineering Interactive Systems Embedding AI Technologies (EISEAIT workshop at EICS 2025)2025.05
arXiv(v1) 2025A Survey on Personalized and Pluralistic Preference Alignment in Large Language ModelsAuthors TBDpersonalization, preference alignment, survey, pluralisticcs.CL, cs.AITraining and inference-time methods2025.04
arXiv(v2) 2025Differential Privacy Personalized Federated Learning Based on Dynamically Sparsified Client UpdatesChuanyin Wang, et al.differential privacy, federated learning, personalization, LLMcs.LG, cs.CR10 pages,2 figures2025.04
arXiv(v1) 2025Know Me, Respond to Me: Benchmarking LLMs for Dynamic User ProfilingAuthors TBDpersonalization, benchmark, user profiling, dynamiccs.CL, cs.AI2025.04
arXiv(v1) 2025LoRe: Personalizing LLMs via Low-Rank Reward ModelingAvinandan Bose, et al.low-rank reward modeling, personalization, LLM, RLHFcs.LG, cs.AI, cs.CL2025.04
arXiv(v1) 2025PaRT: Enhancing Proactive Social Chatbots with Personalized Real-Time RetrievalAuthors TBDpersonalization, chatbot, retrieval, real-timecs.CL, cs.AI21.77% dialogue improvement, production deployed2025.04
arXiv(v3) 2025Tuning-Free Personalized Alignment via Trial-Error-Explain In-Context LearningHyundong Cho, et al.in-context learning, trial-error-explain, personalization, LLMcs.CL, cs.AINAACL 2025 Findings2025.04
arXiv(v1) 2025User Feedback Alignment for LLM-powered Exploration in Large-scale RecommendationAuthors TBDpersonalization, recommendation, feedback, explorationcs.IR, cs.AIClick and dwell time signals2025.04
arXiv(v1) 2025A Shared Low-Rank Adaptation Approach to Personalized RLHFRenpu Liu, et al.RLHF, low-rank adaptation, personalization, LLMcs.LG, cs.AIPublished as a conference paper at AISTATS 20252025.03
arXiv(v1) 2025Agentic Recommender Systems in the Era of Multimodal LLMs: SurveyAuthors TBDpersonalization, recommendation, agent, MLLM, surveycs.IR, cs.AIUser agent simulation2025.03
arXiv(v1) 2025BAHE: LLM-Enhanced CTR Prediction in Long Textual User BehaviorsAuthors TBDpersonalization, CTR, user behavior, industrycs.IR, cs.AIDeployed on 50M daily data2025.03
arXiv(v1) 2025Can LLM Agents Simulate Multi-Turn Human Behavior? Evidence from Online ShoppingAuthors TBDpersonalization, agent, user simulation, behaviorcs.AI, cs.HC31,865 shopping sessions2025.03
arXiv(v1) 2025Language Model Personalization via Reward FactorizationIdan Shenfeld, et al.reward factorization, RLHF, personalization, LLMcs.LG2025.03
arXiv(v1) 2025Measuring What Makes You Unique: Difference-Aware User Modeling for LLM PersonalizationAuthors TBDpersonalization, user modeling, difference-awarecs.CL, cs.AI2025.03
arXiv(v7) 2025PAD: Personalized Alignment of LLMs at Decoding-TimeRuizhe Chen, et al.decoding-time alignment, personalization, LLMcs.CL, cs.AIICLR 20252025.03
arXiv(v1) 2025PersonaX: A Recommendation Agent Oriented User Modeling FrameworkAuthors TBDpersonalization, recommendation, user modeling, agentcs.IR, cs.AI3-11% improvement on AgentCF2025.03
arXiv(v1) 2025A Survey of Personalized Large Language Models: Progress and Future DirectionsAuthors TBDpersonalization, survey, LLM, prompting, finetuningcs.CL, cs.AIInput/model/objective level2025.02
arXiv(v1) 2025FSPO: Few-Shot Preference Optimization of Synthetic Preference Data in LLMs Elicits Effective Personalization to Real UsersAnikait Singh, et al.few-shot, preference optimization, personalization, LLMcs.LG, cs.AI, cs.CL, cs.HC, stat.MLWebsite: https://fewshot-preference-optimization.github.io/2025.02
arXiv(v1) 2025LoCoMo: Evaluating Very Long-Term Conversational Memory of LLM AgentsAuthors TBDpersonalization, memory, conversation, benchmarkcs.CL, cs.AI300 turns, 9K tokens, 35 sessions2025.02
arXiv(v1) 2025PrefEval: Do LLMs Recognize Your Preferences? Evaluating Personalized Preference FollowingAuthors TBDpersonalization, benchmark, preference, evaluationcs.CL3000 preference-query pairs, 20 topics2025.02
arXiv(v3) 2025Privacy-Preserving Personalized Federated Prompt Learning for Multimodal Large Language ModelsLinh Tran, et al.federated learning, privacy-preserving, personalized prompt, MLLMcs.LG2025.02
arXiv(v1) 2025RLTHF: Targeted Human Feedback for LLM AlignmentAuthors TBDpersonalization, RLHF, targeted feedback, efficientcs.CL, cs.AI6-7% human annotation effort2025.02
arXiv(v3) 2025SmartAgent: Chain-of-User-Thought for Embodied Personalized Agent in Cyber WorldJiaqi Zhang, et al.personalization, agent, user modeling, embodiedcs.AI2025.02
arXiv(v1) 2025User Profile Construction and Updating with LLMs: BenchmarkAuthors TBDpersonalization, user profile, construction, updatingcs.CL, cs.AIStatic and dynamic profiling2025.02
arXiv(v1) 2025When Personalization Meets Reality: Multi-Faceted Analysis of Personalized Preference LearningAuthors TBDpersonalization, preference learning, fairness, evaluationcs.CL, cs.AI2025.02
arXiv(v1) 2025Advancing Personalized Federated Learning: Integrative Approaches with AI for Enhanced Privacy and CustomizationKevin Cooper, et al.federated learning, privacy, personalization, LLMcs.LG, eess.SParXiv admin note: substantial text overlap with arXiv:2501.167582025.01
arXiv(v2) 2025Identifying and Manipulating Personality Traits in LLMs Through Activation EngineeringRumi Allbert, et al.personalization, personality, activation engineering, steeringcs.CL, cs.AI2025.01
arXiv(v1) 2025PerRecBench: Can LLMs Understand Preferences in Personalized Recommendation?Authors TBDpersonalization, benchmark, recommendation, preferencecs.IR, cs.CL19 LLMs evaluated2025.01
arXiv(v2) 2025PsychAdapter: Adapting LLM Transformers to Reflect Traits, Personality and Mental HealthHuy Vu, et al.personalization, personality, traits, mental healthcs.AI, cs.CL2025.01
arXiv(v1) 2024AI PERSONA: Towards Life-long Personalization of LLMsTiannan Wang, et al.personalization, persona, lifelong, LLMcs.CL, cs.AIWork in progress2024.12
arXiv(v1) 2024Beyond Discrete Personas: Personality Modeling Through Journal Intensive ConversationsSayantan Pal, et al.personalization, persona, personality modeling, long-termcs.CL, cs.AIAccepted in COLING 20252024.12
arXiv(v1) 2024Can Large Language Models Understand You Better? An MBTI Personality Detection Dataset Aligned with Population TraitsBohan Li, et al.personalization, personality, dataset, MBTIcs.CL, cs.CYAccepted by COLING 2025. 28 papges, 20 figures, 10 tables2024.12
arXiv(v1) 2024Disentangling Preference Representation and Text Generation for Efficient Individual Preference AlignmentJianfei Zhang, et al.personalization, preference alignment, efficiency, LLMcs.CL, cs.AIColing 20252024.12
arXiv(v2) 2024Molar: Multimodal LLMs with Collaborative Filtering Alignment for Enhanced Sequential RecommendationYucong Luo, et al.collaborative filtering, sequential recommendation, MLLMcs.IR, cs.AI2024.12
arXiv(v1) 2024Semantic Convergence: Harmonizing Recommender Systems via Two-Stage Alignment and Behavioral Semantic TokenizationGuanghan Li, et al.recommendation system, alignment, behavioral semantic, LLMcs.IR, cs.AI, cs.CL7 pages, 3 figures, AAAI 20252024.12
arXiv(v1) 2024ULMRec: User-centric Large Language Model for Sequential RecommendationMinglai Shao, et al.personalization, recommendation, sequential, LLMcs.IR2024.12
arXiv(v1) 2024Personalizing Reinforcement Learning from Human Feedback with Variational Preference LearningSriyash Poddar, et al.RLHF, variational preference learning, personalization, LLMcs.LG, cs.AI, cs.CL, cs.ROweirdlabuw.github.io/vpl2024.08
arXiv(v2) 2024RLHF from Heterogeneous Feedback via Personalization and Preference AggregationChanwoo Park, et al.RLHF, heterogeneous feedback, preference aggregation, personalization, LLMcs.AI, cs.LGAdded experiments2024.05
agent
awesome
awesome-list
embodied-ai
hci
human-computer-interaction
imu
llm
mllm
multimodal
rag
rl
rlhf
sensor
sensors
survey
ubiquitous
vision
vllm
vlm

Languages

Python

100.0%