{"id":1990,"date":"2024-09-30T21:50:35","date_gmt":"2024-09-30T16:20:35","guid":{"rendered":"https:\/\/icmi.acm.org\/2024\/?page_id=1990"},"modified":"2026-09-09T14:48:07","modified_gmt":"2026-09-09T09:18:07","slug":"sessions","status":"publish","type":"page","link":"https:\/\/icmi.acm.org\/2026\/sessions\/","title":{"rendered":"Sessions"},"content":{"rendered":"<p>[et_pb_section fb_built=&#8221;1&#8243; admin_label=&#8221;section&#8221; _builder_version=&#8221;4.14.4&#8243; background_enable_image=&#8221;off&#8221; custom_padding=&#8221;3px||0px|||&#8221; global_colors_info=&#8221;{}&#8221;][et_pb_row admin_label=&#8221;row&#8221; _builder_version=&#8221;4.14.4&#8243; background_size=&#8221;initial&#8221; background_position=&#8221;top_left&#8221; background_repeat=&#8221;repeat&#8221; width=&#8221;90%&#8221; min_height=&#8221;1612.7px&#8221; custom_margin=&#8221;|auto|221px|auto||&#8221; custom_padding=&#8221;4px|||||&#8221; global_colors_info=&#8221;{}&#8221;][et_pb_column type=&#8221;4_4&#8243; _builder_version=&#8221;3.25&#8243; custom_padding=&#8221;|||&#8221; global_colors_info=&#8221;{}&#8221; custom_padding__hover=&#8221;|||&#8221;][et_pb_text _builder_version=&#8221;4.14.4&#8243; _module_preset=&#8221;default&#8221; text_font=&#8221;||||||||&#8221; text_text_color=&#8221;#4f4f4f&#8221; text_font_size=&#8221;13px&#8221; header_4_text_color=&#8221;#282562&#8243; header_4_line_height=&#8221;2em&#8221; header_5_text_color=&#8221;#6292C2&#8243; header_5_line_height=&#8221;1.6em&#8221; custom_margin=&#8221;||0px|||&#8221; custom_padding=&#8221;||0px|||&#8221; global_colors_info=&#8221;{}&#8221;]<\/p>\n<h4><b>Sessions<\/b><\/h4>\n<p>&nbsp;<\/p>\n<p>[\/et_pb_text][et_pb_text module_id=&#8221;AEDC&#8221; _builder_version=&#8221;4.14.4&#8243; _module_preset=&#8221;default&#8221; header_text_color=&#8221;#282562&#8243; header_4_text_color=&#8221;#672B83&#8243; min_height=&#8221;1354.4px&#8221; custom_padding=&#8221;||0px|||&#8221; hover_enabled=&#8221;0&#8243; global_colors_info=&#8221;{}&#8221; sticky_enabled=&#8221;0&#8243;]<\/p>\n<div id=\"doctoral-consortium\"><\/div>\n<h4><strong>Main Conference Day 1 &#8211; Tuesday, 6 October 2026<\/strong><\/h4>\n<div id=\"poster-session-1\"><\/div>\n<p><strong>10:00-11:00 Poster Session #1: Affective Computing, Health &amp; Wellbeing<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Affect-Aware Game Personalization with Reinforcement Learning: How to Improve Players&#8217; Engagement, Performance, and Emotions<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Mahyar Tourchi Moghaddam, Tiziano Santilli, Mina Alipour<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">A Screening Method for Children with Autism Spectrum Disorder Based on a Dual-Stream, Multi-Scale, Cross-Modal Attention Mechanism<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Haoran Sun, Bang Li, Xiaoqing Jiang, Liu Chen, Shengduo Hu, Chenyang Liang, Jianwei Gu, Kaiyun Li, Zhenxiang Chen<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Multimodal Biomarkers of Dysarthria: Severity-Conditioned Contrastive Alignment of Speech, Text, and Facial Asymmetry<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Karen Rosero, Chi-Chun Lee, Rami R. Hallac, Carlos Busso<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Self-Supervised Dual-Stream Temporal Learning for Neonatal Pain Estimation<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Muhammad Suhaib Shahid, Michel Valstar, Don Sharkey, Joy Egede<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">From Feedback to Neural Dynamics: Network-Based Modeling of EEG in Interactive Tasks<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Behdokht Kiafar, Mohammad Fahim Abrar, Roghayeh Leila Barmaki<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Multi-Scale Spatiotemporal EEG and Self-Supervised Audio Fusion: A Mixture-of-Experts Approach to Continuous Affect<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ali Amini, Sarmad Maqsood, Irfan Abbas, Muhammad Abdullah Sarwar, Rytis Maskeli\u016bnas, Robertas Dama\u0161evi\u010dius<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">From Facial Expressions to Emotional Health: Deploying Affective Computing in Call Centers<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Lesly Nzeusseu Kouamou, Pierrich Plusquellec<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Correcting in the Moment: Evaluating Real-Time AI Postural Feedback for Cellists<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Paolo Wang, Kexin Sha, Kunzhu Xie, Michael Zhang, Shrinand Perumal, Ekaterina Tszyao, Jackson Shields, Luke Choi, Sivamurugan Velmurugan, Preston Mo, Daniel Chindris, Wangyue Xue, Mohammad Rahman, Ming Yin, Zhenyu Cheryl Qian, Yingjie Victor Chen, Ka-wai Yu, Yung-Hsiang Lu, Kristen Yeon-Ji Yun<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">FAIR_XAI: Improving Multimodal Foundation Model Fairness via Explainability for Wellbeing Assessment<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Sophie Chiang, Tom Brennan, Fethiye Irmak Dogan, Jiaee Cheong, Hatice Gunes<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Breathing Signal Prediction from Speech: Toward Reproducible and Comparable Models<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Francisca Pessanha, Glenn de Wildt, Alexis Deighton MacIntyre, Heysem Kaya, Almila Akdag<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Speech Signals Complement LLMs for Predicting Interpersonal Attraction in Speed Dating<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Yuriko Kikuchi, Takato Hayashi, Ryusei Kimura, Naoya Inoue, Ryo Ishii, Shogo Okada<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">From Signals to Connection: Exploring Physiological Markers of Human\u2013Nature Connectedness<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Olivia Brunet, Romain Grandchamp, Beno\u00eet Fernandez, Gladys Barragan-Jason, Axel Carlier<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Light-ED: Lightweight Multimodal Emotion Detection using Enhanced EfficientNet<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Wamika Jha, Mea Wang, Usman Alim, Zoe Kirsman<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Unmasking Suppressed Emotion: A Multi-Agent Approach to Affective Dissonance<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Chenghong Lin, Tochukwu Eze, Dawei Xie, Khalil J Anderson, Bookyung Shin, Marcelo Worsley<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Multimodal Behavioral Typicality as a Training-Free Screening Signal for Dementia<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Leticia Pinto-Alva, Gale Lucas, Maja J Matari\u0107, Jesse Thomason<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Who, Where, How Matters: Evaluating Emotion Recognition under Context Shifts<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Sayak Mukherjee, Tom Viering, Bernd Dudzik<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">A Unified Tokenization Framework for Pain Recognition using Heterogeneous 3D Modalities<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Stefanos Gkikas, Christian Arzate Cruz, Valentina Becchetti, Muhammad Umar Khan, Alessandro Giuseppi, Raul Fernandez Rojas<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Autism Screening via Pose Kinematics and Shallow Graph Embedding for Small-Sample Dyadic Imitative Interaction<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Bang Li, Zhenxiang Chen, Haoran Sun, Xiaoqing Jiang, Jianwei Gu, Chenyang Liang, Shengduo Hu, Kaiyun Li<\/span><\/em><\/p>\n<div id=\"oral-session-1\"><\/div>\n<p><strong>11:00-12:30 Oral Session #1: Conversational AI, Speech &amp; Avatars<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><strong>11:00<\/strong> <span style=\"font-weight: 400;\">HAAS: Holistic Attention-free Animation from Speech using Mamba<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Laxmi Narayen Nagarajan Venkatesan, Harsh Vardhan Singh, Vansh Sinha, Sai Madhavan G, Subhajeet Lahiri, Rahulraj B R, Vaishnavi Josyula, Dinesh Babu Jayagopi, Raj Tumuluri, Magnus Revang, Ajai Devanathan<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>11:15<\/strong> <span style=\"font-weight: 400;\">TANDE: Disentangling Verbal and Nonverbal Backchannels in Emotional AI-Avatar Conversations With Young Adults<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ann-Kareen Gedeus, Jack Good, Nadine Wagener, Angelique Taylor<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>11:30<\/strong> <span style=\"font-weight: 400;\">Unleashing Chain-of-Thought Style Reasoning for Text-to-Speech Systems<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ziyue Jiang, Zhou Zhao<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>11:45<\/strong> <span style=\"font-weight: 400;\">Conversational Facial Dynamics as Behavioral Identity Signals: From Naturalistic Conversations to Avatar-Mediated Communication<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Masoumeh Chapariniya, Pierre Vuillecard, Jean-Marc Odobez, Volker Dellwo, Teodora Vukovic<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>12:00<\/strong> <span style=\"font-weight: 400;\">AFA: Identity-Aware Memory for Preventing Persona Confusion in Multi-User Dialogue<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Mohammad Al-Ratrout, Pavan Uttej Ravva, Shayla Sharmin, Aditya Raikwar, Ju Young Shin, Roghayeh Leila Barmaki<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>12:15<\/strong> <span style=\"font-weight: 400;\">Towards a Voice Design Tool to Support Game Masters of Tabletop Role-Playing Games<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Laura Kanschat, Johanna Kuch, Daksitha Senel Withanage Don, Niklas Heimerl, Silvan Mertes, Elisabeth Andr\u00e9<\/span><\/em><\/p>\n<div id=\"special-session\"><\/div>\n<p><strong>14:00-15:00 Special Session: Multimodal Sensing and Interventions for Mental Well-being<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><strong>14:00<\/strong> <span style=\"font-weight: 400;\">Multimodal Anxiety Detection with Adaptive Hyperbolic Few-Shot Learning<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Aditya Sneh, Nilesh Kumar Sahu, Anushka Sanjay Shelke, Arya Adyasha, Haroon R. Lone<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>14:15<\/strong> <span style=\"font-weight: 400;\">Modeling Affective Distress from Video: A Modality-Aware Multi-Branch Learning Approach<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Om Govind Jha, Nilesh Kumar Sahu, Haroon R. Lone<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>14:30<\/strong> <span style=\"font-weight: 400;\">Multimodal features for multi-perspective assessment of working alliance in therapist-patient psychotherapy<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Rivka Vollebregt, Lennard Ren\u00e9 Bornemann, Sanne J.E. Bruijniks, Albert Ali Salah<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>14:45<\/strong> <span style=\"font-weight: 400;\">\u201cWhy does the computer say I am depressed?\u201d Towards a better understanding of concepts in automatic depression detection<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Sytse Backx, Stan Meyberg, Gizem Sogancioglu, Heysem Kaya<\/span><\/em><\/p>\n<div id=\"poster-session-2\"><\/div>\n<p><strong>15:00-16:30 Poster Session #2: Social Interaction, Dialogue &amp; Human\u2013AI Interaction<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Takes Two to Know One: An Approach for Personality Prediction from Dyadic Conversation Based on Text<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Zijie (Jack) Zhou, Siyuan Chen<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Real-time Multimodal Addressee Detection in Multi-party Dialogue<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Divesh Lala, Koji Inoue, Taiga Mori, Tatsuya Kawahara<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">I Don\u2019t Always Do What I Am Told: The Case for Task Incongruence-Aware Systems for Technology-Supported Reflection on Social Interactions<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Chenxu Hao, Tiffany Matej Hrkalovic, Bernd Dudzik, Hayley Hung<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">The Role of Language Understanding in Speech-Based Multimodal Automatic Personality Perception<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Nisreen Alshubaily, Emily O&#8217;Hara, Tanaya Guha, Alessandro Vinciarelli<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Binaural Speech and Distribution Learning for Predicting Auditory Distance Perception in Audio Augmented Reality<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Sahar Altalhi, Tanaya Guha, Alessandro Vinciarelli<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Multimodal Rapport Estimation in Real-World HRI<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Akihiro Sakuramoto, Takato Hayashi, Ryo Miyoshi, Yuki Okafuji, Shogo Okada<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Automated Behavioural Analysis in Parent-Infant Interactions using Multimodal-Large-Language Models<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Daksitha Senel Withanage Don, Tobias Hallmen, Mitho M\u00fcller, Lea Kaubisch, Linda St\u00fcrmlinger, Yaren G\u00fcnay, Anna-Lena Zietlow, Beate Ditzen, Johannes C. Ehrenthal, Cristina Luna-Jim\u00e9nez, Prof. Dr. Corinna Reck, Elisabeth Andr\u00e9<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">HARMONI: Multimodal Personalization of Multi-User Human-Robot Interactions with LLMs<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Jeanne Malecot, Hamed Rahimi, Jeanne Cattoni, Marie Samson, Mouad Abrini, Mahdi Khoramshahi, Maribel Pino, Mohamed Chetouani<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Multimodal Behavioural Indicators of Social Performance in Collaborative Problem Solving<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Jennifer Hamet Bagnou, Amine Benamara, C\u00e9line Clavel, Jean-Claude Martin, Elise Prigent<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">&#8220;Great! The Next Step Is&#8230;&#8221;: In Pursuit of Proactive Assistance with Multimodal Foundation Models<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Sean Andrist, Dan Bohus, Maia Stiber, Yuwei Bao, Tim Schoonbeek, Eric Horvitz<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Beyond Rhythm and Prosody: Inter-Utterance Silence Predicts Drone Movement Duration in Spontaneous Vocal Guidance<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Allan Henry, Solange Rossato, Christian Graff, Sylvain Huet, Jose-Ernesto Gomez-Balderas<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">From Speech to Interaction: Analyzing Multimodal Systems in Cocktail-Party Scenarios<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Thai Binh Nguyen, Zhaolin Li, Jan Niehues, Alexander Waibel<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Using Machine Mental Imagery for Representing Common Ground in Situated Dialogue<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Biswesh Mohapatra, Giovanni Duca, Laurent Romary, Justine Cassell<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Impact of Adaptive Feedback in Multimodal Human-Agent Interaction<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Shaid Hasan, Sujan Sarker, Tariq Iqbal<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Big5-HCD: A Human-Computer Dialogue Benchmark with Personality Annotations via Semi-Supervised Adaptation<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Weidong Zhang, Haifeng Guo, Weiming Wang, Haoran Xie, Fu Lee Wang<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Clustering Co-Speech Gestures Using Morphological Representations<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Simbarashe Nyatsanga, Huanjie Dong, Ikhsanul Habibie, Christian Theobalt, Michael Neff<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Do Visual Features Improve Other-Initiated Repair Detection? A Dyadic Multimodal Approach<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ha Anh Ngo, Nicolas Rollet, Catherine Pelachaud, Chloe Clavel<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Joint Attention Modeling in Naturalistic Interactions with Social Gaze and Proximity Cues<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Miranda Dickerman, Michael Villamizar, Sabine Stoll, Jean-Marc Odobez<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Facial Expressions and Gaze Behavior Patterns During a Collaborative Board Game<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Amine Benamara, C\u00e9line Clavel, Brian Ravenet, Nicolas Sabouret, Julien Saunier<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Speaker Diarization in Static and Egocentric Videos via Speech-Face Alignment<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Zeyang Zhang, Xavier Yin, Zhaobo Zheng, Kumar Akash, Teruhisa Misu, Carlos Busso<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Deception Detection Across Diverse Interaction Contexts: A Progressive Transfer Learning Approach<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Abdullah Alzahrani, Eve Gittins, Muneeb Ahmad<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Validation of a Multi-level Self-Report Rapport Scale and its Impact on Multimodal Modeling of Small Group Interaction<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Justine Reverdy, Oussama Silem, Justine Cassell<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Real-time Generation of Listener Nodding via Prediction of Kinematic Parameters for Avatar Dialogue Systems<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Kazushi Kato, Koji Inoue, Taiga Mori, Divesh Lala, Tatsuya Kawahara<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Defining \u201cNeutral\u201d Agents: A Critical Review and Future Guidelines<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Emma Jane Pretty, Weichen Li, Anna-Leena Macey, Nannan Xi, Juho Hamari<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Code-to-Speech: An Exploratory Study of AI Text-to-Speech for Source Code Readability<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Marco Cardia, Letizia Angileri, Marina Buzzi, Barbara Leporini<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">CoRead: Using Implicit Behavioural Signals for Mixed-Initiative AI Support in Collaborative Academic Reading<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Sakil Sarker, Yasaman A.Basti, Heidar Davoudi, Ali Neshati<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">NVFCorpus: A Corpus for Analyzing Multimodal and Multifunctional Nonverbal Behavior in Multiparty Conversations<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Kazuhiro Otsuka, Takumi Nishihira, Yuya Nishimura, Issa Tamura<\/span><\/em><\/p>\n<div id=\"oral-session-2\"><\/div>\n<p><strong>16:30-18:00 Oral Session #2: Emotion, Affect &amp; Well-being<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><strong>16:30<\/strong> <span style=\"font-weight: 400;\">Smiling Regulates Emotion During Traumatic Recollection<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Marcus Ma, Emily Zhou, Leonard Ludwig, Julia H\u00f6rath, Christina Winkler, Kleanthis Avramidis, Tiantian Feng, Gabor Mihaly Toth, Alina Bothe, Shrikanth Narayanan<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>16:45<\/strong> <span style=\"font-weight: 400;\">Hierarchical Quantized Cross-modal Masked Autoencoders for Audio-Visual Emotion Recognition<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Zeyang Zhang, Carlos Busso<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>17:00<\/strong> <span style=\"font-weight: 400;\">MindBeat: A Multimodal Interface for Stress Recovery through Rhythmic Fidgeting<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Wenjie Xu, Ziang Xu, Rui Zhang, Qinjie Wang, Fangtian Ying<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>17:15<\/strong> <span style=\"font-weight: 400;\">Beyond Surveys: Predicting Student Well-being from Open-ended Video Responses using LLMs<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Thu Bui, Minjun Choi, Shifa Somji, Hillary Merzdorf, Nan Kong, Louis Tay, Sooyeon Jeong<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>17:30<\/strong> <span style=\"font-weight: 400;\">Interoception-Inspired Emotion Estimation via Intrinsic Diffuseness-Guided Learning<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Haifeng Zhang, Von Ralph Dane Marquez Herbuela, Yukie Nagai<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>17:45<\/strong> <span style=\"font-weight: 400;\">EmotiStage: Designing an Emotion-Driven Performance System with Multi-modal Generative AI to Support Emotional Awareness<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ziying Wang, Yihong Lin, Xianglin Zhao, Yirui Huang, Ziya Zhou, Wei Xue, Wanling Cai, Yucheng Jin<\/span><\/em><\/p>\n<div id=\"challenge-overview\"><\/div>\n<p><strong>18:00-18:30 Challenge Overview<\/strong><\/p>\n<h4><strong>Main Conference Day 2 &#8211; Wednesday, 7 October 2026<\/strong><\/h4>\n<div id=\"demo-dc\"><\/div>\n<p><strong>10:00-11:00 Demo, DC<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><strong>Demonstrations<\/strong><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">AUGI: Multimodal Gestural and Tactile Control for Acoustic Wind and Brass Instruments<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Susan E Green-Mateu, Jackson Tallamy, Kyle Hutchins<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">BuddyBack: A Multimodal Smart Posture Correction System<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Federico Raponi, Marco Realacci, Lorenzo Spataro, Danilo Avola, Maurizio Mancini, Emanuele Panizzi<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">ConceptLens: Interactive Concept Bottlenecks for Black-Box Models<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Hasan Md Tusfiqur Alam, Abdulrahman Mohamed Selim, L\u00e1szl\u00f3 Kop\u00e1csi, Daniel Sonntag<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">Embodying Luca Giordano: A RAG-Grounded Context-Aware Robot for Cultural Heritage Storytelling<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Mario Barbato, Marco Grazioso, Martina Di Bratto, Azzurra Mancini, Valentina Russo, Silvia Rossi<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">Emotional Echoes: A Spatialized VR Gallery for Revisiting and Reinterpreting Emotional Memory<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Yao Liu, Marina Carulli, Monica Bordegoni<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">FLIDIS: An Experimental Infrastructure for Dual-Task Research in Continuous Control and Multimodal Interaction<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Charles O Njoku, James Blundell<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">MULTIDATA: An Interactive Web Platform for Speech-Gesture Data Extraction and Visualization<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ra\u00fal S\u00e1nchez, Daniel Alcaraz Carri\u00f3n, Irene Bolumar Mart\u00ednez, Armine Garibyan, Yassine Iabdounane, Mark Turner, Peter Uhrig, Crist\u00f3bal Pag\u00e1n C\u00e1novas<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">Vibrotactile Feedback for Accessible Audio Production: Demonstrating HAPCI and HFAM<\/span><br \/>\n<em><span style=\"font-weight: 400;\">James Hurley, Jon Drummond, Bert Bongers<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">Windborne: Fostering Nature Connection and Playfulness through Multimodal Interaction in Healthcare Settings<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Borui Wang, Keith Evan Green<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>Doctoral Consortium Posters<\/strong><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">Multimodal Modeling of Microsurgery Training Assessment<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ayobami Oyewole<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">Towards Memory-Aware Adaptive Agents<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Deborah van Sinttruije<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">Buzz, Bid, Body: Multimodal Practices of Distributed Recognition<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Pipob Puthipiroj<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">Generic modeling of Human-Machine Interaction<\/span><br \/>\n<em><span style=\"font-weight: 400;\">lykong un<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">Multimodal detection and classification of humor in human-human and human-machine interactions<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Sofia Callejas<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">Structured Representations For Human-Robot Interaction with Non-Expert Users<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Valerie K. Chen<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">What Matters Over Time: Temporal Properties of Non-Verbal Signals for Intent Recognition in Human\u2013Robot Interaction<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Eunyoung Hwang, Alessandra Rossi, Silvia Rossi<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">Designing Multimodal Conversational AI for Longitudinal Well-Being Monitoring<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Thu Bui<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">Socially Interactive Agent for Group Debate Mediation<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Nahuel G\u00f3mez<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">FAU-Preserving Video Anonymization for Privacy-Aware Reproducible Research<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Sophin Chheng<\/span><\/em><\/p>\n<p style=\"padding-left: 60px;\"><span style=\"font-weight: 400;\">Multimodal Subjective Intention Modeling in Social Interactions<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Arthur Mercier<\/span><\/em><\/p>\n<div id=\"oral-session-3\"><\/div>\n<p><strong>11:00-12:30 Oral Session #3: Neurophysiological &amp; Wearable Multimodal Sensing<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><strong>11:00<\/strong> <span style=\"font-weight: 400;\">Multimodal Learner-State Analytics: Cross-Modal Insights from Motor-expressions and EEG for Instructor-Facing Visualization Platform<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Qi Song, Gaoyuan Zhang, Chengchen Lyu, Xurong Xie, Naiming Yao, Hui Chen, Feng Tian<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>11:15<\/strong> <span style=\"font-weight: 400;\">\u03bcDPad: A Large-Scale Multimodal PPG and IMU Dataset for Wrist-Worn Microgesture Recognition<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Lars Hauptmann, Dominik Hollidt, Xintong Liu, Manuel Meier, Christian Holz<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>11:30<\/strong> <span style=\"font-weight: 400;\">MoDAl: Self-Supervised Neural Modality Discovery via Decorrelation for Speech Neuroprosthesis<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Yuanhao Chen, Peter Chin<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>11:45<\/strong> <span style=\"font-weight: 400;\">SENSE: Efficient EEG-to-Text via Privacy-Preserving Semantic Retrieval<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Akshaj Murhekar, Christina Liu, Abhijit Mishra, Shounak Roychowdhury, Jacek Gwizdka<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>12:00<\/strong> <span style=\"font-weight: 400;\">Investigating Foundation Models, Disentanglement and Latent Alignment for Subject-Independent EEG Learning<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Lia Schmid, Jacopo Burger, Alessandro D&#8217;Amelio, Raffaella Lanzarotti<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>12:15<\/strong> <span style=\"font-weight: 400;\">MIMIC: A Real-Time Multimodal Framework to Provide Fine-Grained Motion Corrections Remotely<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Blain Judkins, Jennifer Yentes, Edgar Javier Rojas-Mu\u00f1oz<\/span><\/em><\/p>\n<div id=\"blue-sky-papers\"><\/div>\n<p><strong>14:00-15:00 Blue Sky Papers<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><strong>14:00<\/strong> <span style=\"font-weight: 400;\">Human Hardware in an Agentic World: Jia\u2019s Journey as a Biological Actuator and the Call for Sovereign Interaction Design<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Lik-Hang Lee<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>14:20<\/strong> <span style=\"font-weight: 400;\">The Person Multimodal AI Cannot Finish<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Mohammad Rashedul Hasan<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>14:40<\/strong> <span style=\"font-weight: 400;\">The Recognition-Consequence Gap: How Gesture Recognition Misses Human Action<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Stephanie A Scopelitis, Dane DeSutter<\/span><\/em><\/p>\n<div id=\"poster-session-3\"><\/div>\n<p><strong>15:00-16:30 Poster Session #3: Multimodal Learning, Vision-Language Models &amp; Responsible AI<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">GRoFA: Noise-Gated Adapters for Jointly Fair and Robust Face Embeddings<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Amu Suemoto, Yutaka Arakawa, Tsunenori Mine<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">LEGO Co-builder: Exploring Fine-Grained Vision-Language Modeling for Multimodal Assembly Assistants<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Haochen Huang, Yue Su, Xin Sun, Moonisa Ahsan, Mohammad Aliannejadi, Irene Viola, Zhaochun Ren, Chuang Yu, Aneta Lisowska, Artem Belopolsky, Koen Hindriks, Pablo Cesar, Junxiao Wang, Jiahuan Pei<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">How YouTube Frames ChatGPT Use in Education: An Epistemic Network Analysis with Supporting Multimodal Metadata<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Shayla Sharmin, Mohammad Al-Ratrout, Mohammad Fahim Abrar, Roghayeh Leila Barmaki<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Bridging the Spatial Blind Spot in Pathology Foundation Models for Tumor Localization<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Guoshuai Xu, Kai Zhang<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">STACK: 3D Spatial Task Analysis in Collaborative Construction Keyframes<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Changsoo Jung, Sheikh Abdul Mannan, Jack Fitzgerald, Ethan Seefried, Videep Venkatesha, Sifatul Anindho, Nathaniel Blanchard<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">NewsSegCap: Efficient Dense Video Captioning for TV News<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Antoine Brimont, Titus Zaharia, Ruxandra Tapu<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">TNG-CLIP:Training-Time Negation Data Generation for Negation Awareness of CLIP<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Yuliang Cai, Jesse Thomason, Mohammad Rostami<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">When Cross-Modal Attention Fails under Channel Degradation: Diagnosing Dominant Interaction Structure for Reliable Multimodal Interaction<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Bin Liu, Zhaoxiang Xiao, Zhengpeng Liu<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Distilling the Noise: Value-Aware Contrastive Hypergraph Learning for Robust Multimodal Rumor Detection<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Heng Zhang, Weiyu Zhang, Chaoqun Zheng, Rui Wang, Jiasheng Si, Wenpeng Lu, Deyu Zhou<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Self-Distillation as a Structure-Dependent Mechanism in Multimodal Dialogue via Multi-Context Modeling<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ryo Ishii, Chihiro Takayama, Jiro Nagao, Toshiki Onishi, Yukiko Nakano, Junichi Sawase<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Human-Centered Metrics for Evaluating Usability in Artificial Intelligence Generated Human Object Interactions<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Shengdi Xiao, Yuto Asano, Jingjing Li, Yoichi Ochiai<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Towards Gaze-Informed AI Disclosure Interfaces: Eye-Tracking Attentional and Cognitive Load While Reading AI-Assisted News<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Pooja Prajod, Hannes Cools, Thomas R\u00f6ggla, Pablo Cesar, Abdallah El Ali<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">SIGMA: Statistical Memory Head for Robust Class-Incremental Vision\u2013Language Learning with CLIP<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Dinh-Dat Nguyen, Duc-Duy Mai, Hung Xuan Ho Mr, Quynh-Trang Pham Thi, Thanh-Hai Dang<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">M3CA: Multi-modal and Multi-view Fusion based on Multi-layer Cross Attention for Isolated Sign Language Recognition<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Jinyang Feng, Zhong Guan, Yongli Hu, Jiahai Yang<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">An Explainable Neurosymbolic Multimodal Deepfake Detection with A Comparative Study of Cross-Dataset Generalization<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Md Shafiqul Baten Sumon, Fatma Najar<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">From Bytes to Semantics: LLM-Agent-Driven Commit Message Generation for Non-Code Assets<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Liran Wang, Zhoujun Li<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">HuGDAT: Guiding Transformer Attention with Human Vision<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ahmad Waseem, Pietro Ruiu, Andrea Lagorio, Seth Nixon, Massimo Tistarelli<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Local Optical Flow for Eye Movement Event Detection in Head-Mounted Setups<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Abdulrahman Mohamed Selim, Omair Shahzad Bhatti, Lukas Wilde, Khue Minh Pham, Michael Barz, Daniel Sonntag<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Towards Responsible AI in Low-Resource Languages: Evaluating LLM Responses with a Culturally Aware Sinhala Dataset<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Liyanaarachchige Dona Rashmi Nirasha Gunawardana, Ayantha Randika, Dushani Perera, Kasun Karunanayaka<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">ADMC: Attention-based Diffusion model for Missing modalities Completion<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Yuhan Li, Wei Zhang, Juan Chen, Jiangjia Yan, Peng Xiangli, En Zhu, Liangze Yin<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">TraceNote: Instructor Cursor Dynamics as Spatiotemporal Signals for Importance-Aware Lecture Note Generation<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Yao Tong, Su Wang<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">ECHO: Explainable Co-editing with Human-in-the-loop Operations for Presentation Slide Refinement<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Yu Fu, Yongqi Kang, Yujia Zhou, Yong Zhao<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">EST-inspired computational approaches to event boundary detection<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Amirarsalan Serajoddin Mirghaed, Gualtiero Volpe, Giovanna Varni<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Pre-processing Techniques for Multimodal Affective Computing: Implications for Data Quality and Multimodal Fusion<\/span><br \/>\n<em><span style=\"font-weight: 400;\">C\u00e9lio Carvalho, Jos\u00e9 Torres, Rui Silva Moreira<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">GroundingGaze: An Interactive Pipeline for Automatic Gaze Annotation<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Mounika Kanakanti, Ferdinand Otmar Paar F.P., Asli Ozyurek, Chinmaya Mishra<\/span><\/em><\/p>\n<div id=\"oral-session-4\"><\/div>\n<p><strong>16:30-18:00 Oral Session #4: Engagement, Behavior &amp; Human State Modeling<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><strong>16:30<\/strong> <span style=\"font-weight: 400;\">Interpretable Multimodal Engagement Prediction with Graph-based Mixture-of-Experts<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Monisha Singh, Abhinav Dhall<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>16:45<\/strong> <span style=\"font-weight: 400;\">From Detection to Mechanism Analysis: Interpretable Multimodal Concept Interactions in Gesture Pragmatics<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Jinqian Zhang, Sixia Li, Candy Olivia Mawalim, Shogo Okada<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>17:00<\/strong> <span style=\"font-weight: 400;\">Self-Supervised Representation Learning for Heterogeneous Behavioral Expressions of Social Engagement<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Naga Venkata Sai Raviteja Chappa, Lisa Yankowitz, Gokul M Nair, Evangelos Sariyanidi, Casey J. Zampella, Kathleen Campbell, Whitney Guthrie, John D. Herrington, Robert T. Schultz, Birkan Tunc<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>17:15<\/strong> <span style=\"font-weight: 400;\">Human-as-a-Sensor: Inferring Individual Engagement from Teammates Alone via Latent Group State Modeling<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Zhaobo Zheng, Navid Salami Pargoo, Kumar Akash, Teruhisa Misu<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>17:30<\/strong> <span style=\"font-weight: 400;\">How Reliable Are Multimodal Signals of Conversational State? Evidence from Remote Dyadic Collaborative Tasks<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Tahiya Chowdhury<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>17:45<\/strong> <span style=\"font-weight: 400;\">SQUAD: A Unified Framework for 3D Dyadic Scene Understanding with Applications to Visual Focus of Attention and Object Use<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Berfu Karaca, Niilo V. Valtakari, Albert Ali Salah, Jaap J. A. Denissen, Sonja M. C. de Zwarte, Ronald Poppe<\/span><\/em><\/p>\n<h4><strong>Main Conference Day 3 &#8211; Thursday, 8 October 2026<\/strong><\/h4>\n<div id=\"poster-session-4\"><\/div>\n<p><strong>10:00-11:00 Poster Session #4: XR, Embodied Interaction &amp; Multimodal Interfaces<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Quantifying 3D Pointing: Characterizing What Happens During Pointing Selection in Virtual Reality<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Junhan Kong, Jacqui Fashimpaur, Michael J Proulx, Hemant Bhaskar Surale, Jacob O. Wobbrock, Amy Karlson<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Did I Just Cheer in a Concert Hall?: Enhancing Remote Participation with Congruent Self-Voice Feedback<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Hikari Kato, Tomoaki Konno, Toshiharu Horiuchi<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">XRPoseSync: A Synchronized Multimodal Dataset for EdgeXR Pose Prediction<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ziyu Zhong, Jens Nirme, Hector A Caltenco, Stefan Lindgren, Bj\u00f6rn Landfeldt<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">TriSelect: Distance Adaptive Multimodal Object Selection in Virtual Reality Using Eye Gaze, Hand Gesture, and Voice<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Jai Sachdeva, Prashant Rawat<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Partner Representation Shapes Coordination Tradeoffs in Shared XR<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ehtisham Ul Haq, Robert S. Allison, Laurie M Wilcox<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Hands-VRee: Towards Hands-Free Gaming in VR via Concurrent Input for Locomotion, Viewport Control, and Pointing<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Adrian Javier Leon Valencia, Pedro Tavares, In\u00eas Alves, Ian Oakley, Augusto Esteves<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">IntentVLM: Open-Vocabulary Intention Recognition through Forward\u2013Inverse Modeling with Video-Language Models<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Hamed Rahimi, Cl\u00e9mence Grislain, Adrien Jacquet Cr\u00e9tides, Olivier Sigaud, Mohamed Chetouani<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Disrupted Harmony: How Visual\u2013Audio Incongruence Influences Creativity in Collage Making<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Leyan Wu, Yalong Luo, Andi Wan, Yucheng Jin<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Han-Vibe: Culturally-Situated Multimodal Interaction for AI-Mediated Live Music Performance<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Zhengyang Ma<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Audiovisual Material Perception in Virtual Reality<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Harshitha Koppisetty, Laurie M Wilcox, Robert S. Allison<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Multimodal User Simulation with Cross-Modal Grounding for Virtual Museum Interaction<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Shreya Sanghamitra, Shihan Wang, Shenghui Wang<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">What you know, How you act: Transactional VR Authentication<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Numan Zafar, Shafique A Chaudhry<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Pose-to-Command: A Personalized Embodied Control Layer for Ordinary Games<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Tung Khau, Matthias Rueger, Gerrit Meixner<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Exploring Thermal Color Cues to Evoke Thermal Illusions in Virtual Object and Hand-to-Hand Interactions in VR<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Artem Sultanov, Alexander Raake, Stephanie Arevalo Arboleda<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Inferring Openness to Experience from Behavioral Signals in a Virtual Reality Art Gallery<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Aaditya Vardhan Narain, Raman Saxena, Y. Raghu Reddy<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Causal Temporal Padding for Low-Latency Real-Time Gesture Generation<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ryo Ishii, Shinichiro Eitoku, Jiro Nagao, Junichi Sawase<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Auditory Feedback as a Communication Modality: Modeling Perceived Semantic Expression of UX Sounds<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Annika Frommholz, Steffen Lepa, Stefan Weinzierl, Johannes Helberger<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">AdaHome: An Adaptive Smart Home Assistant using Local Small Language Models<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Eu Jin Lim, Zhaoxing Li, Sebastian Stein<\/span><\/em><\/p>\n<div id=\"oral-session-5\"><\/div>\n<p><strong>11:00-12:30 Oral Session #5: Multimodal Reasoning, Explainability &amp; Methods<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><strong>11:00<\/strong> <span style=\"font-weight: 400;\">DatasetOS: An Agentic Framework For Multimodal Dataset Construction<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Bashayer Alsreidi, Ghaleb Aldoboni, Lobna Nassar, Fakhreddine Karray<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>11:15<\/strong> <span style=\"font-weight: 400;\">Explanation Needs Emerge Before Explicit Requests in Human-Robot Interaction<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Dimosthenis Kontogiorgos, Joakim Gustafson, Julie Shah<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>11:30<\/strong> <span style=\"font-weight: 400;\">S-AODP++: Structure-Aware AOI-Guided Scanpath Attention for AI-Generated Image Identification<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Jingming Wang, Renwei Meng, Yunjie Si<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>11:45<\/strong> <span style=\"font-weight: 400;\">SPICE: Synergy and Partial Information Based Curriculum Evolution<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ankush Pratap Singh, Houwei Cao, Yong Liu<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>12:00<\/strong> <span style=\"font-weight: 400;\">Scene-Conditioned Capability Mounting as a Visible and Executable Boundary for Situated LLM Agents<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Siyu Jiang, Yuxuan Cheng, Jiayi Wang, Sanshuai Cui<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>12:15<\/strong> <span style=\"font-weight: 400;\">MapQA: A Map-Question-Answering Benchmark for Visual Language Model Reasoning<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Christian Arnold, Andrew Alini, Abdulrahman Alabdulkareem, Jonathan Wang, Pieter Feenstra, Conner Arnold, Jan Kenneth DeWitt, Natalie Christina Ritsema, Jung Hyun Yae, Boris Katz, Andrei Barbu, Brian Cheung<\/span><\/em><\/p>\n<div id=\"lbr\"><\/div>\n<p><strong>15:00-16:30 LBR<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">How Local AI Framing Shapes User Experience and Perceived Data Security with an Embodied AI Psychotherapist in VR<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Paula Friedrich, Lukas Polifke, David Obremski, Marc Erich Latoschik, Carolin Wienrich<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">A Clinical Decision Support System Prototype for Multimodal Youth Mental Health Monitoring<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Carlo Mazzola, Helen Stocks, Ula Kolinska, Izidor Mlakar, Gwendolyn Mayer, Zouhair Haddi<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Designing a Functional Social Robot for Blood Test Procedures in the Emergency Department: A Human-Centered Approach<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Yuval Rubin Kopelman, Dikla Dahan Shriki, Zohar Fein, Mahmod Hamdan, Rony Ben Ami, Daniel Armoni, Andrey Grishko, Nevo Heimann Saadon, Hadas Erel<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">&#8220;Nobody Told me Farming is not Only Farming&#8221;: A Conversational AI Agent for Causal Loop Reasoning with Japan Farmers<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Tomohide Mizuuchi, Eduardo Feteira, Liuru Nan, Mami Yoda, Wan-Jou She, Panote Siriaraya<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Structured Patterns of Dyadic Motion Coordination Predict Communicative Success<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Aditya Mutharasu, Elayna Bowe, James Benjamin Falandays<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Benchmarking Prompt Extension and Image Generation Models for Automotive Instructional Imagery<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Yuval Zak, Eli Tzirkel<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">CEMAS &#8211; a Configurable Embodied Multi-Agent System<\/span><br \/>\n<em><span style=\"font-weight: 400;\">David Obremski, Lukas Polifke, Marc Erich Latoschik, Carolin Wienrich<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Emerging Agency-Aware Interface Patterns in AI-in-the-Loop Platforms<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Massimo Bertolotti, Silvia Di Giulio, Genoveffa Tortora, Luigi Troiano<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">How Humans Evaluate Headline-Map Climate Change Misinformation with Multi-Agent (In)consistent Guidance?<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Nianhua Liu<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Aligning Gaze, Space, and Cognition: A Multimodal Triangulation Framework for Visitor Interactions in Cultural Heritage (An Expert vs. Non-Expert Comparative Study)<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Zihe Wei, Yufei Liu, Dan Zhang, Xin Zhang, Miao Sun<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">How to Greet Humans: Designing Opening Encounters for Functional Robots with Naive Users<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Tom Hitron, Nevo Heimann Saadon, Guy Doron, Thomas H. Weisswange, Hadas Erel<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Implied Expressive Motions in Paintings: An Early Study on the Limits of Text-to-Motion Generation<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Mohammadreza Mahmoudi, Nicola Corbellini, Cora Gasparotti, Nicola Ferrari, Antonio Camurri<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Do Emotional Cues Matter? Exploring Support Strategy Selection in LLM-Based Supportive Conversations<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Weiyi Tian, Safak Dogan, Jie Meng<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Beyond the Vibration Channel: A Signal-Level Benchmark for Dry-Sand Touch on Phones<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Takahiro Yamamoto, Toshiharu Igarashi<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Continuous wrist kinematics carry lexical cues for temporal deixis<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ra\u00fal S\u00e1nchez S\u00e1nchez, Crist\u00f3bal Pag\u00e1n C\u00e1novas<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Does Listening Matter? Backchanneling and Nodding in AI Clone<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Koji Inoue, Kazushi Kato, Tatsuya Kawahara, Shunichi Kasahara<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">An Exploratory Multimodal Analysis of Digital Handwriting and Spontaneous Speech Features to Distinguish Early Cognitive Impairment from Normal Aging<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Maria Santina Ler, Maria Sarno, Miriam Veneziano, Gennaro Cordasco, Anna Esposito, Antonio Perna<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">3D Gaze Ray Dynamics During Palpation-Centered Nursing Assessment of Edema<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Masato Fukuda, Mitsuhiro Goto, Koki Ebina, Aya Saitoh, Noriko Yagyuda, Shigekuni Kondo, Sayuri Sakai<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Modeling Participant Behavior Across Utterances and Silent Intervals for Engagement Estimation of Individuals with ASD<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Hiyori Toda, Jie Zeng, Fumio Nihei, Chihiro Takayama, Ryo Ishii, Masatsugu Tsujii, Yukiko Nakano<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">A Multimodal Latent State-Space Framework Reveals Spontaneous Interpersonal Synchrony Dynamics<\/span><br \/>\n<em><span style=\"font-weight: 400;\">PhD Yubraj Gupta, PhD Stefanie Suttkus, PhD Steffen Schulz, PhD Feliberto de la Cruz, PhD Maria Geisler, Dr Karl-Juergen Baer, PhD Andy Schumann<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">A Multilingual Speech-Driven Semantic Input Module for University Indoor Wayfinding<\/span><br \/>\n<em><span style=\"font-weight: 400;\">G\u00f6khan Ceylan, Gabriella Trasciatti, Emanuele Panizzi<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">How AI Experiences Art: Emergent Aesthetic Structure in a Self-Supervised Multimodal Embedding Space<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Corey DC Heath<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Low-Latency Turn-Taking via Context-Aware Preface Generation in a Real-World Dialogue Robot<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Yuki Okafuji, Koji Inoue, Yoshiki Ohira<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Explainable AI Insights into Gender-Associated Facial and Upper-Body Patterns in Public Speaking: An Exploratory Study<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Nesrine Fourati, Kamel Madi<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Is Bad Handwriting Really Bad? Multimodal Human\u2013AI Interpretation of Illegible Handwriting in Palliative Care<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Zhigang Ni, Lei Yang<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">DyadicTickTacking: Multimodal Co-Evolution of Joint Strategies and Turn-Taking in Expert-Novice Scaffolding<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ali KhaliliKandjani, Gabriele Romano, Nicola Corbellini, Gualtiero Volpe<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">When Coordination Signals Can Mislead: Conversational and Facial Markers of Transactive Functioning in Newly-Formed Virtual-Reality Teams<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Tristan Lannuzel, Beatrice Biancardi, Mukesh Barange, St\u00e9phanie Buisine<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">A Modular Open-Source Augmented Reality Drone Environment for Human\u2013Drone Interaction Research<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Vilmer Malm, Hugo Raimer, Javier Jim\u00e9nez Ruescas, Fanta Camara, Mohammad Obaid<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">A Reproducible Multimodal LLM-as-a-Judge Workflow for Exploring Epistemic Dissonance<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Rei Yuda<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Can Vision-Language Models Extract Relevant Multimodal Cues for Post-Inference Cue Integration? A Comparison of Affective Judgments on Adult and Child Interaction Data<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Xiajie Zhang, Sharifa Alghowinem, Hae Won Park, Cynthia Breazeal<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">CreaRobot &#8211; Improving creativity using a social robot<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Salvatore Maria Anzalone<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Towards a Multimodal Video-Based Approach for Editable Piano Score Transcription<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Silvia Di Giulio, Luigi Troiano, Genoveffa Tortora<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">SMILE: A Hybrid Memory-Driven Dialogue Manager for Social Companion Robots<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Mariacarla Staffa, Salvatore Maria Anzalone, Lorenzo D&#8217;Errico<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">LLMs with a Heart: Using Contextual Heartbeat Haptics to Add Feeling to Text and Audio LLM Interactions<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Prithiv Premkumar, Shuyuan Liu, Ryan Clark, Ken Shibata, Christopher Vaughn Casarez, Bruce N. Walker<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">Evaluating the Impact of Large Language Model Assistance on Human Attention Paradigm in Human\u2013Robot Collaboration<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Tajbeed A. Chowdhury, Jose A. Trapero, Farah I. Corona-Strauss, Daniel J. Strauss<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">From Plots to Words: Model-Aware Multimodal Explanations as a Foundation for Accessible, Non-Visual Interaction<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Nur Kele\u015fo\u011flu, \u0141ukasz Sobczak, Joanna Doma\u0144ska<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><span style=\"font-weight: 400;\">From Blind Edits to Verified Repair: Building Trustworthy User-Side LLM Agents for Web Accessibility<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Lily Bundgaard Wanscher, Markus Heidemann Lorensen, Mohammed Ammad Shafiq, Mahyar Tourchi Moghaddam, Mina Alipour<\/span><\/em><\/p>\n<div id=\"oral-session-6\"><\/div>\n<p><strong>16:30-18:00 Oral Session #6: Human\u2013AI Collaboration, Co-Creation &amp; Simulation<\/strong><\/p>\n<p style=\"padding-left: 40px;\"><strong>16:30<\/strong> <span style=\"font-weight: 400;\">Ethical Guidelines and Stakeholder Perspectives on the Use of Conversational AI in Psychotherapy &#8211; A Scoping Review<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Paula Friedrich, David Obremski, Marc Erich Latoschik, Carolin Wienrich<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>16:45<\/strong> <span style=\"font-weight: 400;\">CHORUS: Designing Human\u2013AI Multi-Agent Collaboration for Professional Translators<\/span><br \/>\n<em><span style=\"font-weight: 400;\">George Xi Wang, Jiaqian Hu, Guande Wu, Jing Qian<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>17:00<\/strong> <span style=\"font-weight: 400;\">Framing the Formless: A Xing\u2013Qi\u2013Shen Framework for Culturally Aware Multimodal Human\u2013AI Co-Creation<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Conggang Yu, Shudan Tan, Yuxing Zhang, Lin Yu<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>17:15<\/strong> <span style=\"font-weight: 400;\">Memory-Driven Self-Disclosure and Relational Turning Points: A Longitudinal Multimodal Study of Human-AI Interaction<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Ryuichi Sumida, Mao Saeki, Masaki Eguchi, Sadahiro Yoshikawa, Koji Inoue, Tatsuya Kawahara, Yoichi Matsuyama<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>17:30<\/strong> <span style=\"font-weight: 400;\">Context-Aware Multimodal AI System for Wildfire Decision Support<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Shaurya Mathur, Shreyas Bellary Manjunath, Nitin Kulkarni, Alina Vereshchaka<\/span><\/em><\/p>\n<p style=\"padding-left: 40px;\"><strong>17:45<\/strong> <span style=\"font-weight: 400;\">A Methodology to Compare Real and LLM-Simulated Multimodal Interactions: The Case of Gender Bias<\/span><br \/>\n<em><span style=\"font-weight: 400;\">Elodie Etienne, Magalie Ochs, Jean-Michel Loubes, Chlo\u00e9 Clavel<\/span><\/em><\/p>\n<p>[\/et_pb_text][et_pb_text module_id=&#8221;AEGC1&#8243; _builder_version=&#8221;4.14.4&#8243; _module_preset=&#8221;default&#8221; header_text_color=&#8221;#282562&#8243; header_4_text_color=&#8221;#672B83&#8243; custom_padding=&#8221;||0px|||&#8221; global_colors_info=&#8221;{}&#8221;][\/et_pb_text][et_pb_text module_id=&#8221;OS1&#8243; _builder_version=&#8221;4.14.4&#8243; _module_preset=&#8221;default&#8221; header_text_color=&#8221;#282562&#8243; header_4_text_color=&#8221;#672B83&#8243; custom_padding=&#8221;||0px|||&#8221; global_colors_info=&#8221;{}&#8221;][\/et_pb_text][\/et_pb_column][\/et_pb_row][\/et_pb_section]<\/p>\n","protected":false},"excerpt":{"rendered":"<p>Sessions &nbsp; Main Conference Day 1 &#8211; Tuesday, 6 October 2026 10:00-11:00 Poster Session #1: Affective Computing, Health &amp; Wellbeing Affect-Aware Game Personalization with Reinforcement Learning: How to Improve Players&#8217; Engagement, Performance, and Emotions Mahyar Tourchi Moghaddam, Tiziano Santilli, Mina Alipour A Screening Method for Children with Autism Spectrum Disorder Based on a Dual-Stream, Multi-Scale, [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":0,"parent":0,"menu_order":0,"comment_status":"closed","ping_status":"closed","template":"","meta":{"_et_pb_use_builder":"on","_et_pb_old_content":"<!-- wp:paragraph -->\n<p>This is an example page. It's different from a blog post because it will stay in one place and will show up in your site navigation (in most themes). Most people start with an About page that introduces them to potential site visitors. It might say something like this:<\/p>\n<!-- \/wp:paragraph -->\n\n<!-- wp:quote -->\n<blockquote class=\"wp-block-quote\"><p>Hi there! I'm a bike messenger by day, aspiring actor by night, and this is my website. I live in Los Angeles, have a great dog named Jack, and I like pi\u00f1a coladas. (And gettin' caught in the rain.)<\/p><\/blockquote>\n<!-- \/wp:quote -->\n\n<!-- wp:paragraph -->\n<p>...or something like this:<\/p>\n<!-- \/wp:paragraph -->\n\n<!-- wp:quote -->\n<blockquote class=\"wp-block-quote\"><p>The XYZ Doohickey Company was founded in 1971, and has been providing quality doohickeys to the public ever since. Located in Gotham City, XYZ employs over 2,000 people and does all kinds of awesome things for the Gotham community.<\/p><\/blockquote>\n<!-- \/wp:quote -->\n\n<!-- wp:paragraph -->\n<p>As a new WordPress user, you should go to <a href=\"https:\/\/icmi.acm.org\/2026\/wp-admin\/\">your dashboard<\/a> to delete this page and create new pages for your content. Have fun!<\/p>\n<!-- \/wp:paragraph -->","_et_gb_content_width":"","inline_featured_image":false,"footnotes":""},"class_list":["post-1990","page","type-page","status-publish","hentry"],"_links":{"self":[{"href":"https:\/\/icmi.acm.org\/2026\/wp-json\/wp\/v2\/pages\/1990","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/icmi.acm.org\/2026\/wp-json\/wp\/v2\/pages"}],"about":[{"href":"https:\/\/icmi.acm.org\/2026\/wp-json\/wp\/v2\/types\/page"}],"author":[{"embeddable":true,"href":"https:\/\/icmi.acm.org\/2026\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/icmi.acm.org\/2026\/wp-json\/wp\/v2\/comments?post=1990"}],"version-history":[{"count":49,"href":"https:\/\/icmi.acm.org\/2026\/wp-json\/wp\/v2\/pages\/1990\/revisions"}],"predecessor-version":[{"id":2916,"href":"https:\/\/icmi.acm.org\/2026\/wp-json\/wp\/v2\/pages\/1990\/revisions\/2916"}],"wp:attachment":[{"href":"https:\/\/icmi.acm.org\/2026\/wp-json\/wp\/v2\/media?parent=1990"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}