Multimodal Behavioral Patterns Analysis with Eye-Tracking and LLM-Based Reasoning

cs.AI updates on arXiv.org 07月25日 12:28

Multimodal Behavioral Patterns Analysis with Eye-Tracking and LLM-Based Reasoning

本文提出一种多模态人机协作框架，通过多阶段管道、专家-模型协同评分模块和混合异常检测模块，提升眼动信号中的认知模式提取，应用于认知建模、自适应学习等领域。

arXiv:2507.18252v1 Announce Type: cross Abstract: Eye-tracking data reveals valuable insights into users' cognitive states but is difficult to analyze due to its structured, non-linguistic nature. While large language models (LLMs) excel at reasoning over text, they struggle with temporal and numerical data. This paper presents a multimodal human-AI collaborative framework designed to enhance cognitive pattern extraction from eye-tracking signals. The framework includes: (1) a multi-stage pipeline using horizontal and vertical segmentation alongside LLM reasoning to uncover latent gaze patterns; (2) an Expert-Model Co-Scoring Module that integrates expert judgment with LLM output to generate trust scores for behavioral interpretations; and (3) a hybrid anomaly detection module combining LSTM-based temporal modeling with LLM-driven semantic analysis. Our results across several LLMs and prompt strategies show improvements in consistency, interpretability, and performance, with up to 50% accuracy in difficulty prediction tasks. This approach offers a scalable, interpretable solution for cognitive modeling and has broad potential in adaptive learning, human-computer interaction, and educational analytics.

Fish AI Reader

AI辅助创作，多种专业模板，深度分析，高质量内容生成。从观点提取到深度思考，FishAI为您提供全方位的创作支持。新版本引入自定义参数，让您的创作更加个性化和精准。

FishAI

鱼阅，AI 时代的下一个智能信息助手，助你摆脱信息焦虑

联系邮箱 441953276@qq.com

相关标签

眼动分析认知模式人机协作多模态认知建模

相关文章

Human-in-the-Loop AI for Emergency Response & More w/ Robert Munro - TWiML Talk #125

Greg 录制了新的ChatGPT实时语音和多模态的演示。最后ChatGPT还即兴创作了一首短歌,歌词涵盖了房间的装饰风格、人物的穿着特点、期间发生的趣味插曲等。真的这...

和@歸藏一起视频会议看完 OpenAI 的发布，讨论了一会，背脊发凉… 1️⃣ 没想到卷推理卷到了这种程度? 现实交流场景下300ms 左右的体验奇点真没想到就这样被...

OpenAI 很鸡贼，提前一天开发布会，让 Google I/O 的气势弱了很多。再加上 Ilya 的官宣离职又分走了不少流量。果然今早一早起来，媒体的报道和用户的关注相比昨...

This AI newsletter is all you need #99

中信建投：OpenAI发布GPT-4o，AGI向前一步

XGen-MM: A Series of Large Multimodal Models (LMMS) Developed by Salesforce Al Research

周鸿祎：留给谷歌的时间不多了，建议把所有产品都开源

Researchers at Stanford Propose TRANSIC: A Human-in-the-Loop Method to Handle the Sim-to-Real Transfer of Policies for Contact-Rich Manipulation Tasks

Lumina-T2X: A Unified AI Framework for Text to Any Modality Generation