Physical Sciences › Computer Science › Human-Computer Interaction
Gaze Tracking and Assistive Technology
154 artículos indexados
Este asunto y su jerarquía proceden de la clasificación OpenAlex, el catálogo abierto de la investigación científica mundial.
Volumen mensual - últimos 12 meses
Países de los laboratorios
- Estados Unidos31 % · 31 artículos
- China25 % · 25 artículos
- Reino Unido14 % · 14 artículos
- Alemania7,9 % · 8 artículos
- Corea del Sur5,9 % · 6 artículos
- Japón5 % · 5 artículos
- Suiza5 % · 5 artículos
- Italia5 % · 5 artículos
Sobre 101 artículos de este tema con al menos un laboratorio localizado. 31 países representados.
Se trata del país del laboratorio, nunca de la nacionalidad de las personas. Un artículo firmado desde varios países cuenta para cada uno de ellos, por lo que las partes suman más del 100 %. La cobertura es parcial y el vacío no es aleatorio: un investigador cuya institución se desconoce suele publicar poco, lo que sobrerrepresenta a los laboratorios consolidados.
Últimos artículos
- Gaze responses to false-positive computer-aided detection prompts during colonoscopy: a paired-video and real-time eye-tracking study
Te Luo, Yan Zhu, Peiyao Fu, Ruijie Yang, Xian Yang, Quanlin Li, Pinghong Zhou, Shuo Wang · 23 de septiembre de 2026
False-positive computer-aided detection (CADe) prompts may divert endoscopists' attention during colonoscopy, yet the attentional impact of individual prompts remains unclear. We used event-locked eye tracking to quantify gaze attraction and attention occupation in complementary retrospective and pr…
- OpenSAL360: Open-Source Crowdsourcing Platform for Omnidirectional Video Saliency Collection
Alexey Bryncev, Andrey Moskalenko, Kira Shilovskaya, Ivan Kosmynin, Dmitriy Vatolin · 21 de septiembre de 2026
Omnidirectional video saliency prediction plays an important role in many immersive multimedia applications, including viewport-adaptive streaming and compression, foveated rendering, mesh simplification, perceptual quality assessment. Yet progress in this area remains constrained by the cost and co…
- AI Smart Glasses for Wearable Intelligence: From Egocentric Sensing to Agentic Personalization
Xu Yuan, Yi Wang, Zhuohang Jiang, Haohao Qu, Yujuan Ding, Shanru Lin, Guoliang Xing, Hongxia Yang, Jiannong Cao, Qing Li, Wenqi Fan · 18 de septiembre de 2026
Recent advances in artificial intelligence (AI) are reshaping smart glasses from egocentric capture and display devices into platforms for wearable intelligence. Smart glasses increasingly serve as wearable AI systems that connect first-person observation with real-time assistance under strict form-…
- SeetaPsych v1.0: An Open-source Computer Vision Toolkit for Behavior-based Psychological Measurement
Jiabei Zeng, Chiqin Li, Kaizhou Li, Fei Chang, Yong Li, Yuanhao Zhao, Dan Han, Wenqiang Yang, Xilin Chen, Shiguang Shan · 18 de septiembre de 2026
Automated visual analysis opens new avenues for behavior--based psychological measurement. Nevertheless, existing technological modules are typically scattered across task specific systems with heterogeneous interfaces and disparate deployment requirements. In this work, we present SeetaPsych v1.0, …
- OptoAgent: A Trustworthy Multi-Agent Framework for Opportunistic Vision Micro-Screening in Classroom Environments
Toqeer Ali Syed, Ali Akarma, Adeel Ahmad, Hammad Muneer · 15 de septiembre de 2026
A child with reduced distance vision often does not know that anything is wrong. Children adapt, move closer, and rarely report the difficulty, so the problem can survive years of schooling before an adult notices. School screening addresses part of this, but it runs on a schedule, depends on staffi…
- Context-Aware Causal Gaze Forecasting for Human-Vehicle Interaction During In-Cabin Tracking Dropouts
Shabnam Shabani, Ghazal Farhani · 14 de septiembre de 2026
Dashboard-mounted gaze trackers often lose sight of the driver's eyes during large head rotations, including shoulder checks, mirror glances, and intersection scanning. These maneuvers occur when information about the driver's visual attention is most useful. Offline gap-filling methods may reconstr…
- TransGaze-Object: Transformer Based Driver Gaze Object Prediction Framework in Real Driving
Pavan Kumar Sharma, Ayush Pande, Pranamesh Chakraborty · 10 de septiembre de 2026
Driver gaze provides information regarding driver visual attention and situational awareness to the surrounding traffic. Existing driver gaze estimation studies represent gaze in terms of gaze zone or gaze vector/point-of-gaze (PoG). However, object-level gaze information provides a more semanticall…
- Marker-free eye-gaze estimation using a single image and depth from defocus
David Hurtubise-Martin, Feriel Fass, Djemel Ziou, Marie-Flavie Auclair-Fortier · 10 de septiembre de 2026
This paper presents a marker-free eye-gaze estimation approach using a single 2D camera, such as an integrated laptop webcam. The gaze-related features are estimated from iris localization and head pose estimated by using depth from defocus. A variational Bayesian multinomial logistic regression fra…
- Diffusion models for eye-gaze trajectory generation using position and velocity representations
Laxman Basnet, Alexander Szorkovszky, Pedro G. Lind, Anis Yazidi, Shailendra Bhandari · 9 de septiembre de 2026
Eye-tracking data are expensive to collect, requiring specialized hardware and controlled laboratory conditions, and difficult to share because of privacy constraints. We address this using two complementary denoising diffusion probabilistic models (DDPMs) for unconditional generation of eye-gaze dy…
- Emergent Goal-Directed Attention in Large Vision-Language Models
Han Zhang · 9 de septiembre de 2026
Human observers prioritize visual information according to task goals. Most computational models of naturalistic viewing are gaze-trained for free viewing, leaving open whether goal-directed attention can emerge in systems without gaze supervision. We tested two off-the-shelf vision-language models …
- EyeMakeYou: Identity-, Task-, and Subjective-State-Conditioned Diffusion for High-Frequency Gaze Synthesis
Kamrul Hasan, Mehedi Hasan Raju, Oleg V. Komogortsev · 7 de septiembre de 2026
Eye movement biometrics (EMB) is an emerging behavioral modality for user authentication, particularly in virtual- and augmented-reality systems, where gaze dynamics contain distinctive subject-specific features. However, robust EMB systems require diverse, high-quality gaze recordings that are expe…
- Hidden In Plain Gaze: Gaze Representations as Privacy Controls for Utility and Re-identification Risk in XR
Cory Ilo, Brendan-David John, Doug A. Bowman · 7 de septiembre de 2026
Intelligent extended reality (XR) systems increasingly use eye and head tracking to infer user intent, task, and attention, but the same signals can also reveal biometric identity. We study whether gaze data representation choice can serve as a lightweight privacy control at feature extraction, befo…
- GazeFS: Target-Centered Gaze-Trajectory Forecasting and Stabilization from Gaze-Head History
Yaozheng Xia, Zaiping Zhu, Bo Pang, Minghao Xie, Hui Li, Shaorong Wang, Sheng Li · 4 de septiembre de 2026
Target-centered gaze interaction requires more than suppressing frame-to-frame fluctuations: target acquisition produces task-aligned changes in gaze-head dynamics, while a gaze trace may retain a persistent target-relative residual direction. We formulate gaze correction as online target-centered g…
- GazeRefine: Expert Gaze as a Test-Time Prompt for Training-Free Medical Image Segmentation
Mohammed Oussama Benyahia, Marouane Tliba, Mohamed Amine Kerkouri, Taifour Yousra, Bin Wang, Max Bengtsson, Gorkem Durak, Elif Keles, Zuheng Ming, Marek Penhaker, Azeddine Beghdadi, Ulas Bagci, Aladine Chetouani · 2 de septiembre de 2026
Medical image segmentation remains difficult to scale because high-performing methods typically rely on dense expert annotations and task-specific training. We introduce GazeRefine, a training-free framework that uses gaze as an inference-time prompt for zero-shot medical image segmentation. Sparse,…
- From Seeing to Acting: Smart Glasses as First-Person Intelligence Platforms
Jiangning Zhang, Haojun Chen, Yong Liu · 26 de agosto de 2026
Smart glasses are evolving from capture and display accessories into first-person intelligence platforms that connect human perception, persistent context, and digital or physical action. Their on-body viewpoint aligns with the wearer's vision, audition, motion, and hand-object interaction, but must…
- Human-Inspired Social Engagement Analysis via Interpretable Mutual Visual Attention
Urwa Fatima, Mohammad Zohaib, Francesca Odone, Nicoletta Noceti · 26 de agosto de 2026
Understanding social interactions from non-verbal visual data is important for behavior analysis and activity monitoring. We propose an interpretable computational model of social engagement inspired by psychological theories of mutual visual attention. Rather than learning interaction patterns end-…
- Predicting Radiologist Expertise from 3D Gaze Patterns During CT Interpretation
Leila Khaertdinova, Anna Anikina, Claudia Mello-Thoms, Bulat Ibragimov · 26 de agosto de 2026
Accurate interpretation of volumetric CT requires efficient navigation of 3D image volumes and attention to diagnostically relevant regions. While eye-tracking has been widely studied in 2D medical imaging, its use for expertise assessment in CT settings remains limited. We propose a gaze-informed t…
- G3Ego: Gaze-Guided Graphs for Egocentric Action Understanding
Marko Haralovi\'c, Akash Ramakrishnan, Estefania Talavera Martinez · 21 de agosto de 2026
Egocentric action understanding is often addressed using large video models pretrained on extensive exocentric datasets. However, many first-person actions depend on a small number of hand-object interactions involving only a few relevant entities. We propose G3Ego, a graph-based framework for ego…
- EgoGazeLite: On-Device Egocentric Gaze Prediction for Token-Efficient Multimodal LLM Video Input
Matteo Stoiber, Niels Buus Lassen · 18 de agosto de 2026
The use of multimodal LLMs (MLLMs) for egocentric video understanding with wearable devices is constrained by the token budget. Memory and compute cost scale with the number of visual tokens, and high-resolution video quickly becomes expensive to transmit and process at scale. Prior work (GazeLLM) a…
- Measuring Browser Webcam Gaze Honestly: A Capture-Clock Methodology and Open Reference Implementation
Chi-Sheng Chen, Gabriel A. Brat · 13 de agosto de 2026
Browser-based webcam gaze trackers are increasingly used for crowd-scale data collection and in clinical settings where lab eye trackers are impractical, but the reported latency numbers may not represent real world functionality. The common practice of timestamping each gaze sample when it is emitt…
- Gaze Target Estimation Anywhere with Concepts
Xu Cao, Houze Yang, Vipin Gunda, Zhongyi Zhou, Tianyu Xu, Adarsh Kowdle, Inki Kim, James M. Rehg · 13 de agosto de 2026
Estimating human gaze targets from images in-the-wild is an important and formidable task. Existing approaches primarily employ brittle, multi-stage pipelines that require explicit inputs, like head bounding boxes and human pose, in order to identify the subject of gaze analysis. As a result, detect…
- Human versus Computer Vision
Elena Sirotkina · 12 de agosto de 2026
Computer vision saliency models predict where people will look, one map per image, and a billion-dollar predicted-attention industry sells those maps in place of measuring real viewers. I test the leading models from the audience side, against 11.4 million webcam gaze points from 3,023 US adults rec…
- Can Webcam Gaze Constrain Mesa-Objectives in Driving Models? An Instrument Precision Analysis
Lennox Anderson, Ahmed Boutar, Jonah Mulcrone, Tal Erez · 11 de agosto de 2026
Current hazard detection systems in autonomous driving may develop mesa objectives, learned internal goals that achieve high training performance through spurious correlations rather than genuine hazard recognition. We investigate whether human gaze patterns, captured via webcam-based eye tracking (…
- From Spatial Semantics to Temporal Context: Leveraging Gaze Trajectory for Weakly Supervised Medical Image Segmentation
Shaoxuan Wu, Xiao Zhang, Xiaodi Zhao, Yunzhi Tian, Yilin Tang, Jun Feng · 30 de julio de 2026
Medical image segmentation heavily depends on labor-intensive and time-consuming pixel-level annotations. Eye tracking offers a cost-effective solution that can be naturally integrated into clinical workflows. Recorded by eye trackers, gaze conveys the spatial regions of clinicians' attention throug…
- Gaze-Anchored Social Net: Decoding Implicit Relations via Joint Modeling
Yuqi Hou, Zhuo Chen, Han Hu, Je Woo Kim, Jianbo Jiao, Hyung Jin Chang · 28 de julio de 2026
Human gaze does more than point to visual targets; it serves as a subtle indicator of social intent within static images, whereas standard models typically process individuals independently, treating gaze as an i.i.d. quantity or predicting social semantics in isolation. Recent multi-person methods …
