arXiv Science⌕ Search

arXiv subjects

Mehedi Hasan Raju

Publications and source records attributed to Mehedi Hasan Raju.

9 recordsLinked to original sources

Privacy-Preserving Gaze Interaction: Reducing Re-Identification Without Degrading Utility

Gaze-based interaction is emerging as a standard input modality on consumer extended-reality (XR) devices. Yet each gaze input constitutes an involuntary biometric disclosure, as the same signal that selects a button can also identify the user who produced it. Reducing identity information without degrading interaction is difficult: most prior methods are validated on small datasets or optimized for only one of the two competing objectives. We introduce a dual-assessment framework that scores any real-time gaze transformation on two axes at once: interaction utility, measured through an offline gaze-interaction simulation and an operational spatial-accuracy metric, and privacy preservation, measured through the Rank-1 Identification Rate (IR) of a state-of-the-art re-identification adversary. We test the framework on 23 conditions: one raw baseline and 22 privacy variants from eight lightweight signal-processing families using two publicly available datasets (GazeBase and GazeBaseVR). Deterministic operators produce no significant change in identity on either dataset, because a subject-invariant mapping preserves the relative geometry a biometric embedding exploits. Identity drops sharply only when identity-uncorrelated randomness is injected per sample. Smoothing applied afterwards to recover signal quality partially restores identifiability on both datasets. The best-balanced configuration reduces the Rank-1 IR by 64.3 percentage points on GazeBase and 67.9 points on GazeBaseVR, while leaving target-selection success essentially unchanged on both datasets. Privacy-preserving gaze interaction is therefore not zero-sum. However, the reduction in identifiability is achieved by injecting randomness rather than by improving signal fidelity.

cs.HC↗

EyeMakeYou: Identity-, Task-, and Subjective-State-Conditioned Diffusion for High-Frequency Gaze Synthesis

Eye movement biometrics (EMB) is an emerging behavioral modality for user authentication, particularly in virtual- and augmented-reality systems, where gaze dynamics contain distinctive subject-specific features. However, robust EMB systems require diverse, high-quality gaze recordings that are expensive to collect and often unavailable at the scale needed for model development. Generative models can mitigate data scarcity, but existing methods either synthesize generic gaze behavior or personalize signals primarily by identity, without jointly representing the user's task and subjective state. Consequently, generated signals may appear visually realistic while failing to retain the behavioral properties required for biometric applications. To address this limitation, we propose EyeMakeYou, a multi-conditional denoising diffusion framework for subject-specific, high-frequency gaze synthesis. EyeMakeYou generates 5-s, 1000-Hz bivariate gaze-velocity sequences from an identity-removed reference trajectory and conditions the denoising process on an identity embedding, a task embedding, and self-reported ratings of overall difficulty, mental tiredness, and eye tiredness. Its objective combines diffusion noise prediction and identity preservation with multi-resolution spectral, drift-consistency, and event-weighted local-smoothness losses. Experiments on GazeBase show that EyeMakeYou achieves higher median spatial accuracy and greater real--synthetic similarity in the embedding feature space than the existing generative approaches, while retaining selected task-dependent associations between subjective reports and oculomotor features. These findings support conditional diffusion as a practical approach for augmenting gaze datasets for biometric and interactive applications.

cs.HC↗

Gaze Authentication: Factors Influencing Authentication Performance

This paper examines the key factors that influence the performance of state-of-the-art gaze-based authentication. Experiments were conducted on a large-scale, in-house dataset comprising 8,849 subjects collected with Meta Quest Pro equivalent hardware running a video oculography-driven gaze estimation pipeline at 72~Hz. State of the neural network architecture was employed to study the influence of the following factors on authentication performance: eye tracking signal quality, various aspects of eye tracking calibration, and simple filtering on estimated raw gaze. This report provides performance results and their analysis.

cs.CV↗

Quantitative and Qualitative Comparison of Generative Models for Subject-Specific Gaze Synthesis: Diffusion vs GANs

Gaze-based biometrics has emerged as a promising approach for user authentication, but advances in this area are constrained by the limited availability of high-quality, subject-specific gaze recordings. Recent generative models have shown promise for synthesizing gaze data, yet most existing approaches rely on random noise distributions or global, predefined latent embeddings and do not explicitly model subject-specific gaze characteristics. To address this limitation, we revisit two recent generative models, diffusion and generative adversarial networks (GANs), and modify both to support subject-aware gaze synthesis. For the diffusion-based approach, we incorporate compact user embeddings to capture subject-level gaze traits. For the GAN-based approach, we introduce a subject-specific conditioning module that guides the generator to preserve idiosyncratic gaze patterns. Later, we evaluate both approaches using standard eye-movement signal quality metrics, including spatial accuracy and precision, and assess whether the generated sequences retain identity-related features relevant to biometric applications. Experimental results show that the diffusion-based approach produces more realistic, identity-preserving gaze sequences than the GAN-based approach. Overall, this work advances the understanding of synthetic gaze quality, realism, and subject specificity and supports the development of gaze-based biometric applications.

cs.HC↗

Enhancing Eye Movement Biometrics for User Authentication via Continuous Gaze Offset Score Fusion

Eye movement biometrics (EMB) use subject-specific gaze dynamics for user authentication and identification. Recent deep learning-based EMB systems achieve strong performance by modeling temporal eye movement behavior. However, these systems typically overlook continuous gaze offset, despite prior evidence that it contains user-discriminative information. This work examines whether continuous gaze offset can improve biometric performance when combined with existing biometric features. We evaluate linear and nonlinear fusion methods on two publicly available datasets, collected via the lab-grade eye tracker and virtual reality headset across multiple tasks and observation durations. Results indicate that fusion offers performance benefits on both datasets, particularly when using nonlinear fusion. Additionally, fusing biometric information across multiple tasks further improves authentication performance. These findings support the hypothesis that continuous gaze offset may serve as useful auxiliary information under conditions of degraded or noisy eye tracking.

cs.HC↗

Ocular Authentication: Fusion of Gaze and Periocular Modalities

This paper investigates the feasibility of fusing two eye-centric authentication modalities-eye movements and periocular images-within a calibration-free authentication system. While each modality has independently shown promise for user authentication, their combination within a unified gaze-estimation pipeline has not been thoroughly explored at scale. In this report, we propose a multimodal authentication system and evaluate it using a large-scale in-house dataset comprising 9202 subjects with an eye tracking (ET) signal quality equivalent to a consumer-facing virtual reality (VR) device. Our results show that the multimodal approach consistently outperforms both unimodal systems across all scenarios, surpassing the FIDO benchmark. The integration of a state-of-the-art machine learning architecture contributed significantly to the overall authentication performance at scale, driven by the model's ability to capture authentication representations and the complementary discriminative characteristics of the fused modalities.

cs.CV↗

Evaluating Eye Tracking Signal Quality with Real-time Gaze Interaction Simulation

We present a real-time gaze-based interaction simulation methodology using an offline dataset to evaluate the eye-tracking signal quality. This study employs three fundamental eye-movement classification algorithms to identify physiological fixations from the eye-tracking data. We introduce the Rank-1 fixation selection approach to identify the most stable fixation period nearest to a target, referred to as the trigger-event. Our evaluation explores how varying constraints impact the definition of trigger-events and evaluates the eye-tracking signal quality of defined trigger-events. Results show that while the dispersion threshold-based algorithm identifies trigger-events more accurately, the Kalman filter-based classification algorithm performs better in eye-tracking signal quality, as demonstrated through a user-centric quality assessment using user- and error-percentile tiers. Despite median user-level performance showing minor differences across algorithms, significant variability in signal quality across participants highlights the importance of algorithm selection to ensure system reliability.

cs.HC↗

Temporal Persistence and Intercorrelation of Embeddings Learned by an End-to-End Deep Learning Eye Movement-driven Biometrics Pipeline

What qualities make a feature useful for biometric performance? In prior research, pre-dating the advent of deep learning (DL) approaches to biometric analysis, a strong relationship between temporal persistence, as indexed by the intraclass correlation coefficient (ICC), and biometric performance (Equal Error Rate, EER) was noted. More generally, the claim was made that good biometric performance resulted from a relatively large set of weakly intercorrelated features with high ICC. The present study aimed to determine whether the same relationships are found in a state-of-the-art DL-based eye movement biometric system (``Eye-Know-You-Too''), as applied to two publicly available eye movement datasets. To this end, we manipulate various aspects of eye-tracking signal quality, which produces variation in biometric performance, and relate that performance to the temporal persistence and intercorrelation of the resulting embeddings. Data quality indices were related to EER with either linear or logarithmic fits, and the resulting model R^2 was noted. As a general matter, we found that temporal persistence was an important predictor of DL-based biometric performance, and also that DL-learned embeddings were generally weakly intercorrelated.

cs.CV↗

Evaluating Eye Movement Biometrics in Virtual Reality: A Comparative Analysis of VR Headset and High-End Eye-Tracker Collected Dataset

Previous studies have shown that eye movement data recorded at 1000 Hz can be used to authenticate individuals. This study explores the effectiveness of eye movement-based biometrics (EMB) by utilizing data from an eye-tracking (ET)-enabled virtual reality (VR) headset (GazeBaseVR) and compares it to the performance using data from a high-end eye tracker (GazeBase) that has been downsampled to 250 Hz. The research also aims to assess the biometric potential of both binocular and monocular eye movement data. GazeBaseVR dataset achieves an equal error rate (EER) of 1.67% and a false rejection rate (FRR) at 10^-4 false acceptance rate (FAR) of 22.73% in a binocular configuration. This study underscores the biometric viability of data obtained from eye-tracking-enabled VR headset.

cs.HC↗