Pith. sign in

REVIEW 3 cited by

L2CS-Net: Fine-Grained Gaze Estimation in Unconstrained Environments

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2203.03339 v1 pith:KKEPDOTE submitted 2022-03-07 cs.CV cs.LGcs.RO

classification cs.CVcs.LGcs.RO
keywords gazemodelunconstrainedaccuracyangledatasetsimprovel2cs-net
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Human gaze is a crucial cue used in various applications such as human-robot interaction and virtual reality. Recently, convolution neural network (CNN) approaches have made notable progress in predicting gaze direction. However, estimating gaze in-the-wild is still a challenging problem due to the uniqueness of eye appearance, lightning conditions, and the diversity of head pose and gaze directions. In this paper, we propose a robust CNN-based model for predicting gaze in unconstrained settings. We propose to regress each gaze angle separately to improve the per-angel prediction accuracy, which will enhance the overall gaze performance. In addition, we use two identical losses, one for each angle, to improve network learning and increase its generalization. We evaluate our model with two popular datasets collected with unconstrained settings. Our proposed model achieves state-of-the-art accuracy of 3.92{\deg} and 10.41{\deg} on MPIIGaze and Gaze360 datasets, respectively. We make our code open source at https://github.com/Ahmednull/L2CS-Net.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Camera-based implicit mind reading by capturing higher-order semantic dynamics of human gaze within environmental context

    cs.CV 2025-07 conditional novelty 6.0 of 10

    Ordered sequences of gaze fixations mapped to semantic objects, encoded by the new SIO representation and EmoGazeNet, are reported to recognize six emotions with accuracy close to EEG-based methods on self-collected data.

  2. GoHD: Gaze-oriented and Highly Disentangled Portrait Animation with Rhythmic Poses and Realistic Expression

    cs.CV 2024-12 conditional novelty 5.0 of 10

    An audio-driven portrait animation framework that adds gaze control, prosody-aware head poses, and two-stage lip versus eye motion distillation to a latent-navigation animator.

  3. Learning Nonverbal Cues in Multiparty Social Interactions for Robotic Facilitators

    cs.RO 2025-01 conditional novelty 4.0 of 10

    Using Implicit Behavior Cloning, the authors report a gaze-generation model for robotic facilitators that achieves higher target-reaching success (96% vs 93%) and smoother trajectories than an MSE behavior cloning baseline.

Pith tools