arXiv:2411.01969cs.CVcs.AI2024-11被引 1

婴儿主动看物的视线行为帮助无监督学习物体不变特征。

Toddlers' Active Gaze Behavior Supports Self-Supervised Object Learning

  • 用眼动追踪记录婴儿注视点,构建视觉流输入无监督模型。
  • 婴儿的注视策略能促进形成视角不变的物体表征。
  • 高清晰度中央视野范围小是关键,适合做自监督学习。

婴儿在极少监督下即可从不同视角识别物体。在此过程中,他们频繁进行眼动与头部运动,塑造了视觉体验。当前尚不清楚这些行为如何影响其物体识别能力的发展。本研究结合头戴式眼动追踪与双人互动游戏,通过眼动估计实时定位注视点,裁剪头戴摄像头图像中的中心视觉区域,构建视觉流输入无监督计算模型,该模型生成随时间缓慢变化的视觉表征。实验表明,婴儿的注视策略有助于学习视角不变的物体表征;分析还揭示,高视觉敏锐度的有限中央视野范围对这一过程至关重要。研究揭示了婴儿注视行为如何支持其视角不变物体识别能力的发展。

原文摘要 · Abstract (English)

Toddlers learn to recognize objects from different viewpoints with almost no supervision. During this learning, they execute frequent eye and head movements that shape their visual experience. It is presently unclear if and how these behaviors contribute to toddlers' emerging object recognition abilities. To answer this question, we here combine head-mounted eye tracking during dyadic play with unsupervised machine learning. We approximate toddlers' central visual field experience by cropping image regions from a head-mounted camera centered on the current gaze location estimated via eye tracking. This visual stream feeds an unsupervised computational model of toddlers' learning, which constructs visual representations that slowly change over time. Our experiments demonstrate that toddlers' gaze strategy supports the learning of invariant object representations. Our analysis also shows that the limited size of the central visual field where acuity is high is crucial for this. Overall, our work reveals how toddlers' gaze behavior may support their development of view-invariant object recognition.

自监督学习婴儿认知视觉表征

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。