Google Introduces Promptable Gaze Target Estimation with Text and Visual Prompts
google · hf · 2026-08-14
Google released Gaze Target Estimation Anywhere with Concepts on Hugging Face.
This new promptable paradigm integrates subject localization and gaze estimation into an end-to-end transformer model. Without relying on multi-stage pipelines, the model uses text or visual prompts to identify subjects and predict their gaze targets directly.
More from Research
- OptiPrime in Nature Biotech: ML Model Accurately Predicts Prime Editing Efficiencies — anshulkundaje · 2026-08-14
- Nature Study Challenged: Decline in Disruptive Science Blamed on Dataset Artefacts — Robert_Palgrave · 2026-08-14
- Exploring Why Recent AI Models Are Suddenly Hacking Into Things — xuanalogue · 2026-08-14
- Deep Dive: Why Attention Mechanisms Are Hard to Replace — akbirthko · 2026-08-14
- Trainable Langton's ant borrows from NCA architecture, dubbed Mordvintsev's ant — max_romana · 2026-08-14
- WorldFM Open Source: Real-Time Multi-View Diffusion from Target Poses — tom_doerr · 2026-08-14