3D Prior is All You Need: Cross-Task Few-shot 2D Gaze Estimation
Fuente:
arXiv
Salvato in:
| Autori principali: | Cheng, Yihua, Wang, Hengfei, Zhang, Zhongqun, Yue, Yang, Kim, Bo Eun, Lu, Feng, Chang, Hyung Jin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
RTGaze: Real-Time 3D-Aware Gaze Redirection from a Single Image
di: Wang, Hengfei, et al.
Pubblicazione: (2025)
di: Wang, Hengfei, et al.
Pubblicazione: (2025)
TextGaze: Gaze-Controllable Face Generation with Natural Language
di: Wang, Hengfei, et al.
Pubblicazione: (2024)
di: Wang, Hengfei, et al.
Pubblicazione: (2024)
NL2Contact: Natural Language Guided 3D Hand-Object Contact Modeling with Diffusion Model
di: Zhang, Zhongqun, et al.
Pubblicazione: (2024)
di: Zhang, Zhongqun, et al.
Pubblicazione: (2024)
Multi-Modal Gaze Following in Conversational Scenarios
di: Hou, Yuqi, et al.
Pubblicazione: (2023)
di: Hou, Yuqi, et al.
Pubblicazione: (2023)
Force-Aware 3D Contact Modeling for Stable Grasp Generation
di: Chen, Zhuo, et al.
Pubblicazione: (2025)
di: Chen, Zhuo, et al.
Pubblicazione: (2025)
What Do You See in Vehicle? Comprehensive Vision Solution for In-Vehicle Gaze Estimation
di: Cheng, Yihua, et al.
Pubblicazione: (2024)
di: Cheng, Yihua, et al.
Pubblicazione: (2024)
VL4Gaze: Unleashing Vision-Language Models for Gaze Following
di: Wang, Shijing, et al.
Pubblicazione: (2025)
di: Wang, Shijing, et al.
Pubblicazione: (2025)
Search is All You Need for Few-shot Anomaly Detection
di: Wang, Qishan, et al.
Pubblicazione: (2025)
di: Wang, Qishan, et al.
Pubblicazione: (2025)
Enhancing Gaze Reasoning in Vision Foundation Models for Gaze Following
di: Wang, Shijing, et al.
Pubblicazione: (2026)
di: Wang, Shijing, et al.
Pubblicazione: (2026)
Appearance-based Gaze Estimation With Deep Learning: A Review and Benchmark
di: Cheng, Yihua, et al.
Pubblicazione: (2021)
di: Cheng, Yihua, et al.
Pubblicazione: (2021)
Attn-Adapter: Attention Is All You Need for Online Few-shot Learner of Vision-Language Model
di: Bui, Phuoc-Nguyen, et al.
Pubblicazione: (2025)
di: Bui, Phuoc-Nguyen, et al.
Pubblicazione: (2025)
Predictable Emergent Abilities of LLMs: Proxy Tasks Are All You Need
di: Zhang, Bo-Wen, et al.
Pubblicazione: (2024)
di: Zhang, Bo-Wen, et al.
Pubblicazione: (2024)
All You Need Is Synthetic Task Augmentation
di: Godin, Guillaume
Pubblicazione: (2025)
di: Godin, Guillaume
Pubblicazione: (2025)
Eyes on VLM: Benchmarking Gaze Following and Social Gaze Prediction in Vision Language Models
di: Wang, Hengfei, et al.
Pubblicazione: (2026)
di: Wang, Hengfei, et al.
Pubblicazione: (2026)
Roll Your Eyes: Gaze Redirection via Explicit 3D Eyeball Rotation
di: Choi, YoungChan, et al.
Pubblicazione: (2025)
di: Choi, YoungChan, et al.
Pubblicazione: (2025)
CrossGaze: A Strong Method for 3D Gaze Estimation in the Wild
di: Cătrună, Andy, et al.
Pubblicazione: (2024)
di: Cătrună, Andy, et al.
Pubblicazione: (2024)
EV-CLIP: Efficient Visual Prompt Adaptation for CLIP in Few-shot Action Recognition under Visual Challenges
di: Jon, Hyo Jin, et al.
Pubblicazione: (2026)
di: Jon, Hyo Jin, et al.
Pubblicazione: (2026)
SynthDST: Synthetic Data is All You Need for Few-Shot Dialog State Tracking
di: Kulkarni, Atharva, et al.
Pubblicazione: (2024)
di: Kulkarni, Atharva, et al.
Pubblicazione: (2024)
Perception Is All You Need: A Neuroscience Framework for Low Cost Sensorless Gaze in HRI
di: Kadem, Mason
Pubblicazione: (2026)
di: Kadem, Mason
Pubblicazione: (2026)
Element-wise Attention Is All You Need
di: Feng, Guoxin
Pubblicazione: (2025)
di: Feng, Guoxin
Pubblicazione: (2025)
Similarity Memory Prior is All You Need for Medical Image Segmentation
di: Tang, Hao, et al.
Pubblicazione: (2025)
di: Tang, Hao, et al.
Pubblicazione: (2025)
Cross-Validation Is All You Need: A Statistical Approach To Label Noise Estimation
di: Chen, Jianan, et al.
Pubblicazione: (2023)
di: Chen, Jianan, et al.
Pubblicazione: (2023)
Improve Meta-learning for Few-Shot Text Classification with All You Can Acquire from the Tasks
di: Liu, Xinyue, et al.
Pubblicazione: (2024)
di: Liu, Xinyue, et al.
Pubblicazione: (2024)
Block Rotation is All You Need for MXFP4 Quantization
di: Shao, Yuantian, et al.
Pubblicazione: (2025)
di: Shao, Yuantian, et al.
Pubblicazione: (2025)
TAVP: Task-Adaptive Visual Prompt for Cross-domain Few-shot Segmentation
di: Yang, Jiaqi, et al.
Pubblicazione: (2024)
di: Yang, Jiaqi, et al.
Pubblicazione: (2024)
Unlearnable 3D Point Clouds: Class-wise Transformation Is All You Need
di: Wang, Xianlong, et al.
Pubblicazione: (2024)
di: Wang, Xianlong, et al.
Pubblicazione: (2024)
Synthetic Data RL: Task Definition Is All You Need
di: Guo, Yiduo, et al.
Pubblicazione: (2025)
di: Guo, Yiduo, et al.
Pubblicazione: (2025)
Few-shot Personalized Saliency Prediction Based on Interpersonal Gaze Patterns
di: Moroto, Yuya, et al.
Pubblicazione: (2023)
di: Moroto, Yuya, et al.
Pubblicazione: (2023)
Are Foundation Models All You Need for Zero-shot Face Presentation Attack Detection?
di: Gonzalez-Sole, Lazaro Janier, et al.
Pubblicazione: (2025)
di: Gonzalez-Sole, Lazaro Janier, et al.
Pubblicazione: (2025)
GazeD: Context-Aware Diffusion for Accurate 3D Gaze Estimation
di: Catalini, Riccardo, et al.
Pubblicazione: (2026)
di: Catalini, Riccardo, et al.
Pubblicazione: (2026)
Attention is All You Want: Machinic Gaze and the Anthropocene
di: Magee, Liam, et al.
Pubblicazione: (2024)
di: Magee, Liam, et al.
Pubblicazione: (2024)
Image is All You Need to Empower Large-scale Diffusion Models for In-Domain Generation
di: Cao, Pu, et al.
Pubblicazione: (2023)
di: Cao, Pu, et al.
Pubblicazione: (2023)
Confidence Is All You Need: Few-Shot RL Fine-Tuning of Language Models
di: Li, Pengyi, et al.
Pubblicazione: (2025)
di: Li, Pengyi, et al.
Pubblicazione: (2025)
Is Diversity All You Need for Scalable Robotic Manipulation?
di: Shi, Modi, et al.
Pubblicazione: (2025)
di: Shi, Modi, et al.
Pubblicazione: (2025)
Synthesis4AD: Synthetic Anomalies are All You Need for 3D Anomaly Detection
di: Sun, Yihan, et al.
Pubblicazione: (2026)
di: Sun, Yihan, et al.
Pubblicazione: (2026)
Pricing is All You Need to Improve Traffic Routing
di: Tang, Yu, et al.
Pubblicazione: (2025)
di: Tang, Yu, et al.
Pubblicazione: (2025)
Attention is All You Need Until You Need Retention
di: Yaslioglu, M. Murat
Pubblicazione: (2025)
di: Yaslioglu, M. Murat
Pubblicazione: (2025)
MOKD: Cross-domain Finetuning for Few-shot Classification via Maximizing Optimized Kernel Dependence
di: Tian, Hongduan, et al.
Pubblicazione: (2024)
di: Tian, Hongduan, et al.
Pubblicazione: (2024)
3D-U-SAM Network For Few-shot Tooth Segmentation in CBCT Images
di: Zhang, Yifu, et al.
Pubblicazione: (2023)
di: Zhang, Yifu, et al.
Pubblicazione: (2023)
Don't Do RAG: When Cache-Augmented Generation is All You Need for Knowledge Tasks
di: Chan, Brian J, et al.
Pubblicazione: (2024)
di: Chan, Brian J, et al.
Pubblicazione: (2024)
Documenti analoghi
-
RTGaze: Real-Time 3D-Aware Gaze Redirection from a Single Image
di: Wang, Hengfei, et al.
Pubblicazione: (2025) -
TextGaze: Gaze-Controllable Face Generation with Natural Language
di: Wang, Hengfei, et al.
Pubblicazione: (2024) -
NL2Contact: Natural Language Guided 3D Hand-Object Contact Modeling with Diffusion Model
di: Zhang, Zhongqun, et al.
Pubblicazione: (2024) -
Multi-Modal Gaze Following in Conversational Scenarios
di: Hou, Yuqi, et al.
Pubblicazione: (2023) -
Force-Aware 3D Contact Modeling for Stable Grasp Generation
di: Chen, Zhuo, et al.
Pubblicazione: (2025)