Tokenization of Gaze Data
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Rolff, Tim, Karimian, Jurik, Hypki, Niklas, Schmidt, Susanne, Lappe, Markus, Steinicke, Frank |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
A Toolkit for Virtual Reality Data Collection
par: Rolff, Tim, et autres
Publié: (2024)
par: Rolff, Tim, et autres
Publié: (2024)
A Hands-free Spatial Selection and Interaction Technique using Gaze and Blink Input with Blink Prediction for Extended Reality
par: Rolff, Tim, et autres
Publié: (2025)
par: Rolff, Tim, et autres
Publié: (2025)
A Review of Driver Gaze Estimation and Application in Gaze Behavior Understanding
par: Sharma, Pavan Kumar, et autres
Publié: (2023)
par: Sharma, Pavan Kumar, et autres
Publié: (2023)
Gaze-Guided Graph Neural Network for Action Anticipation Conditioned on Intention
par: Ozdel, Suleyman, et autres
Publié: (2024)
par: Ozdel, Suleyman, et autres
Publié: (2024)
A Transformer-Based Model for the Prediction of Human Gaze Behavior on Videos
par: Ozdel, Suleyman, et autres
Publié: (2024)
par: Ozdel, Suleyman, et autres
Publié: (2024)
Gaze4HRI: Zero-shot Benchmarking Gaze Estimation Neural-Networks for Human-Robot Interaction
par: Sezer, Berk, et autres
Publié: (2026)
par: Sezer, Berk, et autres
Publié: (2026)
Unsupervised visualization of image datasets using contrastive learning
par: Böhm, Jan Niklas, et autres
Publié: (2022)
par: Böhm, Jan Niklas, et autres
Publié: (2022)
Chronotome: Real-Time Topic Modeling for Streaming Embedding Spaces
par: Lim, Matte, et autres
Publié: (2025)
par: Lim, Matte, et autres
Publié: (2025)
Voting-based Multimodal Automatic Deception Detection
par: Touma, Lana, et autres
Publié: (2023)
par: Touma, Lana, et autres
Publié: (2023)
Android in the Zoo: Chain-of-Action-Thought for GUI Agents
par: Zhang, Jiwen, et autres
Publié: (2024)
par: Zhang, Jiwen, et autres
Publié: (2024)
Generalization in Online Reinforcement Learning for Mobile Agents
par: Gu, Li, et autres
Publié: (2026)
par: Gu, Li, et autres
Publié: (2026)
Magic NeRF Lens: Interactive Fusion of Neural Radiance Fields for Virtual Facility Inspection
par: Li, Ke, et autres
Publié: (2023)
par: Li, Ke, et autres
Publié: (2023)
Realtime Dynamic Gaze Target Tracking and Depth-Level Estimation
par: Seraj, Esmaeil, et autres
Publié: (2024)
par: Seraj, Esmaeil, et autres
Publié: (2024)
GazeTrack: High-Precision Eye Tracking Based on Regularization and Spatial Computing
par: Yang, Xiaoyin
Publié: (2025)
par: Yang, Xiaoyin
Publié: (2025)
No Need to Sacrifice Data Quality for Quantity: Crowd-Informed Machine Annotation for Cost-Effective Understanding of Visual Data
par: Klugmann, Christopher, et autres
Publié: (2024)
par: Klugmann, Christopher, et autres
Publié: (2024)
Signformer is all you need: Towards Edge AI for Sign Language
par: Yang, Eta
Publié: (2024)
par: Yang, Eta
Publié: (2024)
Privacy Enhancement for Gaze Data Using a Noise-Infused Autoencoder
par: Aziz, Samantha, et autres
Publié: (2025)
par: Aziz, Samantha, et autres
Publié: (2025)
Minority Reports: Balancing Cost and Quality in Ground Truth Data Annotation
par: Liao, Hsuan Wei, et autres
Publié: (2025)
par: Liao, Hsuan Wei, et autres
Publié: (2025)
A Rigorous Behavior Assessment of CNNs Using a Data-Domain Sampling Regime
par: Jiang, Shuning, et autres
Publié: (2025)
par: Jiang, Shuning, et autres
Publié: (2025)
GazeTarget360: Towards Gaze Target Estimation in 360-Degree for Robot Perception
par: Dai, Zhuangzhuang, et autres
Publié: (2025)
par: Dai, Zhuangzhuang, et autres
Publié: (2025)
DiffGaze: A Diffusion Model for Continuous Gaze Sequence Generation on 360° Images
par: Jiao, Chuhan, et autres
Publié: (2024)
par: Jiao, Chuhan, et autres
Publié: (2024)
GazeGPT: Augmenting Human Capabilities using Gaze-contingent Contextual AI for Smart Eyewear
par: Konrad, Robert, et autres
Publié: (2024)
par: Konrad, Robert, et autres
Publié: (2024)
When Incentives Backfire, Data Stops Being Human
par: Santy, Sebastin, et autres
Publié: (2025)
par: Santy, Sebastin, et autres
Publié: (2025)
How Do LLMs Acquire New Knowledge? A Knowledge Circuits Perspective on Continual Pre-Training
par: Ou, Yixin, et autres
Publié: (2025)
par: Ou, Yixin, et autres
Publié: (2025)
UISim: An Interactive Image-Based UI Simulator for Dynamic Mobile Environments
par: Xiang, Jiannan, et autres
Publié: (2025)
par: Xiang, Jiannan, et autres
Publié: (2025)
AIN: The Arabic INclusive Large Multimodal Model
par: Heakl, Ahmed, et autres
Publié: (2025)
par: Heakl, Ahmed, et autres
Publié: (2025)
EasyEdit2: An Easy-to-use Steering Framework for Editing Large Language Models
par: Xu, Ziwen, et autres
Publié: (2025)
par: Xu, Ziwen, et autres
Publié: (2025)
GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding
par: Tang, Fei, et autres
Publié: (2025)
par: Tang, Fei, et autres
Publié: (2025)
Optimizing Small Language Models for In-Vehicle Function-Calling
par: Khiabani, Yahya Sowti, et autres
Publié: (2025)
par: Khiabani, Yahya Sowti, et autres
Publié: (2025)
KITAB-Bench: A Comprehensive Multi-Domain Benchmark for Arabic OCR and Document Understanding
par: Heakl, Ahmed, et autres
Publié: (2025)
par: Heakl, Ahmed, et autres
Publié: (2025)
Developing a Dyslexia Indicator Using Eye Tracking
par: Cogan, Kevin, et autres
Publié: (2025)
par: Cogan, Kevin, et autres
Publié: (2025)
ChainReaction: Causal Chain-Guided Reasoning for Modular and Explainable Causal-Why Video Question Answering
par: Parmar, Paritosh, et autres
Publié: (2025)
par: Parmar, Paritosh, et autres
Publié: (2025)
NeuralOS: Towards Simulating Operating Systems via Neural Generative Models
par: Rivard, Luke, et autres
Publié: (2025)
par: Rivard, Luke, et autres
Publié: (2025)
GUI-AIMA: Aligning Intrinsic Multimodal Attention with a Context Anchor for GUI Grounding
par: Zhou, Shijie, et autres
Publié: (2025)
par: Zhou, Shijie, et autres
Publié: (2025)
ReLearn: Unlearning via Learning for Large Language Models
par: Xu, Haoming, et autres
Publié: (2025)
par: Xu, Haoming, et autres
Publié: (2025)
Human-like object concept representations emerge naturally in multimodal large language models
par: Du, Changde, et autres
Publié: (2024)
par: Du, Changde, et autres
Publié: (2024)
Gemini Goes to Med School: Exploring the Capabilities of Multimodal Large Language Models on Medical Challenge Problems & Hallucinations
par: Pal, Ankit, et autres
Publié: (2024)
par: Pal, Ankit, et autres
Publié: (2024)
Detoxifying Large Language Models via Knowledge Editing
par: Wang, Mengru, et autres
Publié: (2024)
par: Wang, Mengru, et autres
Publié: (2024)
Knowledge Mechanisms in Large Language Models: A Survey and Perspective
par: Wang, Mengru, et autres
Publié: (2024)
par: Wang, Mengru, et autres
Publié: (2024)
DOTA: Distributional Test-Time Adaptation of Vision-Language Models
par: Han, Zongbo, et autres
Publié: (2024)
par: Han, Zongbo, et autres
Publié: (2024)
Documents similaires
-
A Toolkit for Virtual Reality Data Collection
par: Rolff, Tim, et autres
Publié: (2024) -
A Hands-free Spatial Selection and Interaction Technique using Gaze and Blink Input with Blink Prediction for Extended Reality
par: Rolff, Tim, et autres
Publié: (2025) -
A Review of Driver Gaze Estimation and Application in Gaze Behavior Understanding
par: Sharma, Pavan Kumar, et autres
Publié: (2023) -
Gaze-Guided Graph Neural Network for Action Anticipation Conditioned on Intention
par: Ozdel, Suleyman, et autres
Publié: (2024) -
A Transformer-Based Model for the Prediction of Human Gaze Behavior on Videos
par: Ozdel, Suleyman, et autres
Publié: (2024)