Unified Dynamic Scanpath Predictors Outperform Individually Trained Neural Models
Fuente:
arXiv
Saved in:
| Main Authors: | Abawi, Fares, Fu, Di, Wermter, Stefan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Balancing long- and short-term dynamics for the modeling of saliency in videos
by: Wulff, Theodor, et al.
Published: (2025)
by: Wulff, Theodor, et al.
Published: (2025)
Modeling Human Gaze Behavior with Diffusion Models for Unified Scanpath Prediction
by: Cartella, Giuseppe, et al.
Published: (2025)
by: Cartella, Giuseppe, et al.
Published: (2025)
Unifying Top-down and Bottom-up Scanpath Prediction Using Transformers
by: Yang, Zhibo, et al.
Published: (2023)
by: Yang, Zhibo, et al.
Published: (2023)
Unconstrained Open Vocabulary Image Classification: Zero-Shot Transfer from Text to Image via CLIP Inversion
by: Allgeuer, Philipp, et al.
Published: (2024)
by: Allgeuer, Philipp, et al.
Published: (2024)
From Neural Activations to Concepts: A Survey on Explaining Concepts in Neural Networks
by: Lee, Jae Hee, et al.
Published: (2023)
by: Lee, Jae Hee, et al.
Published: (2023)
A Robotics-Inspired Scanpath Model Reveals the Importance of Uncertainty and Semantic Object Cues for Gaze Guidance in Dynamic Scenes
by: Mengers, Vito, et al.
Published: (2024)
by: Mengers, Vito, et al.
Published: (2024)
Debiasing Central Fixation Confounds Reveals a Peripheral "Sweet Spot" for Human-like Scanpaths in Hard-Attention Vision
by: Pan, Pengcheng, et al.
Published: (2026)
by: Pan, Pengcheng, et al.
Published: (2026)
Sparsity Outperforms Low-Rank Projections in Few-Shot Adaptation
by: Mrabah, Nairouz, et al.
Published: (2025)
by: Mrabah, Nairouz, et al.
Published: (2025)
Simple Agents Outperform Experts in Biomedical Imaging Workflow Optimization
by: Xuefei, et al.
Published: (2025)
by: Xuefei, et al.
Published: (2025)
Beyond Average: Individualized Visual Scanpath Prediction
by: Chen, Xianyu, et al.
Published: (2024)
by: Chen, Xianyu, et al.
Published: (2024)
Snapture -- A Novel Neural Architecture for Combined Static and Dynamic Hand Gesture Recognition
by: Ali, Hassan, et al.
Published: (2022)
by: Ali, Hassan, et al.
Published: (2022)
Chipmunk: Training-Free Acceleration of Diffusion Transformers with Dynamic Column-Sparse Deltas
by: Silveria, Austin, et al.
Published: (2025)
by: Silveria, Austin, et al.
Published: (2025)
EyeFormer: Predicting Personalized Scanpaths with Transformer-Guided Reinforcement Learning
by: Jiang, Yue, et al.
Published: (2024)
by: Jiang, Yue, et al.
Published: (2024)
When CNNs Outperform Transformers and Mambas: Revisiting Deep Architectures for Dental Caries Segmentation
by: Ghimire, Aashish, et al.
Published: (2025)
by: Ghimire, Aashish, et al.
Published: (2025)
Lance: Unified Multimodal Modeling by Multi-Task Synergy
by: Fu, Fengyi, et al.
Published: (2026)
by: Fu, Fengyi, et al.
Published: (2026)
Grids Often Outperform Implicit Neural Representation at Compressing Dense Signals
by: Kim, Namhoon, et al.
Published: (2025)
by: Kim, Namhoon, et al.
Published: (2025)
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision
by: Liu, Zeyu, et al.
Published: (2026)
by: Liu, Zeyu, et al.
Published: (2026)
Unified Text-Image Generation with Weakness-Targeted Post-Training
by: Chen, Jiahui, et al.
Published: (2026)
by: Chen, Jiahui, et al.
Published: (2026)
BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset
by: Chen, Jiuhai, et al.
Published: (2025)
by: Chen, Jiuhai, et al.
Published: (2025)
GenDDS: Generating Diverse Driving Video Scenarios with Prompt-to-Video Generative Model
by: Fu, Yongjie, et al.
Published: (2024)
by: Fu, Yongjie, et al.
Published: (2024)
SOMA: Unifying Parametric Human Body Models
by: Saito, Jun, et al.
Published: (2026)
by: Saito, Jun, et al.
Published: (2026)
GPT-NAS: Evolutionary Neural Architecture Search with the Generative Pre-Trained Model
by: Yu, Caiyang, et al.
Published: (2023)
by: Yu, Caiyang, et al.
Published: (2023)
UniPTS: A Unified Framework for Proficient Post-Training Sparsity
by: Xie, Jingjing, et al.
Published: (2024)
by: Xie, Jingjing, et al.
Published: (2024)
Efficient Quantization-Aware Training on Segment Anything Model in Medical Images and Its Deployment
by: Lu, Haisheng, et al.
Published: (2024)
by: Lu, Haisheng, et al.
Published: (2024)
Flexiffusion: Training-Free Segment-Wise Neural Architecture Search for Efficient Diffusion Models
by: Huang, Hongtao, et al.
Published: (2025)
by: Huang, Hongtao, et al.
Published: (2025)
Keypoints as Dynamic Centroids for Unified Human Pose and Segmentation
by: Ahmad, Niaz, et al.
Published: (2025)
by: Ahmad, Niaz, et al.
Published: (2025)
PriorProbe: Recovering Individual-Level Priors for Personalizing Neural Networks in Facial Expression Recognition
by: Yan, Haijiang, et al.
Published: (2026)
by: Yan, Haijiang, et al.
Published: (2026)
Synthesizing Efficient Data with Diffusion Models for Person Re-Identification Pre-Training
by: Niu, Ke, et al.
Published: (2024)
by: Niu, Ke, et al.
Published: (2024)
DUET-VLM: Dual stage Unified Efficient Token reduction for VLM Training and Inference
by: Singh, Aditya Kumar, et al.
Published: (2026)
by: Singh, Aditya Kumar, et al.
Published: (2026)
Loss Functions for Predictor-based Neural Architecture Search
by: Ji, Han, et al.
Published: (2025)
by: Ji, Han, et al.
Published: (2025)
Trusted Unified Feature-Neighborhood Dynamics for Multi-View Classification
by: Huang, Haojian, et al.
Published: (2024)
by: Huang, Haojian, et al.
Published: (2024)
Unified and Dynamic Graph for Temporal Character Grouping in Long Videos
by: Shu, Xiujun, et al.
Published: (2023)
by: Shu, Xiujun, et al.
Published: (2023)
Training Convolutional Neural Networks with the Forward-Forward algorithm
by: Scodellaro, Riccardo, et al.
Published: (2023)
by: Scodellaro, Riccardo, et al.
Published: (2023)
Luminark: Training-free, Probabilistically-Certified Watermarking for General Vision Generative Models
by: Xu, Jiayi, et al.
Published: (2026)
by: Xu, Jiayi, et al.
Published: (2026)
Med-UniC: Unifying Cross-Lingual Medical Vision-Language Pre-Training by Diminishing Bias
by: Wan, Zhongwei, et al.
Published: (2023)
by: Wan, Zhongwei, et al.
Published: (2023)
Bootstrapping Action-Grounded Visual Dynamics in Unified Vision-Language Models
by: Qiu, Yifu, et al.
Published: (2025)
by: Qiu, Yifu, et al.
Published: (2025)
MRI Embeddings Complement Clinical Predictors for Cognitive Decline Modeling in Alzheimer's Disease Cohorts
by: Putera, Nathaniel, et al.
Published: (2025)
by: Putera, Nathaniel, et al.
Published: (2025)
Flow-GRPO: Training Flow Matching Models via Online RL
by: Liu, Jie, et al.
Published: (2025)
by: Liu, Jie, et al.
Published: (2025)
Geodesics with Unified Tangent-constrained Priors and Curvature Regularization
by: Di, Chong, et al.
Published: (2026)
by: Di, Chong, et al.
Published: (2026)
Dynamic Sparse Training versus Dense Training: The Unexpected Winner in Image Corruption Robustness
by: Wu, Boqian, et al.
Published: (2024)
by: Wu, Boqian, et al.
Published: (2024)
Similar Items
-
Balancing long- and short-term dynamics for the modeling of saliency in videos
by: Wulff, Theodor, et al.
Published: (2025) -
Modeling Human Gaze Behavior with Diffusion Models for Unified Scanpath Prediction
by: Cartella, Giuseppe, et al.
Published: (2025) -
Unifying Top-down and Bottom-up Scanpath Prediction Using Transformers
by: Yang, Zhibo, et al.
Published: (2023) -
Unconstrained Open Vocabulary Image Classification: Zero-Shot Transfer from Text to Image via CLIP Inversion
by: Allgeuer, Philipp, et al.
Published: (2024) -
From Neural Activations to Concepts: A Survey on Explaining Concepts in Neural Networks
by: Lee, Jae Hee, et al.
Published: (2023)