Saved in:
| Main Authors: | Wilson, Ethan, Shic, Frederick, Jörg, Sophie, Jain, Eakta |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2402.03188 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Gaze Behavior During a Long-Term, In-Home, Social Robot Intervention for Children with ASD
by: Ramnauth, Rebecca, et al.
Published: (2025)
by: Ramnauth, Rebecca, et al.
Published: (2025)
RadEyeVideo: Enhancing general-domain Large Vision Language Model for chest X-ray analysis with video representations of eye gaze
by: Kim, Yunsoo, et al.
Published: (2025)
by: Kim, Yunsoo, et al.
Published: (2025)
Certified vs. Empirical Adversarial Robust-ness via Hybrid Convolutions with Attention Stochasticity
by: Dhar, Joy, et al.
Published: (2026)
by: Dhar, Joy, et al.
Published: (2026)
End-to-end Video Gaze Estimation via Capturing Head-face-eye Spatial-temporal Interaction Context
by: Guan, Yiran, et al.
Published: (2023)
by: Guan, Yiran, et al.
Published: (2023)
DeepFaceLab: Integrated, flexible and extensible face-swapping framework
by: Perov, Ivan, et al.
Published: (2020)
by: Perov, Ivan, et al.
Published: (2020)
Evaluation of neural network algorithms for atmospheric turbulence mitigation
by: Jain, Tushar, et al.
Published: (2024)
by: Jain, Tushar, et al.
Published: (2024)
On mitigating stability-plasticity dilemma in CLIP-guided image morphing via geodesic distillation loss
by: Oh, Yeongtak, et al.
Published: (2024)
by: Oh, Yeongtak, et al.
Published: (2024)
Spotting tell-tale visual artifacts in face swapping videos: strengths and pitfalls of CNN detectors
by: Ziglio, Riccardo, et al.
Published: (2025)
by: Ziglio, Riccardo, et al.
Published: (2025)
Towards LLM-centric Affective Visual Customization via Efficient and Precise Emotion Manipulating
by: Luo, Jiamin, et al.
Published: (2026)
by: Luo, Jiamin, et al.
Published: (2026)
Robust face recognition based on the wing loss and the $\ell_1$ regularization
by: Yun, Yaoyao, et al.
Published: (2025)
by: Yun, Yaoyao, et al.
Published: (2025)
FakeTracer: Catching Face-swap DeepFakes via Implanting Traces in Training
by: Sun, Pu, et al.
Published: (2023)
by: Sun, Pu, et al.
Published: (2023)
Towards Privacy-preserving Photorealistic Self-avatars in Mixed Reality
by: Wilson, Ethan, et al.
Published: (2025)
by: Wilson, Ethan, et al.
Published: (2025)
Enhancing accuracy of uncertainty estimation in appearance-based gaze tracking with probabilistic evaluation and calibration
by: Zheng, Qiaojie, et al.
Published: (2025)
by: Zheng, Qiaojie, et al.
Published: (2025)
Improving saliency models' predictions of the next fixation with humans' intrinsic cost of gaze shifts
by: Kadner, Florian, et al.
Published: (2022)
by: Kadner, Florian, et al.
Published: (2022)
EgoExo-Gen: Ego-centric Video Prediction by Watching Exo-centric Videos
by: Xu, Jilan, et al.
Published: (2025)
by: Xu, Jilan, et al.
Published: (2025)
Learning to mask: Towards generalized face forgery detection
by: Fei, Jianwei, et al.
Published: (2022)
by: Fei, Jianwei, et al.
Published: (2022)
Camera-based implicit mind reading by capturing higher-order semantic dynamics of human gaze within environmental context
by: Song, Mengke, et al.
Published: (2025)
by: Song, Mengke, et al.
Published: (2025)
AdvSwap: Covert Adversarial Perturbation with High Frequency Info-swapping for Autonomous Driving Perception
by: Huang, Yuanhao, et al.
Published: (2025)
by: Huang, Yuanhao, et al.
Published: (2025)
High-frequency near-eye ground truth for event-based eye tracking
by: Simpsi, Andrea, et al.
Published: (2025)
by: Simpsi, Andrea, et al.
Published: (2025)
Caption Anything in Video: Fine-grained Object-centric Captioning via Spatiotemporal Multimodal Prompting
by: Tang, Yunlong, et al.
Published: (2025)
by: Tang, Yunlong, et al.
Published: (2025)
Data-centric Prediction Explanation via Kernelized Stein Discrepancy
by: Sarvmaili, Mahtab, et al.
Published: (2024)
by: Sarvmaili, Mahtab, et al.
Published: (2024)
Rethinking Image-to-Video Adaptation: An Object-centric Perspective
by: Qian, Rui, et al.
Published: (2024)
by: Qian, Rui, et al.
Published: (2024)
Efficient Multimodal Learning from Data-centric Perspective
by: He, Muyang, et al.
Published: (2024)
by: He, Muyang, et al.
Published: (2024)
Controllable Human-centric Keyframe Interpolation with Generative Prior
by: Guo, Zujin, et al.
Published: (2025)
by: Guo, Zujin, et al.
Published: (2025)
Ego-centric Predictive Model Conditioned on Hand Trajectories
by: Zhang, Binjie, et al.
Published: (2025)
by: Zhang, Binjie, et al.
Published: (2025)
Situational Scene Graph for Structured Human-centric Situation Understanding
by: Sugandhika, Chinthani, et al.
Published: (2024)
by: Sugandhika, Chinthani, et al.
Published: (2024)
UniScene: Unified Occupancy-centric Driving Scene Generation
by: Li, Bohan, et al.
Published: (2024)
by: Li, Bohan, et al.
Published: (2024)
Style-Editor: Text-driven object-centric style editing
by: Park, Jihun, et al.
Published: (2024)
by: Park, Jihun, et al.
Published: (2024)
ORV: 4D Occupancy-centric Robot Video Generation
by: Yang, Xiuyu, et al.
Published: (2025)
by: Yang, Xiuyu, et al.
Published: (2025)
Object-centric Video Question Answering with Visual Grounding and Referring
by: Wang, Haochen, et al.
Published: (2025)
by: Wang, Haochen, et al.
Published: (2025)
ORIDa: Object-centric Real-world Image Composition Dataset
by: Kim, Jinwoo, et al.
Published: (2025)
by: Kim, Jinwoo, et al.
Published: (2025)
Dynamic Avatar-Scene Rendering from Human-centric Context
by: Wang, Wenqing, et al.
Published: (2025)
by: Wang, Wenqing, et al.
Published: (2025)
Deep Learning on Object-centric 3D Neural Fields
by: Ramirez, Pierluigi Zama, et al.
Published: (2023)
by: Ramirez, Pierluigi Zama, et al.
Published: (2023)
Vision-centric Token Compression in Large Language Model
by: Xing, Ling, et al.
Published: (2025)
by: Xing, Ling, et al.
Published: (2025)
A Unified Framework for Human-centric Point Cloud Video Understanding
by: Xu, Yiteng, et al.
Published: (2024)
by: Xu, Yiteng, et al.
Published: (2024)
WcDT: World-centric Diffusion Transformer for Traffic Scene Generation
by: Yang, Chen, et al.
Published: (2024)
by: Yang, Chen, et al.
Published: (2024)
OCSU: Optical Chemical Structure Understanding for Molecule-centric Scientific Discovery
by: Fan, Siqi, et al.
Published: (2025)
by: Fan, Siqi, et al.
Published: (2025)
Improving Vision-language Models with Perception-centric Process Reward Models
by: Min, Yingqian, et al.
Published: (2026)
by: Min, Yingqian, et al.
Published: (2026)
Wavelet Burst Accumulation for turbulence mitigation
by: Gilles, Jerome, et al.
Published: (2024)
by: Gilles, Jerome, et al.
Published: (2024)
Bias mitigation in graph diffusion models
by: Yu, Meng, et al.
Published: (2026)
by: Yu, Meng, et al.
Published: (2026)
Similar Items
-
Gaze Behavior During a Long-Term, In-Home, Social Robot Intervention for Children with ASD
by: Ramnauth, Rebecca, et al.
Published: (2025) -
RadEyeVideo: Enhancing general-domain Large Vision Language Model for chest X-ray analysis with video representations of eye gaze
by: Kim, Yunsoo, et al.
Published: (2025) -
Certified vs. Empirical Adversarial Robust-ness via Hybrid Convolutions with Attention Stochasticity
by: Dhar, Joy, et al.
Published: (2026) -
End-to-end Video Gaze Estimation via Capturing Head-face-eye Spatial-temporal Interaction Context
by: Guan, Yiran, et al.
Published: (2023) -
DeepFaceLab: Integrated, flexible and extensible face-swapping framework
by: Perov, Ivan, et al.
Published: (2020)