Saved in:
| Main Authors: | De Mathia, Joseph, Moreno-García, Carlos Francisco |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2507.16330 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reading in the Dark: Low-light Scene Text Recognition
by: Fu, Xuanshuo, et al.
Published: (2026)
by: Fu, Xuanshuo, et al.
Published: (2026)
Aria-NeRF: Multimodal Egocentric View Synthesis
by: Sun, Jiankai, et al.
Published: (2023)
by: Sun, Jiankai, et al.
Published: (2023)
Vision-Based Robust Lane Detection and Tracking under Different Challenging Environmental Conditions
by: Sultana, Samia, et al.
Published: (2022)
by: Sultana, Samia, et al.
Published: (2022)
SVIPTR: Fast and Efficient Scene Text Recognition with Vision Permutable Extractor
by: Cheng, Xianfu, et al.
Published: (2024)
by: Cheng, Xianfu, et al.
Published: (2024)
The First Swahili Language Scene Text Detection and Recognition Dataset
by: Douamba, Fadila Wendigoundi, et al.
Published: (2024)
by: Douamba, Fadila Wendigoundi, et al.
Published: (2024)
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering
by: Zhou, Sheng, et al.
Published: (2025)
by: Zhou, Sheng, et al.
Published: (2025)
Anomaly Detection for People with Visual Impairments Using an Egocentric 360-Degree Camera
by: Song, Inpyo, et al.
Published: (2024)
by: Song, Inpyo, et al.
Published: (2024)
Challenges and Trends in Egocentric Vision: A Survey
by: Li, Xiang, et al.
Published: (2025)
by: Li, Xiang, et al.
Published: (2025)
Pandora: Articulated 3D Scene Graphs from Egocentric Vision
by: Yu, Alan, et al.
Published: (2026)
by: Yu, Alan, et al.
Published: (2026)
Recognition-Synergistic Scene Text Editing
by: Fang, Zhengyao, et al.
Published: (2025)
by: Fang, Zhengyao, et al.
Published: (2025)
Instruction-Guided Scene Text Recognition
by: Du, Yongkun, et al.
Published: (2024)
by: Du, Yongkun, et al.
Published: (2024)
ViewDelta: Scaling Scene Change Detection through Text-Conditioning
by: Varghese, Subin, et al.
Published: (2024)
by: Varghese, Subin, et al.
Published: (2024)
Egocentric Gaze Estimation via Neck-Mounted Camera
by: Huang, Haoyu, et al.
Published: (2026)
by: Huang, Haoyu, et al.
Published: (2026)
Text-driven Affordance Learning from Egocentric Vision
by: Yoshida, Tomoya, et al.
Published: (2024)
by: Yoshida, Tomoya, et al.
Published: (2024)
Domain Generalization using Action Sequences for Egocentric Action Recognition
by: Nasirimajd, Amirshayan, et al.
Published: (2025)
by: Nasirimajd, Amirshayan, et al.
Published: (2025)
DarkShake-DVS: Event-based Human Action Recognition under Low-light andShaking Camera Conditions
by: Chen, Jiaqi, et al.
Published: (2026)
by: Chen, Jiaqi, et al.
Published: (2026)
Technical Report for Egocentric Mistake Detection for the HoloAssist Challenge
by: Patsch, Constantin, et al.
Published: (2025)
by: Patsch, Constantin, et al.
Published: (2025)
Decoder Pre-Training with only Text for Scene Text Recognition
by: Zhao, Shuai, et al.
Published: (2024)
by: Zhao, Shuai, et al.
Published: (2024)
TEACH: Text Encoding as Curriculum Hints for Scene Text Recognition
by: Yang, Xiahan, et al.
Published: (2025)
by: Yang, Xiahan, et al.
Published: (2025)
Improved Scene Landmark Detection for Camera Localization
by: Do, Tien, et al.
Published: (2024)
by: Do, Tien, et al.
Published: (2024)
Adaptive Geodesic Conformal Prediction for Egocentric Camera Pose Estimation
by: Pathak, Aishani, et al.
Published: (2026)
by: Pathak, Aishani, et al.
Published: (2026)
EgoEvGesture: Gesture Recognition Based on Egocentric Event Camera
by: Wang, Luming, et al.
Published: (2025)
by: Wang, Luming, et al.
Published: (2025)
An Outlook into the Future of Egocentric Vision
by: Plizzari, Chiara, et al.
Published: (2023)
by: Plizzari, Chiara, et al.
Published: (2023)
Egocentric Vision Language Planning
by: Fang, Zhirui, et al.
Published: (2024)
by: Fang, Zhirui, et al.
Published: (2024)
Dataset and Benchmark for Urdu Natural Scenes Text Detection, Recognition and Visual Question Answering
by: Maryam, Hiba, et al.
Published: (2024)
by: Maryam, Hiba, et al.
Published: (2024)
KhmerST: A Low-Resource Khmer Scene Text Detection and Recognition Benchmark
by: Nom, Vannkinh, et al.
Published: (2024)
by: Nom, Vannkinh, et al.
Published: (2024)
TextSSR: Diffusion-based Data Synthesis for Scene Text Recognition
by: Ye, Xingsong, et al.
Published: (2024)
by: Ye, Xingsong, et al.
Published: (2024)
Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis
by: Yuan, Yu, et al.
Published: (2024)
by: Yuan, Yu, et al.
Published: (2024)
egoPPG: Heart Rate Estimation from Eye-Tracking Cameras in Egocentric Systems to Benefit Downstream Vision Tasks
by: Braun, Björn, et al.
Published: (2025)
by: Braun, Björn, et al.
Published: (2025)
Aggregated Text Transformer for Scene Text Detection
by: Zhou, Zhao, et al.
Published: (2022)
by: Zhou, Zhao, et al.
Published: (2022)
Efficient and Accurate Scene Text Recognition with Cascaded-Transformers
by: Ozkan, Savas, et al.
Published: (2025)
by: Ozkan, Savas, et al.
Published: (2025)
SimpleEgo: Predicting Probabilistic Body Pose from Egocentric Cameras
by: Cuevas-Velasquez, Hanz, et al.
Published: (2024)
by: Cuevas-Velasquez, Hanz, et al.
Published: (2024)
Hybrid Structure-from-Motion and Camera Relocalization for Enhanced Egocentric Localization
by: Mai, Jinjie, et al.
Published: (2024)
by: Mai, Jinjie, et al.
Published: (2024)
StyleTextGen: Style-Conditioned Multilingual Scene Text Generation
by: Chen, Zeyu, et al.
Published: (2026)
by: Chen, Zeyu, et al.
Published: (2026)
EgoAdapt: A Multi-Scene Egocentric Adaptation Method for CVPR 2026 HD-EPIC VQA Challenge
by: Chen, Zhiwei, et al.
Published: (2026)
by: Chen, Zhiwei, et al.
Published: (2026)
PEDESTRIAN: An Egocentric Vision Dataset for Obstacle Detection on Pavements
by: Thoma, Marios, et al.
Published: (2025)
by: Thoma, Marios, et al.
Published: (2025)
Measuring Natural Scenes SFR of Automotive Fisheye Cameras
by: Jakab, Daniel, et al.
Published: (2024)
by: Jakab, Daniel, et al.
Published: (2024)
CLIP4STR: A Simple Baseline for Scene Text Recognition with Pre-trained Vision-Language Model
by: Zhao, Shuai, et al.
Published: (2023)
by: Zhao, Shuai, et al.
Published: (2023)
AnimateScene: Camera-controllable Animation in Any Scene
by: Liu, Qingyang, et al.
Published: (2025)
by: Liu, Qingyang, et al.
Published: (2025)
Spatial-Conditioned Reasoning in Long-Egocentric Videos
by: Tribble, James, et al.
Published: (2026)
by: Tribble, James, et al.
Published: (2026)
Similar Items
-
Reading in the Dark: Low-light Scene Text Recognition
by: Fu, Xuanshuo, et al.
Published: (2026) -
Aria-NeRF: Multimodal Egocentric View Synthesis
by: Sun, Jiankai, et al.
Published: (2023) -
Vision-Based Robust Lane Detection and Tracking under Different Challenging Environmental Conditions
by: Sultana, Samia, et al.
Published: (2022) -
SVIPTR: Fast and Efficient Scene Text Recognition with Vision Permutable Extractor
by: Cheng, Xianfu, et al.
Published: (2024) -
The First Swahili Language Scene Text Detection and Recognition Dataset
by: Douamba, Fadila Wendigoundi, et al.
Published: (2024)