Synthetic data augmentation for robotic mobility aids to support blind and low vision people
Fuente:
arXiv
Guardado en:
| Autores principales: | Hwang, Hochul, Adhikari, Krisha, Shodhaka, Satya, Kim, Donghyun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Is it safe to cross? Interpretable Risk Assessment with GPT-4V for Safety-Aware Street Crossing
por: Hwang, Hochul, et al.
Publicado: (2024)
por: Hwang, Hochul, et al.
Publicado: (2024)
An evaluation of CNN models and data augmentation techniques in hierarchical localization of mobile robots
por: Cabrera, J. J., et al.
Publicado: (2024)
por: Cabrera, J. J., et al.
Publicado: (2024)
Explaining CLIP's performance disparities on data from blind/low vision users
por: Massiceti, Daniela, et al.
Publicado: (2023)
por: Massiceti, Daniela, et al.
Publicado: (2023)
Rethinking Saliency-Guided Weakly-Supervised Semantic Segmentation
por: Kim, Beomyoung, et al.
Publicado: (2024)
por: Kim, Beomyoung, et al.
Publicado: (2024)
GuideNav: User-Informed Development of a Vision-Only Robotic Navigation Assistant For Blind Travelers
por: Hwang, Hochul, et al.
Publicado: (2025)
por: Hwang, Hochul, et al.
Publicado: (2025)
LoRA-Loop: Closing the Synthetic Replay Cycle for Continual VLM Learning
por: Wang, Kaihong, et al.
Publicado: (2025)
por: Wang, Kaihong, et al.
Publicado: (2025)
FALCON: Frequency Adjoint Link with CONtinuous Density Mask for Fast Single Image Dehazing
por: Kim, Donghyun, et al.
Publicado: (2024)
por: Kim, Donghyun, et al.
Publicado: (2024)
PLATYPUS: Progressive Local Surface Estimator for Arbitrary-Scale Point Cloud Upsampling
por: Kim, Donghyun, et al.
Publicado: (2024)
por: Kim, Donghyun, et al.
Publicado: (2024)
Delaunay Canopy: Building Wireframe Reconstruction from Airborne LiDAR Point Clouds via Delaunay Graph
por: Kim, Donghyun, et al.
Publicado: (2026)
por: Kim, Donghyun, et al.
Publicado: (2026)
GenQ: Quantization in Low Data Regimes with Generative Synthetic Data
por: Li, Yuhang, et al.
Publicado: (2023)
por: Li, Yuhang, et al.
Publicado: (2023)
Physics-consistent deep learning for blind aberration recovery in mobile optics
por: Jhawar, Kartik, et al.
Publicado: (2026)
por: Jhawar, Kartik, et al.
Publicado: (2026)
Rethinking Graph Convolution for 2D-to-3D Hand Pose Lifting
por: Kim, Chanyoung, et al.
Publicado: (2026)
por: Kim, Chanyoung, et al.
Publicado: (2026)
Difference Inversion: Interpolate and Isolate the Difference with Token Consistency for Image Analogy Generation
por: Kim, Hyunsoo, et al.
Publicado: (2025)
por: Kim, Hyunsoo, et al.
Publicado: (2025)
SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data
por: Kim, Dong-Hee, et al.
Publicado: (2025)
por: Kim, Dong-Hee, et al.
Publicado: (2025)
Fourier Decomposition for Explicit Representation of 3D Point Cloud Attributes
por: Kim, Donghyun, et al.
Publicado: (2025)
por: Kim, Donghyun, et al.
Publicado: (2025)
A survey of synthetic data augmentation methods in computer vision
por: Mumuni, Alhassan, et al.
Publicado: (2024)
por: Mumuni, Alhassan, et al.
Publicado: (2024)
FastSHADE: Fast Self-augmented Hierarchical Asymmetric Denoising for Efficient inference on mobile devices
por: Falaleev, Nikolay
Publicado: (2026)
por: Falaleev, Nikolay
Publicado: (2026)
ContextMix: A context-aware data augmentation method for industrial visual inspection systems
por: Kim, Hyungmin, et al.
Publicado: (2024)
por: Kim, Hyungmin, et al.
Publicado: (2024)
Crafting Query-Aware Selective Attention for Single Image Super-Resolution
por: Kim, Junyoung, et al.
Publicado: (2025)
por: Kim, Junyoung, et al.
Publicado: (2025)
Think as Needed: Geometry-Driven Adaptive Perception for Autonomous Driving
por: Kim, Donghyun, et al.
Publicado: (2026)
por: Kim, Donghyun, et al.
Publicado: (2026)
SYNAPSE: Synergizing an Adapter and Finetuning for High-Fidelity EEG Synthesis from a CLIP-Aligned Encoder
por: Lee, Jeyoung, et al.
Publicado: (2025)
por: Lee, Jeyoung, et al.
Publicado: (2025)
METER: a mobile vision transformer architecture for monocular depth estimation
por: Papa, L., et al.
Publicado: (2024)
por: Papa, L., et al.
Publicado: (2024)
Are We Truly Forgetting? A Critical Re-examination of Machine Unlearning Evaluation Protocols
por: Kim, Yongwoo, et al.
Publicado: (2025)
por: Kim, Yongwoo, et al.
Publicado: (2025)
Erase at the Core: Representation Unlearning for Machine Unlearning
por: Lee, Jaewon, et al.
Publicado: (2026)
por: Lee, Jaewon, et al.
Publicado: (2026)
Distilling Vision-Language Pretraining for Efficient Cross-Modal Retrieval
por: Jang, Young Kyun, et al.
Publicado: (2024)
por: Jang, Young Kyun, et al.
Publicado: (2024)
On the power of data augmentation for head pose estimation
por: Welter, Michael
Publicado: (2024)
por: Welter, Michael
Publicado: (2024)
Synthetic images aid the recognition of human-made art forgeries
por: Ostmeyer, Johann, et al.
Publicado: (2023)
por: Ostmeyer, Johann, et al.
Publicado: (2023)
Hyperspectral data augmentation with transformer-based diffusion models
por: Ferrari, Mattia, et al.
Publicado: (2025)
por: Ferrari, Mattia, et al.
Publicado: (2025)
Diversity Over Frequency: Rethinking Tool Use in Visual Chain-of-Thought Agents
por: Kim, Dong-Hee, et al.
Publicado: (2026)
por: Kim, Dong-Hee, et al.
Publicado: (2026)
Interpreting vision transformers via residual replacement model
por: Kim, Jinyeong, et al.
Publicado: (2025)
por: Kim, Jinyeong, et al.
Publicado: (2025)
RA-SGG: Retrieval-Augmented Scene Graph Generation Framework via Multi-Prototype Learning
por: Yoon, Kanghoon, et al.
Publicado: (2024)
por: Yoon, Kanghoon, et al.
Publicado: (2024)
Grid-augmented vision: A simple yet effective approach for enhanced spatial understanding in multi-modal agents
por: Chae, Joongwon, et al.
Publicado: (2024)
por: Chae, Joongwon, et al.
Publicado: (2024)
A data-centric approach to class-specific bias in image data augmentation
por: Angelakis, Athanasios, et al.
Publicado: (2024)
por: Angelakis, Athanasios, et al.
Publicado: (2024)
CaptionSmiths: Flexibly Controlling Language Pattern in Image Captioning
por: Saito, Kuniaki, et al.
Publicado: (2025)
por: Saito, Kuniaki, et al.
Publicado: (2025)
Efficient and Versatile Robust Fine-Tuning of Zero-shot Models
por: Kim, Sungyeon, et al.
Publicado: (2024)
por: Kim, Sungyeon, et al.
Publicado: (2024)
Single-image driven 3d viewpoint training data augmentation for effective wine label recognition
por: Huang, Yueh-Cheng, et al.
Publicado: (2024)
por: Huang, Yueh-Cheng, et al.
Publicado: (2024)
Computer vision training dataset generation for robotic environments using Gaussian splatting
por: Niżeniec, Patryk, et al.
Publicado: (2025)
por: Niżeniec, Patryk, et al.
Publicado: (2025)
Quantitative evaluation of brain-inspired vision sensors in high-speed robotic perception
por: Wang, Taoyi, et al.
Publicado: (2025)
por: Wang, Taoyi, et al.
Publicado: (2025)
OxML Challenge 2023: Carcinoma classification using data augmentation
por: Raj, Kislay, et al.
Publicado: (2024)
por: Raj, Kislay, et al.
Publicado: (2024)
Weak-to-Strong Compositional Learning from Generative Models for Language-based Object Detection
por: Park, Kwanyong, et al.
Publicado: (2024)
por: Park, Kwanyong, et al.
Publicado: (2024)
Ejemplares similares
-
Is it safe to cross? Interpretable Risk Assessment with GPT-4V for Safety-Aware Street Crossing
por: Hwang, Hochul, et al.
Publicado: (2024) -
An evaluation of CNN models and data augmentation techniques in hierarchical localization of mobile robots
por: Cabrera, J. J., et al.
Publicado: (2024) -
Explaining CLIP's performance disparities on data from blind/low vision users
por: Massiceti, Daniela, et al.
Publicado: (2023) -
Rethinking Saliency-Guided Weakly-Supervised Semantic Segmentation
por: Kim, Beomyoung, et al.
Publicado: (2024) -
GuideNav: User-Informed Development of a Vision-Only Robotic Navigation Assistant For Blind Travelers
por: Hwang, Hochul, et al.
Publicado: (2025)