RA-Touch: Retrieval-Augmented Touch Understanding with Enriched Visual Data
Fuente:
arXiv
Saved in:
| Main Authors: | Cho, Yoorhim, Kim, Hongyeob, Kim, Semin, Zhang, Youjia, Choi, Yunseok, Hong, Sungeun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Question-Aware Gaussian Experts for Audio-Visual Question Answering
by: Kim, Hongyeob, et al.
Published: (2025)
by: Kim, Hongyeob, et al.
Published: (2025)
Backpropagation-Free Test-Time Adaptation via Probabilistic Gaussian Alignment
by: Zhang, Youjia, et al.
Published: (2025)
by: Zhang, Youjia, et al.
Published: (2025)
Chameleon: A Data-Efficient Generalist for Dense Visual Prediction in the Wild
by: Kim, Donggyun, et al.
Published: (2024)
by: Kim, Donggyun, et al.
Published: (2024)
ZOO-Prune: Training-Free Token Pruning via Zeroth-Order Gradient Estimation in Vision-Language Models
by: Kim, Youngeun, et al.
Published: (2025)
by: Kim, Youngeun, et al.
Published: (2025)
Seeing Through Touch: Tactile-Driven Visual Localization of Material Regions
by: Kim, Seongyu, et al.
Published: (2026)
by: Kim, Seongyu, et al.
Published: (2026)
Touch-R1: Reinforcing Touch Reasoning in MLLMs
by: Lai, Yingxin, et al.
Published: (2026)
by: Lai, Yingxin, et al.
Published: (2026)
Touch2Shape: Touch-Conditioned 3D Diffusion for Shape Exploration and Reconstruction
by: Wang, Yuanbo, et al.
Published: (2025)
by: Wang, Yuanbo, et al.
Published: (2025)
The Midas Touch for Metric Depth
by: Ma, Yu, et al.
Published: (2026)
by: Ma, Yu, et al.
Published: (2026)
Retrieval-Augmented Open-Vocabulary Object Detection
by: Kim, Jooyeon, et al.
Published: (2024)
by: Kim, Jooyeon, et al.
Published: (2024)
TouchAnything: Diffusion-Guided 3D Reconstruction from Sparse Robot Touches
by: Gu, Langzhe, et al.
Published: (2026)
by: Gu, Langzhe, et al.
Published: (2026)
Cross-Sensor Touch Generation
by: Rodriguez, Samanta, et al.
Published: (2025)
by: Rodriguez, Samanta, et al.
Published: (2025)
Data Augmentation For Small Object using Fast AutoAugment
by: Yoon, DaeEun, et al.
Published: (2025)
by: Yoon, DaeEun, et al.
Published: (2025)
RA-RRG: Multimodal Retrieval-Augmented Radiology Report Generation with Key Phrase Extraction
by: Park, Jonggwon, et al.
Published: (2025)
by: Park, Jonggwon, et al.
Published: (2025)
RA-SGG: Retrieval-Augmented Scene Graph Generation Framework via Multi-Prototype Learning
by: Yoon, Kanghoon, et al.
Published: (2024)
by: Yoon, Kanghoon, et al.
Published: (2024)
EgoTouch: On-Body Touch Input Using AR/VR Headset Cameras
by: Mollyn, Vimal, et al.
Published: (2025)
by: Mollyn, Vimal, et al.
Published: (2025)
Touch-GS: Visual-Tactile Supervised 3D Gaussian Splatting
by: Swann, Aiden, et al.
Published: (2024)
by: Swann, Aiden, et al.
Published: (2024)
Bridging the gap to real-world language-grounded visual concept learning
by: Jung, Whie, et al.
Published: (2025)
by: Jung, Whie, et al.
Published: (2025)
TextToucher: Fine-Grained Text-to-Touch Generation
by: Tu, Jiahang, et al.
Published: (2024)
by: Tu, Jiahang, et al.
Published: (2024)
Retrieval-Augmented Natural Language Reasoning for Explainable Visual Question Answering
by: Lim, Su Hyeon, et al.
Published: (2024)
by: Lim, Su Hyeon, et al.
Published: (2024)
FaceGCD: Generalized Face Discovery via Dynamic Prefix Generation
by: Oh, Yunseok, et al.
Published: (2025)
by: Oh, Yunseok, et al.
Published: (2025)
Detecting Precise Hand Touch Moments in Egocentric Video
by: Nguyen, Huy Anh, et al.
Published: (2026)
by: Nguyen, Huy Anh, et al.
Published: (2026)
Object Pose Estimation through Dexterous Touch
by: Shahidzadeh, Amir-Hossein, et al.
Published: (2025)
by: Shahidzadeh, Amir-Hossein, et al.
Published: (2025)
IFQA: Interpretable Face Quality Assessment
by: Jo, Byungho, et al.
Published: (2022)
by: Jo, Byungho, et al.
Published: (2022)
QualiRAG: Retrieval-Augmented Generation for Visual Quality Understanding
by: Cao, Linhan, et al.
Published: (2026)
by: Cao, Linhan, et al.
Published: (2026)
PerTouch: VLM-Driven Agent for Personalized and Semantic Image Retouching
by: Chang, Zewei, et al.
Published: (2025)
by: Chang, Zewei, et al.
Published: (2025)
EclipseTouch: Touch Segmentation on Ad Hoc Surfaces using Worn Infrared Shadow Casting
by: Mollyn, Vimal, et al.
Published: (2025)
by: Mollyn, Vimal, et al.
Published: (2025)
A Review of Image Retrieval Techniques: Data Augmentation and Adversarial Learning Approaches
by: Jinwoo, Kim
Published: (2024)
by: Jinwoo, Kim
Published: (2024)
IAM: Enhancing RGB-D Instance Segmentation with New Benchmarks
by: Jung, Aecheon, et al.
Published: (2025)
by: Jung, Aecheon, et al.
Published: (2025)
Seeing Cells Clearly: Evaluating Machine Vision Strategies for Microglia Centroid Detection in 3D Images
by: Zhang, Youjia
Published: (2025)
by: Zhang, Youjia
Published: (2025)
A Touch, Vision, and Language Dataset for Multimodal Alignment
by: Fu, Letian, et al.
Published: (2024)
by: Fu, Letian, et al.
Published: (2024)
MI-CXR: A Benchmark for Longitudinal Reasoning over Multi-Interval Chest X-rays
by: Cho, Sunghwan Steve, et al.
Published: (2026)
by: Cho, Sunghwan Steve, et al.
Published: (2026)
BiFingerPose: Bimodal Finger Pose Estimation for Touch Devices
by: Guan, Xiongjun, et al.
Published: (2025)
by: Guan, Xiongjun, et al.
Published: (2025)
Learning an Ensemble Token from Task-driven Priors in Facial Analysis
by: Seo, Sunyong, et al.
Published: (2025)
by: Seo, Sunyong, et al.
Published: (2025)
Hearing Touch: Audio-Visual Pretraining for Contact-Rich Manipulation
by: Mejia, Jared, et al.
Published: (2024)
by: Mejia, Jared, et al.
Published: (2024)
Binding Touch to Everything: Learning Unified Multimodal Tactile Representations
by: Yang, Fengyu, et al.
Published: (2024)
by: Yang, Fengyu, et al.
Published: (2024)
TouchMap-OR: Multi-View 3D Mapping of Hand-Surface Contacts
by: Ktistakis, Sophokles, et al.
Published: (2026)
by: Ktistakis, Sophokles, et al.
Published: (2026)
DALDA: Data Augmentation Leveraging Diffusion Model and LLM with Adaptive Guidance Scaling
by: Jung, Kyuheon, et al.
Published: (2024)
by: Jung, Kyuheon, et al.
Published: (2024)
PseudoTouch: Efficiently Imaging the Surface Feel of Objects for Robotic Manipulation
by: Röfer, Adrian, et al.
Published: (2024)
by: Röfer, Adrian, et al.
Published: (2024)
Towards Comprehensive Multimodal Perception: Introducing the Touch-Language-Vision Dataset
by: Cheng, Ning, et al.
Published: (2024)
by: Cheng, Ning, et al.
Published: (2024)
The Language of Touch: Translating Vibrations into Text with Dual-Branch Learning
by: Chen, Jin, et al.
Published: (2026)
by: Chen, Jin, et al.
Published: (2026)
Similar Items
-
Question-Aware Gaussian Experts for Audio-Visual Question Answering
by: Kim, Hongyeob, et al.
Published: (2025) -
Backpropagation-Free Test-Time Adaptation via Probabilistic Gaussian Alignment
by: Zhang, Youjia, et al.
Published: (2025) -
Chameleon: A Data-Efficient Generalist for Dense Visual Prediction in the Wild
by: Kim, Donggyun, et al.
Published: (2024) -
ZOO-Prune: Training-Free Token Pruning via Zeroth-Order Gradient Estimation in Vision-Language Models
by: Kim, Youngeun, et al.
Published: (2025) -
Seeing Through Touch: Tactile-Driven Visual Localization of Material Regions
by: Kim, Seongyu, et al.
Published: (2026)