Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision
Fuente:
arXiv
Saved in:
| Main Authors: | Kansana, Manish, Penchala, Sindhuja, Rahimi, Shahram, Golilarz, Noorbakhsh Amiri |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Surformer v1: Transformer-Based Surface Classification Using Tactile and Vision Features
by: Kansana, Manish, et al.
Published: (2025)
by: Kansana, Manish, et al.
Published: (2025)
Bridging the Gap: Toward Cognitive Autonomy in Artificial Intelligence
by: Golilarz, Noorbakhsh Amiri, et al.
Published: (2025)
by: Golilarz, Noorbakhsh Amiri, et al.
Published: (2025)
Where to Bind Matters: Hebbian Fast Weights in Vision Transformers for Few-Shot Character Recognition
by: Money, Gavin, et al.
Published: (2026)
by: Money, Gavin, et al.
Published: (2026)
One Patch is All You Need: Joint Surface Material Reconstruction and Classification from Minimal Visual Cues
by: Penchala, Sindhuja, et al.
Published: (2025)
by: Penchala, Sindhuja, et al.
Published: (2025)
Edge-Based Learning for Improved Classification Under Adversarial Noise
by: Kansana, Manish, et al.
Published: (2025)
by: Kansana, Manish, et al.
Published: (2025)
Towards Neurocognitive-Inspired Intelligence: From AI's Structural Mimicry to Human-Like Functional Cognition
by: Golilarz, Noorbakhsh Amiri, et al.
Published: (2025)
by: Golilarz, Noorbakhsh Amiri, et al.
Published: (2025)
Learning in Focus: Detecting Behavioral and Collaborative Engagement Using Vision Transformers
by: Penchala, Sindhuja, et al.
Published: (2025)
by: Penchala, Sindhuja, et al.
Published: (2025)
Advancing Generative Model Evaluation: A Novel Algorithm for Realistic Image Synthesis and Comparison in OCR System
by: Memari, Majid, et al.
Published: (2024)
by: Memari, Majid, et al.
Published: (2024)
R-GAT: Cancer Document Classification Leveraging Graph-Based Residual Network for Scenarios with Limited Data
by: Hossain, Elias, et al.
Published: (2024)
by: Hossain, Elias, et al.
Published: (2024)
MedInsight: A Multi-Source Context Augmentation Framework for Generating Patient-Centric Medical Responses using Large Language Models
by: Neupane, Subash, et al.
Published: (2024)
by: Neupane, Subash, et al.
Published: (2024)
Estimating Reliability of Electric Vehicle Charging Ecosystem using the Principle of Maximum Entropy
by: Tripathi, Himanshu, et al.
Published: (2025)
by: Tripathi, Himanshu, et al.
Published: (2025)
Gamma2Patterns: Deep Cognitive Attention Region Identification and Gamma-Alpha Pattern Analysis
by: Jahan, Sobhana, et al.
Published: (2026)
by: Jahan, Sobhana, et al.
Published: (2026)
AI Learning Algorithms: Deep Learning, Hybrid Models, and Large-Scale Model Integration
by: Golilarz, Noorbakhsh Amiri, et al.
Published: (2024)
by: Golilarz, Noorbakhsh Amiri, et al.
Published: (2024)
Towards Secure MLOps: Surveying Attacks, Mitigation Strategies, and Research Challenges
by: Patel, Raj, et al.
Published: (2025)
by: Patel, Raj, et al.
Published: (2025)
Touch100k: A Large-Scale Touch-Language-Vision Dataset for Touch-Centric Multimodal Representation
by: Cheng, Ning, et al.
Published: (2024)
by: Cheng, Ning, et al.
Published: (2024)
Patient-Centric Knowledge Graphs: A Survey of Current Methods, Challenges, and Applications
by: Khatib, Hassan S. Al, et al.
Published: (2024)
by: Khatib, Hassan S. Al, et al.
Published: (2024)
A Touch, Vision, and Language Dataset for Multimodal Alignment
by: Fu, Letian, et al.
Published: (2024)
by: Fu, Letian, et al.
Published: (2024)
Touch Speaks, Sound Feels: A Multimodal Approach to Affective and Social Touch from Robots to Humans
by: Ren, Qiaoqiao, et al.
Published: (2025)
by: Ren, Qiaoqiao, et al.
Published: (2025)
From Questions to Insightful Answers: Building an Informed Chatbot for University Resources
by: Neupane, Subash, et al.
Published: (2024)
by: Neupane, Subash, et al.
Published: (2024)
Towards Comprehensive Multimodal Perception: Introducing the Touch-Language-Vision Dataset
by: Cheng, Ning, et al.
Published: (2024)
by: Cheng, Ning, et al.
Published: (2024)
Analysing the Interplay of Vision and Touch for Dexterous Insertion Tasks
by: Lenz, Janis, et al.
Published: (2024)
by: Lenz, Janis, et al.
Published: (2024)
Contrastive Touch-to-Touch Pretraining
by: Rodriguez, Samanta, et al.
Published: (2024)
by: Rodriguez, Samanta, et al.
Published: (2024)
HydroelasticTouch: Simulation of Tactile Sensors with Hydroelastic Contact Surfaces
by: Leins, David P., et al.
Published: (2025)
by: Leins, David P., et al.
Published: (2025)
Touch2Touch: Cross-Modal Tactile Generation for Object Manipulation
by: Rodriguez, Samanta, et al.
Published: (2024)
by: Rodriguez, Samanta, et al.
Published: (2024)
Gaussian Process-Based Active Exploration Strategies in Vision and Touch
by: Choi, Ho Jin, et al.
Published: (2025)
by: Choi, Ho Jin, et al.
Published: (2025)
ViTacGen: Robotic Pushing with Vision-to-Touch Generation
by: Wu, Zhiyuan, et al.
Published: (2025)
by: Wu, Zhiyuan, et al.
Published: (2025)
Dance2Hesitate: A Multi-Modal Dataset of Dancer-Taught Hesitancy for Understandable Robot Motion
by: Raghu, Srikrishna Bangalore, et al.
Published: (2026)
by: Raghu, Srikrishna Bangalore, et al.
Published: (2026)
Vi-TacMan: Articulated Object Manipulation via Vision and Touch
by: Cui, Leiyao, et al.
Published: (2025)
by: Cui, Leiyao, et al.
Published: (2025)
Touch2Insert: Zero-Shot Peg Insertion by Touching Intersections of Peg and Hole
by: Yajima, Masaru, et al.
Published: (2026)
by: Yajima, Masaru, et al.
Published: (2026)
When Vision Meets Touch: A Contemporary Review for Visuotactile Sensors from the Signal Processing Perspective
by: Li, Shoujie, et al.
Published: (2024)
by: Li, Shoujie, et al.
Published: (2024)
VinT-6D: A Large-Scale Object-in-hand Dataset from Vision, Touch and Proprioception
by: Wan, Zhaoliang, et al.
Published: (2024)
by: Wan, Zhaoliang, et al.
Published: (2024)
Tabero: Learning Gentle Manipulation with Closed-Loop Force Feedback from Vision, Touch, and Language
by: Wu, Qiwei, et al.
Published: (2026)
by: Wu, Qiwei, et al.
Published: (2026)
Look-to-Touch: A Vision-Enhanced Proximity and Tactile Sensor for Distance and Geometry Perception in Robotic Manipulation
by: Dong, Yueshi, et al.
Published: (2025)
by: Dong, Yueshi, et al.
Published: (2025)
Touch and Tell: Multimodal Decoding of Human Emotions and Social Gestures for Robots
by: Ren, Qiaoqiao, et al.
Published: (2024)
by: Ren, Qiaoqiao, et al.
Published: (2024)
TouchGuide: Inference-Time Steering of Visuomotor Policies via Touch Guidance
by: Zhang, Zhemeng, et al.
Published: (2026)
by: Zhang, Zhemeng, et al.
Published: (2026)
Touch-to-Touch Translation -- Learning the Mapping Between Heterogeneous Tactile Sensing Technologies
by: Grella, Francesco, et al.
Published: (2024)
by: Grella, Francesco, et al.
Published: (2024)
TactEx: An Explainable Multimodal Robotic Interaction Framework for Human-Like Touch and Hardness Estimation
by: Verstraete, Felix, et al.
Published: (2026)
by: Verstraete, Felix, et al.
Published: (2026)
Skin-Machine Interface with Multimodal Contact Motion Classifier
by: Confente, Alberto, et al.
Published: (2025)
by: Confente, Alberto, et al.
Published: (2025)
Beyond the Manual Touch: Situational-aware Force Control for Increased Safety in Robot-assisted Skullbase Surgery
by: Ishida, Hisashi, et al.
Published: (2024)
by: Ishida, Hisashi, et al.
Published: (2024)
A Grasp Pose is All You Need: Learning Multi-fingered Grasping with Deep Reinforcement Learning from Vision and Touch
by: Ceola, Federico, et al.
Published: (2023)
by: Ceola, Federico, et al.
Published: (2023)
Similar Items
-
Surformer v1: Transformer-Based Surface Classification Using Tactile and Vision Features
by: Kansana, Manish, et al.
Published: (2025) -
Bridging the Gap: Toward Cognitive Autonomy in Artificial Intelligence
by: Golilarz, Noorbakhsh Amiri, et al.
Published: (2025) -
Where to Bind Matters: Hebbian Fast Weights in Vision Transformers for Few-Shot Character Recognition
by: Money, Gavin, et al.
Published: (2026) -
One Patch is All You Need: Joint Surface Material Reconstruction and Classification from Minimal Visual Cues
by: Penchala, Sindhuja, et al.
Published: (2025) -
Edge-Based Learning for Improved Classification Under Adversarial Noise
by: Kansana, Manish, et al.
Published: (2025)