Saved in:
| Main Authors: | Zhao, Fei, Zhang, Runlin, Zhang, Chengcui, Saxena, Nitesh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2505.07032 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mitigating Image Captioning Hallucinations in Vision-Language Models
by: Zhao, Fei, et al.
Published: (2025)
by: Zhao, Fei, et al.
Published: (2025)
BeyondPixels: A Comprehensive Review of the Evolution of Neural Radiance Fields
by: Rabby, AKM Shahariar Azad, et al.
Published: (2023)
by: Rabby, AKM Shahariar Azad, et al.
Published: (2023)
Translation-based Video-to-Video Synthesis
by: Saha, Pratim, et al.
Published: (2024)
by: Saha, Pratim, et al.
Published: (2024)
Multi-Granularity Hand Action Detection
by: Zhe, Ting, et al.
Published: (2023)
by: Zhe, Ting, et al.
Published: (2023)
Driving with Context: Online Map Matching for Complex Roads Using Lane Markings and Scenario Recognition
by: Bi, Xin, et al.
Published: (2025)
by: Bi, Xin, et al.
Published: (2025)
HSS-IAD: A Heterogeneous Same-Sort Industrial Anomaly Detection Dataset
by: Wang, Qishan, et al.
Published: (2025)
by: Wang, Qishan, et al.
Published: (2025)
mmEgoHand: Egocentric Hand Pose Estimation and Gesture Recognition with Head-mounted Millimeter-wave Radar and IMU
by: Lv, Yizhe, et al.
Published: (2025)
by: Lv, Yizhe, et al.
Published: (2025)
Same Answer, Different Representations: Hidden instability in VLMs
by: Wani, Farooq Ahmad, et al.
Published: (2026)
by: Wani, Farooq Ahmad, et al.
Published: (2026)
StableHand: Quality-Aware Flow Matching for World-Space Dual-Hand Motion Estimation from Egocentric Video
by: Zeng, Huajian, et al.
Published: (2026)
by: Zeng, Huajian, et al.
Published: (2026)
LEAF: Unveiling Two Sides of the Same Coin in Semi-supervised Facial Expression Recognition
by: Zhang, Fan, et al.
Published: (2024)
by: Zhang, Fan, et al.
Published: (2024)
PAD-Hand: Physics-Aware Diffusion for Hand Motion Recovery
by: Ismayilzada, Elkhan, et al.
Published: (2026)
by: Ismayilzada, Elkhan, et al.
Published: (2026)
Are They the Same? Exploring Visual Correspondence Shortcomings of Multimodal LLMs
by: Zhou, Yikang, et al.
Published: (2025)
by: Zhou, Yikang, et al.
Published: (2025)
LampMark: Proactive Deepfake Detection via Training-Free Landmark Perceptual Watermarks
by: Wang, Tianyi, et al.
Published: (2024)
by: Wang, Tianyi, et al.
Published: (2024)
VRM: Knowledge Distillation via Virtual Relation Matching
by: Zhang, Weijia, et al.
Published: (2025)
by: Zhang, Weijia, et al.
Published: (2025)
MESA: Matching Everything by Segmenting Anything
by: Zhang, Yesheng, et al.
Published: (2024)
by: Zhang, Yesheng, et al.
Published: (2024)
MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation
by: Zhou, Bohan, et al.
Published: (2025)
by: Zhou, Bohan, et al.
Published: (2025)
HandCraft: Anatomically Correct Restoration of Malformed Hands in Diffusion Generated Images
by: Qin, Zhenyue, et al.
Published: (2024)
by: Qin, Zhenyue, et al.
Published: (2024)
LandMarkSystem Technical Report
by: Ma, Zhenxiang, et al.
Published: (2025)
by: Ma, Zhenxiang, et al.
Published: (2025)
BOTH2Hands: Inferring 3D Hands from Both Text Prompts and Body Dynamics
by: Zhang, Wenqian, et al.
Published: (2023)
by: Zhang, Wenqian, et al.
Published: (2023)
GigaHands: A Massive Annotated Dataset of Bimanual Hand Activities
by: Fu, Rao, et al.
Published: (2024)
by: Fu, Rao, et al.
Published: (2024)
MT-Mark: Rethinking Image Watermarking via Mutual-Teacher Collaboration with Adaptive Feature Modulation
by: Ge, Fei, et al.
Published: (2025)
by: Ge, Fei, et al.
Published: (2025)
HandRefiner: Refining Malformed Hands in Generated Images by Diffusion-based Conditional Inpainting
by: Lu, Wenquan, et al.
Published: (2023)
by: Lu, Wenquan, et al.
Published: (2023)
Open-Vocabulary Animal Keypoint Detection with Semantic-feature Matching
by: Zhang, Hao, et al.
Published: (2023)
by: Zhang, Hao, et al.
Published: (2023)
SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics
by: O, Hayeon, et al.
Published: (2024)
by: O, Hayeon, et al.
Published: (2024)
Task-Specific Distance Correlation Matching for Few-Shot Action Recognition
by: Long, Fei, et al.
Published: (2025)
by: Long, Fei, et al.
Published: (2025)
Aligning What EEG Can See: Structural Representations for Brain-Vision Matching
by: Tang, Jingyi, et al.
Published: (2026)
by: Tang, Jingyi, et al.
Published: (2026)
FoundHand: Large-Scale Domain-Specific Learning for Controllable Hand Image Generation
by: Chen, Kefan, et al.
Published: (2024)
by: Chen, Kefan, et al.
Published: (2024)
FlowCast: Trajectory Forecasting for Scalable Zero-Cost Speculative Flow Matching
by: Bajpai, Divya Jyoti, et al.
Published: (2026)
by: Bajpai, Divya Jyoti, et al.
Published: (2026)
Ensemble Quadratic Assignment Network for Graph Matching
by: Tan, Haoru, et al.
Published: (2024)
by: Tan, Haoru, et al.
Published: (2024)
HO-Flow: Generalizable Hand-Object Interaction Generation with Latent Flow Matching
by: Chen, Zerui, et al.
Published: (2026)
by: Chen, Zerui, et al.
Published: (2026)
Homography Guided Temporal Fusion for Road Line and Marking Segmentation
by: Wang, Shan, et al.
Published: (2024)
by: Wang, Shan, et al.
Published: (2024)
HandOS: 3D Hand Reconstruction in One Stage
by: Chen, Xingyu, et al.
Published: (2024)
by: Chen, Xingyu, et al.
Published: (2024)
EgoHandICL: Egocentric 3D Hand Reconstruction with In-Context Learning
by: Xie, Binzhu, et al.
Published: (2026)
by: Xie, Binzhu, et al.
Published: (2026)
Detecting Precise Hand Touch Moments in Egocentric Video
by: Nguyen, Huy Anh, et al.
Published: (2026)
by: Nguyen, Huy Anh, et al.
Published: (2026)
HandGCAT: Occlusion-Robust 3D Hand Mesh Reconstruction from Monocular Images
by: Wang, Shuaibing, et al.
Published: (2024)
by: Wang, Shuaibing, et al.
Published: (2024)
MSN: Multi-directional Similarity Network for Hand-crafted and Deep-synthesized Copy-Move Forgery Detection
by: Jiang, Liangwei, et al.
Published: (2025)
by: Jiang, Liangwei, et al.
Published: (2025)
MESA: Effective Matching Redundancy Reduction by Semantic Area Segmentation
by: Zhang, Yesheng, et al.
Published: (2024)
by: Zhang, Yesheng, et al.
Published: (2024)
Decomposed Global Optimization for Robust Point Matching with Low-Dimensional Branching
by: Lian, Wei, et al.
Published: (2024)
by: Lian, Wei, et al.
Published: (2024)
Ego-centric Predictive Model Conditioned on Hand Trajectories
by: Zhang, Binjie, et al.
Published: (2025)
by: Zhang, Binjie, et al.
Published: (2025)
Hand-Centric Motion Refinement for 3D Hand-Object Interaction via Hierarchical Spatial-Temporal Modeling
by: Hao, Yuze, et al.
Published: (2024)
by: Hao, Yuze, et al.
Published: (2024)
Similar Items
-
Mitigating Image Captioning Hallucinations in Vision-Language Models
by: Zhao, Fei, et al.
Published: (2025) -
BeyondPixels: A Comprehensive Review of the Evolution of Neural Radiance Fields
by: Rabby, AKM Shahariar Azad, et al.
Published: (2023) -
Translation-based Video-to-Video Synthesis
by: Saha, Pratim, et al.
Published: (2024) -
Multi-Granularity Hand Action Detection
by: Zhe, Ting, et al.
Published: (2023) -
Driving with Context: Online Map Matching for Complex Roads Using Lane Markings and Scenario Recognition
by: Bi, Xin, et al.
Published: (2025)