HOIST-Former: Hand-held Objects Identification, Segmentation, and Tracking in the Wild
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Narasimhaswamy, Supreeth, Nguyen, Huy Anh, Huang, Lihan, Hoai, Minh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HanDiffuser: Text-to-Image Generation With Realistic Hand Appearances
von: Narasimhaswamy, Supreeth, et al.
Veröffentlicht: (2024)
von: Narasimhaswamy, Supreeth, et al.
Veröffentlicht: (2024)
Detecting Precise Hand Touch Moments in Egocentric Video
von: Nguyen, Huy Anh, et al.
Veröffentlicht: (2026)
von: Nguyen, Huy Anh, et al.
Veröffentlicht: (2026)
Driver Attention Tracking and Analysis
von: Nguyen, Dat Viet Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen, Dat Viet Thanh, et al.
Veröffentlicht: (2024)
Count What You Want: Exemplar Identification and Few-shot Counting of Human Actions in the Wild
von: Huang, Yifeng, et al.
Veröffentlicht: (2023)
von: Huang, Yifeng, et al.
Veröffentlicht: (2023)
SwiftPie: Lightning-fast Subject-driven Image Personalization via One step Diffusion
von: Duong, Huy, et al.
Veröffentlicht: (2026)
von: Duong, Huy, et al.
Veröffentlicht: (2026)
ShapeFormer: Shape Prior Visible-to-Amodal Transformer-based Amodal Instance Segmentation
von: Tran, Minh, et al.
Veröffentlicht: (2024)
von: Tran, Minh, et al.
Veröffentlicht: (2024)
HORT: Monocular Hand-held Objects Reconstruction with Transformers
von: Chen, Zerui, et al.
Veröffentlicht: (2025)
von: Chen, Zerui, et al.
Veröffentlicht: (2025)
Detecting Omissions in Geographic Maps through Computer Vision
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2024)
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2024)
Any3DIS: Class-Agnostic 3D Instance Segmentation by 2D Mask Tracking
von: Nguyen, Phuc, et al.
Veröffentlicht: (2024)
von: Nguyen, Phuc, et al.
Veröffentlicht: (2024)
CountEx: Fine-Grained Counting via Exemplars and Exclusion
von: Huang, Yifeng, et al.
Veröffentlicht: (2026)
von: Huang, Yifeng, et al.
Veröffentlicht: (2026)
Can Current AI Models Count What We Mean, Not What They See? A Benchmark and Systematic Evaluation
von: Nguyen, Gia Khanh, et al.
Veröffentlicht: (2025)
von: Nguyen, Gia Khanh, et al.
Veröffentlicht: (2025)
MetaFormer-driven Encoding Network for Robust Medical Semantic Segmentation
von: Tran, Le-Anh, et al.
Veröffentlicht: (2026)
von: Tran, Le-Anh, et al.
Veröffentlicht: (2026)
MOHO: Learning Single-view Hand-held Object Reconstruction with Multi-view Occlusion-Aware Supervision
von: Zhang, Chenyangguang, et al.
Veröffentlicht: (2023)
von: Zhang, Chenyangguang, et al.
Veröffentlicht: (2023)
MacFormer: Semantic Segmentation with Fine Object Boundaries
von: Xu, Guoan, et al.
Veröffentlicht: (2024)
von: Xu, Guoan, et al.
Veröffentlicht: (2024)
Learning Explicit Contact for Implicit Reconstruction of Hand-held Objects from Monocular Images
von: Hu, Junxing, et al.
Veröffentlicht: (2023)
von: Hu, Junxing, et al.
Veröffentlicht: (2023)
Scribble-Supervised Medical Image Segmentation with Dynamic Teacher Switching and Hierarchical Consistency
von: Nguyen, Thanh-Huy, et al.
Veröffentlicht: (2026)
von: Nguyen, Thanh-Huy, et al.
Veröffentlicht: (2026)
Revisit Visual Prompt Tuning: The Expressiveness of Prompt Experts
von: Le, Minh, et al.
Veröffentlicht: (2025)
von: Le, Minh, et al.
Veröffentlicht: (2025)
BluRef: Unsupervised Image Deblurring with Dense-Matching References
von: Pham, Bang-Dang, et al.
Veröffentlicht: (2026)
von: Pham, Bang-Dang, et al.
Veröffentlicht: (2026)
ShapeGraFormer: GraFormer-Based Network for Hand-Object Reconstruction from a Single Depth Map
von: Aboukhadra, Ahmed Tawfik, et al.
Veröffentlicht: (2023)
von: Aboukhadra, Ahmed Tawfik, et al.
Veröffentlicht: (2023)
Blur2Blur: Blur Conversion for Unsupervised Image Deblurring on Unknown Domains
von: Pham, Bang-Dang, et al.
Veröffentlicht: (2024)
von: Pham, Bang-Dang, et al.
Veröffentlicht: (2024)
Dual Strategies for Test-Time Adaptation
von: Phuong, Nam Nguyen, et al.
Veröffentlicht: (2026)
von: Phuong, Nam Nguyen, et al.
Veröffentlicht: (2026)
GraspMamba: A Mamba-based Language-driven Grasp Detection Framework with Hierarchical Feature Learning
von: Nguyen, Huy Hoang, et al.
Veröffentlicht: (2024)
von: Nguyen, Huy Hoang, et al.
Veröffentlicht: (2024)
MoBind: Motion Binding for Fine-Grained IMU-Video Pose Alignment
von: Nguyen, Duc Duy, et al.
Veröffentlicht: (2026)
von: Nguyen, Duc Duy, et al.
Veröffentlicht: (2026)
CaptionFormer: Unified Segmentation, Tracking, and Captioning for Spatio-Temporal Objects
von: Fiastre, Gabriel, et al.
Veröffentlicht: (2025)
von: Fiastre, Gabriel, et al.
Veröffentlicht: (2025)
Label-Efficient Cross-Modality Generalization for Liver Segmentation in Multi-Phase MRI
von: Bui-Tran, Quang-Khai, et al.
Veröffentlicht: (2025)
von: Bui-Tran, Quang-Khai, et al.
Veröffentlicht: (2025)
ScriptHOI: Learning Scripted State Transitions for Open-Vocabulary Human-Object Interaction Detection
von: Nguyen, Minh Anh, et al.
Veröffentlicht: (2026)
von: Nguyen, Minh Anh, et al.
Veröffentlicht: (2026)
QORT-Former: Query-optimized Real-time Transformer for Understanding Two Hands Manipulating Objects
von: Ismayilzada, Elkhan, et al.
Veröffentlicht: (2025)
von: Ismayilzada, Elkhan, et al.
Veröffentlicht: (2025)
Improving Zero-Shot Object-Level Change Detection by Incorporating Visual Correspondence
von: Nguyen, Hung Huy, et al.
Veröffentlicht: (2025)
von: Nguyen, Hung Huy, et al.
Veröffentlicht: (2025)
PRS-Med: Position Reasoning Segmentation in Medical Imaging
von: Trinh, Quoc-Huy, et al.
Veröffentlicht: (2025)
von: Trinh, Quoc-Huy, et al.
Veröffentlicht: (2025)
YOLO-Former: YOLO Shakes Hand With ViT
von: Khoramdel, Javad, et al.
Veröffentlicht: (2024)
von: Khoramdel, Javad, et al.
Veröffentlicht: (2024)
SAM-EG: Segment Anything Model with Egde Guidance framework for efficient Polyp Segmentation
von: Trinh, Quoc-Huy, et al.
Veröffentlicht: (2024)
von: Trinh, Quoc-Huy, et al.
Veröffentlicht: (2024)
Online Adaptation for Implicit Object Tracking and Shape Reconstruction in the Wild
von: Ye, Jianglong, et al.
Veröffentlicht: (2021)
von: Ye, Jianglong, et al.
Veröffentlicht: (2021)
Find First, Track Next: Decoupling Identification and Propagation in Referring Video Object Segmentation
von: Cho, Suhwan, et al.
Veröffentlicht: (2025)
von: Cho, Suhwan, et al.
Veröffentlicht: (2025)
EasyHOI: Unleashing the Power of Large Models for Reconstructing Hand-Object Interactions in the Wild
von: Liu, Yumeng, et al.
Veröffentlicht: (2024)
von: Liu, Yumeng, et al.
Veröffentlicht: (2024)
Hand Held Multi-Object Tracking Dataset in American Football
von: Otsubo, Rintaro, et al.
Veröffentlicht: (2025)
von: Otsubo, Rintaro, et al.
Veröffentlicht: (2025)
ComPose: When to Trust Hands for Object Pose Tracking
von: Shin, Jisu, et al.
Veröffentlicht: (2026)
von: Shin, Jisu, et al.
Veröffentlicht: (2026)
Ego-Exo 3D Hand Tracking in the Wild with a Mobile Multi-Camera Rig
von: Rim, Patrick, et al.
Veröffentlicht: (2025)
von: Rim, Patrick, et al.
Veröffentlicht: (2025)
Beyond Traditional Approaches: Multi-Task Network for Breast Ultrasound Diagnosis
von: Chung, Dat T., et al.
Veröffentlicht: (2024)
von: Chung, Dat T., et al.
Veröffentlicht: (2024)
ChildPlay-Hand: A Dataset of Hand Manipulations in the Wild
von: Farkhondeh, Arya, et al.
Veröffentlicht: (2024)
von: Farkhondeh, Arya, et al.
Veröffentlicht: (2024)
SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA)
von: Nguyen, Trong-Thuan, et al.
Veröffentlicht: (2025)
von: Nguyen, Trong-Thuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
HanDiffuser: Text-to-Image Generation With Realistic Hand Appearances
von: Narasimhaswamy, Supreeth, et al.
Veröffentlicht: (2024) -
Detecting Precise Hand Touch Moments in Egocentric Video
von: Nguyen, Huy Anh, et al.
Veröffentlicht: (2026) -
Driver Attention Tracking and Analysis
von: Nguyen, Dat Viet Thanh, et al.
Veröffentlicht: (2024) -
Count What You Want: Exemplar Identification and Few-shot Counting of Human Actions in the Wild
von: Huang, Yifeng, et al.
Veröffentlicht: (2023) -
SwiftPie: Lightning-fast Subject-driven Image Personalization via One step Diffusion
von: Duong, Huy, et al.
Veröffentlicht: (2026)