PALM: A Dataset and Baseline for Learning Multi-subject Hand Prior
Fuente:
arXiv
Guardado en:
| Autores principales: | Fan, Zicong, Remelli, Edoardo, Dimond, David, Sener, Fadime, Ge, Liuhao, Tekin, Bugra, Keskin, Cem, Hampali, Shreyas |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DiffH2O: Diffusion-Based Synthesis of Hand-Object Interactions from Textual Descriptions
por: Christen, Sammy, et al.
Publicado: (2024)
por: Christen, Sammy, et al.
Publicado: (2024)
X-MIC: Cross-Modal Instance Conditioning for Egocentric Action Generalization
por: Kukleva, Anna, et al.
Publicado: (2024)
por: Kukleva, Anna, et al.
Publicado: (2024)
Memory-efficient Streaming VideoLLMs for Real-time Procedural Video Understanding
por: Chatterjee, Dibyadip, et al.
Publicado: (2025)
por: Chatterjee, Dibyadip, et al.
Publicado: (2025)
FoundHand: Large-Scale Domain-Specific Learning for Controllable Hand Image Generation
por: Chen, Kefan, et al.
Publicado: (2024)
por: Chen, Kefan, et al.
Publicado: (2024)
EgoPoseFormer: A Simple Baseline for Stereo Egocentric 3D Human Pose Estimation
por: Yang, Chenhongyi, et al.
Publicado: (2024)
por: Yang, Chenhongyi, et al.
Publicado: (2024)
GoTrack: Generic 6DoF Object Pose Refinement and Tracking
por: Nguyen, Van Nguyen, et al.
Publicado: (2025)
por: Nguyen, Van Nguyen, et al.
Publicado: (2025)
On the Utility of 3D Hand Poses for Action Recognition
por: Shamil, Md Salman, et al.
Publicado: (2024)
por: Shamil, Md Salman, et al.
Publicado: (2024)
FoundPose: Unseen Object Pose Estimation with Foundation Features
por: Örnek, Evin Pınar, et al.
Publicado: (2023)
por: Örnek, Evin Pınar, et al.
Publicado: (2023)
Context-Enhanced Memory-Refined Transformer for Online Action Detection
por: Pang, Zhanzhong, et al.
Publicado: (2025)
por: Pang, Zhanzhong, et al.
Publicado: (2025)
Cost-Sensitive Learning for Long-Tailed Temporal Action Segmentation
por: Pang, Zhanzhong, et al.
Publicado: (2025)
por: Pang, Zhanzhong, et al.
Publicado: (2025)
Introducing HOT3D: An Egocentric Dataset for 3D Hand and Object Tracking
por: Banerjee, Prithviraj, et al.
Publicado: (2024)
por: Banerjee, Prithviraj, et al.
Publicado: (2024)
Decouple and Cache: KV Cache Construction for Streaming Video Understanding
por: Pang, Zhanzhong, et al.
Publicado: (2026)
por: Pang, Zhanzhong, et al.
Publicado: (2026)
Long-Tail Temporal Action Segmentation with Group-wise Temporal Logit Adjustment
por: Pang, Zhanzhong, et al.
Publicado: (2024)
por: Pang, Zhanzhong, et al.
Publicado: (2024)
On Discriminative vs. Generative classifiers: Rethinking MLLMs for Action Understanding
por: Pang, Zhanzhong, et al.
Publicado: (2026)
por: Pang, Zhanzhong, et al.
Publicado: (2026)
MagicHOI: Leveraging 3D Priors for Accurate Hand-object Reconstruction from Short Monocular Video Clips
por: Wang, Shibo, et al.
Publicado: (2025)
por: Wang, Shibo, et al.
Publicado: (2025)
Geometric Neural Distance Fields for Learning Human Motion Priors
por: Yu, Zhengdi, et al.
Publicado: (2025)
por: Yu, Zhengdi, et al.
Publicado: (2025)
Don't Pause! Every prediction matters in a streaming video
por: Chatterjee, Dibyadip, et al.
Publicado: (2026)
por: Chatterjee, Dibyadip, et al.
Publicado: (2026)
HOT3D: Hand and Object Tracking in 3D from Egocentric Multi-View Videos
por: Banerjee, Prithviraj, et al.
Publicado: (2024)
por: Banerjee, Prithviraj, et al.
Publicado: (2024)
HOIGPT: Learning Long Sequence Hand-Object Interaction with Language Models
por: Huang, Mingzhen, et al.
Publicado: (2025)
por: Huang, Mingzhen, et al.
Publicado: (2025)
CigTime: Corrective Instruction Generation Through Inverse Motion Editing
por: Fang, Qihang, et al.
Publicado: (2024)
por: Fang, Qihang, et al.
Publicado: (2024)
HuMoCon: Concept Discovery for Human Motion Understanding
por: Fang, Qihang, et al.
Publicado: (2025)
por: Fang, Qihang, et al.
Publicado: (2025)
A Simple Baseline for Efficient Hand Mesh Reconstruction
por: Zhou, Zhishan, et al.
Publicado: (2024)
por: Zhou, Zhishan, et al.
Publicado: (2024)
Oral Imaging for Malocclusion Issues Assessments: OMNI Dataset, Deep Learning Baselines and Benchmarking
por: Xue, Pujun, et al.
Publicado: (2025)
por: Xue, Pujun, et al.
Publicado: (2025)
VTON-HandFit: Virtual Try-on for Arbitrary Hand Pose Guided by Hand Priors Embedding
por: Liang, Yujie, et al.
Publicado: (2024)
por: Liang, Yujie, et al.
Publicado: (2024)
SneakPeek: Future-Guided Instructional Streaming Video Generation
por: Hong, Cheeun, et al.
Publicado: (2025)
por: Hong, Cheeun, et al.
Publicado: (2025)
Enhancing Monocular 3D Hand Reconstruction with Learned Texture Priors
por: Karvounas, Giorgos, et al.
Publicado: (2025)
por: Karvounas, Giorgos, et al.
Publicado: (2025)
PALM: Predicting Actions through Language Models
por: Kim, Sanghwan, et al.
Publicado: (2023)
por: Kim, Sanghwan, et al.
Publicado: (2023)
Hand Held Multi-Object Tracking Dataset in American Football
por: Otsubo, Rintaro, et al.
Publicado: (2025)
por: Otsubo, Rintaro, et al.
Publicado: (2025)
GeoHand: Unlocking Prior Geometry Knowledge for Monocular 3D Hand Reconstruction
por: Lin, Weiquan, et al.
Publicado: (2026)
por: Lin, Weiquan, et al.
Publicado: (2026)
ChildPlay-Hand: A Dataset of Hand Manipulations in the Wild
por: Farkhondeh, Arya, et al.
Publicado: (2024)
por: Farkhondeh, Arya, et al.
Publicado: (2024)
Cost-aware LLM-based Online Dataset Annotation
por: Elumar, Eray Can, et al.
Publicado: (2025)
por: Elumar, Eray Can, et al.
Publicado: (2025)
PALM: Pushing Adaptive Learning Rate Mechanisms for Continual Test-Time Adaptation
por: Maharana, Sarthak Kumar, et al.
Publicado: (2024)
por: Maharana, Sarthak Kumar, et al.
Publicado: (2024)
Anymate: A Dataset and Baselines for Learning 3D Object Rigging
por: Deng, Yufan, et al.
Publicado: (2025)
por: Deng, Yufan, et al.
Publicado: (2025)
GigaHands: A Massive Annotated Dataset of Bimanual Hand Activities
por: Fu, Rao, et al.
Publicado: (2024)
por: Fu, Rao, et al.
Publicado: (2024)
TUMTraf EMOT: Event-Based Multi-Object Tracking Dataset and Baseline for Traffic Scenarios
por: Li, Mengyu, et al.
Publicado: (2025)
por: Li, Mengyu, et al.
Publicado: (2025)
Compressed Feature Quality Assessment: Dataset and Baselines
por: Gao, Changsheng, et al.
Publicado: (2025)
por: Gao, Changsheng, et al.
Publicado: (2025)
Power Line Aerial Image Restoration under dverse Weather: Datasets and Baselines
por: Yang, Sai, et al.
Publicado: (2024)
por: Yang, Sai, et al.
Publicado: (2024)
A Dataset and Baseline for Deep Learning-Based Visual Quality Inspection in Remanufacturing
por: Bauer, Johannes C., et al.
Publicado: (2025)
por: Bauer, Johannes C., et al.
Publicado: (2025)
Affordance-Guided Diffusion Prior for 3D Hand Reconstruction
por: Suzuki, Naru, et al.
Publicado: (2025)
por: Suzuki, Naru, et al.
Publicado: (2025)
Quantifying Nematodes through Images: Datasets, Models, and Baselines of Deep Learning
por: Yuan, Zhipeng, et al.
Publicado: (2024)
por: Yuan, Zhipeng, et al.
Publicado: (2024)
Ejemplares similares
-
DiffH2O: Diffusion-Based Synthesis of Hand-Object Interactions from Textual Descriptions
por: Christen, Sammy, et al.
Publicado: (2024) -
X-MIC: Cross-Modal Instance Conditioning for Egocentric Action Generalization
por: Kukleva, Anna, et al.
Publicado: (2024) -
Memory-efficient Streaming VideoLLMs for Real-time Procedural Video Understanding
por: Chatterjee, Dibyadip, et al.
Publicado: (2025) -
FoundHand: Large-Scale Domain-Specific Learning for Controllable Hand Image Generation
por: Chen, Kefan, et al.
Publicado: (2024) -
EgoPoseFormer: A Simple Baseline for Stereo Egocentric 3D Human Pose Estimation
por: Yang, Chenhongyi, et al.
Publicado: (2024)