Less is More: Decoder-Free Masked Modeling for Efficient Skeleton Representation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Do, Jeonghyeok, Chen, Yun, Youk, Geunhyuk, Kim, Munchurl |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PAN-Crafter: Learning Modality-Consistent Alignment for PAN-Sharpening
by: Do, Jeonghyeok, et al.
Published: (2025)
by: Do, Jeonghyeok, et al.
Published: (2025)
Bridging the Skeleton-Text Modality Gap: Diffusion-Powered Modality Alignment for Zero-shot Skeleton-based Action Recognition
by: Do, Jeonghyeok, et al.
Published: (2024)
by: Do, Jeonghyeok, et al.
Published: (2024)
FMA-Net: Flow-Guided Dynamic Filtering and Iterative Feature Refinement with Multi-Attention for Joint Video Super-Resolution and Deblurring
by: Youk, Geunhyuk, et al.
Published: (2024)
by: Youk, Geunhyuk, et al.
Published: (2024)
FMA-Net++: Motion- and Exposure-Aware Real-World Joint Video Super-Resolution and Deblurring
by: Youk, Geunhyuk, et al.
Published: (2025)
by: Youk, Geunhyuk, et al.
Published: (2025)
SkateFormer: Skeletal-Temporal Transformer for Human Action Recognition
by: Do, Jeonghyeok, et al.
Published: (2024)
by: Do, Jeonghyeok, et al.
Published: (2024)
C-DiffSET: Leveraging Latent Diffusion for SAR-to-EO Image Translation with Confidence-Guided Reliable Object Generation
by: Do, Jeonghyeok, et al.
Published: (2024)
by: Do, Jeonghyeok, et al.
Published: (2024)
U-Know-DiffPAN: An Uncertainty-aware Knowledge Distillation Diffusion Framework with Details Enhancement for PAN-Sharpening
by: Kim, Sungpyo, et al.
Published: (2024)
by: Kim, Sungpyo, et al.
Published: (2024)
OmniText: A Training-Free Generalist for Controllable Text-Image Manipulation
by: Gunawan, Agus, et al.
Published: (2025)
by: Gunawan, Agus, et al.
Published: (2025)
Less is More: Efficient Point Cloud Reconstruction via Multi-Head Decoders
by: Alonso, Pedro, et al.
Published: (2025)
by: Alonso, Pedro, et al.
Published: (2025)
Towards Efficient General Feature Prediction in Masked Skeleton Modeling
by: Sun, Shengkai, et al.
Published: (2025)
by: Sun, Shengkai, et al.
Published: (2025)
One Look is Enough: Seamless Patchwise Refinement for Zero-Shot Monocular Depth Estimation on High-Resolution Images
by: Kwon, Byeongjun, et al.
Published: (2025)
by: Kwon, Byeongjun, et al.
Published: (2025)
Achieving More with Less: Additive Prompt Tuning for Rehearsal-Free Class-Incremental Learning
by: Chen, Haoran, et al.
Published: (2025)
by: Chen, Haoran, et al.
Published: (2025)
See More, Store Less: Memory-Efficient Resolution for Video Moment Retrieval
by: Jeon, Mingyu, et al.
Published: (2026)
by: Jeon, Mingyu, et al.
Published: (2026)
ASMa: Asymmetric Spatio-temporal Masking for Skeleton Action Representation Learning
by: Anand, Aman, et al.
Published: (2026)
by: Anand, Aman, et al.
Published: (2026)
Less is More: Masking Elements in Image Condition Features Avoids Content Leakages in Style Transfer Diffusion Models
by: Zhu, Lin, et al.
Published: (2025)
by: Zhu, Lin, et al.
Published: (2025)
Mask and Compress: Efficient Skeleton-based Action Recognition in Continual Learning
by: Mosconi, Matteo, et al.
Published: (2024)
by: Mosconi, Matteo, et al.
Published: (2024)
PropFly: Learning to Propagate via On-the-Fly Supervision from Pre-trained Video Diffusion Models
by: Seo, Wonyong, et al.
Published: (2026)
by: Seo, Wonyong, et al.
Published: (2026)
BiM-VFI: Bidirectional Motion Field-Guided Frame Interpolation for Video with Non-uniform Motions
by: Seo, Wonyong, et al.
Published: (2024)
by: Seo, Wonyong, et al.
Published: (2024)
WebSpline: Structure-Informed Splines for Real-Time 3D Gaussians from Monocular Videos
by: Park, Jongmin, et al.
Published: (2026)
by: Park, Jongmin, et al.
Published: (2026)
MotionGrounder: Grounded Multi-Object Motion Transfer via Diffusion Transformer
by: Teodoro, Samuel, et al.
Published: (2026)
by: Teodoro, Samuel, et al.
Published: (2026)
Less is More: Improving Motion Diffusion Models with Sparse Keyframes
by: Bae, Jinseok, et al.
Published: (2025)
by: Bae, Jinseok, et al.
Published: (2025)
Learning More by Seeing Less: Structure First Learning for Efficient, Transferable, and Human-Aligned Vision
by: Li, Tianqin, et al.
Published: (2025)
by: Li, Tianqin, et al.
Published: (2025)
Send Less, Perceive More: Masked Quantized Point Cloud Communication for Loss-Tolerant Collaborative Perception
by: Xu, Sheng, et al.
Published: (2026)
by: Xu, Sheng, et al.
Published: (2026)
Heterogeneous Skeleton-Based Action Representation Learning
by: Wang, Hongsong, et al.
Published: (2025)
by: Wang, Hongsong, et al.
Published: (2025)
AA-Splat: Anti-Aliased Feed-forward Gaussian Splatting
by: Suh, Taewoo, et al.
Published: (2026)
by: Suh, Taewoo, et al.
Published: (2026)
Technical Report: Masked Skeleton Sequence Modeling for Learning Larval Zebrafish Behavior Latent Embeddings
by: Xu, Lanxin, et al.
Published: (2024)
by: Xu, Lanxin, et al.
Published: (2024)
Diffusion-based Data Augmentation and Knowledge Distillation with Generated Soft Labels Solving Data Scarcity Problems of SAR Oil Spill Segmentation
by: Moon, Jaeho, et al.
Published: (2024)
by: Moon, Jaeho, et al.
Published: (2024)
Skeleton-to-Image Encoding: Enabling Skeleton Representation Learning via Vision-Pretrained Models
by: Yang, Siyuan, et al.
Published: (2026)
by: Yang, Siyuan, et al.
Published: (2026)
Less is More: Token Context-aware Learning for Object Tracking
by: Xu, Chenlong, et al.
Published: (2025)
by: Xu, Chenlong, et al.
Published: (2025)
Multi-Modality Co-Learning for Efficient Skeleton-based Action Recognition
by: Liu, Jinfu, et al.
Published: (2024)
by: Liu, Jinfu, et al.
Published: (2024)
Wavelet-Driven Masked Image Modeling: A Path to Efficient Visual Representation
by: Xiang, Wenzhao, et al.
Published: (2025)
by: Xiang, Wenzhao, et al.
Published: (2025)
Skeleton2vec: A Self-supervised Learning Framework with Contextualized Target Representations for Skeleton Sequence
by: Xu, Ruizhuo, et al.
Published: (2024)
by: Xu, Ruizhuo, et al.
Published: (2024)
Skeleton-in-Context: Unified Skeleton Sequence Modeling with In-Context Learning
by: Wang, Xinshun, et al.
Published: (2023)
by: Wang, Xinshun, et al.
Published: (2023)
SECOND: Mitigating Perceptual Hallucination in Vision-Language Models via Selective and Contrastive Decoding
by: Park, Woohyeon, et al.
Published: (2025)
by: Park, Woohyeon, et al.
Published: (2025)
AirSplat: Alignment and Rating for Robust Feed-Forward 3D Gaussian Splatting
by: Bui, Minh-Quan Viet, et al.
Published: (2026)
by: Bui, Minh-Quan Viet, et al.
Published: (2026)
Sense Less, Generate More: Pre-training LiDAR Perception with Masked Autoencoders for Ultra-Efficient 3D Sensing
by: Tayebati, Sina, et al.
Published: (2024)
by: Tayebati, Sina, et al.
Published: (2024)
MaskDiff: Modeling Mask Distribution with Diffusion Probabilistic Model for Few-Shot Instance Segmentation
by: Le, Minh-Quan, et al.
Published: (2023)
by: Le, Minh-Quan, et al.
Published: (2023)
Seeing More with Less: Video Capsule Endoscopy with Multi-Task Learning
by: Werner, Julia, et al.
Published: (2025)
by: Werner, Julia, et al.
Published: (2025)
Less is More: Token-Efficient Video-QA via Adaptive Frame-Pruning and Semantic Graph Integration
by: Wang, Shaoguang, et al.
Published: (2025)
by: Wang, Shaoguang, et al.
Published: (2025)
Less is More: Discovering Concise Network Explanations
by: Kondapaneni, Neehar, et al.
Published: (2024)
by: Kondapaneni, Neehar, et al.
Published: (2024)
Similar Items
-
PAN-Crafter: Learning Modality-Consistent Alignment for PAN-Sharpening
by: Do, Jeonghyeok, et al.
Published: (2025) -
Bridging the Skeleton-Text Modality Gap: Diffusion-Powered Modality Alignment for Zero-shot Skeleton-based Action Recognition
by: Do, Jeonghyeok, et al.
Published: (2024) -
FMA-Net: Flow-Guided Dynamic Filtering and Iterative Feature Refinement with Multi-Attention for Joint Video Super-Resolution and Deblurring
by: Youk, Geunhyuk, et al.
Published: (2024) -
FMA-Net++: Motion- and Exposure-Aware Real-World Joint Video Super-Resolution and Deblurring
by: Youk, Geunhyuk, et al.
Published: (2025) -
SkateFormer: Skeletal-Temporal Transformer for Human Action Recognition
by: Do, Jeonghyeok, et al.
Published: (2024)