Enregistré dans:
| Auteurs principaux: | Sun, Shengkai, Zhang, Zefan, Dong, Jianfeng, Cheng, Zhiyong, Chang, Xiaojun, Wang, Meng |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2509.03609 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
VERHallu: Evaluating and Mitigating Event Relation Hallucination in Video Large Language Models
par: Zhang, Zefan, et autres
Publié: (2026)
par: Zhang, Zefan, et autres
Publié: (2026)
Less is More: Decoder-Free Masked Modeling for Efficient Skeleton Representation Learning
par: Do, Jeonghyeok, et autres
Publié: (2026)
par: Do, Jeonghyeok, et autres
Publié: (2026)
Reinforcement Learning Meets Masked Generative Models: Mask-GRPO for Text-to-Image Generation
par: Luo, Yifu, et autres
Publié: (2025)
par: Luo, Yifu, et autres
Publié: (2025)
Motion Keyframe Interpolation for Any Human Skeleton via Temporally Consistent Point Cloud Sampling and Reconstruction
par: Mo, Clinton, et autres
Publié: (2024)
par: Mo, Clinton, et autres
Publié: (2024)
WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens
par: Wang, Xiaofeng, et autres
Publié: (2024)
par: Wang, Xiaofeng, et autres
Publié: (2024)
SkeletonX: Data-Efficient Skeleton-based Action Recognition via Cross-sample Feature Aggregation
par: Zhang, Zongye, et autres
Publié: (2025)
par: Zhang, Zongye, et autres
Publié: (2025)
MacDiff: Unified Skeleton Modeling with Masked Conditional Diffusion
par: Wu, Lehong, et autres
Publié: (2024)
par: Wu, Lehong, et autres
Publié: (2024)
Plug-and-Play Context Feature Reuse for Efficient Masked Generation
par: Liu, Xuejie, et autres
Publié: (2025)
par: Liu, Xuejie, et autres
Publié: (2025)
Technical Report: Masked Skeleton Sequence Modeling for Learning Larval Zebrafish Behavior Latent Embeddings
par: Xu, Lanxin, et autres
Publié: (2024)
par: Xu, Lanxin, et autres
Publié: (2024)
Mask and Compress: Efficient Skeleton-based Action Recognition in Continual Learning
par: Mosconi, Matteo, et autres
Publié: (2024)
par: Mosconi, Matteo, et autres
Publié: (2024)
From Prototypes to General Distributions: An Efficient Curriculum for Masked Image Modeling
par: Lin, Jinhong, et autres
Publié: (2024)
par: Lin, Jinhong, et autres
Publié: (2024)
PUMPS: Skeleton-Agnostic Point-based Universal Motion Pre-Training for Synthesis in Human Motion Tasks
par: Mo, Clinton Ansun, et autres
Publié: (2025)
par: Mo, Clinton Ansun, et autres
Publié: (2025)
Skeleton and Font Generation Network for Zero-shot Chinese Character Generation
par: Xue, Mobai, et autres
Publié: (2025)
par: Xue, Mobai, et autres
Publié: (2025)
X-WIN: Building Chest Radiograph World Model via Predictive Sensing
par: Yang, Zefan, et autres
Publié: (2025)
par: Yang, Zefan, et autres
Publié: (2025)
Deep Feature Surgery: Towards Accurate and Efficient Multi-Exit Networks
par: Gong, Cheng, et autres
Publié: (2024)
par: Gong, Cheng, et autres
Publié: (2024)
SkeletonAgent: An Agentic Interaction Framework for Skeleton-based Action Recognition
par: Liu, Hongda, et autres
Publié: (2025)
par: Liu, Hongda, et autres
Publié: (2025)
Skeleton-Snippet Contrastive Learning with Multiscale Feature Fusion for Action Localization
par: Cheng, Qiushuo, et autres
Publié: (2025)
par: Cheng, Qiushuo, et autres
Publié: (2025)
Siformer: Feature-isolated Transformer for Efficient Skeleton-based Sign Language Recognition
par: Pu, Muxin, et autres
Publié: (2025)
par: Pu, Muxin, et autres
Publié: (2025)
Generative Data Augmentation for Skeleton Action Recognition
par: Dong, Xu, et autres
Publié: (2026)
par: Dong, Xu, et autres
Publié: (2026)
Towards Universal Skeleton-Based Action Recognition
par: Kuang, Jidong, et autres
Publié: (2026)
par: Kuang, Jidong, et autres
Publié: (2026)
Skeleton-in-Context: Unified Skeleton Sequence Modeling with In-Context Learning
par: Wang, Xinshun, et autres
Publié: (2023)
par: Wang, Xinshun, et autres
Publié: (2023)
MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation
par: Chen, Wenchao, et autres
Publié: (2024)
par: Chen, Wenchao, et autres
Publié: (2024)
Significance of Skeleton-based Features in Virtual Try-On
par: Roy, Debapriya, et autres
Publié: (2022)
par: Roy, Debapriya, et autres
Publié: (2022)
Towards Fine-Grained Emotion Understanding via Skeleton-Based Micro-Gesture Recognition
par: Xu, Hao, et autres
Publié: (2025)
par: Xu, Hao, et autres
Publié: (2025)
Rethinking UMM Visual Generation: Masked Modeling for Efficient Image-Only Pre-training
par: Sun, Peng, et autres
Publié: (2026)
par: Sun, Peng, et autres
Publié: (2026)
Boosting Adversarial Transferability for Skeleton-based Action Recognition via Exploring the Model Posterior Space
par: Diao, Yunfeng, et autres
Publié: (2024)
par: Diao, Yunfeng, et autres
Publié: (2024)
Memory Efficient Transformer Adapter for Dense Predictions
par: Zhang, Dong, et autres
Publié: (2025)
par: Zhang, Dong, et autres
Publié: (2025)
ALTo: Adaptive-Length Tokenizer for Autoregressive Mask Generation
par: Wang, Lingfeng, et autres
Publié: (2025)
par: Wang, Lingfeng, et autres
Publié: (2025)
Soften the Mask: Adaptive Temporal Soft Mask for Efficient Dynamic Facial Expression Recognition
par: Li, Meng-zhu, et autres
Publié: (2025)
par: Li, Meng-zhu, et autres
Publié: (2025)
Effective and Efficient Masked Image Generation Models
par: You, Zebin, et autres
Publié: (2025)
par: You, Zebin, et autres
Publié: (2025)
Language-Assisted Human Part Motion Learning for Skeleton-Based Temporal Action Segmentation
par: Chen, Bowen, et autres
Publié: (2024)
par: Chen, Bowen, et autres
Publié: (2024)
Unified Static and Dynamic Network: Efficient Temporal Filtering for Video Grounding
par: Hu, Jingjing, et autres
Publié: (2024)
par: Hu, Jingjing, et autres
Publié: (2024)
InterMask: 3D Human Interaction Generation via Collaborative Masked Modeling
par: Javed, Muhammad Gohar, et autres
Publié: (2024)
par: Javed, Muhammad Gohar, et autres
Publié: (2024)
Token Painter: Training-Free Text-Guided Image Inpainting via Mask Autoregressive Models
par: Jiang, Longtao, et autres
Publié: (2025)
par: Jiang, Longtao, et autres
Publié: (2025)
TopoVST: Toward Topology-fidelitous Vessel Skeleton Tracking
par: Liu, Yaoyu, et autres
Publié: (2026)
par: Liu, Yaoyu, et autres
Publié: (2026)
A Unified Masked Autoencoder with Patchified Skeletons for Motion Synthesis
par: Mascaro, Esteve Valls, et autres
Publié: (2023)
par: Mascaro, Esteve Valls, et autres
Publié: (2023)
OmniMotion: Multimodal Motion Generation with Continuous Masked Autoregression
par: Li, Zhe, et autres
Publié: (2025)
par: Li, Zhe, et autres
Publié: (2025)
From Ideal to Real: Unified and Data-Efficient Dense Prediction for Real-World Scenarios
par: Xia, Changliang, et autres
Publié: (2025)
par: Xia, Changliang, et autres
Publié: (2025)
Delta Attention Residuals
par: Luo, Cheng, et autres
Publié: (2026)
par: Luo, Cheng, et autres
Publié: (2026)
MENTOR: Efficient Multimodal-Conditioned Tuning for Autoregressive Vision Generation Models
par: Zhao, Haozhe, et autres
Publié: (2025)
par: Zhao, Haozhe, et autres
Publié: (2025)
Documents similaires
-
VERHallu: Evaluating and Mitigating Event Relation Hallucination in Video Large Language Models
par: Zhang, Zefan, et autres
Publié: (2026) -
Less is More: Decoder-Free Masked Modeling for Efficient Skeleton Representation Learning
par: Do, Jeonghyeok, et autres
Publié: (2026) -
Reinforcement Learning Meets Masked Generative Models: Mask-GRPO for Text-to-Image Generation
par: Luo, Yifu, et autres
Publié: (2025) -
Motion Keyframe Interpolation for Any Human Skeleton via Temporally Consistent Point Cloud Sampling and Reconstruction
par: Mo, Clinton, et autres
Publié: (2024) -
WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens
par: Wang, Xiaofeng, et autres
Publié: (2024)