Detecting Informative Channels: ActionFormer
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhao, Kunpeng, Miyazaki, Asahi, Okita, Tsuyoshi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Thoughts on Objectives of Sparse and Hierarchical Masked Image Model
por: Miyazaki, Asahi, et al.
Publicado: (2025)
por: Miyazaki, Asahi, et al.
Publicado: (2025)
Multi-instance Learning as Downstream Task of Self-Supervised Learning-based Pre-trained Model
por: Matsuishi, Koki, et al.
Publicado: (2025)
por: Matsuishi, Koki, et al.
Publicado: (2025)
Brain Hematoma Marker Recognition Using Multitask Learning: SwinTransformer and Swin-Unet
por: Hirata, Kodai, et al.
Publicado: (2025)
por: Hirata, Kodai, et al.
Publicado: (2025)
Diffusion Model-based Activity Completion for AI Motion Capture from Videos
por: Huayu, Gao, et al.
Publicado: (2025)
por: Huayu, Gao, et al.
Publicado: (2025)
Multimodal Foundation Model for Cross-Modal Retrieval and Activity Recognition Tasks
por: Matsuishi, Koki, et al.
Publicado: (2025)
por: Matsuishi, Koki, et al.
Publicado: (2025)
Image Classification Using a Diffusion Model as a Pre-Training Model
por: Ukita, Kosuke, et al.
Publicado: (2025)
por: Ukita, Kosuke, et al.
Publicado: (2025)
Window to Wall Ratio Detection using SegFormer
por: De Simone, Zoe, et al.
Publicado: (2024)
por: De Simone, Zoe, et al.
Publicado: (2024)
ParFormer: A Vision Transformer with Parallel Mixer and Sparse Channel Attention Patch Embedding
por: Setyawan, Novendra, et al.
Publicado: (2024)
por: Setyawan, Novendra, et al.
Publicado: (2024)
AnchorFormer: Differentiable Anchor Attention for Efficient Vision Transformer
por: Shan, Jiquan, et al.
Publicado: (2025)
por: Shan, Jiquan, et al.
Publicado: (2025)
Towards Robust Nonlinear Subspace Clustering: A Kernel Learning Approach
por: Xu, Kunpeng, et al.
Publicado: (2025)
por: Xu, Kunpeng, et al.
Publicado: (2025)
Uncertainty-Aware Global-View Reconstruction for Multi-View Multi-Label Feature Selection
por: Hao, Pingting, et al.
Publicado: (2025)
por: Hao, Pingting, et al.
Publicado: (2025)
NavFormer: IGRF Forecasting in Moving Coordinate Frames
por: Hwang, Yoontae, et al.
Publicado: (2026)
por: Hwang, Yoontae, et al.
Publicado: (2026)
Group Relative Augmentation for Data Efficient Action Detection
por: Patel, Deep Anil, et al.
Publicado: (2025)
por: Patel, Deep Anil, et al.
Publicado: (2025)
ChromaFormer: A Scalable and Accurate Transformer Architecture for Land Cover Classification
por: Li, Mingshi, et al.
Publicado: (2025)
por: Li, Mingshi, et al.
Publicado: (2025)
BiEquiFormer: Bi-Equivariant Representations for Global Point Cloud Registration
por: Pertigkiozoglou, Stefanos, et al.
Publicado: (2024)
por: Pertigkiozoglou, Stefanos, et al.
Publicado: (2024)
MatFormer: Nested Transformer for Elastic Inference
por: Devvrit, et al.
Publicado: (2023)
por: Devvrit, et al.
Publicado: (2023)
StruSR: Structure-Aware Symbolic Regression with Physics-Informed Taylor Guidance
por: Gong, Yunpeng, et al.
Publicado: (2025)
por: Gong, Yunpeng, et al.
Publicado: (2025)
PerFormer: A Permutation Based Vision Transformer for Remaining Useful Life Prediction
por: Fan, Zhengyang, et al.
Publicado: (2025)
por: Fan, Zhengyang, et al.
Publicado: (2025)
FixationFormer: Direct Utilization of Expert Gaze Trajectories for Chest X-Ray Classification
por: Beckmann, Daniel, et al.
Publicado: (2026)
por: Beckmann, Daniel, et al.
Publicado: (2026)
DA-SegFormer: Damage-Aware Semantic Segmentation for Fine-Grained Disaster Assessment
por: Zhu, Kevin, et al.
Publicado: (2026)
por: Zhu, Kevin, et al.
Publicado: (2026)
SpecFormer: Guarding Vision Transformer Robustness via Maximum Singular Value Penalization
por: Hu, Xixu, et al.
Publicado: (2024)
por: Hu, Xixu, et al.
Publicado: (2024)
GeoFormer: A Vision and Sequence Transformer-based Approach for Greenhouse Gas Monitoring
por: Khirwar, Madhav, et al.
Publicado: (2024)
por: Khirwar, Madhav, et al.
Publicado: (2024)
X-Former: Unifying Contrastive and Reconstruction Learning for MLLMs
por: Swetha, Sirnam, et al.
Publicado: (2024)
por: Swetha, Sirnam, et al.
Publicado: (2024)
P2ANet: A Dataset and Benchmark for Dense Action Detection from Table Tennis Match Broadcasting Videos
por: Bian, Jiang, et al.
Publicado: (2022)
por: Bian, Jiang, et al.
Publicado: (2022)
MetaFormer Baselines for Vision
por: Yu, Weihao, et al.
Publicado: (2022)
por: Yu, Weihao, et al.
Publicado: (2022)
SegFormer Fine-Tuning with Dropout: Advancing Hair Artifact Removal in Skin Lesion Analysis
por: Saad, Asif Mohammed, et al.
Publicado: (2025)
por: Saad, Asif Mohammed, et al.
Publicado: (2025)
Task Relevance Is Not Local Replaceability: A Two-Axis View of Channel Information
por: Safaai, Houman, et al.
Publicado: (2026)
por: Safaai, Houman, et al.
Publicado: (2026)
Estimating Physical Information Consistency of Channel Data Augmentation for Remote Sensing Images
por: Burgert, Tom, et al.
Publicado: (2024)
por: Burgert, Tom, et al.
Publicado: (2024)
JARViS: Detecting Actions in Video Using Unified Actor-Scene Context Relation Modeling
por: Lee, Seok Hwan, et al.
Publicado: (2024)
por: Lee, Seok Hwan, et al.
Publicado: (2024)
Activation-Free Backbones for Image Recognition: Polynomial Alternatives within MetaFormer-Style Vision Models
por: Wang, Jeffrey, et al.
Publicado: (2026)
por: Wang, Jeffrey, et al.
Publicado: (2026)
Channel-Aware Probing for Multi-Channel Imaging
por: Marikkar, Umar, et al.
Publicado: (2026)
por: Marikkar, Umar, et al.
Publicado: (2026)
Soft-TransFormers for Continual Learning
por: Kang, Haeyong, et al.
Publicado: (2024)
por: Kang, Haeyong, et al.
Publicado: (2024)
Action-Agnostic Point-Level Supervision for Temporal Action Detection
por: Yoshida, Shuhei M., et al.
Publicado: (2024)
por: Yoshida, Shuhei M., et al.
Publicado: (2024)
Future-Proofing Class-Incremental Learning
por: Jodelet, Quentin, et al.
Publicado: (2024)
por: Jodelet, Quentin, et al.
Publicado: (2024)
Selective, Interpretable, and Motion Consistent Privacy Attribute Obfuscation for Action Recognition
por: Ilic, Filip, et al.
Publicado: (2024)
por: Ilic, Filip, et al.
Publicado: (2024)
Action Dubber: Timing Audible Actions via Inflectional Flow
por: Wan, Wenlong, et al.
Publicado: (2025)
por: Wan, Wenlong, et al.
Publicado: (2025)
AdaTreeFormer: Few Shot Domain Adaptation for Tree Counting from a Single High-Resolution Image
por: Amirkolaee, Hamed Amini, et al.
Publicado: (2024)
por: Amirkolaee, Hamed Amini, et al.
Publicado: (2024)
RenderFormer: Transformer-based Neural Rendering of Triangle Meshes with Global Illumination
por: Zeng, Chong, et al.
Publicado: (2025)
por: Zeng, Chong, et al.
Publicado: (2025)
ChA-MAEViT: Unifying Channel-Aware Masked Autoencoders and Multi-Channel Vision Transformers for Improved Cross-Channel Learning
por: Pham, Chau, et al.
Publicado: (2025)
por: Pham, Chau, et al.
Publicado: (2025)
Multi-State-Action Tokenisation in Decision Transformers for Multi-Discrete Action Spaces
por: Moodley, Perusha, et al.
Publicado: (2024)
por: Moodley, Perusha, et al.
Publicado: (2024)
Ejemplares similares
-
Thoughts on Objectives of Sparse and Hierarchical Masked Image Model
por: Miyazaki, Asahi, et al.
Publicado: (2025) -
Multi-instance Learning as Downstream Task of Self-Supervised Learning-based Pre-trained Model
por: Matsuishi, Koki, et al.
Publicado: (2025) -
Brain Hematoma Marker Recognition Using Multitask Learning: SwinTransformer and Swin-Unet
por: Hirata, Kodai, et al.
Publicado: (2025) -
Diffusion Model-based Activity Completion for AI Motion Capture from Videos
por: Huayu, Gao, et al.
Publicado: (2025) -
Multimodal Foundation Model for Cross-Modal Retrieval and Activity Recognition Tasks
por: Matsuishi, Koki, et al.
Publicado: (2025)