MPT-PAR:Mix-Parameters Transformer for Panoramic Activity Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gan, Wenqing, Sun, Yan, Liu, Feiran, Luo, Xiangfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MPT: A Large-scale Multi-Phytoplankton Tracking Benchmark
von: Yu, Yang, et al.
Veröffentlicht: (2024)
von: Yu, Yang, et al.
Veröffentlicht: (2024)
UniPAR: A Unified Framework for Pedestrian Attribute Recognition
von: Xu, Minghe, et al.
Veröffentlicht: (2026)
von: Xu, Minghe, et al.
Veröffentlicht: (2026)
Logi-PAR: Logic-Infused Patient Activity Recognition via Differentiable Rule
von: Zarar, Muhammad, et al.
Veröffentlicht: (2026)
von: Zarar, Muhammad, et al.
Veröffentlicht: (2026)
Feature Fusion for Human Activity Recognition using Parameter-Optimized Multi-Stage Graph Convolutional Network and Transformer Models
von: Belal, Mohammad, et al.
Veröffentlicht: (2024)
von: Belal, Mohammad, et al.
Veröffentlicht: (2024)
VLM-PAR: A Vision Language Model for Pedestrian Attribute Recognition
von: Sellam, Abdellah Zakaria, et al.
Veröffentlicht: (2025)
von: Sellam, Abdellah Zakaria, et al.
Veröffentlicht: (2025)
A Lightweight Transformer for Pain Recognition from Brain Activity
von: Gkikas, Stefanos, et al.
Veröffentlicht: (2026)
von: Gkikas, Stefanos, et al.
Veröffentlicht: (2026)
Hybrid Generative Fusion for Efficient and Privacy-Preserving Face Recognition Dataset Generation
von: Li, Feiran, et al.
Veröffentlicht: (2025)
von: Li, Feiran, et al.
Veröffentlicht: (2025)
PanoTPS-Net: Panoramic Room Layout Estimation via Thin Plate Spline Transformation
von: Ibrahem, Hatem, et al.
Veröffentlicht: (2025)
von: Ibrahem, Hatem, et al.
Veröffentlicht: (2025)
ViTA-PAR: Visual and Textual Attribute Alignment with Attribute Prompting for Pedestrian Attribute Recognition
von: Park, Minjeong, et al.
Veröffentlicht: (2025)
von: Park, Minjeong, et al.
Veröffentlicht: (2025)
Non-stationary BERT: Exploring Augmented IMU Data For Robust Human Activity Recognition
von: Sun, Ning, et al.
Veröffentlicht: (2024)
von: Sun, Ning, et al.
Veröffentlicht: (2024)
From Narrow to Panoramic Vision: Attention-Guided Cold-Start Reshapes Multimodal Reasoning
von: Luo, Ruilin, et al.
Veröffentlicht: (2026)
von: Luo, Ruilin, et al.
Veröffentlicht: (2026)
Enhancing Floor Plan Recognition: A Hybrid Mix-Transformer and U-Net Approach for Precise Wall Segmentation
von: Parashchuk, Dmitriy, et al.
Veröffentlicht: (2025)
von: Parashchuk, Dmitriy, et al.
Veröffentlicht: (2025)
Fraesormer: Learning Adaptive Sparse Transformer for Efficient Food Recognition
von: Zou, Shun, et al.
Veröffentlicht: (2025)
von: Zou, Shun, et al.
Veröffentlicht: (2025)
PanoWorld: Geometry-Consistent Panoramic Video World Modeling
von: Jiang, Le, et al.
Veröffentlicht: (2026)
von: Jiang, Le, et al.
Veröffentlicht: (2026)
OPGAgent: An Agent for Auditable Dental Panoramic X-ray Interpretation
von: Yu, Zhaolin, et al.
Veröffentlicht: (2026)
von: Yu, Zhaolin, et al.
Veröffentlicht: (2026)
Unified Panoramic Geometry Estimation via Multi-View Foundation Models
von: Bozic, Vukasin, et al.
Veröffentlicht: (2026)
von: Bozic, Vukasin, et al.
Veröffentlicht: (2026)
Texo: Formula Recognition within 20M Parameters
von: Mao, Sicheng
Veröffentlicht: (2026)
von: Mao, Sicheng
Veröffentlicht: (2026)
TransFace++: Rethinking the Face Recognition Paradigm with a Focus on Accuracy, Efficiency, and Security
von: Dan, Jun, et al.
Veröffentlicht: (2023)
von: Dan, Jun, et al.
Veröffentlicht: (2023)
PAR: Prompt-Aware Token Reduction Method for Efficient Large Multimodal Models
von: Liu, Yingen, et al.
Veröffentlicht: (2024)
von: Liu, Yingen, et al.
Veröffentlicht: (2024)
CMD-HAR: Cross-Modal Disentanglement for Wearable Human Activity Recognition
von: Yu, Ying, et al.
Veröffentlicht: (2025)
von: Yu, Ying, et al.
Veröffentlicht: (2025)
Panoramic Distortion-Aware Tokenization for Person Detection and Localization in Overhead Fisheye Images
von: Wakai, Nobuhiko, et al.
Veröffentlicht: (2025)
von: Wakai, Nobuhiko, et al.
Veröffentlicht: (2025)
PanoLora: Bridging Perspective and Panoramic Video Generation with LoRA Adaptation
von: Dong, Zeyu, et al.
Veröffentlicht: (2025)
von: Dong, Zeyu, et al.
Veröffentlicht: (2025)
TexLiDAR: Automated Text Understanding for Panoramic LiDAR Data
von: Cohen, Naor, et al.
Veröffentlicht: (2025)
von: Cohen, Naor, et al.
Veröffentlicht: (2025)
SNN-PAR: Energy Efficient Pedestrian Attribute Recognition via Spiking Neural Networks
von: Wang, Haiyang, et al.
Veröffentlicht: (2024)
von: Wang, Haiyang, et al.
Veröffentlicht: (2024)
Decoupled Prompt-Adapter Tuning for Continual Activity Recognition
von: Fu, Di, et al.
Veröffentlicht: (2024)
von: Fu, Di, et al.
Veröffentlicht: (2024)
Transformer-based Models to Deal with Heterogeneous Environments in Human Activity Recognition
von: EK, Sannara, et al.
Veröffentlicht: (2022)
von: EK, Sannara, et al.
Veröffentlicht: (2022)
AUFormer: Vision Transformers are Parameter-Efficient Facial Action Unit Detectors
von: Yuan, Kaishen, et al.
Veröffentlicht: (2024)
von: Yuan, Kaishen, et al.
Veröffentlicht: (2024)
Integrating Features for Recognizing Human Activities through Optimized Parameters in Graph Convolutional Networks and Transformer Architectures
von: Belal, Mohammad, et al.
Veröffentlicht: (2024)
von: Belal, Mohammad, et al.
Veröffentlicht: (2024)
GI-Bench: A Panoramic Benchmark Revealing the Knowledge-Experience Dissociation of Multimodal Large Language Models in Gastrointestinal Endoscopy Against Clinical Standards
von: Zhu, Yan, et al.
Veröffentlicht: (2026)
von: Zhu, Yan, et al.
Veröffentlicht: (2026)
Mixed Non-linear Quantization for Vision Transformers
von: Kim, Gihwan, et al.
Veröffentlicht: (2024)
von: Kim, Gihwan, et al.
Veröffentlicht: (2024)
LoViT: Long Video Transformer for Surgical Phase Recognition
von: Liu, Yang, et al.
Veröffentlicht: (2023)
von: Liu, Yang, et al.
Veröffentlicht: (2023)
Unified-EGformer: Exposure Guided Lightweight Transformer for Mixed-Exposure Image Enhancement
von: Adhikarla, Eashan, et al.
Veröffentlicht: (2024)
von: Adhikarla, Eashan, et al.
Veröffentlicht: (2024)
Toward Optimal Sampling Rate Selection and Unbiased Classification for Precise Animal Activity Recognition
von: Mao, Axiu, et al.
Veröffentlicht: (2026)
von: Mao, Axiu, et al.
Veröffentlicht: (2026)
USAD: End-to-End Human Activity Recognition via Diffusion Model with Spatiotemporal Attention
von: Xiao, Hang, et al.
Veröffentlicht: (2025)
von: Xiao, Hang, et al.
Veröffentlicht: (2025)
Emotion Recognition Using Transformers with Masked Learning
von: Min, Seongjae, et al.
Veröffentlicht: (2024)
von: Min, Seongjae, et al.
Veröffentlicht: (2024)
Beyond Motion Primitives: Behavioral Activity Recognition from Head-Mounted IMU
von: Huang, Chung-Ta, et al.
Veröffentlicht: (2026)
von: Huang, Chung-Ta, et al.
Veröffentlicht: (2026)
DCMM-Transformer: Degree-Corrected Mixed-Membership Attention for Medical Imaging
von: Cheng, Huimin, et al.
Veröffentlicht: (2025)
von: Cheng, Huimin, et al.
Veröffentlicht: (2025)
VideoPanda: Video Panoramic Diffusion with Multi-view Attention
von: Xie, Kevin, et al.
Veröffentlicht: (2025)
von: Xie, Kevin, et al.
Veröffentlicht: (2025)
TxP: Reciprocal Generation of Ground Pressure Dynamics and Activity Descriptions for Improving Human Activity Recognition
von: Ray, Lala Shakti Swarup, et al.
Veröffentlicht: (2025)
von: Ray, Lala Shakti Swarup, et al.
Veröffentlicht: (2025)
DreamScene360: Unconstrained Text-to-3D Scene Generation with Panoramic Gaussian Splatting
von: Zhou, Shijie, et al.
Veröffentlicht: (2024)
von: Zhou, Shijie, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MPT: A Large-scale Multi-Phytoplankton Tracking Benchmark
von: Yu, Yang, et al.
Veröffentlicht: (2024) -
UniPAR: A Unified Framework for Pedestrian Attribute Recognition
von: Xu, Minghe, et al.
Veröffentlicht: (2026) -
Logi-PAR: Logic-Infused Patient Activity Recognition via Differentiable Rule
von: Zarar, Muhammad, et al.
Veröffentlicht: (2026) -
Feature Fusion for Human Activity Recognition using Parameter-Optimized Multi-Stage Graph Convolutional Network and Transformer Models
von: Belal, Mohammad, et al.
Veröffentlicht: (2024) -
VLM-PAR: A Vision Language Model for Pedestrian Attribute Recognition
von: Sellam, Abdellah Zakaria, et al.
Veröffentlicht: (2025)