Feature Hallucination for Self-supervised Action Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Lei, Koniusz, Piotr |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Video Understanding by Design: How Datasets Shape Architectures and Insights
von: Wang, Lei, et al.
Veröffentlicht: (2025)
von: Wang, Lei, et al.
Veröffentlicht: (2025)
Learnable Expansion of Graph Operators for Multi-Modal Feature Fusion
von: Ding, Dexuan, et al.
Veröffentlicht: (2024)
von: Ding, Dexuan, et al.
Veröffentlicht: (2024)
Graph Your Own Prompt
von: Ding, Xi, et al.
Veröffentlicht: (2025)
von: Ding, Xi, et al.
Veröffentlicht: (2025)
Learning Time in Static Classifiers
von: Ding, Xi, et al.
Veröffentlicht: (2025)
von: Ding, Xi, et al.
Veröffentlicht: (2025)
Subspace Kernel Learning on Tensor Sequences
von: Wang, Lei, et al.
Veröffentlicht: (2026)
von: Wang, Lei, et al.
Veröffentlicht: (2026)
Motion meets Attention: Video Motion Prompts
von: Chen, Qixiang, et al.
Veröffentlicht: (2024)
von: Chen, Qixiang, et al.
Veröffentlicht: (2024)
Adaptive Multi-head Contrastive Learning
von: Wang, Lei, et al.
Veröffentlicht: (2023)
von: Wang, Lei, et al.
Veröffentlicht: (2023)
Uncertainty-DTW for Sequences and Visual Tokens
von: Wang, Lei, et al.
Veröffentlicht: (2026)
von: Wang, Lei, et al.
Veröffentlicht: (2026)
Meet JEANIE: a Similarity Measure for 3D Skeleton Sequences via Temporal-Viewpoint Alignment
von: Wang, Lei, et al.
Veröffentlicht: (2024)
von: Wang, Lei, et al.
Veröffentlicht: (2024)
When Spatial meets Temporal in Action Recognition
von: Chen, Huilin, et al.
Veröffentlicht: (2024)
von: Chen, Huilin, et al.
Veröffentlicht: (2024)
Pre-training with Random Orthogonal Projection Image Modeling
von: Haghighat, Maryam, et al.
Veröffentlicht: (2023)
von: Haghighat, Maryam, et al.
Veröffentlicht: (2023)
Possibilistic Predictive Uncertainty for Deep Learning
von: Ni, Yao, et al.
Veröffentlicht: (2026)
von: Ni, Yao, et al.
Veröffentlicht: (2026)
ICED: Concept-level Machine Unlearning via Interpretable Concept Decomposition
von: Lin, Shen, et al.
Veröffentlicht: (2026)
von: Lin, Shen, et al.
Veröffentlicht: (2026)
Evolving Skeletons: Motion Dynamics in Action Recognition
von: Qiu, Jushang, et al.
Veröffentlicht: (2025)
von: Qiu, Jushang, et al.
Veröffentlicht: (2025)
Hierarchically Robust Zero-shot Vision-language Models
von: Dong, Junhao, et al.
Veröffentlicht: (2026)
von: Dong, Junhao, et al.
Veröffentlicht: (2026)
Federated Learning for Face Recognition via Intra-subject Self-supervised Learning
von: Kim, Hansol, et al.
Veröffentlicht: (2024)
von: Kim, Hansol, et al.
Veröffentlicht: (2024)
Mitigating Object and Action Hallucinations in Multimodal LLMs via Self-Augmented Contrastive Alignment
von: Chang, Kai-Po, et al.
Veröffentlicht: (2025)
von: Chang, Kai-Po, et al.
Veröffentlicht: (2025)
SST: Self-training with Self-adaptive Thresholding for Semi-supervised Learning
von: Zhao, Shuai, et al.
Veröffentlicht: (2025)
von: Zhao, Shuai, et al.
Veröffentlicht: (2025)
Noise Consistency Regularization for Improved Subject-Driven Image Synthesis
von: Ni, Yao, et al.
Veröffentlicht: (2025)
von: Ni, Yao, et al.
Veröffentlicht: (2025)
CHAIN: Enhancing Generalization in Data-Efficient GANs via lipsCHitz continuity constrAIned Normalization
von: Ni, Yao, et al.
Veröffentlicht: (2024)
von: Ni, Yao, et al.
Veröffentlicht: (2024)
Real-Time Human Action Recognition on Embedded Platforms
von: Wang, Ruiqi, et al.
Veröffentlicht: (2024)
von: Wang, Ruiqi, et al.
Veröffentlicht: (2024)
Self-supervised Learning for Hyperspectral Images of Trees
von: Rahman, Moqsadur, et al.
Veröffentlicht: (2025)
von: Rahman, Moqsadur, et al.
Veröffentlicht: (2025)
Self-supervised Transformation Learning for Equivariant Representations
von: Yu, Jaemyung, et al.
Veröffentlicht: (2025)
von: Yu, Jaemyung, et al.
Veröffentlicht: (2025)
Self-supervised Pre-training of Text Recognizers
von: Kišš, Martin, et al.
Veröffentlicht: (2024)
von: Kišš, Martin, et al.
Veröffentlicht: (2024)
SHAMISA: SHAped Modeling of Implicit Structural Associations for Self-supervised No-Reference Image Quality Assessment
von: Naseri, Mahdi, et al.
Veröffentlicht: (2026)
von: Naseri, Mahdi, et al.
Veröffentlicht: (2026)
Identify, Isolate, and Purge: Mitigating Hallucinations in LVLMs via Self-Evolving Distillation
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
Towards Multimodal Open-Set Domain Generalization and Adaptation through Self-supervision
von: Dong, Hao, et al.
Veröffentlicht: (2024)
von: Dong, Hao, et al.
Veröffentlicht: (2024)
MOCA: Self-supervised Representation Learning by Predicting Masked Online Codebook Assignments
von: Gidaris, Spyros, et al.
Veröffentlicht: (2023)
von: Gidaris, Spyros, et al.
Veröffentlicht: (2023)
Hypergraph-based Multi-View Action Recognition using Event Cameras
von: Gao, Yue, et al.
Veröffentlicht: (2024)
von: Gao, Yue, et al.
Veröffentlicht: (2024)
One-Frame Calibration with Siamese Network in Facial Action Unit Recognition
von: Feng, Shuangquan, et al.
Veröffentlicht: (2024)
von: Feng, Shuangquan, et al.
Veröffentlicht: (2024)
EITNet: An IoT-Enhanced Framework for Real-Time Basketball Action Recognition
von: Liu, Jingyu, et al.
Veröffentlicht: (2024)
von: Liu, Jingyu, et al.
Veröffentlicht: (2024)
Learning Contrastive Feature Representations for Facial Action Unit Detection
von: Shang, Ziqiao, et al.
Veröffentlicht: (2024)
von: Shang, Ziqiao, et al.
Veröffentlicht: (2024)
Understanding the Role of Equivariance in Self-supervised Learning
von: Wang, Yifei, et al.
Veröffentlicht: (2024)
von: Wang, Yifei, et al.
Veröffentlicht: (2024)
Wildlife Target Re-Identification Using Self-supervised Learning in Non-Urban Settings
von: Muthivhi, Mufhumudzi, et al.
Veröffentlicht: (2025)
von: Muthivhi, Mufhumudzi, et al.
Veröffentlicht: (2025)
HiT-JEPA: A Hierarchical Self-supervised Trajectory Embedding Framework for Similarity Computation
von: Li, Lihuan, et al.
Veröffentlicht: (2025)
von: Li, Lihuan, et al.
Veröffentlicht: (2025)
Self-supervised video pretraining yields robust and more human-aligned visual representations
von: Parthasarathy, Nikhil, et al.
Veröffentlicht: (2022)
von: Parthasarathy, Nikhil, et al.
Veröffentlicht: (2022)
HalluRNN: Mitigating Hallucinations via Recurrent Cross-Layer Reasoning in Large Vision-Language Models
von: Yu, Le, et al.
Veröffentlicht: (2025)
von: Yu, Le, et al.
Veröffentlicht: (2025)
PACE: Marrying generalization in PArameter-efficient fine-tuning with Consistency rEgularization
von: Ni, Yao, et al.
Veröffentlicht: (2024)
von: Ni, Yao, et al.
Veröffentlicht: (2024)
Self-supervised Learning on Camera Trap Footage Yields a Strong Universal Face Embedder
von: Iashin, Vladimir, et al.
Veröffentlicht: (2025)
von: Iashin, Vladimir, et al.
Veröffentlicht: (2025)
RNNs, CNNs and Transformers in Human Action Recognition: A Survey and a Hybrid Model
von: Alomar, Khaled, et al.
Veröffentlicht: (2024)
von: Alomar, Khaled, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Video Understanding by Design: How Datasets Shape Architectures and Insights
von: Wang, Lei, et al.
Veröffentlicht: (2025) -
Learnable Expansion of Graph Operators for Multi-Modal Feature Fusion
von: Ding, Dexuan, et al.
Veröffentlicht: (2024) -
Graph Your Own Prompt
von: Ding, Xi, et al.
Veröffentlicht: (2025) -
Learning Time in Static Classifiers
von: Ding, Xi, et al.
Veröffentlicht: (2025) -
Subspace Kernel Learning on Tensor Sequences
von: Wang, Lei, et al.
Veröffentlicht: (2026)