Saved in:
| Main Authors: | Prakash, Pritesh, Rai, Anoop Kumar |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2412.02198 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Context-Aware Network Based on Multi-scale Spatio-temporal Attention for Action Recognition in Videos
by: Li, Xiaoyang, et al.
Published: (2025)
by: Li, Xiaoyang, et al.
Published: (2025)
SpATr: MoCap 3D Human Action Recognition based on Spiral Auto-encoder and Transformer Network
by: Bouzid, Hamza, et al.
Published: (2023)
by: Bouzid, Hamza, et al.
Published: (2023)
Distinguishing Visually Similar Actions: Prompt-Guided Semantic Prototype Modulation for Few-Shot Action Recognition
by: Li, Xiaoyang, et al.
Published: (2025)
by: Li, Xiaoyang, et al.
Published: (2025)
Robust Palm-Vein Recognition Using the MMD Filter: Improving SIFT-Based Feature Matching
by: Perera, Kaveen, et al.
Published: (2025)
by: Perera, Kaveen, et al.
Published: (2025)
ViBED-Net: Video Based Engagement Detection Network Using Face-Aware and Scene-Aware Spatiotemporal Cues
by: Gothwal, Prateek, et al.
Published: (2025)
by: Gothwal, Prateek, et al.
Published: (2025)
Enhancing Explainable AI: A Hybrid Approach Combining GradCAM and LRP for CNN Interpretability
by: Dhore, Vaibhav, et al.
Published: (2024)
by: Dhore, Vaibhav, et al.
Published: (2024)
MSPCaps: A Multi-Scale Patchify Capsule Network with Cross-Agreement Routing for Visual Recognition
by: Hu, Yudong, et al.
Published: (2025)
by: Hu, Yudong, et al.
Published: (2025)
WatchHAR: Real-time On-device Human Activity Recognition System for Smartwatches
by: Yeon, Taeyoung, et al.
Published: (2025)
by: Yeon, Taeyoung, et al.
Published: (2025)
Pan-Arctic Permafrost Landform and Human-built Infrastructure Feature Detection with Vision Transformers and Location Embeddings
by: Perera, Amal S., et al.
Published: (2025)
by: Perera, Amal S., et al.
Published: (2025)
Mamba-VA: A Mamba-based Approach for Continuous Emotion Recognition in Valence-Arousal Space
by: Liang, Yuheng, et al.
Published: (2025)
by: Liang, Yuheng, et al.
Published: (2025)
DistillMatch: Leveraging Knowledge Distillation from Vision Foundation Model for Multimodal Image Matching
by: Yang, Meng, et al.
Published: (2025)
by: Yang, Meng, et al.
Published: (2025)
Adaptive Cascading Network for Continual Test-Time Adaptation
by: Nguyen, Kien X., et al.
Published: (2024)
by: Nguyen, Kien X., et al.
Published: (2024)
Guidelines for External Disturbance Factors in the Use of OCR in Real-World Environments
by: Iwata, Kenji, et al.
Published: (2025)
by: Iwata, Kenji, et al.
Published: (2025)
SymFace: Additional Facial Symmetry Loss for Deep Face Recognition
by: Prakash, Pritesh, et al.
Published: (2024)
by: Prakash, Pritesh, et al.
Published: (2024)
ExpressNet-MoE: A Hybrid Deep Neural Network for Emotion Recognition
by: Banerjee, Deeptimaan, et al.
Published: (2025)
by: Banerjee, Deeptimaan, et al.
Published: (2025)
FerretNet: Efficient Synthetic Image Detection via Local Pixel Dependencies
by: Liang, Shuqiao, et al.
Published: (2025)
by: Liang, Shuqiao, et al.
Published: (2025)
Dual-Teacher Ensemble Models with Double-Copy-Paste for 3D Semi-Supervised Medical Image Segmentation
by: Fa, Zhan, et al.
Published: (2024)
by: Fa, Zhan, et al.
Published: (2024)
Sit-to-Stand Transitions Detection and Duration Measurement Using Smart Lacelock Sensor
by: Islam, Md Rafi, et al.
Published: (2026)
by: Islam, Md Rafi, et al.
Published: (2026)
When Labels Have Structure: Improving Image Classification with Hierarchy-Aware Cross-Entropy
by: Chan, April, et al.
Published: (2026)
by: Chan, April, et al.
Published: (2026)
Pairwise Spatiotemporal Partial Trajectory Matching for Co-movement Analysis
by: Cardei, Maria, et al.
Published: (2024)
by: Cardei, Maria, et al.
Published: (2024)
Multi-stage Bridge Inspection System: Integrating Foundation Models with Location Anonymization
by: Yasuno, Takato
Published: (2026)
by: Yasuno, Takato
Published: (2026)
LiqD: A Dynamic Liquid Level Detection Model under Tricky Small Containers
by: Ma, Yukun, et al.
Published: (2024)
by: Ma, Yukun, et al.
Published: (2024)
Structured Analytic Coherent Point Drift for Non-Rigid Point Set Registration
by: Feng, Wei, et al.
Published: (2026)
by: Feng, Wei, et al.
Published: (2026)
Precision at Scale: Domain-Specific Datasets On-Demand
by: Rodríguez-de-Vera, Jesús M, et al.
Published: (2024)
by: Rodríguez-de-Vera, Jesús M, et al.
Published: (2024)
Accurate Thyroid Cancer Classification using a Novel Binary Pattern Driven Local Discrete Cosine Transform Descriptor
by: Saini, Saurabh, et al.
Published: (2025)
by: Saini, Saurabh, et al.
Published: (2025)
On the Equivalence of Regression and Classification
by: Jayadeva, et al.
Published: (2025)
by: Jayadeva, et al.
Published: (2025)
MM-Conv: A Multimodal Dataset and Benchmark for Context-Aware Grounding in 3D Dialogue
by: Deichler, Anna, et al.
Published: (2026)
by: Deichler, Anna, et al.
Published: (2026)
treeX: Unsupervised Tree Instance Segmentation in Dense Forest Point Clouds
by: Burmeister, Josafat-Mattias, et al.
Published: (2025)
by: Burmeister, Josafat-Mattias, et al.
Published: (2025)
CoCoA-Mix: Confusion-and-Confidence-Aware Mixture Model for Context Optimization
by: Hong, Dasol, et al.
Published: (2025)
by: Hong, Dasol, et al.
Published: (2025)
CUBIC: Concept Embeddings for Unsupervised Bias Identification using VLMs
by: Méndez, David, et al.
Published: (2025)
by: Méndez, David, et al.
Published: (2025)
When Interpretability Becomes a Liability: Adversarial Attacks on CBM Concept Layers
by: Sridhar, Aditya
Published: (2026)
by: Sridhar, Aditya
Published: (2026)
AI Assisted AR Assembly: Object Recognition and Computer Vision for Augmented Reality Assisted Assembly
by: Kyaw, Alexander Htet, et al.
Published: (2025)
by: Kyaw, Alexander Htet, et al.
Published: (2025)
U-R-VEDA: Integrating UNET, Residual Links, Edge and Dual Attention, and Vision Transformer for Accurate Semantic Segmentation of CMRs
by: Mukisa, Racheal, et al.
Published: (2025)
by: Mukisa, Racheal, et al.
Published: (2025)
Automatic Dance Video Segmentation for Understanding Choreography
by: Endo, Koki, et al.
Published: (2024)
by: Endo, Koki, et al.
Published: (2024)
Lotus: Creating Short Videos From Long Videos With Abstractive and Extractive Summarization
by: Barua, Aadit, et al.
Published: (2025)
by: Barua, Aadit, et al.
Published: (2025)
MIRAGE: A Micro-Interaction Relational Architecture for Grounded Exploration in Multi-Figure Artworks
by: Chiu, Jui-Cheng, et al.
Published: (2026)
by: Chiu, Jui-Cheng, et al.
Published: (2026)
MDA: An Interpretable and Scalable Multi-Modal Fusion under Missing Modalities and Intrinsic Noise Conditions
by: Fan, Lin, et al.
Published: (2024)
by: Fan, Lin, et al.
Published: (2024)
Block-Fused Attention-Driven Adaptively-Pooled ResNet Model for Improved Cervical Cancer Classification
by: Saini, Saurabh, et al.
Published: (2024)
by: Saini, Saurabh, et al.
Published: (2024)
Interpretable Machine Learning-Derived Spectral Indices for Vegetation Monitoring
by: Lotfi, Ali, et al.
Published: (2025)
by: Lotfi, Ali, et al.
Published: (2025)
CAL-RAG: Retrieval-Augmented Multi-Agent Generation for Content-Aware Layout Design
by: Forouzandehmehr, Najmeh, et al.
Published: (2025)
by: Forouzandehmehr, Najmeh, et al.
Published: (2025)
Similar Items
-
Context-Aware Network Based on Multi-scale Spatio-temporal Attention for Action Recognition in Videos
by: Li, Xiaoyang, et al.
Published: (2025) -
SpATr: MoCap 3D Human Action Recognition based on Spiral Auto-encoder and Transformer Network
by: Bouzid, Hamza, et al.
Published: (2023) -
Distinguishing Visually Similar Actions: Prompt-Guided Semantic Prototype Modulation for Few-Shot Action Recognition
by: Li, Xiaoyang, et al.
Published: (2025) -
Robust Palm-Vein Recognition Using the MMD Filter: Improving SIFT-Based Feature Matching
by: Perera, Kaveen, et al.
Published: (2025) -
ViBED-Net: Video Based Engagement Detection Network Using Face-Aware and Scene-Aware Spatiotemporal Cues
by: Gothwal, Prateek, et al.
Published: (2025)