Saved in:
| Main Authors: | Singh, Alakhsimar, Goyal, Kanav, Verma, Nischay, Kumar, Puneet, Li, Xiaobai, Singh, Amritpal |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2409.16126 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TCCT-Net: Two-Stream Network Architecture for Fast and Efficient Engagement Estimation via Behavioral Feature Signals
by: Vedernikov, Alexander, et al.
Published: (2024)
by: Vedernikov, Alexander, et al.
Published: (2024)
Vision Large Language Models Are Good Noise Handlers in Engagement Analysis
by: Vedernikov, Alexander, et al.
Published: (2025)
by: Vedernikov, Alexander, et al.
Published: (2025)
GAViD: A Large-Scale Multimodal Dataset for Context-Aware Group Affect Recognition from Videos
by: Kumar, Deepak, et al.
Published: (2026)
by: Kumar, Deepak, et al.
Published: (2026)
VISTANet: VIsual Spoken Textual Additive Net for Interpretable Multimodal Emotion Recognition
by: Kumar, Puneet, et al.
Published: (2022)
by: Kumar, Puneet, et al.
Published: (2022)
Contrast-Phys: Unsupervised Video-based Remote Physiological Measurement via Spatiotemporal Contrast
by: Sun, Zhaodong, et al.
Published: (2022)
by: Sun, Zhaodong, et al.
Published: (2022)
Contrast-Phys+: Unsupervised and Weakly-supervised Video-based Remote Physiological Measurement via Spatiotemporal Contrast
by: Sun, Zhaodong, et al.
Published: (2023)
by: Sun, Zhaodong, et al.
Published: (2023)
ENet-21: An Optimized light CNN Structure for Lane Detection
by: Hosseini, Seyed Rasoul, et al.
Published: (2024)
by: Hosseini, Seyed Rasoul, et al.
Published: (2024)
PhysioSync: Temporal and Cross-Modal Contrastive Learning Inspired by Physiological Synchronization for EEG-Based Emotion Recognition
by: Cui, Kai, et al.
Published: (2025)
by: Cui, Kai, et al.
Published: (2025)
Beyond Images: Adaptive Fusion of Visual and Textual Data for Food Classification
by: Mittal, Prateek, et al.
Published: (2023)
by: Mittal, Prateek, et al.
Published: (2023)
Analyzing Participants' Engagement during Online Meetings Using Unsupervised Remote Photoplethysmography with Behavioral Features
by: Vedernikov, Alexander, et al.
Published: (2024)
by: Vedernikov, Alexander, et al.
Published: (2024)
LightMedSeg: Lightweight 3D Medical Image Segmentation with Learned Spatial Anchors
by: Tyagi, Kavyansh, et al.
Published: (2026)
by: Tyagi, Kavyansh, et al.
Published: (2026)
VisioMath: Benchmarking Figure-based Mathematical Reasoning in LMMs
by: Li, Can, et al.
Published: (2025)
by: Li, Can, et al.
Published: (2025)
A Diffusion-Driven Fine-Grained Nodule Synthesis Framework for Enhanced Lung Nodule Detection from Chest Radiographs
by: Goyal, Aryan, et al.
Published: (2026)
by: Goyal, Aryan, et al.
Published: (2026)
SciFigDetect: A Benchmark for AI-Generated Scientific Figure Detection
by: Hu, You, et al.
Published: (2026)
by: Hu, You, et al.
Published: (2026)
E2ENet: Dynamic Sparse Feature Fusion for Accurate and Efficient 3D Medical Image Segmentation
by: Wu, Boqian, et al.
Published: (2023)
by: Wu, Boqian, et al.
Published: (2023)
Machine Learning-Based Classification of Jhana Advanced Concentrative Absorption Meditation (ACAM-J) using 7T fMRI
by: Kumar, Puneet, et al.
Published: (2026)
by: Kumar, Puneet, et al.
Published: (2026)
Reliable or Deceptive? Investigating Gated Features for Smooth Visual Explanations in CNNs
by: Mitra, Soham, et al.
Published: (2024)
by: Mitra, Soham, et al.
Published: (2024)
VigilEye -- Artificial Intelligence-based Real-time Driver Drowsiness Detection
by: Sengar, Sandeep Singh, et al.
Published: (2024)
by: Sengar, Sandeep Singh, et al.
Published: (2024)
SDLNet: Statistical Deep Learning Network for Co-Occurring Object Detection and Identification
by: Singh, Binay Kumar, et al.
Published: (2024)
by: Singh, Binay Kumar, et al.
Published: (2024)
Exploring the Spectrum of Visio-Linguistic Compositionality and Recognition
by: Oh, Youngtaek, et al.
Published: (2024)
by: Oh, Youngtaek, et al.
Published: (2024)
Class-Agnostic Visio-Temporal Scene Sketch Semantic Segmentation
by: Kütük, Aleyna, et al.
Published: (2024)
by: Kütük, Aleyna, et al.
Published: (2024)
SOAR: Advancements in Small Body Object Detection for Aerial Imagery Using State Space Models and Programmable Gradients
by: Verma, Tushar, et al.
Published: (2024)
by: Verma, Tushar, et al.
Published: (2024)
Multimodal Fusion Learning with Dual Attention for Medical Imaging
by: Dhar, Joy, et al.
Published: (2024)
by: Dhar, Joy, et al.
Published: (2024)
DDEvENet: Evidence-based Ensemble Learning for Uncertainty-aware Brain Parcellation Using Diffusion MRI
by: Li, Chenjun, et al.
Published: (2024)
by: Li, Chenjun, et al.
Published: (2024)
A Multimodal Dataset for Enhancing Industrial Task Monitoring and Engagement Prediction
by: Mehta, Naval Kishore, et al.
Published: (2025)
by: Mehta, Naval Kishore, et al.
Published: (2025)
Distilling Knowledge from Text-to-Image Generative Models Improves Visio-Linguistic Reasoning in CLIP
by: Basu, Samyadeep, et al.
Published: (2023)
by: Basu, Samyadeep, et al.
Published: (2023)
Weed Detection using Convolutional Neural Network
by: Tripathi, Santosh Kumar, et al.
Published: (2025)
by: Tripathi, Santosh Kumar, et al.
Published: (2025)
Benchmarking Joint Face Spoofing and Forgery Detection with Visual and Physiological Cues
by: Yu, Zitong, et al.
Published: (2022)
by: Yu, Zitong, et al.
Published: (2022)
Supervised Multilabel Image Classification Using Residual Networks with Probabilistic Reasoning
by: Singh, Lokender, et al.
Published: (2025)
by: Singh, Lokender, et al.
Published: (2025)
OmniFD: A Unified Model for Versatile Face Forgery Detection
by: Liu, Haotian, et al.
Published: (2025)
by: Liu, Haotian, et al.
Published: (2025)
Contrasting Intra-Modal and Ranking Cross-Modal Hard Negatives to Enhance Visio-Linguistic Compositional Understanding
by: Zhang, Le, et al.
Published: (2023)
by: Zhang, Le, et al.
Published: (2023)
VisioFirm: Cross-Platform AI-assisted Annotation Tool for Computer Vision
by: Ghazouali, Safouane El, et al.
Published: (2025)
by: Ghazouali, Safouane El, et al.
Published: (2025)
Leveraging Language Prior for Infrared Small Target Detection
by: Singh, Pranav, et al.
Published: (2025)
by: Singh, Pranav, et al.
Published: (2025)
A Hybrid Transformer-Sequencer approach for Age and Gender classification from in-wild facial images
by: Singh, Aakash, et al.
Published: (2024)
by: Singh, Aakash, et al.
Published: (2024)
Semi-supervised Active Learning for Video Action Detection
by: Singh, Ayush, et al.
Published: (2023)
by: Singh, Ayush, et al.
Published: (2023)
Unboxing Engagement in YouTube Influencer Videos: An Attention-Based Approach
by: Rajaram, Prashant, et al.
Published: (2020)
by: Rajaram, Prashant, et al.
Published: (2020)
Impact of Financial Literacy on Investment Decisions and Stock Market Participation using Extreme Learning Machines
by: Baveja, Gunbir Singh, et al.
Published: (2024)
by: Baveja, Gunbir Singh, et al.
Published: (2024)
Are Data Augmentation and Segmentation Always Necessary? Insights from COVID-19 X-Rays and a Methodology Thereof
by: Swaraj, Aman, et al.
Published: (2026)
by: Swaraj, Aman, et al.
Published: (2026)
Theoretical Analysis of Power-law Transformation on Images for Text Polarity Detection
by: Yadav, Narendra Singh, et al.
Published: (2025)
by: Yadav, Narendra Singh, et al.
Published: (2025)
DeiTFake: Deepfake Detection Model using DeiT Multi-Stage Training
by: Kumar, Saksham, et al.
Published: (2025)
by: Kumar, Saksham, et al.
Published: (2025)
Similar Items
-
TCCT-Net: Two-Stream Network Architecture for Fast and Efficient Engagement Estimation via Behavioral Feature Signals
by: Vedernikov, Alexander, et al.
Published: (2024) -
Vision Large Language Models Are Good Noise Handlers in Engagement Analysis
by: Vedernikov, Alexander, et al.
Published: (2025) -
GAViD: A Large-Scale Multimodal Dataset for Context-Aware Group Affect Recognition from Videos
by: Kumar, Deepak, et al.
Published: (2026) -
VISTANet: VIsual Spoken Textual Additive Net for Interpretable Multimodal Emotion Recognition
by: Kumar, Puneet, et al.
Published: (2022) -
Contrast-Phys: Unsupervised Video-based Remote Physiological Measurement via Spatiotemporal Contrast
by: Sun, Zhaodong, et al.
Published: (2022)