Salvato in:
| Autori principali: | Pathak, Gourang, Kumar, Abhay, Rawat, Sannidhya, Gupta, Shikha |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2412.02127 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Colors See Colors Ignore: Clothes Changing ReID with Color Disentanglement
di: Pathak, Priyank, et al.
Pubblicazione: (2025)
di: Pathak, Priyank, et al.
Pubblicazione: (2025)
Coarse Attribute Prediction with Task Agnostic Distillation for Real World Clothes Changing ReID
di: Pathak, Priyank, et al.
Pubblicazione: (2025)
di: Pathak, Priyank, et al.
Pubblicazione: (2025)
Stable Mean Teacher for Semi-supervised Video Action Detection
di: Kumar, Akash, et al.
Pubblicazione: (2024)
di: Kumar, Akash, et al.
Pubblicazione: (2024)
A Large-Scale Analysis on Contextual Self-Supervised Video Representation Learning
di: Kumar, Akash, et al.
Pubblicazione: (2025)
di: Kumar, Akash, et al.
Pubblicazione: (2025)
CoSPlan: Corrective Sequential Planning via Scene Graph Incremental Updates
di: Grover, Shresth, et al.
Pubblicazione: (2025)
di: Grover, Shresth, et al.
Pubblicazione: (2025)
Semi-supervised Active Learning for Video Action Detection
di: Singh, Ayush, et al.
Pubblicazione: (2023)
di: Singh, Ayush, et al.
Pubblicazione: (2023)
Contextual Self-paced Learning for Weakly Supervised Spatio-Temporal Video Grounding
di: Kumar, Akash, et al.
Pubblicazione: (2025)
di: Kumar, Akash, et al.
Pubblicazione: (2025)
LR0.FM: Low-Res Benchmark and Improving Robustness for Zero-Shot Classification in Foundation Models
di: Pathak, Priyank, et al.
Pubblicazione: (2025)
di: Pathak, Priyank, et al.
Pubblicazione: (2025)
Fast and Memory-Efficient Video Diffusion Using Streamlined Inference
di: Zhan, Zheng, et al.
Pubblicazione: (2024)
di: Zhan, Zheng, et al.
Pubblicazione: (2024)
EZ-CLIP: Efficient Zeroshot Video Action Recognition
di: Ahmad, Shahzad, et al.
Pubblicazione: (2023)
di: Ahmad, Shahzad, et al.
Pubblicazione: (2023)
DashCop: Automated E-ticket Generation for Two-Wheeler Traffic Violations Using Dashcam Videos
di: Rawat, Deepti, et al.
Pubblicazione: (2025)
di: Rawat, Deepti, et al.
Pubblicazione: (2025)
VISTA: Video Interaction Spatio-Temporal Analysis Benchmark
di: Aparcedo, Alejandro, et al.
Pubblicazione: (2026)
di: Aparcedo, Alejandro, et al.
Pubblicazione: (2026)
SynthForge: Synthesizing High-Quality Face Dataset with Controllable 3D Generative Models
di: Rawat, Abhay, et al.
Pubblicazione: (2024)
di: Rawat, Abhay, et al.
Pubblicazione: (2024)
Magic Fixup: Streamlining Photo Editing by Watching Dynamic Videos
di: Alzayer, Hadi, et al.
Pubblicazione: (2024)
di: Alzayer, Hadi, et al.
Pubblicazione: (2024)
Comparative Analysis: Violence Recognition from Videos using Transfer Learning
di: Dashdamirov, Dursun
Pubblicazione: (2024)
di: Dashdamirov, Dursun
Pubblicazione: (2024)
Streamlined Open-Vocabulary Human-Object Interaction Detection
di: Sun, Chang, et al.
Pubblicazione: (2026)
di: Sun, Chang, et al.
Pubblicazione: (2026)
On Occlusions in Video Action Detection: Benchmark Datasets And Training Recipes
di: Modi, Rajat, et al.
Pubblicazione: (2024)
di: Modi, Rajat, et al.
Pubblicazione: (2024)
Tool-Augmented Spatiotemporal Reasoning for Streamlining Video Question Answering Task
di: Fan, Sunqi, et al.
Pubblicazione: (2025)
di: Fan, Sunqi, et al.
Pubblicazione: (2025)
ScaleLSD: Scalable Deep Line Segment Detection Streamlined
di: Ke, Zeran, et al.
Pubblicazione: (2025)
di: Ke, Zeran, et al.
Pubblicazione: (2025)
Beyond Euclidean: Dual-Space Representation Learning for Weakly Supervised Video Violence Detection
di: Leng, Jiaxu, et al.
Pubblicazione: (2024)
di: Leng, Jiaxu, et al.
Pubblicazione: (2024)
PiercingEye: Dual-Space Video Violence Detection with Hyperbolic Vision-Language Guidance
di: Leng, Jiaxu, et al.
Pubblicazione: (2025)
di: Leng, Jiaxu, et al.
Pubblicazione: (2025)
RobustGait: Robustness Analysis for Appearance Based Gait Recognition
di: Sayera, Reeshoon, et al.
Pubblicazione: (2025)
di: Sayera, Reeshoon, et al.
Pubblicazione: (2025)
Scaling Open-Vocabulary Action Detection
di: Sia, Zhen Hao, et al.
Pubblicazione: (2025)
di: Sia, Zhen Hao, et al.
Pubblicazione: (2025)
Self-Consistency in Vision-Language Models for Precision Agriculture: Multi-Response Consensus for Crop Disease Management
di: Gupta, Mihir, et al.
Pubblicazione: (2025)
di: Gupta, Mihir, et al.
Pubblicazione: (2025)
STPro: Spatial and Temporal Progressive Learning for Weakly Supervised Spatio-Temporal Grounding
di: Garg, Aaryan, et al.
Pubblicazione: (2025)
di: Garg, Aaryan, et al.
Pubblicazione: (2025)
StreamReady: Learning What to Answer and When in Long Streaming Videos
di: Azad, Shehreen, et al.
Pubblicazione: (2026)
di: Azad, Shehreen, et al.
Pubblicazione: (2026)
Optimizing Violence Detection in Video Classification Accuracy through 3D Convolutional Neural Networks
di: Kavathia, Aarjav, et al.
Pubblicazione: (2024)
di: Kavathia, Aarjav, et al.
Pubblicazione: (2024)
Joint Audio-Visual Idling Vehicle Detection with Streamlined Input Dependencies
di: Li, Xiwen, et al.
Pubblicazione: (2024)
di: Li, Xiwen, et al.
Pubblicazione: (2024)
Streamlining the Development of Active Learning Methods in Real-World Object Detection
di: Sbeyti, Moussa Kassem, et al.
Pubblicazione: (2025)
di: Sbeyti, Moussa Kassem, et al.
Pubblicazione: (2025)
HierarQ: Task-Aware Hierarchical Q-Former for Enhanced Video Understanding
di: Azad, Shehreen, et al.
Pubblicazione: (2025)
di: Azad, Shehreen, et al.
Pubblicazione: (2025)
Exploring Personalized Federated Learning Architectures for Violence Detection in Surveillance Videos
di: Kassir, Mohammad, et al.
Pubblicazione: (2025)
di: Kassir, Mohammad, et al.
Pubblicazione: (2025)
DENSER: Depth-Guided Ensemble with Staged EFA-GS Reconstruction for Soccer Novel View Synthesis
di: Rawat, Parthsarthi
Pubblicazione: (2026)
di: Rawat, Parthsarthi
Pubblicazione: (2026)
SMART: SMPLest-X Mesh Adaptation and RAFT Tracking for Soccer Pose Estimation
di: Rawat, Parthsarthi
Pubblicazione: (2026)
di: Rawat, Parthsarthi
Pubblicazione: (2026)
GETReason: Enhancing Image Context Extraction through Hierarchical Multi-Agent Reasoning
di: Siingh, Shikhhar, et al.
Pubblicazione: (2025)
di: Siingh, Shikhhar, et al.
Pubblicazione: (2025)
MolSight: Molecular Property Prediction with Images
di: Baranwal, Aaditya, et al.
Pubblicazione: (2026)
di: Baranwal, Aaditya, et al.
Pubblicazione: (2026)
Exploring Remote Photoplethysmography for Neonatal Pain Detection from Facial Videos
di: Dhamaniya, Ashutosh, et al.
Pubblicazione: (2026)
di: Dhamaniya, Ashutosh, et al.
Pubblicazione: (2026)
EfficientGS: Streamlining Gaussian Splatting for Large-Scale High-Resolution Scene Representation
di: Liu, Wenkai, et al.
Pubblicazione: (2024)
di: Liu, Wenkai, et al.
Pubblicazione: (2024)
RoadSocial: A Diverse VideoQA Dataset and Benchmark for Road Event Understanding from Social Video Narratives
di: Parikh, Chirag, et al.
Pubblicazione: (2025)
di: Parikh, Chirag, et al.
Pubblicazione: (2025)
Asynchronous Perception Machine For Efficient Test-Time-Training
di: Modi, Rajat, et al.
Pubblicazione: (2024)
di: Modi, Rajat, et al.
Pubblicazione: (2024)
Foundation Models for Video Understanding: A Survey
di: Madan, Neelu, et al.
Pubblicazione: (2024)
di: Madan, Neelu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Colors See Colors Ignore: Clothes Changing ReID with Color Disentanglement
di: Pathak, Priyank, et al.
Pubblicazione: (2025) -
Coarse Attribute Prediction with Task Agnostic Distillation for Real World Clothes Changing ReID
di: Pathak, Priyank, et al.
Pubblicazione: (2025) -
Stable Mean Teacher for Semi-supervised Video Action Detection
di: Kumar, Akash, et al.
Pubblicazione: (2024) -
A Large-Scale Analysis on Contextual Self-Supervised Video Representation Learning
di: Kumar, Akash, et al.
Pubblicazione: (2025) -
CoSPlan: Corrective Sequential Planning via Scene Graph Incremental Updates
di: Grover, Shresth, et al.
Pubblicazione: (2025)