An Attribute-Based Measure of Video Complexity
Fuente:
arXiv
Salvato in:
| Autori principali: | Sarkar, Aditya, Li, Yi, Wang, Zihao, Cheng, Jiacheng, Nuthalapati, Sai Vidyaranya, Singh, Aashu, Mishra, Shlok Kumar, Jacobs, David, Vasconcelos, Nuno |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Leveraging Data to Say No: Memory Augmented Plug-and-Play Selective Prediction
di: Sarkar, Aditya, et al.
Pubblicazione: (2026)
di: Sarkar, Aditya, et al.
Pubblicazione: (2026)
Unifying Contrastive and Generative Objectives for Visual Understanding and Text-to-Image Generation
di: Li, Chao, et al.
Pubblicazione: (2026)
di: Li, Chao, et al.
Pubblicazione: (2026)
StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding
di: Yang, Yanlai, et al.
Pubblicazione: (2025)
di: Yang, Yanlai, et al.
Pubblicazione: (2025)
IntroStyle: Training-Free Introspective Style Attribution using Diffusion Features
di: Kumar, Anand, et al.
Pubblicazione: (2024)
di: Kumar, Anand, et al.
Pubblicazione: (2024)
Transfer between Modalities with MetaQueries
di: Pan, Xichen, et al.
Pubblicazione: (2025)
di: Pan, Xichen, et al.
Pubblicazione: (2025)
Adapting Dual-encoder Vision-language Models for Paraphrased Retrieval
di: Cheng, Jiacheng, et al.
Pubblicazione: (2024)
di: Cheng, Jiacheng, et al.
Pubblicazione: (2024)
Matching High-Dimensional Geometric Quantiles for Test-Time Adaptation of Transformers and Convolutional Networks Alike
di: Danda, Sravan, et al.
Pubblicazione: (2026)
di: Danda, Sravan, et al.
Pubblicazione: (2026)
Prompt Sliders for Fine-Grained Control, Editing and Erasing of Concepts in Diffusion Models
di: Sridhar, Deepak, et al.
Pubblicazione: (2024)
di: Sridhar, Deepak, et al.
Pubblicazione: (2024)
Improving image synthesis with diffusion-negative sampling
di: Desai, Alakh, et al.
Pubblicazione: (2024)
di: Desai, Alakh, et al.
Pubblicazione: (2024)
Diffusion Models with Adaptive Negative Sampling Without External Resources
di: Desai, Alakh, et al.
Pubblicazione: (2025)
di: Desai, Alakh, et al.
Pubblicazione: (2025)
Learning Skill-Attributes for Transferable Assessment in Video
di: Ashutosh, Kumar, et al.
Pubblicazione: (2025)
di: Ashutosh, Kumar, et al.
Pubblicazione: (2025)
Atmospheric Noise-Resilient Image Classification in a Real-World Scenario: Using Hybrid CNN and Pin-GTSVM
di: Mehendale, Shlok, et al.
Pubblicazione: (2025)
di: Mehendale, Shlok, et al.
Pubblicazione: (2025)
Ego-VPA: Egocentric Video Understanding with Parameter-efficient Adaptation
di: Wu, Tz-Ying, et al.
Pubblicazione: (2024)
di: Wu, Tz-Ying, et al.
Pubblicazione: (2024)
SCHEME: Scalable Channel Mixer for Vision Transformers
di: Sridhar, Deepak, et al.
Pubblicazione: (2023)
di: Sridhar, Deepak, et al.
Pubblicazione: (2023)
EditAR: Unified Conditional Generation with Autoregressive Models
di: Mu, Jiteng, et al.
Pubblicazione: (2025)
di: Mu, Jiteng, et al.
Pubblicazione: (2025)
Generating HDR Video from SDR Video
di: Tedla, SaiKiran, et al.
Pubblicazione: (2026)
di: Tedla, SaiKiran, et al.
Pubblicazione: (2026)
MeasureNet: Measurement Based Celiac Disease Identification
di: Tyagi, Aayush Kumar, et al.
Pubblicazione: (2024)
di: Tyagi, Aayush Kumar, et al.
Pubblicazione: (2024)
VideoAVE: A Multi-Attribute Video-to-Text Attribute Value Extraction Dataset and Benchmark Models
di: Cheng, Ming, et al.
Pubblicazione: (2025)
di: Cheng, Ming, et al.
Pubblicazione: (2025)
EgoPrivacy: What Your First-Person Camera Says About You?
di: Li, Yijiang, et al.
Pubblicazione: (2025)
di: Li, Yijiang, et al.
Pubblicazione: (2025)
EVA: Zero-shot Accurate Attributes and Multi-Object Video Editing
di: Yang, Xiangpeng, et al.
Pubblicazione: (2024)
di: Yang, Xiangpeng, et al.
Pubblicazione: (2024)
ProTeCt: Prompt Tuning for Taxonomic Open Set Classification
di: Wu, Tz-Ying, et al.
Pubblicazione: (2023)
di: Wu, Tz-Ying, et al.
Pubblicazione: (2023)
Long-Tailed Anomaly Detection with Learnable Class Names
di: Ho, Chih-Hui, et al.
Pubblicazione: (2024)
di: Ho, Chih-Hui, et al.
Pubblicazione: (2024)
Adapting Diffusion Models for Improved Prompt Compliance and Controllable Image Synthesis
di: Sridhar, Deepak, et al.
Pubblicazione: (2024)
di: Sridhar, Deepak, et al.
Pubblicazione: (2024)
Diffusion-based Data Augmentation for Object Counting Problems
di: Wang, Zhen, et al.
Pubblicazione: (2024)
di: Wang, Zhen, et al.
Pubblicazione: (2024)
Temporal Object-Aware Vision Transformer for Few-Shot Video Object Detection
di: Kumar, Yogesh, et al.
Pubblicazione: (2025)
di: Kumar, Yogesh, et al.
Pubblicazione: (2025)
Video Reasoning without Training
di: Sridhar, Deepak, et al.
Pubblicazione: (2025)
di: Sridhar, Deepak, et al.
Pubblicazione: (2025)
Aligning Moments in Time using Video Queries
di: Kumar, Yogesh, et al.
Pubblicazione: (2025)
di: Kumar, Yogesh, et al.
Pubblicazione: (2025)
Beyond Simple Edits: X-Planner for Complex Instruction-Based Image Editing
di: Yeh, Chun-Hsiao, et al.
Pubblicazione: (2025)
di: Yeh, Chun-Hsiao, et al.
Pubblicazione: (2025)
Overcoming Small Data Limitations in Video-Based Infant Respiration Estimation
di: Song, Liyang, et al.
Pubblicazione: (2025)
di: Song, Liyang, et al.
Pubblicazione: (2025)
Xray-Visual Models: Scaling Vision models on Industry Scale Data
di: Mishra, Shlok, et al.
Pubblicazione: (2026)
di: Mishra, Shlok, et al.
Pubblicazione: (2026)
"Previously on ..." From Recaps to Story Summarization
di: Singh, Aditya Kumar, et al.
Pubblicazione: (2024)
di: Singh, Aditya Kumar, et al.
Pubblicazione: (2024)
From Prompt to Progression: Taming Video Diffusion Models for Seamless Attribute Transition
di: Lo, Ling, et al.
Pubblicazione: (2025)
di: Lo, Ling, et al.
Pubblicazione: (2025)
ViBe: Ultra-High-Resolution Video Synthesis Born from Pure Images
di: Wu, Yunfeng, et al.
Pubblicazione: (2026)
di: Wu, Yunfeng, et al.
Pubblicazione: (2026)
Light Interaction: Training-Free Inference Acceleration for Interactive Video World Models
di: Lu, Jiacheng, et al.
Pubblicazione: (2026)
di: Lu, Jiacheng, et al.
Pubblicazione: (2026)
Mitigating the Impact of Attribute Editing on Face Recognition
di: Banerjee, Sudipta, et al.
Pubblicazione: (2024)
di: Banerjee, Sudipta, et al.
Pubblicazione: (2024)
Towards Universal Video MLLMs with Attribute-Structured and Quality-Verified Instructions
di: Li, Yunheng, et al.
Pubblicazione: (2026)
di: Li, Yunheng, et al.
Pubblicazione: (2026)
Video-KTR: Reinforcing Video Reasoning via Key Token Attribution
di: Wang, Ziyue, et al.
Pubblicazione: (2026)
di: Wang, Ziyue, et al.
Pubblicazione: (2026)
Plug-and-Play Versatile Compressed Video Enhancement
di: Zeng, Huimin, et al.
Pubblicazione: (2025)
di: Zeng, Huimin, et al.
Pubblicazione: (2025)
Advanced Gesture Recognition for Autism Spectrum Disorder Detection: Integrating YOLOv7, Video Augmentation, and VideoMAE for Naturalistic Video Analysis
di: Singh, Amit Kumar, et al.
Pubblicazione: (2024)
di: Singh, Amit Kumar, et al.
Pubblicazione: (2024)
Attribution as Retrieval: Model-Agnostic AI-Generated Image Attribution
di: Wang, Hongsong, et al.
Pubblicazione: (2026)
di: Wang, Hongsong, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Leveraging Data to Say No: Memory Augmented Plug-and-Play Selective Prediction
di: Sarkar, Aditya, et al.
Pubblicazione: (2026) -
Unifying Contrastive and Generative Objectives for Visual Understanding and Text-to-Image Generation
di: Li, Chao, et al.
Pubblicazione: (2026) -
StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding
di: Yang, Yanlai, et al.
Pubblicazione: (2025) -
IntroStyle: Training-Free Introspective Style Attribution using Diffusion Features
di: Kumar, Anand, et al.
Pubblicazione: (2024) -
Transfer between Modalities with MetaQueries
di: Pan, Xichen, et al.
Pubblicazione: (2025)