Less is More - diveXplore 5.0 at VBS 2021
Fuente:
arXiv
Salvato in:
| Autori principali: | Leibetseder, Andreas, Schoeffmann, Klaus |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
diveXplore 6.0: ITEC's Interactive Video Exploration System at VBS 2022
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025)
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025)
diveXplore at the Video Browser Showdown 2024
di: Schoeffmann, Klaus, et al.
Pubblicazione: (2025)
di: Schoeffmann, Klaus, et al.
Pubblicazione: (2025)
lifeXplore at the Lifelog Search Challenge 2021
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025)
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025)
lifeXplore at the Lifelog Search Challenge 2020
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025)
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025)
Post-surgical Endometriosis Segmentation in Laparoscopic Videos
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025)
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025)
GLENDA: Gynecologic Laparoscopy Endometriosis Dataset
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025)
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025)
Video Quality Assessment with Texture Information Fusion for Streaming Applications
di: Menon, Vignesh V, et al.
Pubblicazione: (2023)
di: Menon, Vignesh V, et al.
Pubblicazione: (2023)
Results of the 2025 Video Browser Showdown
di: Rossetto, Luca, et al.
Pubblicazione: (2025)
di: Rossetto, Luca, et al.
Pubblicazione: (2025)
Results of the 2024 Video Browser Showdown
di: Rossetto, Luca, et al.
Pubblicazione: (2024)
di: Rossetto, Luca, et al.
Pubblicazione: (2024)
Identifying Surgical Instruments in Laparoscopy Using Deep Learning Instance Segmentation
di: Kletz, Sabrina, et al.
Pubblicazione: (2025)
di: Kletz, Sabrina, et al.
Pubblicazione: (2025)
Optimal Quality and Efficiency in Adaptive Live Streaming with JND-Aware Low latency Encoding
di: Menon, Vignesh V, et al.
Pubblicazione: (2024)
di: Menon, Vignesh V, et al.
Pubblicazione: (2024)
Less for More: Enhanced Feedback-aligned Mixed LLMs for Molecule Caption Generation and Fine-Grained NLI Evaluation
di: Gkoumas, Dimitris, et al.
Pubblicazione: (2024)
di: Gkoumas, Dimitris, et al.
Pubblicazione: (2024)
The State-of-the-Art in Lifelog Retrieval: A Review of Progress at the ACM Lifelog Search Challenge Workshop 2022-24
di: Tran, Allie, et al.
Pubblicazione: (2025)
di: Tran, Allie, et al.
Pubblicazione: (2025)
Less is More: A Simple yet Effective Token Reduction Method for Efficient Multi-modal LLMs
di: Song, Dingjie, et al.
Pubblicazione: (2024)
di: Song, Dingjie, et al.
Pubblicazione: (2024)
MCPNS: A Macropixel Collocated Position and Its Neighbors Search for Plenoptic 2.0 Video Coding
di: Van Duong, Vinh, et al.
Pubblicazione: (2023)
di: Van Duong, Vinh, et al.
Pubblicazione: (2023)
Getting More for Less: Using Weak Labels and AV-Mixup for Robust Audio-Visual Speaker Verification
di: Selvakumar, Anith, et al.
Pubblicazione: (2023)
di: Selvakumar, Anith, et al.
Pubblicazione: (2023)
Design of a 5G Multimedia Broadcast Application Function Supporting Adaptive Error Recovery
di: Lentisco, C. M., et al.
Pubblicazione: (2024)
di: Lentisco, C. M., et al.
Pubblicazione: (2024)
Unlearning the Noisy Correspondence Makes CLIP More Robust
di: Han, Haochen, et al.
Pubblicazione: (2025)
di: Han, Haochen, et al.
Pubblicazione: (2025)
MultiMediate'24: Multi-Domain Engagement Estimation
di: Müller, Philipp, et al.
Pubblicazione: (2024)
di: Müller, Philipp, et al.
Pubblicazione: (2024)
An Experimental Study of Low-Latency Video Streaming over 5G
di: Khan, Imran, et al.
Pubblicazione: (2024)
di: Khan, Imran, et al.
Pubblicazione: (2024)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
di: Zhang, Zhenxing, et al.
Pubblicazione: (2024)
di: Zhang, Zhenxing, et al.
Pubblicazione: (2024)
"The Intangible Victory", Interactive Audiovisual Installation
di: Tsioutas, Konstantinos, et al.
Pubblicazione: (2026)
di: Tsioutas, Konstantinos, et al.
Pubblicazione: (2026)
MIntRec2.0: A Large-scale Benchmark Dataset for Multimodal Intent Recognition and Out-of-scope Detection in Conversations
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
Inferencias causales durante la comprensión de textos expositivos en formato multimedia
di: Gastón Saux
Pubblicazione: (2012)
di: Gastón Saux
Pubblicazione: (2012)
Machine Learning-Based Prediction of Quality Shifts on Video Streaming Over 5G
di: Mustafa, Raza Ul, et al.
Pubblicazione: (2025)
di: Mustafa, Raza Ul, et al.
Pubblicazione: (2025)
Contribution-Guided Asymmetric Learning for Robust Multimodal Fusion under Imbalance and Noise
di: Xu, Zijing, et al.
Pubblicazione: (2025)
di: Xu, Zijing, et al.
Pubblicazione: (2025)
Learning Quality from Complexity and Structure: A Feature-Fused XGBoost Model for Video Quality Assessment
di: Premkumar, Amritha, et al.
Pubblicazione: (2025)
di: Premkumar, Amritha, et al.
Pubblicazione: (2025)
ISMAF: Intrinsic-Social Modality Alignment and Fusion for Multimodal Rumor Detection
di: Yu, Zihao, et al.
Pubblicazione: (2025)
di: Yu, Zihao, et al.
Pubblicazione: (2025)
Merge Mode for Template-based Intra Mode Derivation (TIMD) in ECM
di: Abdoli, Mohsen, et al.
Pubblicazione: (2025)
di: Abdoli, Mohsen, et al.
Pubblicazione: (2025)
MMC: Iterative Refinement of VLM Reasoning via MCTS-based Multimodal Critique
di: Liu, Shuhang, et al.
Pubblicazione: (2025)
di: Liu, Shuhang, et al.
Pubblicazione: (2025)
FakeSV-VLM: Taming VLM for Detecting Fake Short-Video News via Progressive Mixture-Of-Experts Adapter
di: Wang, Junxi, et al.
Pubblicazione: (2025)
di: Wang, Junxi, et al.
Pubblicazione: (2025)
Evaluation of Objective Image Quality Metrics for High-Fidelity Image Compression
di: Mohammadi, Shima, et al.
Pubblicazione: (2025)
di: Mohammadi, Shima, et al.
Pubblicazione: (2025)
Harnessing Multimodal Large Language Models for Personalized Product Search with Query-aware Refinement
di: Zhang, Beibei, et al.
Pubblicazione: (2025)
di: Zhang, Beibei, et al.
Pubblicazione: (2025)
VRAgent-R1: Boosting Video Recommendation with MLLM-based Agents via Reinforcement Learning
di: Chen, Siran, et al.
Pubblicazione: (2025)
di: Chen, Siran, et al.
Pubblicazione: (2025)
Orthogonal Disentanglement with Projected Feature Alignment for Multimodal Emotion Recognition in Conversation
di: Che, Xinyi, et al.
Pubblicazione: (2025)
di: Che, Xinyi, et al.
Pubblicazione: (2025)
Towards Structure-aware Model for Multi-modal Knowledge Graph Completion
di: Li, Linyu, et al.
Pubblicazione: (2025)
di: Li, Linyu, et al.
Pubblicazione: (2025)
Challenging Dataset and Multi-modal Gated Mixture of Experts Model for Remote Sensing Copy-Move Forgery Understanding
di: Zhang, Ze, et al.
Pubblicazione: (2025)
di: Zhang, Ze, et al.
Pubblicazione: (2025)
Efficient and Accurate Image Provenance Analysis: A Scalable Pipeline for Large-scale Images
di: Lai, Jiewei, et al.
Pubblicazione: (2025)
di: Lai, Jiewei, et al.
Pubblicazione: (2025)
Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency
di: Zarghani, Abolfazl, et al.
Pubblicazione: (2025)
di: Zarghani, Abolfazl, et al.
Pubblicazione: (2025)
Latent Feature-Guided Conditional Diffusion for Generative Image Semantic Communication
di: Chen, Zehao, et al.
Pubblicazione: (2025)
di: Chen, Zehao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
diveXplore 6.0: ITEC's Interactive Video Exploration System at VBS 2022
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025) -
diveXplore at the Video Browser Showdown 2024
di: Schoeffmann, Klaus, et al.
Pubblicazione: (2025) -
lifeXplore at the Lifelog Search Challenge 2021
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025) -
lifeXplore at the Lifelog Search Challenge 2020
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025) -
Post-surgical Endometriosis Segmentation in Laparoscopic Videos
di: Leibetseder, Andreas, et al.
Pubblicazione: (2025)