U-CESE: Unified Clip-based Event Search Engine for AI Challenge HCMC 2025
Fuente:
arXiv
Saved in:
| Main Authors: | Le, Duc-Nhuan, Nguyen, Hoang-Phuc, Lam, Thanh-Duy, Dang, Minh-Nhut, Le, Minh-Hoang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Event-Enriched Image Analysis Grand Challenge at ACM Multimedia 2025
by: Tran, Thien-Phuc, et al.
Published: (2025)
by: Tran, Thien-Phuc, et al.
Published: (2025)
N-EIoU-YOLOv9: A Signal-Aware Bounding Box Regression Loss for Lightweight Mobile Detection of Rice Leaf Diseases
by: Duc, Dung Ta Nguyen, et al.
Published: (2026)
by: Duc, Dung Ta Nguyen, et al.
Published: (2026)
GeoSearch: Augmenting Worldwide Geolocalization with Web-Scale Reverse Image Search and Image Matching
by: Le-Duc, Tung-Duong, et al.
Published: (2026)
by: Le-Duc, Tung-Duong, et al.
Published: (2026)
AgriKD: Cross-Architecture Knowledge Distillation for Efficient Leaf Disease Classification
by: Le, Minh-Dung, et al.
Published: (2026)
by: Le, Minh-Dung, et al.
Published: (2026)
OpenEvents V1: Large-Scale Benchmark Dataset for Multimodal Event Grounding
by: Nguyen, Hieu, et al.
Published: (2025)
by: Nguyen, Hieu, et al.
Published: (2025)
Leveraging knowledge distillation for partial multi-task learning from multiple remote sensing datasets
by: Lê, Hoàng-Ân, et al.
Published: (2024)
by: Lê, Hoàng-Ân, et al.
Published: (2024)
SADL: An Effective In-Context Learning Method for Compositional Visual QA
by: Dang, Long Hoang, et al.
Published: (2024)
by: Dang, Long Hoang, et al.
Published: (2024)
ReCap: Event-Aware Image Captioning with Article Retrieval and Semantic Gaussian Normalization
by: Nguyen, Thinh-Phuc, et al.
Published: (2025)
by: Nguyen, Thinh-Phuc, et al.
Published: (2025)
VietMEAgent: Culturally-Aware Few-Shot Multimodal Explanation for Vietnamese Visual Question Answering
by: Nguyen, Hai-Dang, et al.
Published: (2025)
by: Nguyen, Hai-Dang, et al.
Published: (2025)
MedXplain-VQA: Multi-Component Explainable Medical Visual Question Answering
by: Nguyen, Hai-Dang, et al.
Published: (2025)
by: Nguyen, Hai-Dang, et al.
Published: (2025)
Robust Deepfake Detection: Mitigating Spatial Attention Drift via Calibrated Complementary Ensembles
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
EDGER: EDge-Guided with HEatmap Refinement for Generalizable Image Forgery Localization
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
Unified Interactive Multimodal Moment Retrieval via Cascaded Embedding-Reranking and Temporal-Aware Score Fusion
by: Thanh, Toan Le Ngo, et al.
Published: (2025)
by: Thanh, Toan Le Ngo, et al.
Published: (2025)
SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA)
by: Nguyen, Trong-Thuan, et al.
Published: (2025)
by: Nguyen, Trong-Thuan, et al.
Published: (2025)
Universal Multi-Domain Translation via Diffusion Routers
by: Kieu, Duc, et al.
Published: (2025)
by: Kieu, Duc, et al.
Published: (2025)
Fact4ac at the Financial Misinformation Detection Challenge Task: Reference-Free Financial Misinformation Detection via Fine-Tuning and Few-Shot Prompting of Large Language Models
by: Hoang, Cuong, et al.
Published: (2026)
by: Hoang, Cuong, et al.
Published: (2026)
From Prompting to Preference Optimization: A Comparative Study of LLM-based Automated Essay Scoring
by: Nguyen, Minh Hoang, et al.
Published: (2026)
by: Nguyen, Minh Hoang, et al.
Published: (2026)
Dietary Alcalase‐Derived Fish Protein Hydrolysate Supplementation Enhanced Growth Performance, Amino Acid Profiles, and Gut Health in Giant Trevally ( Caranx ignobilis ) Fed Low Fishmeal‐Based Diets
by: Hung Duc Pham, et al.
Published: (2025)
by: Hung Duc Pham, et al.
Published: (2025)
Rethinking Top Probability from Multi-view for Distracted Driver Behaviour Localization
by: Nguyen, Quang Vinh, et al.
Published: (2024)
by: Nguyen, Quang Vinh, et al.
Published: (2024)
A conversational gesture synthesis system based on emotions and semantics
by: Hoang-Minh, Thanh
Published: (2025)
by: Hoang-Minh, Thanh
Published: (2025)
Persistent Test-time Adaptation in Recurring Testing Scenarios
by: Hoang, Trung-Hieu, et al.
Published: (2023)
by: Hoang, Trung-Hieu, et al.
Published: (2023)
Towards Signboard-Oriented Visual Question Answering: ViSignVQA Dataset, Method and Benchmark
by: Nguyen, Hieu Minh, et al.
Published: (2025)
by: Nguyen, Hieu Minh, et al.
Published: (2025)
Box for Mask and Mask for Box: weak losses for multi-task partially supervised learning
by: Lê, Hoàng-Ân, et al.
Published: (2024)
by: Lê, Hoàng-Ân, et al.
Published: (2024)
FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation
by: Le, Minh Khoa, et al.
Published: (2026)
by: Le, Minh Khoa, et al.
Published: (2026)
STER-VLM: Spatio-Temporal With Enhanced Reference Vision-Language Models
by: Nguyen-Nhu, Tinh-Anh, et al.
Published: (2025)
by: Nguyen-Nhu, Tinh-Anh, et al.
Published: (2025)
A Source Identification Problem for the Bi-Parabolic Equation Containing a Poly-harmonic Operator
by: Trong, Dang Duc, et al.
Published: (2025)
by: Trong, Dang Duc, et al.
Published: (2025)
Agentic Design Patterns: A System-Theoretic Framework
by: Dao, Minh-Dung, et al.
Published: (2026)
by: Dao, Minh-Dung, et al.
Published: (2026)
Progressive Multi-granular Alignments for Grounded Reasoning in Large Vision-Language Models
by: Le, Quang-Hung, et al.
Published: (2024)
by: Le, Quang-Hung, et al.
Published: (2024)
ToolBrain: A Flexible Reinforcement Learning Framework for Agentic Tools
by: Le, Quy Minh, et al.
Published: (2025)
by: Le, Quy Minh, et al.
Published: (2025)
GenFlow: Interactive Modular System for Image Generation
by: Nguyen, Duc-Hung, et al.
Published: (2025)
by: Nguyen, Duc-Hung, et al.
Published: (2025)
Navigating knowledge in the age of generative AI
by: Minh-Hoang Nguyen
Published: (2025)
by: Minh-Hoang Nguyen
Published: (2025)
MambaU-Lite: A Lightweight Model based on Mamba and Integrated Channel-Spatial Attention for Skin Lesion Segmentation
by: Nguyen, Thi-Nhu-Quynh, et al.
Published: (2024)
by: Nguyen, Thi-Nhu-Quynh, et al.
Published: (2024)
Scribble-Supervised Medical Image Segmentation with Dynamic Teacher Switching and Hierarchical Consistency
by: Nguyen, Thanh-Huy, et al.
Published: (2026)
by: Nguyen, Thanh-Huy, et al.
Published: (2026)
Count What You Want: Exemplar Identification and Few-shot Counting of Human Actions in the Wild
by: Huang, Yifeng, et al.
Published: (2023)
by: Huang, Yifeng, et al.
Published: (2023)
ARtVista: Gateway To Empower Anyone Into Artist
by: Hoang, Trong-Vu, et al.
Published: (2024)
by: Hoang, Trong-Vu, et al.
Published: (2024)
TP-GMOT: Tracking Generic Multiple Object by Textual Prompt with Motion-Appearance Cost (MAC) SORT
by: Anh, Duy Le Dinh, et al.
Published: (2024)
by: Anh, Duy Le Dinh, et al.
Published: (2024)
AHMsys: An Automated HVAC Modeling System for BIM Project
by: Dang, Long Hoang, et al.
Published: (2024)
by: Dang, Long Hoang, et al.
Published: (2024)
VisionGuard: Synergistic Framework for Helmet Violation Detection
by: Nguyen, Lam-Huy, et al.
Published: (2025)
by: Nguyen, Lam-Huy, et al.
Published: (2025)
MSC: A Marine Wildlife Video Dataset with Grounded Segmentation and Clip-Level Captioning
by: Truong, Quang-Trung, et al.
Published: (2025)
by: Truong, Quang-Trung, et al.
Published: (2025)
CamoFA: A Learnable Fourier-based Augmentation for Camouflage Segmentation
by: Le, Minh-Quan, et al.
Published: (2023)
by: Le, Minh-Quan, et al.
Published: (2023)
Similar Items
-
Event-Enriched Image Analysis Grand Challenge at ACM Multimedia 2025
by: Tran, Thien-Phuc, et al.
Published: (2025) -
N-EIoU-YOLOv9: A Signal-Aware Bounding Box Regression Loss for Lightweight Mobile Detection of Rice Leaf Diseases
by: Duc, Dung Ta Nguyen, et al.
Published: (2026) -
GeoSearch: Augmenting Worldwide Geolocalization with Web-Scale Reverse Image Search and Image Matching
by: Le-Duc, Tung-Duong, et al.
Published: (2026) -
AgriKD: Cross-Architecture Knowledge Distillation for Efficient Leaf Disease Classification
by: Le, Minh-Dung, et al.
Published: (2026) -
OpenEvents V1: Large-Scale Benchmark Dataset for Multimodal Event Grounding
by: Nguyen, Hieu, et al.
Published: (2025)