EAGLE: Efficient Adaptive Geometry-based Learning in Cross-view Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Truong, Thanh-Dat, Prabhu, Utsav, Wang, Dongyi, Raj, Bhiksha, Gauch, Susan, Subbiah, Jeyamkondan, Luu, Khoa |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FALCON: Fairness Learning via Contrastive Attention Approach to Continual Semantic Scene Understanding
by: Truong, Thanh-Dat, et al.
Published: (2023)
by: Truong, Thanh-Dat, et al.
Published: (2023)
Cross-view Action Recognition Understanding From Exocentric to Egocentric Perspective
by: Truong, Thanh-Dat, et al.
Published: (2023)
by: Truong, Thanh-Dat, et al.
Published: (2023)
ED-SAM: An Efficient Diffusion Sampling Approach to Domain Generalization in Vision-Language Foundation Models
by: Truong, Thanh-Dat, et al.
Published: (2024)
by: Truong, Thanh-Dat, et al.
Published: (2024)
$ϕ$-DPO: Fairness Direct Preference Optimization Approach to Continual Learning in Large Multimodal Models
by: Truong, Thanh-Dat, et al.
Published: (2026)
by: Truong, Thanh-Dat, et al.
Published: (2026)
Multi-view Action Recognition via Directed Gromov-Wasserstein Discrepancy
by: Nguyen, Hoang-Quan, et al.
Published: (2024)
by: Nguyen, Hoang-Quan, et al.
Published: (2024)
Directed-Tokens: A Robust Multi-Modality Alignment Approach to Large Language-Vision Models
by: Truong, Thanh-Dat, et al.
Published: (2025)
by: Truong, Thanh-Dat, et al.
Published: (2025)
MANGO: Multimodal Attention-based Normalizing Flow Approach to Fusion Learning
by: Truong, Thanh-Dat, et al.
Published: (2025)
by: Truong, Thanh-Dat, et al.
Published: (2025)
BRAIN: Bias-Mitigation Continual Learning Approach to Vision-Brain Understanding
by: Nguyen, Xuan-Bac, et al.
Published: (2025)
by: Nguyen, Xuan-Bac, et al.
Published: (2025)
BIMA: Bijective Maximum Likelihood Learning Approach to Hallucination Prediction and Mitigation in Large Vision-Language Models
by: Tran, Huu-Thien, et al.
Published: (2025)
by: Tran, Huu-Thien, et al.
Published: (2025)
CONDA: Continual Unsupervised Domain Adaptation Learning in Visual Perception for Self-Driving Cars
by: Truong, Thanh-Dat, et al.
Published: (2022)
by: Truong, Thanh-Dat, et al.
Published: (2022)
FLAASH: Flow-Attention Adaptive Semantic Hierarchical Fusion for Multi-Modal Tobacco Content Analysis
by: Chappa, Naga VS Raviteja, et al.
Published: (2024)
by: Chappa, Naga VS Raviteja, et al.
Published: (2024)
Towards Robust and Fair Vision Learning in Open-World Environments
by: Truong, Thanh-Dat
Published: (2024)
by: Truong, Thanh-Dat
Published: (2024)
Insect-Foundation: A Foundation Model and Large Multimodal Dataset for Vision-Language Insect Understanding
by: Truong, Thanh-Dat, et al.
Published: (2025)
by: Truong, Thanh-Dat, et al.
Published: (2025)
Insect-Foundation: A Foundation Model and Large-scale 1M Dataset for Visual Insect Understanding
by: Nguyen, Hoang-Quan, et al.
Published: (2023)
by: Nguyen, Hoang-Quan, et al.
Published: (2023)
TinyBEV: Cross Modal Knowledge Distillation for Efficient Multi Task Bird's Eye View Perception and Planning
by: Khan, Reeshad, et al.
Published: (2025)
by: Khan, Reeshad, et al.
Published: (2025)
HIG: Hierarchical Interlacement Graph Approach to Scene Graph Generation in Video Understanding
by: Nguyen, Trong-Thuan, et al.
Published: (2023)
by: Nguyen, Trong-Thuan, et al.
Published: (2023)
GIIM: Graph-based Learning of Inter- and Intra-view Dependencies for Multi-view Medical Image Diagnosis
by: Sam, Tran Bao, et al.
Published: (2026)
by: Sam, Tran Bao, et al.
Published: (2026)
NICO-RAG: Multimodal Hypergraph Retrieval-Augmented Generation for Understanding the Nicotine Public Health Crisis
by: Serna-Aguilera, Manuel, et al.
Published: (2026)
by: Serna-Aguilera, Manuel, et al.
Published: (2026)
COBRA: A Continual Learning Approach to Vision-Brain Understanding
by: Nguyen, Xuan-Bac, et al.
Published: (2024)
by: Nguyen, Xuan-Bac, et al.
Published: (2024)
Geometry-guided Cross-view Diffusion for One-to-many Cross-view Image Synthesis
by: Lin, Tao Jun, et al.
Published: (2024)
by: Lin, Tao Jun, et al.
Published: (2024)
VideoJudge: Bootstrapping Enables Scalable Supervision of MLLM-as-a-Judge for Video Understanding
by: Waheed, Abdul, et al.
Published: (2025)
by: Waheed, Abdul, et al.
Published: (2025)
Learning to Sense for Driving: Joint Optics-Sensor-Model Co-Design for Semantic Segmentation
by: Khan, Reeshad, et al.
Published: (2025)
by: Khan, Reeshad, et al.
Published: (2025)
Multi-Perspective Data Augmentation for Few-shot Object Detection
by: Vu, Anh-Khoa Nguyen, et al.
Published: (2025)
by: Vu, Anh-Khoa Nguyen, et al.
Published: (2025)
DINTR: Tracking via Diffusion-based Interpolation
by: Nguyen, Pha, et al.
Published: (2024)
by: Nguyen, Pha, et al.
Published: (2024)
LiGAR: LiDAR-Guided Hierarchical Transformer for Multi-Modal Group Activity Recognition
by: Chappa, Naga Venkata Sai Raviteja, et al.
Published: (2024)
by: Chappa, Naga Venkata Sai Raviteja, et al.
Published: (2024)
Towards multi-modal forgery representation learning for AI-generated video detection and localization
by: Le, Dat, et al.
Published: (2026)
by: Le, Dat, et al.
Published: (2026)
NeIn: Telling What You Don't Want
by: Bui, Nhat-Tan, et al.
Published: (2024)
by: Bui, Nhat-Tan, et al.
Published: (2024)
Learning Multi-view Anomaly Detection with Efficient Adaptive Selection
by: He, Haoyang, et al.
Published: (2024)
by: He, Haoyang, et al.
Published: (2024)
Hierarchical Quantum Control Gates for Functional MRI Understanding
by: Nguyen, Xuan-Bac, et al.
Published: (2024)
by: Nguyen, Xuan-Bac, et al.
Published: (2024)
Cross-view and Cross-pose Completion for 3D Human Understanding
by: Armando, Matthieu, et al.
Published: (2023)
by: Armando, Matthieu, et al.
Published: (2023)
Guiding Noisy Label Conditional Diffusion Models with Score-based Discriminator Correction
by: Cong, Dat Nguyen, et al.
Published: (2025)
by: Cong, Dat Nguyen, et al.
Published: (2025)
FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation
by: Le, Minh Khoa, et al.
Published: (2026)
by: Le, Minh Khoa, et al.
Published: (2026)
Adaptive Extensions of Unbiased Risk Estimators for Unsupervised Magnetic Resonance Image Denoising
by: Khan, Reeshad, et al.
Published: (2024)
by: Khan, Reeshad, et al.
Published: (2024)
QLAM: A Quantum Long-Attention Memory Approach to Long-Sequence Token Modeling
by: Nguyen, Hoang-Quan, et al.
Published: (2026)
by: Nguyen, Hoang-Quan, et al.
Published: (2026)
View-aware Cross-modal Distillation for Multi-view Action Recognition
by: Nguyen, Trung Thanh, et al.
Published: (2025)
by: Nguyen, Trung Thanh, et al.
Published: (2025)
EAGLE: Eigen Aggregation Learning for Object-Centric Unsupervised Semantic Segmentation
by: Kim, Chanyoung, et al.
Published: (2024)
by: Kim, Chanyoung, et al.
Published: (2024)
ML-CrAIST: Multi-scale Low-high Frequency Information-based Cross black Attention with Image Super-resolving Transformer
by: Pramanick, Alik, et al.
Published: (2024)
by: Pramanick, Alik, et al.
Published: (2024)
QDFormer: Towards Robust Audiovisual Segmentation in Complex Environments with Quantization-based Semantic Decomposition
by: Li, Xiang, et al.
Published: (2023)
by: Li, Xiang, et al.
Published: (2023)
Geo$^\textbf{2}$: Geometry-Guided Cross-view Geo-Localization and Image Synthesis
by: Zhang, Yancheng, et al.
Published: (2026)
by: Zhang, Yancheng, et al.
Published: (2026)
Learning Intra-view and Cross-view Geometric Knowledge for Stereo Matching
by: Gong, Rui, et al.
Published: (2024)
by: Gong, Rui, et al.
Published: (2024)
Similar Items
-
FALCON: Fairness Learning via Contrastive Attention Approach to Continual Semantic Scene Understanding
by: Truong, Thanh-Dat, et al.
Published: (2023) -
Cross-view Action Recognition Understanding From Exocentric to Egocentric Perspective
by: Truong, Thanh-Dat, et al.
Published: (2023) -
ED-SAM: An Efficient Diffusion Sampling Approach to Domain Generalization in Vision-Language Foundation Models
by: Truong, Thanh-Dat, et al.
Published: (2024) -
$ϕ$-DPO: Fairness Direct Preference Optimization Approach to Continual Learning in Large Multimodal Models
by: Truong, Thanh-Dat, et al.
Published: (2026) -
Multi-view Action Recognition via Directed Gromov-Wasserstein Discrepancy
by: Nguyen, Hoang-Quan, et al.
Published: (2024)