OneEncoder: A Lightweight Framework for Progressive Alignment of Modalities
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Faye, Bilal, Azzag, Hanane, Lebbah, Mustapha |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Lightweight Modular Parameter-Efficient Tuning for Open-Vocabulary Object Detection
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
Adaptative Context Normalization: A Boost for Deep Learning in Image Processing
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
Energy-Regularized Spatial Masking: A Novel Approach to Enhancing Robustness and Interpretability in Vision Models
von: Devynck, Tom, et al.
Veröffentlicht: (2026)
von: Devynck, Tom, et al.
Veröffentlicht: (2026)
Context Normalization Layer with Applications
von: Faye, Bilal, et al.
Veröffentlicht: (2023)
von: Faye, Bilal, et al.
Veröffentlicht: (2023)
Lightweight Cross-Modal Representation Learning
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
Supervised Batch Normalization
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
Prototype-Guided Diffusion: Visual Conditioning without External Memory
von: Faye, Bilal, et al.
Veröffentlicht: (2025)
von: Faye, Bilal, et al.
Veröffentlicht: (2025)
Value-Free Policy Optimization via Reward Partitioning
von: Faye, Bilal, et al.
Veröffentlicht: (2025)
von: Faye, Bilal, et al.
Veröffentlicht: (2025)
Enhancing Neural Network Representations with Prior Knowledge-Based Normalization
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
Unsupervised Adaptive Normalization
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
von: Faye, Bilal, et al.
Veröffentlicht: (2024)
Adaptive Head Budgeting for Efficient Multi-Head Attention
von: Faye, Bilal, et al.
Veröffentlicht: (2026)
von: Faye, Bilal, et al.
Veröffentlicht: (2026)
MB-ORES: A Multi-Branch Object Reasoner for Visual Grounding in Remote Sensing
von: Radouane, Karim, et al.
Veröffentlicht: (2025)
von: Radouane, Karim, et al.
Veröffentlicht: (2025)
Game Theory Meets Statistical Mechanics in Deep Learning Design
von: Bouchaffra, Djamel, et al.
Veröffentlicht: (2024)
von: Bouchaffra, Djamel, et al.
Veröffentlicht: (2024)
Tuning Just Enough: Lightweight Backdoor Attacks on Multi-Encoder Diffusion Models
von: Chen, Ziyuan, et al.
Veröffentlicht: (2026)
von: Chen, Ziyuan, et al.
Veröffentlicht: (2026)
TinyAlign: Boosting Lightweight Vision-Language Models by Mitigating Modal Alignment Bottlenecks
von: Hu, Yuanze, et al.
Veröffentlicht: (2025)
von: Hu, Yuanze, et al.
Veröffentlicht: (2025)
Distributed MCMC inference for Bayesian Non-Parametric Latent Block Model
von: Khoufache, Reda, et al.
Veröffentlicht: (2024)
von: Khoufache, Reda, et al.
Veröffentlicht: (2024)
DiA-gnostic VLVAE: Disentangled Alignment-Constrained Vision Language Variational AutoEncoder for Robust Radiology Reporting with Missing Modalities
von: Shaik, Nagur Shareef, et al.
Veröffentlicht: (2025)
von: Shaik, Nagur Shareef, et al.
Veröffentlicht: (2025)
Modality Alignment across Trees on Heterogeneous Hyperbolic Manifolds
von: Wu, Wei, et al.
Veröffentlicht: (2025)
von: Wu, Wei, et al.
Veröffentlicht: (2025)
Improved Alignment of Modalities in Large Vision Language Models
von: Jangra, Kartik, et al.
Veröffentlicht: (2025)
von: Jangra, Kartik, et al.
Veröffentlicht: (2025)
Progressive Local Alignment for Medical Multimodal Pre-training
von: Yan, Huimin, et al.
Veröffentlicht: (2025)
von: Yan, Huimin, et al.
Veröffentlicht: (2025)
AFN: Adaptive Fusion Normalization via an Encoder-Decoder Framework
von: Zhou, Zikai, et al.
Veröffentlicht: (2023)
von: Zhou, Zikai, et al.
Veröffentlicht: (2023)
Efficient Remote Sensing with Harmonized Transfer Learning and Modality Alignment
von: Huang, Tengjun
Veröffentlicht: (2024)
von: Huang, Tengjun
Veröffentlicht: (2024)
Text-centric Alignment for Multi-Modality Learning
von: Tsai, Yun-Da, et al.
Veröffentlicht: (2024)
von: Tsai, Yun-Da, et al.
Veröffentlicht: (2024)
Lightweight Facial Landmark Detection in Thermal Images via Multi-Level Cross-Modal Knowledge Transfer
von: Tong, Qiyi, et al.
Veröffentlicht: (2025)
von: Tong, Qiyi, et al.
Veröffentlicht: (2025)
Bridging Modalities via Progressive Re-alignment for Multimodal Test-Time Adaptation
von: Li, Jiacheng, et al.
Veröffentlicht: (2025)
von: Li, Jiacheng, et al.
Veröffentlicht: (2025)
OneLLM: One Framework to Align All Modalities with Language
von: Han, Jiaming, et al.
Veröffentlicht: (2023)
von: Han, Jiaming, et al.
Veröffentlicht: (2023)
Bayesian Cross-Modal Alignment Learning for Few-Shot Out-of-Distribution Generalization
von: Zhu, Lin, et al.
Veröffentlicht: (2025)
von: Zhu, Lin, et al.
Veröffentlicht: (2025)
Progressive Alignment with VLM-LLM Feature to Augment Defect Classification for the ASE Dataset
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
EMMA: Efficient Visual Alignment in Multi-Modal LLMs
von: Ghazanfari, Sara, et al.
Veröffentlicht: (2024)
von: Ghazanfari, Sara, et al.
Veröffentlicht: (2024)
DUNIA: Pixel-Sized Embeddings via Cross-Modal Alignment for Earth Observation Applications
von: Fayad, Ibrahim, et al.
Veröffentlicht: (2025)
von: Fayad, Ibrahim, et al.
Veröffentlicht: (2025)
JANUS: A Lightweight Framework for Jailbreaking Text-to-Image Models via Distribution Optimization
von: Zheng, Haolun, et al.
Veröffentlicht: (2026)
von: Zheng, Haolun, et al.
Veröffentlicht: (2026)
Exploiting Lightweight Hierarchical ViT and Dynamic Framework for Efficient Visual Tracking
von: Kang, Ben, et al.
Veröffentlicht: (2025)
von: Kang, Ben, et al.
Veröffentlicht: (2025)
An Attentive Dual-Encoder Framework Leveraging Multimodal Visual and Semantic Information for Automatic OSAHS Diagnosis
von: Wei, Yingchen, et al.
Veröffentlicht: (2024)
von: Wei, Yingchen, et al.
Veröffentlicht: (2024)
Data Selection for Fine-tuning Vision Language Models via Cross Modal Alignment Trajectories
von: Naharas, Nilay, et al.
Veröffentlicht: (2025)
von: Naharas, Nilay, et al.
Veröffentlicht: (2025)
GNSP: Gradient Null Space Projection for Preserving Cross-Modal Alignment in VLMs Continual Learning
von: Peng, Tiantian, et al.
Veröffentlicht: (2025)
von: Peng, Tiantian, et al.
Veröffentlicht: (2025)
Semi-MedRef: Semi-Supervised Medical Referring Image Segmentation with Cross-Modal Alignment
von: Li, Yuchen, et al.
Veröffentlicht: (2026)
von: Li, Yuchen, et al.
Veröffentlicht: (2026)
X-VILA: Cross-Modality Alignment for Large Language Model
von: Ye, Hanrong, et al.
Veröffentlicht: (2024)
von: Ye, Hanrong, et al.
Veröffentlicht: (2024)
Regularization by denoising: Bayesian model and Langevin-within-split Gibbs sampling
von: Faye, Elhadji C., et al.
Veröffentlicht: (2024)
von: Faye, Elhadji C., et al.
Veröffentlicht: (2024)
Multi-Modal Landslide Detection from Sentinel-1 SAR and Sentinel-2 Optical Imagery Using Multi-Encoder Vision Transformers and Ensemble Learning
von: Nasios, Ioannis
Veröffentlicht: (2026)
von: Nasios, Ioannis
Veröffentlicht: (2026)
General-Purpose Multi-Modal OOD Detection Framework
von: Duong, Viet, et al.
Veröffentlicht: (2023)
von: Duong, Viet, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Lightweight Modular Parameter-Efficient Tuning for Open-Vocabulary Object Detection
von: Faye, Bilal, et al.
Veröffentlicht: (2024) -
Adaptative Context Normalization: A Boost for Deep Learning in Image Processing
von: Faye, Bilal, et al.
Veröffentlicht: (2024) -
Energy-Regularized Spatial Masking: A Novel Approach to Enhancing Robustness and Interpretability in Vision Models
von: Devynck, Tom, et al.
Veröffentlicht: (2026) -
Context Normalization Layer with Applications
von: Faye, Bilal, et al.
Veröffentlicht: (2023) -
Lightweight Cross-Modal Representation Learning
von: Faye, Bilal, et al.
Veröffentlicht: (2024)