CausalEmbed: Auto-Regressive Multi-Vector Generation in Latent Space for Visual Document Embedding
Fuente:
arXiv
Saved in:
| Main Authors: | Huo, Jiahao, Huang, Yu, Yan, Yibo, Pan, Ye, Zheng, Kening, Huang, Wei-Chieh, Cao, Yi, Ou, Mingdong, Yu, Philip S., Hu, Xuming |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sculpting the Vector Space: Towards Efficient Multi-Vector Visual Document Retrieval via Prune-then-Merge Framework
by: Yan, Yibo, et al.
Published: (2026)
by: Yan, Yibo, et al.
Published: (2026)
Beyond the Grid: Layout-Informed Multi-Vector Retrieval with Parsed Visual Document Representations
by: Yan, Yibo, et al.
Published: (2026)
by: Yan, Yibo, et al.
Published: (2026)
Visual Late Chunking: An Empirical Study of Contextual Chunking for Efficient Visual Document Retrieval
by: Yan, Yibo, et al.
Published: (2026)
by: Yan, Yibo, et al.
Published: (2026)
Unlocking Multimodal Document Intelligence: From Current Triumphs to Future Frontiers of Visual Document Retrieval
by: Yan, Yibo, et al.
Published: (2026)
by: Yan, Yibo, et al.
Published: (2026)
SAMark: A Self-Anchored Text Watermarking with Paragraph-Level Paraphrase Robustness
by: Huo, Jiahao, et al.
Published: (2026)
by: Huo, Jiahao, et al.
Published: (2026)
DocPruner: A Storage-Efficient Framework for Multi-Vector Visual Document Retrieval via Adaptive Patch-Level Embedding Pruning
by: Yan, Yibo, et al.
Published: (2025)
by: Yan, Yibo, et al.
Published: (2025)
Unveiling Language Routing Isolation in Multilingual MoE Models for Interpretable Subnetwork Adaptation
by: Zheng, Kening, et al.
Published: (2026)
by: Zheng, Kening, et al.
Published: (2026)
MathAgent: Leveraging a Mixture-of-Math-Agent Framework for Real-World Multimodal Mathematical Error Detection
by: Yan, Yibo, et al.
Published: (2025)
by: Yan, Yibo, et al.
Published: (2025)
Can You Trust the Vectors in Your Vector Database? Black-Hole Attack from Embedding Space Defects
by: Li, Hanxi, et al.
Published: (2026)
by: Li, Hanxi, et al.
Published: (2026)
WindVE: Collaborative CPU-NPU Vector Embedding
by: Huang, Jinqi, et al.
Published: (2025)
by: Huang, Jinqi, et al.
Published: (2025)
AKRMap: Adaptive Kernel Regression for Trustworthy Visualization of Cross-Modal Embeddings
by: Ye, Yilin, et al.
Published: (2025)
by: Ye, Yilin, et al.
Published: (2025)
MINER: Mining the Underlying Pattern of Modality-Specific Neurons in Multimodal Large Language Models
by: Huang, Kaichen, et al.
Published: (2024)
by: Huang, Kaichen, et al.
Published: (2024)
Reefknot: A Comprehensive Benchmark for Relation Hallucination Evaluation, Analysis and Mitigation in Multimodal Large Language Models
by: Zheng, Kening, et al.
Published: (2024)
by: Zheng, Kening, et al.
Published: (2024)
Look Twice Before You Answer: Memory-Space Visual Retracing for Hallucination Mitigation in Multimodal Large Language Models
by: Zou, Xin, et al.
Published: (2024)
by: Zou, Xin, et al.
Published: (2024)
MMNeuron: Discovering Neuron-Level Domain-Specific Interpretation in Multimodal Large Language Model
by: Huo, Jiahao, et al.
Published: (2024)
by: Huo, Jiahao, et al.
Published: (2024)
Bayesian Vector AutoRegression with Factorised Granger-Causal Graphs
by: Zhao, He, et al.
Published: (2024)
by: Zhao, He, et al.
Published: (2024)
LatentGandr: Visual Exploration of Generative AI Latent Space via Local Embeddings
by: Li, Mingwei, et al.
Published: (2026)
by: Li, Mingwei, et al.
Published: (2026)
Pierce the Mists, Greet the Sky: Decipher Knowledge Overshadowing via Knowledge Circuit Analysis
by: Huang, Haoming, et al.
Published: (2025)
by: Huang, Haoming, et al.
Published: (2025)
GM-PRM: A Generative Multimodal Process Reward Model for Multimodal Mathematical Reasoning
by: Zhang, Jianghangfan, et al.
Published: (2025)
by: Zhang, Jianghangfan, et al.
Published: (2025)
Exploring Visual Embedding Spaces Induced by Vision Transformers for Online Auto Parts Marketplaces
by: Armijo, Cameron, et al.
Published: (2025)
by: Armijo, Cameron, et al.
Published: (2025)
Educational Cone Model in Embedding Vector Spaces
by: Ehara, Yo
Published: (2025)
by: Ehara, Yo
Published: (2025)
Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning
by: Yan, Yibo, et al.
Published: (2025)
by: Yan, Yibo, et al.
Published: (2025)
MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings
by: Li, Zijie, et al.
Published: (2026)
by: Li, Zijie, et al.
Published: (2026)
Nomic Embed Vision: Expanding the Latent Space
by: Nussbaum, Zach, et al.
Published: (2024)
by: Nussbaum, Zach, et al.
Published: (2024)
TreeHop: Generate and Filter Next Query Embeddings Efficiently for Multi-hop Question Answering
by: Li, Zhonghao, et al.
Published: (2025)
by: Li, Zhonghao, et al.
Published: (2025)
Matrix-Weighted Besov-Triebel-Lizorkin Spaces of Optimal Scale: Real-Variable Characterizations, Invariance on Integrable Index, and Sobolev-Type Embedding
by: Bu, Fan, et al.
Published: (2025)
by: Bu, Fan, et al.
Published: (2025)
Give Me More Details: Improving Fact-Checking with Latent Retrieval
by: Hu, Xuming, et al.
Published: (2023)
by: Hu, Xuming, et al.
Published: (2023)
Network Embedding with Completely-imbalanced Labels
by: Wang, Zheng, et al.
Published: (2020)
by: Wang, Zheng, et al.
Published: (2020)
Unifying Multimodal Retrieval via Document Screenshot Embedding
by: Ma, Xueguang, et al.
Published: (2024)
by: Ma, Xueguang, et al.
Published: (2024)
Influence of Cooling Rate on Residual Strain Evolution and Interlaminar Properties in Thermoplastic Composites With Embedded FBG Monitoring
by: Zijing Qu, et al.
Published: (2025)
by: Zijing Qu, et al.
Published: (2025)
Refiner: Restructure Retrieval Content Efficiently to Advance Question-Answering Capabilities
by: Li, Zhonghao, et al.
Published: (2024)
by: Li, Zhonghao, et al.
Published: (2024)
Disaggregating Embedding Recommendation Systems with FlexEMR
by: Huang, Yibo, et al.
Published: (2024)
by: Huang, Yibo, et al.
Published: (2024)
Hierarchies over Vector Space: Orienting Word and Graph Embeddings
by: Guo, Xingzhi, et al.
Published: (2022)
by: Guo, Xingzhi, et al.
Published: (2022)
Vector-Valued Native Space Embedding for Adaptive State Observation
by: Niu, Shengyuan, et al.
Published: (2025)
by: Niu, Shengyuan, et al.
Published: (2025)
TopoBind: Multi-Modal Prediction of Antibody-Antigen Binding Free Energy via Sequence Embeddings and Structural Topology
by: Yu, Ciyuan, et al.
Published: (2025)
by: Yu, Ciyuan, et al.
Published: (2025)
Poly-Vector Retrieval: Reference and Content Embeddings for Legal Documents
by: Lima, João Alberto de Oliveira
Published: (2025)
by: Lima, João Alberto de Oliveira
Published: (2025)
Towards Storage-Efficient Visual Document Retrieval: An Empirical Study on Reducing Patch-Level Embeddings
by: Ma, Yubo, et al.
Published: (2025)
by: Ma, Yubo, et al.
Published: (2025)
Transferable Embedding Inversion Attack: Uncovering Privacy Risks in Text Embeddings without Model Queries
by: Huang, Yu-Hsiang, et al.
Published: (2024)
by: Huang, Yu-Hsiang, et al.
Published: (2024)
MMUnlearner: Reformulating Multimodal Machine Unlearning in the Era of Multimodal Large Language Models
by: Huo, Jiahao, et al.
Published: (2025)
by: Huo, Jiahao, et al.
Published: (2025)
The Illusion of Role Separation: Hidden Shortcuts in LLM Role Learning (and How to Fix Them)
by: Wang, Zihao, et al.
Published: (2025)
by: Wang, Zihao, et al.
Published: (2025)
Similar Items
-
Sculpting the Vector Space: Towards Efficient Multi-Vector Visual Document Retrieval via Prune-then-Merge Framework
by: Yan, Yibo, et al.
Published: (2026) -
Beyond the Grid: Layout-Informed Multi-Vector Retrieval with Parsed Visual Document Representations
by: Yan, Yibo, et al.
Published: (2026) -
Visual Late Chunking: An Empirical Study of Contextual Chunking for Efficient Visual Document Retrieval
by: Yan, Yibo, et al.
Published: (2026) -
Unlocking Multimodal Document Intelligence: From Current Triumphs to Future Frontiers of Visual Document Retrieval
by: Yan, Yibo, et al.
Published: (2026) -
SAMark: A Self-Anchored Text Watermarking with Paragraph-Level Paraphrase Robustness
by: Huo, Jiahao, et al.
Published: (2026)