Attention Itself Could Retrieve.RetrieveVGGT: Training-Free Long Context Streaming 3D Reconstruction via Query-Key Similarity Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zou, Zichen, Jia, Xiaosong, Wu, Zuxuan, Jiang, Yu-Gang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval
von: Liu, Di, et al.
Veröffentlicht: (2024)
von: Liu, Di, et al.
Veröffentlicht: (2024)
Efficient-LVSM: Faster, Cheaper, and Better Large View Synthesis Model via Decoupled Co-Refinement Attention
von: Jia, Xiaosong, et al.
Veröffentlicht: (2026)
von: Jia, Xiaosong, et al.
Veröffentlicht: (2026)
Attention Reveals More Than Tokens: Training-Free Long-Context Reasoning with Attention-guided Retrieval
von: Zhang, Yuwei, et al.
Veröffentlicht: (2025)
von: Zhang, Yuwei, et al.
Veröffentlicht: (2025)
ZigzagAttention: Efficient Long-Context Inference with Exclusive Retrieval and Streaming Heads
von: Liu, Zhuorui, et al.
Veröffentlicht: (2025)
von: Liu, Zhuorui, et al.
Veröffentlicht: (2025)
DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads
von: Xiao, Guangxuan, et al.
Veröffentlicht: (2024)
von: Xiao, Guangxuan, et al.
Veröffentlicht: (2024)
Activation-aware Probe-Query: Effective Key-Value Retrieval for Long-Context LLMs Inference
von: Xiao, Qingfa, et al.
Veröffentlicht: (2025)
von: Xiao, Qingfa, et al.
Veröffentlicht: (2025)
AttentionRetriever: Attention Layers are Secretly Long Document Retrievers
von: Fu, David Jiahao, et al.
Veröffentlicht: (2026)
von: Fu, David Jiahao, et al.
Veröffentlicht: (2026)
SEAL: Scaling to Emphasize Attention for Long-Context Retrieval
von: Lee, Changhun, et al.
Veröffentlicht: (2025)
von: Lee, Changhun, et al.
Veröffentlicht: (2025)
Efficient Length-Generalizable Attention via Causal Retrieval for Long-Context Language Modeling
von: Hu, Xiang, et al.
Veröffentlicht: (2024)
von: Hu, Xiang, et al.
Veröffentlicht: (2024)
Repeating Words for Video-Language Retrieval with Coarse-to-Fine Objectives
von: Zhao, Haoyu, et al.
Veröffentlicht: (2025)
von: Zhao, Haoyu, et al.
Veröffentlicht: (2025)
Infinite Retrieval: Attention Enhanced LLMs in Long-Context Processing
von: Ye, Xiaoju, et al.
Veröffentlicht: (2025)
von: Ye, Xiaoju, et al.
Veröffentlicht: (2025)
Retrieval, Reward, and Training Protocols: What Matters in Training Search Agents?
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
Query-Focused Retrieval Heads Improve Long-Context Reasoning and Re-ranking
von: Zhang, Wuwei, et al.
Veröffentlicht: (2025)
von: Zhang, Wuwei, et al.
Veröffentlicht: (2025)
Spatial Retrieval Augmented Autonomous Driving
von: Jia, Xiaosong, et al.
Veröffentlicht: (2025)
von: Jia, Xiaosong, et al.
Veröffentlicht: (2025)
SkewRoute: Training-Free LLM Routing for Knowledge Graph Retrieval-Augmented Generation via Score Skewness of Retrieved Context
von: Wang, Hairu, et al.
Veröffentlicht: (2025)
von: Wang, Hairu, et al.
Veröffentlicht: (2025)
You Only Use Reactive Attention Slice For Long Context Retrieval
von: Soh, Yun Joon, et al.
Veröffentlicht: (2024)
von: Soh, Yun Joon, et al.
Veröffentlicht: (2024)
Query Suggestion for Retrieval-Augmented Generation via Dynamic In-Context Learning
von: Spaeh, Fabian, et al.
Veröffentlicht: (2026)
von: Spaeh, Fabian, et al.
Veröffentlicht: (2026)
Retrieval with Learned Similarities
von: Ding, Bailu, et al.
Veröffentlicht: (2024)
von: Ding, Bailu, et al.
Veröffentlicht: (2024)
CTkvr: KV Cache Retrieval for Long-Context LLMs via Centroid then Token Indexing
von: Lu, Kuan, et al.
Veröffentlicht: (2025)
von: Lu, Kuan, et al.
Veröffentlicht: (2025)
S$^3$-Attention:Attention-Aligned Endogenous Retrieval for Memory-Bounded Long-Context Inference
von: Ma, Qingsen, et al.
Veröffentlicht: (2026)
von: Ma, Qingsen, et al.
Veröffentlicht: (2026)
Reason-before-Retrieve: One-Stage Reflective Chain-of-Thoughts for Training-Free Zero-Shot Composed Image Retrieval
von: Tang, Yuanmin, et al.
Veröffentlicht: (2024)
von: Tang, Yuanmin, et al.
Veröffentlicht: (2024)
Training-Free Personalization via Retrieval and Reasoning on Fingerprints
von: Das, Deepayan, et al.
Veröffentlicht: (2025)
von: Das, Deepayan, et al.
Veröffentlicht: (2025)
GeoGS3D: Single-view 3D Reconstruction via Geometric-aware Diffusion Model and Gaussian Splatting
von: Feng, Qijun, et al.
Veröffentlicht: (2024)
von: Feng, Qijun, et al.
Veröffentlicht: (2024)
LongEmbed: Extending Embedding Models for Long Context Retrieval
von: Zhu, Dawei, et al.
Veröffentlicht: (2024)
von: Zhu, Dawei, et al.
Veröffentlicht: (2024)
FreeRet: MLLMs as Training-Free Retrievers
von: Zhu, Yuhan, et al.
Veröffentlicht: (2025)
von: Zhu, Yuhan, et al.
Veröffentlicht: (2025)
Controlled Retrieval-augmented Context Evaluation for Long-form RAG
von: Ju, Jia-Huei, et al.
Veröffentlicht: (2025)
von: Ju, Jia-Huei, et al.
Veröffentlicht: (2025)
Retrieval Head Mechanistically Explains Long-Context Factuality
von: Wu, Wenhao, et al.
Veröffentlicht: (2024)
von: Wu, Wenhao, et al.
Veröffentlicht: (2024)
RAQG-QPP: Query Performance Prediction with Retrieved Query Variants and Retrieval Augmented Query Generation
von: Tian, Fangzheng, et al.
Veröffentlicht: (2026)
von: Tian, Fangzheng, et al.
Veröffentlicht: (2026)
Literary Evidence Retrieval via Long-Context Language Models
von: Thai, Katherine, et al.
Veröffentlicht: (2025)
von: Thai, Katherine, et al.
Veröffentlicht: (2025)
Improving Retrieval in Sponsored Search by Leveraging Query Context Signals
von: Mohankumar, Akash Kumar, et al.
Veröffentlicht: (2024)
von: Mohankumar, Akash Kumar, et al.
Veröffentlicht: (2024)
Attendre: Wait To Attend By Retrieval With Evicted Queries in Memory-Based Transformers for Long Context Processing
von: Yang, Zi, et al.
Veröffentlicht: (2024)
von: Yang, Zi, et al.
Veröffentlicht: (2024)
Training-free Zero-shot Composed Image Retrieval via Weighted Modality Fusion and Similarity
von: Wu, Ren-Di, et al.
Veröffentlicht: (2024)
von: Wu, Ren-Di, et al.
Veröffentlicht: (2024)
Q-RAG: Long Context Multi-step Retrieval via Value-based Embedder Training
von: Sorokin, Artyom, et al.
Veröffentlicht: (2025)
von: Sorokin, Artyom, et al.
Veröffentlicht: (2025)
Retrieval meets Long Context Large Language Models
von: Xu, Peng, et al.
Veröffentlicht: (2023)
von: Xu, Peng, et al.
Veröffentlicht: (2023)
Long$^2$RAG: Evaluating Long-Context & Long-Form Retrieval-Augmented Generation with Key Point Recall
von: Qi, Zehan, et al.
Veröffentlicht: (2024)
von: Qi, Zehan, et al.
Veröffentlicht: (2024)
RARe: Retrieval Augmented Retrieval with In-Context Examples
von: Tejaswi, Atula, et al.
Veröffentlicht: (2024)
von: Tejaswi, Atula, et al.
Veröffentlicht: (2024)
Search and Detect: Training-Free Long Tail Object Detection via Web-Image Retrieval
von: Sidhu, Mankeerat, et al.
Veröffentlicht: (2024)
von: Sidhu, Mankeerat, et al.
Veröffentlicht: (2024)
Decomposing Queries into Tool Calls for Long-Video Keyframe Retrieval
von: Shlapentokh-Rothman, Michal, et al.
Veröffentlicht: (2026)
von: Shlapentokh-Rothman, Michal, et al.
Veröffentlicht: (2026)
Personalize Before Retrieve: LLM-based Personalized Query Expansion for User-Centric Retrieval
von: Zhang, Yingyi, et al.
Veröffentlicht: (2025)
von: Zhang, Yingyi, et al.
Veröffentlicht: (2025)
Diffusion Augmented Retrieval: A Training-Free Approach to Interactive Text-to-Image Retrieval
von: Long, Zijun, et al.
Veröffentlicht: (2025)
von: Long, Zijun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval
von: Liu, Di, et al.
Veröffentlicht: (2024) -
Efficient-LVSM: Faster, Cheaper, and Better Large View Synthesis Model via Decoupled Co-Refinement Attention
von: Jia, Xiaosong, et al.
Veröffentlicht: (2026) -
Attention Reveals More Than Tokens: Training-Free Long-Context Reasoning with Attention-guided Retrieval
von: Zhang, Yuwei, et al.
Veröffentlicht: (2025) -
ZigzagAttention: Efficient Long-Context Inference with Exclusive Retrieval and Streaming Heads
von: Liu, Zhuorui, et al.
Veröffentlicht: (2025) -
DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads
von: Xiao, Guangxuan, et al.
Veröffentlicht: (2024)