LensVLM: Selective Context Expansion for Compressed Visual Representation of Text
Fuente:
arXiv
Salvato in:
| Autori principali: | Xie, Roy, Friedman, Dan, Yu, Donghan, Pan, Bowen, Fifty, Christopher, Kim, Jang-Hyun, Du, Xianzhi, Gan, Zhe, Rathod, Vivek, Dhingra, Bhuwan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GenEOL: Harnessing the Generative Power of LLMs for Training-Free Sentence Embeddings
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024)
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024)
Coding Agents are Effective Long-Context Processors
di: Cao, Weili, et al.
Pubblicazione: (2026)
di: Cao, Weili, et al.
Pubblicazione: (2026)
To Trust or Not to Trust? Enhancing Large Language Models' Situated Faithfulness to External Contexts
di: Huang, Yukun, et al.
Pubblicazione: (2024)
di: Huang, Yukun, et al.
Pubblicazione: (2024)
Adversarial Math Word Problem Generation
di: Xie, Roy, et al.
Pubblicazione: (2024)
di: Xie, Roy, et al.
Pubblicazione: (2024)
Raccoon: Prompt Extraction Benchmark of LLM-Integrated Applications
di: Wang, Junlin, et al.
Pubblicazione: (2024)
di: Wang, Junlin, et al.
Pubblicazione: (2024)
Compressing LLMs: The Truth is Rarely Pure and Never Simple
di: Jaiswal, Ajay, et al.
Pubblicazione: (2023)
di: Jaiswal, Ajay, et al.
Pubblicazione: (2023)
Document-as-Image Representations Fall Short for Scientific Retrieval
di: Khalighinejad, Ghazal, et al.
Pubblicazione: (2026)
di: Khalighinejad, Ghazal, et al.
Pubblicazione: (2026)
Text-Guided Semantic Image Encoder
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2025)
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2025)
Real-time Factuality Assessment from Adversarial Feedback
di: Chen, Sanxing, et al.
Pubblicazione: (2024)
di: Chen, Sanxing, et al.
Pubblicazione: (2024)
Atomic Self-Consistency for Better Long Form Generations
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024)
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024)
RVPO: Risk-Sensitive Alignment via Variance Regularization
di: Montero, Ivan, et al.
Pubblicazione: (2026)
di: Montero, Ivan, et al.
Pubblicazione: (2026)
ChatShop: Interactive Information Seeking with Language Agents
di: Chen, Sanxing, et al.
Pubblicazione: (2024)
di: Chen, Sanxing, et al.
Pubblicazione: (2024)
Fuzzy Speculative Decoding for a Tunable Accuracy-Runtime Tradeoff
di: Holsman, Maximilian, et al.
Pubblicazione: (2025)
di: Holsman, Maximilian, et al.
Pubblicazione: (2025)
Knowing When to Stop: Efficient Context Processing via Latent Sufficiency Signals
di: Xie, Roy, et al.
Pubblicazione: (2025)
di: Xie, Roy, et al.
Pubblicazione: (2025)
When Greedy Wins: Emergent Exploitation Bias in Meta-Bandit LLM Training
di: Chen, Sanxing, et al.
Pubblicazione: (2025)
di: Chen, Sanxing, et al.
Pubblicazione: (2025)
Vision2Code: A Multi-Domain Benchmark for Evaluating Image-to-Code Generation
di: Periasami, Ajay Vikram, et al.
Pubblicazione: (2026)
di: Periasami, Ajay Vikram, et al.
Pubblicazione: (2026)
Hierarchical Multi-Label Classification of Online Vaccine Concerns
di: Zhu, Chloe Qinyu, et al.
Pubblicazione: (2024)
di: Zhu, Chloe Qinyu, et al.
Pubblicazione: (2024)
A Platform for Investigating Public Health Content with Efficient Concern Classification
di: Li, Christopher, et al.
Pubblicazione: (2025)
di: Li, Christopher, et al.
Pubblicazione: (2025)
Generalizability of Large Language Model-Based Agents: A Comprehensive Survey
di: Zhang, Minxing, et al.
Pubblicazione: (2025)
di: Zhang, Minxing, et al.
Pubblicazione: (2025)
Compressed Context Memory For Online Language Model Interaction
di: Kim, Jang-Hyun, et al.
Pubblicazione: (2023)
di: Kim, Jang-Hyun, et al.
Pubblicazione: (2023)
Extracting Polymer Nanocomposite Samples from Full-Length Documents
di: Khalighinejad, Ghazal, et al.
Pubblicazione: (2024)
di: Khalighinejad, Ghazal, et al.
Pubblicazione: (2024)
InData: Towards Secure Multi-Step, Tool-Based Data Analysis
di: K, Karthikeyan, et al.
Pubblicazione: (2025)
di: K, Karthikeyan, et al.
Pubblicazione: (2025)
How Much Backtracking is Enough? Exploring the Interplay of SFT and RL in Enhancing LLM Reasoning
di: Cai, Hongyi James, et al.
Pubblicazione: (2025)
di: Cai, Hongyi James, et al.
Pubblicazione: (2025)
Glyph: Scaling Context Windows via Visual-Text Compression
di: Cheng, Jiale, et al.
Pubblicazione: (2025)
di: Cheng, Jiale, et al.
Pubblicazione: (2025)
Equichordal Points of Convex Bodies
di: Jang, Leo, et al.
Pubblicazione: (2025)
di: Jang, Leo, et al.
Pubblicazione: (2025)
Atomic Consistency Preference Optimization for Long-Form Question Answering
di: Chen, Jingfeng, et al.
Pubblicazione: (2025)
di: Chen, Jingfeng, et al.
Pubblicazione: (2025)
Cite Pretrain: Retrieval-Free Knowledge Attribution for Large Language Models
di: Huang, Yukun, et al.
Pubblicazione: (2025)
di: Huang, Yukun, et al.
Pubblicazione: (2025)
Calibrating Long-form Generations from Large Language Models
di: Huang, Yukun, et al.
Pubblicazione: (2024)
di: Huang, Yukun, et al.
Pubblicazione: (2024)
IsoBench: Benchmarking Multimodal Foundation Models on Isomorphic Representations
di: Fu, Deqing, et al.
Pubblicazione: (2024)
di: Fu, Deqing, et al.
Pubblicazione: (2024)
Over-Searching in Search-Augmented Large Language Models
di: Xie, Roy, et al.
Pubblicazione: (2026)
di: Xie, Roy, et al.
Pubblicazione: (2026)
MatViX: Multimodal Information Extraction from Visually Rich Articles
di: Khalighinejad, Ghazal, et al.
Pubblicazione: (2024)
di: Khalighinejad, Ghazal, et al.
Pubblicazione: (2024)
FlashVLM: Text-Guided Visual Token Selection for Large Multimodal Models
di: Cai, Kaitong, et al.
Pubblicazione: (2025)
di: Cai, Kaitong, et al.
Pubblicazione: (2025)
Rethinking Model Selection in VLM Through the Lens of Gromov-Wasserstein Distance
di: Li, Muyang, et al.
Pubblicazione: (2026)
di: Li, Muyang, et al.
Pubblicazione: (2026)
KVzip: Query-Agnostic KV Cache Compression with Context Reconstruction
di: Kim, Jang-Hyun, et al.
Pubblicazione: (2025)
di: Kim, Jang-Hyun, et al.
Pubblicazione: (2025)
Context-Aware Meta-Learning
di: Fifty, Christopher, et al.
Pubblicazione: (2023)
di: Fifty, Christopher, et al.
Pubblicazione: (2023)
MOFI: Learning Image Representations from Noisy Entity Annotated Images
di: Wu, Wentao, et al.
Pubblicazione: (2023)
di: Wu, Wentao, et al.
Pubblicazione: (2023)
Interleaved Reasoning for Large Language Models via Reinforcement Learning
di: Xie, Roy, et al.
Pubblicazione: (2025)
di: Xie, Roy, et al.
Pubblicazione: (2025)
CStream: Parallel Data Stream Compression on Multicore Edge Devices
di: Zeng, Xianzhi, et al.
Pubblicazione: (2023)
di: Zeng, Xianzhi, et al.
Pubblicazione: (2023)
Poison Once, Control Anywhere: Clean-Text Visual Backdoors in VLM-based Mobile Agents
di: Wang, Xuan, et al.
Pubblicazione: (2025)
di: Wang, Xuan, et al.
Pubblicazione: (2025)
VLM-driven Skill Selection for Robotic Assembly Tasks
di: Kim, Jeong-Jung, et al.
Pubblicazione: (2025)
di: Kim, Jeong-Jung, et al.
Pubblicazione: (2025)
Documenti analoghi
-
GenEOL: Harnessing the Generative Power of LLMs for Training-Free Sentence Embeddings
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024) -
Coding Agents are Effective Long-Context Processors
di: Cao, Weili, et al.
Pubblicazione: (2026) -
To Trust or Not to Trust? Enhancing Large Language Models' Situated Faithfulness to External Contexts
di: Huang, Yukun, et al.
Pubblicazione: (2024) -
Adversarial Math Word Problem Generation
di: Xie, Roy, et al.
Pubblicazione: (2024) -
Raccoon: Prompt Extraction Benchmark of LLM-Integrated Applications
di: Wang, Junlin, et al.
Pubblicazione: (2024)