Retrievit: In-context Retrieval Capabilities of Transformers, State Space Models, and Hybrid Architectures
Fuente:
arXiv
Saved in:
| Main Authors: | Pantazopoulos, Georgios, Nikandrou, Malvina, Konstas, Ioannis, Suglia, Alessandro |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhancing Continual Learning in Visual Question Answering with Modality-Aware Feature Distillation
by: Nikandrou, Malvina, et al.
Published: (2024)
by: Nikandrou, Malvina, et al.
Published: (2024)
Shaking Up VLMs: Comparing Transformers and Structured State Space Models for Vision & Language Modeling
by: Pantazopoulos, Georgios, et al.
Published: (2024)
by: Pantazopoulos, Georgios, et al.
Published: (2024)
CROPE: Evaluating In-Context Adaptation of Vision and Language Models to Culture-Specific Concepts
by: Nikandrou, Malvina, et al.
Published: (2024)
by: Nikandrou, Malvina, et al.
Published: (2024)
Learning To See But Forgetting To Follow: Visual Instruction Tuning Makes LLMs More Prone To Jailbreak Attacks
by: Pantazopoulos, Georgios, et al.
Published: (2024)
by: Pantazopoulos, Georgios, et al.
Published: (2024)
Lost in Space: Probing Fine-grained Spatial Understanding in Vision and Language Resamplers
by: Pantazopoulos, Georgios, et al.
Published: (2024)
by: Pantazopoulos, Georgios, et al.
Published: (2024)
Investigating the Role of Instruction Variety and Task Difficulty in Robotic Manipulation Tasks
by: Parekh, Amit, et al.
Published: (2024)
by: Parekh, Amit, et al.
Published: (2024)
Task Formulation Matters When Learning Continually: A Case Study in Visual Question Answering
by: Nikandrou, Mavina, et al.
Published: (2022)
by: Nikandrou, Mavina, et al.
Published: (2022)
Evaluating Multimodal Language Models as Visual Assistants for Visually Impaired Users
by: Karamolegkou, Antonia, et al.
Published: (2025)
by: Karamolegkou, Antonia, et al.
Published: (2025)
AlanaVLM: A Multimodal Embodied AI Foundation Model for Egocentric Video Understanding
by: Suglia, Alessandro, et al.
Published: (2024)
by: Suglia, Alessandro, et al.
Published: (2024)
When Chain-of-Thought Fails, the Solution Hides in the Hidden States
by: Mehrafarin, Houman, et al.
Published: (2026)
by: Mehrafarin, Houman, et al.
Published: (2026)
An Efficient Training Pipeline for Reasoning Graphical User Interface Agents
by: Pantazopoulos, Georgios, et al.
Published: (2025)
by: Pantazopoulos, Georgios, et al.
Published: (2025)
Towards Understanding Visual Grounding in Visual Language Models
by: Pantazopoulos, Georgios, et al.
Published: (2025)
by: Pantazopoulos, Georgios, et al.
Published: (2025)
AgriPath: A Systematic Exploration of Architectural Trade-offs for Crop Disease Classification
by: Mooraj, Hamza, et al.
Published: (2026)
by: Mooraj, Hamza, et al.
Published: (2026)
VoyagerVision: Investigating the Role of Multi-modal Information for Open-ended Learning Systems
by: Smyth, Ethan, et al.
Published: (2025)
by: Smyth, Ethan, et al.
Published: (2025)
VLM-RobustBench: A Comprehensive Benchmark for Robustness of Vision-Language Models
by: Saxena, Rohit, et al.
Published: (2026)
by: Saxena, Rohit, et al.
Published: (2026)
Understanding In-Context Learning Beyond Transformers: An Investigation of State Space and Hybrid Architectures
by: Wang, Shenran, et al.
Published: (2025)
by: Wang, Shenran, et al.
Published: (2025)
FOSSIL: Harnessing Feedback on Suboptimal Samples for Data-Efficient Generalisation with Imitation Learning for Embodied Vision-and-Language Tasks
by: McCallum, Sabrina, et al.
Published: (2025)
by: McCallum, Sabrina, et al.
Published: (2025)
Priming: Hybrid State Space Models From Pre-trained Transformers
by: Chattopadhyay, Aditya, et al.
Published: (2026)
by: Chattopadhyay, Aditya, et al.
Published: (2026)
Frontier Models are Capable of In-context Scheming
by: Meinke, Alexander, et al.
Published: (2024)
by: Meinke, Alexander, et al.
Published: (2024)
Beyond Semantics: How Temporal Biases Shape Retrieval in Transformer and State-Space Models
by: Bajaj, Anooshka, et al.
Published: (2025)
by: Bajaj, Anooshka, et al.
Published: (2025)
HyMem: Hybrid Memory Architecture with Dynamic Retrieval Scheduling
by: Zhao, Xiaochen, et al.
Published: (2026)
by: Zhao, Xiaochen, et al.
Published: (2026)
RoboSSM: Scalable In-context Imitation Learning via State-Space Models
by: Yoo, Youngju, et al.
Published: (2025)
by: Yoo, Youngju, et al.
Published: (2025)
Retrieval-Augmented Decision Transformer: External Memory for In-context RL
by: Schmied, Thomas, et al.
Published: (2024)
by: Schmied, Thomas, et al.
Published: (2024)
TransXSSM: A Hybrid Transformer State Space Model with Unified Rotary Position Embedding
by: Wu, Bingheng, et al.
Published: (2025)
by: Wu, Bingheng, et al.
Published: (2025)
TACLer: Tailored Curriculum Reinforcement Learning for Efficient Reasoning
by: Lai, Huiyuan, et al.
Published: (2026)
by: Lai, Huiyuan, et al.
Published: (2026)
SegMaFormer: A Hybrid State-Space and Transformer Model for Efficient Segmentation
by: Nguyen, Duy D., et al.
Published: (2026)
by: Nguyen, Duy D., et al.
Published: (2026)
Algorithmic Administration and the EU AI Act: Legal Principles for Public Sector Use of AI
by: Pavlidis, Georgios, et al.
Published: (2026)
by: Pavlidis, Georgios, et al.
Published: (2026)
Motor Imagery Decoding Using Ensemble Curriculum Learning and Collaborative Training
by: Zoumpourlis, Georgios, et al.
Published: (2022)
by: Zoumpourlis, Georgios, et al.
Published: (2022)
Retrieval-Aware Distillation for Transformer-SSM Hybrids
by: Bick, Aviv, et al.
Published: (2026)
by: Bick, Aviv, et al.
Published: (2026)
A Reference Architecture for Agentic Hybrid Retrieval in Dataset Search
by: Terrenzi, Riccardo, et al.
Published: (2026)
by: Terrenzi, Riccardo, et al.
Published: (2026)
White-Basilisk: A Hybrid Model for Code Vulnerability Detection
by: Lamprou, Ioannis, et al.
Published: (2025)
by: Lamprou, Ioannis, et al.
Published: (2025)
State Space Models for Bioacoustics: A Comparative Evaluation with Transformers
by: Tang, Chengyu, et al.
Published: (2025)
by: Tang, Chengyu, et al.
Published: (2025)
STree: Speculative Tree Decoding for Hybrid State-Space Models
by: Wu, Yangchao, et al.
Published: (2025)
by: Wu, Yangchao, et al.
Published: (2025)
Neural Architecture Retrieval
by: Pei, Xiaohuan, et al.
Published: (2023)
by: Pei, Xiaohuan, et al.
Published: (2023)
Multi-property Steering of Large Language Models with Dynamic Activation Composition
by: Scalena, Daniel, et al.
Published: (2024)
by: Scalena, Daniel, et al.
Published: (2024)
Pickup & Delivery with Time Windows and Transfers: combining decomposition with metaheuristics
by: Avgerinos, Ioannis, et al.
Published: (2025)
by: Avgerinos, Ioannis, et al.
Published: (2025)
Retrieval-Infused Reasoning Sandbox: A Benchmark for Decoupling Retrieval and Reasoning Capabilities
by: Ying, Shuangshuang, et al.
Published: (2026)
by: Ying, Shuangshuang, et al.
Published: (2026)
Non Verbis, Sed Rebus: Large Language Models are Weak Solvers of Italian Rebuses
by: Sarti, Gabriele, et al.
Published: (2024)
by: Sarti, Gabriele, et al.
Published: (2024)
Practising responsibility: Ethics in NLP as a hands-on course
by: Nissim, Malvina, et al.
Published: (2025)
by: Nissim, Malvina, et al.
Published: (2025)
Monte-Carlo Tree Search with Neural Network Guidance for Lane-Free Autonomous Driving
by: Peridis, Ioannis, et al.
Published: (2026)
by: Peridis, Ioannis, et al.
Published: (2026)
Similar Items
-
Enhancing Continual Learning in Visual Question Answering with Modality-Aware Feature Distillation
by: Nikandrou, Malvina, et al.
Published: (2024) -
Shaking Up VLMs: Comparing Transformers and Structured State Space Models for Vision & Language Modeling
by: Pantazopoulos, Georgios, et al.
Published: (2024) -
CROPE: Evaluating In-Context Adaptation of Vision and Language Models to Culture-Specific Concepts
by: Nikandrou, Malvina, et al.
Published: (2024) -
Learning To See But Forgetting To Follow: Visual Instruction Tuning Makes LLMs More Prone To Jailbreak Attacks
by: Pantazopoulos, Georgios, et al.
Published: (2024) -
Lost in Space: Probing Fine-grained Spatial Understanding in Vision and Language Resamplers
by: Pantazopoulos, Georgios, et al.
Published: (2024)