Retrievit: In-context Retrieval Capabilities of Transformers, State Space Models, and Hybrid Architectures
Fuente:
arXiv
Salvato in:
| Autori principali: | Pantazopoulos, Georgios, Nikandrou, Malvina, Konstas, Ioannis, Suglia, Alessandro |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Enhancing Continual Learning in Visual Question Answering with Modality-Aware Feature Distillation
di: Nikandrou, Malvina, et al.
Pubblicazione: (2024)
di: Nikandrou, Malvina, et al.
Pubblicazione: (2024)
Shaking Up VLMs: Comparing Transformers and Structured State Space Models for Vision & Language Modeling
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2024)
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2024)
CROPE: Evaluating In-Context Adaptation of Vision and Language Models to Culture-Specific Concepts
di: Nikandrou, Malvina, et al.
Pubblicazione: (2024)
di: Nikandrou, Malvina, et al.
Pubblicazione: (2024)
Learning To See But Forgetting To Follow: Visual Instruction Tuning Makes LLMs More Prone To Jailbreak Attacks
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2024)
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2024)
Lost in Space: Probing Fine-grained Spatial Understanding in Vision and Language Resamplers
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2024)
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2024)
Investigating the Role of Instruction Variety and Task Difficulty in Robotic Manipulation Tasks
di: Parekh, Amit, et al.
Pubblicazione: (2024)
di: Parekh, Amit, et al.
Pubblicazione: (2024)
Task Formulation Matters When Learning Continually: A Case Study in Visual Question Answering
di: Nikandrou, Mavina, et al.
Pubblicazione: (2022)
di: Nikandrou, Mavina, et al.
Pubblicazione: (2022)
Evaluating Multimodal Language Models as Visual Assistants for Visually Impaired Users
di: Karamolegkou, Antonia, et al.
Pubblicazione: (2025)
di: Karamolegkou, Antonia, et al.
Pubblicazione: (2025)
AlanaVLM: A Multimodal Embodied AI Foundation Model for Egocentric Video Understanding
di: Suglia, Alessandro, et al.
Pubblicazione: (2024)
di: Suglia, Alessandro, et al.
Pubblicazione: (2024)
When Chain-of-Thought Fails, the Solution Hides in the Hidden States
di: Mehrafarin, Houman, et al.
Pubblicazione: (2026)
di: Mehrafarin, Houman, et al.
Pubblicazione: (2026)
An Efficient Training Pipeline for Reasoning Graphical User Interface Agents
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2025)
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2025)
Towards Understanding Visual Grounding in Visual Language Models
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2025)
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2025)
AgriPath: A Systematic Exploration of Architectural Trade-offs for Crop Disease Classification
di: Mooraj, Hamza, et al.
Pubblicazione: (2026)
di: Mooraj, Hamza, et al.
Pubblicazione: (2026)
VoyagerVision: Investigating the Role of Multi-modal Information for Open-ended Learning Systems
di: Smyth, Ethan, et al.
Pubblicazione: (2025)
di: Smyth, Ethan, et al.
Pubblicazione: (2025)
VLM-RobustBench: A Comprehensive Benchmark for Robustness of Vision-Language Models
di: Saxena, Rohit, et al.
Pubblicazione: (2026)
di: Saxena, Rohit, et al.
Pubblicazione: (2026)
Understanding In-Context Learning Beyond Transformers: An Investigation of State Space and Hybrid Architectures
di: Wang, Shenran, et al.
Pubblicazione: (2025)
di: Wang, Shenran, et al.
Pubblicazione: (2025)
FOSSIL: Harnessing Feedback on Suboptimal Samples for Data-Efficient Generalisation with Imitation Learning for Embodied Vision-and-Language Tasks
di: McCallum, Sabrina, et al.
Pubblicazione: (2025)
di: McCallum, Sabrina, et al.
Pubblicazione: (2025)
Priming: Hybrid State Space Models From Pre-trained Transformers
di: Chattopadhyay, Aditya, et al.
Pubblicazione: (2026)
di: Chattopadhyay, Aditya, et al.
Pubblicazione: (2026)
Frontier Models are Capable of In-context Scheming
di: Meinke, Alexander, et al.
Pubblicazione: (2024)
di: Meinke, Alexander, et al.
Pubblicazione: (2024)
Beyond Semantics: How Temporal Biases Shape Retrieval in Transformer and State-Space Models
di: Bajaj, Anooshka, et al.
Pubblicazione: (2025)
di: Bajaj, Anooshka, et al.
Pubblicazione: (2025)
HyMem: Hybrid Memory Architecture with Dynamic Retrieval Scheduling
di: Zhao, Xiaochen, et al.
Pubblicazione: (2026)
di: Zhao, Xiaochen, et al.
Pubblicazione: (2026)
RoboSSM: Scalable In-context Imitation Learning via State-Space Models
di: Yoo, Youngju, et al.
Pubblicazione: (2025)
di: Yoo, Youngju, et al.
Pubblicazione: (2025)
Retrieval-Augmented Decision Transformer: External Memory for In-context RL
di: Schmied, Thomas, et al.
Pubblicazione: (2024)
di: Schmied, Thomas, et al.
Pubblicazione: (2024)
TransXSSM: A Hybrid Transformer State Space Model with Unified Rotary Position Embedding
di: Wu, Bingheng, et al.
Pubblicazione: (2025)
di: Wu, Bingheng, et al.
Pubblicazione: (2025)
TACLer: Tailored Curriculum Reinforcement Learning for Efficient Reasoning
di: Lai, Huiyuan, et al.
Pubblicazione: (2026)
di: Lai, Huiyuan, et al.
Pubblicazione: (2026)
SegMaFormer: A Hybrid State-Space and Transformer Model for Efficient Segmentation
di: Nguyen, Duy D., et al.
Pubblicazione: (2026)
di: Nguyen, Duy D., et al.
Pubblicazione: (2026)
Algorithmic Administration and the EU AI Act: Legal Principles for Public Sector Use of AI
di: Pavlidis, Georgios, et al.
Pubblicazione: (2026)
di: Pavlidis, Georgios, et al.
Pubblicazione: (2026)
Motor Imagery Decoding Using Ensemble Curriculum Learning and Collaborative Training
di: Zoumpourlis, Georgios, et al.
Pubblicazione: (2022)
di: Zoumpourlis, Georgios, et al.
Pubblicazione: (2022)
Retrieval-Aware Distillation for Transformer-SSM Hybrids
di: Bick, Aviv, et al.
Pubblicazione: (2026)
di: Bick, Aviv, et al.
Pubblicazione: (2026)
A Reference Architecture for Agentic Hybrid Retrieval in Dataset Search
di: Terrenzi, Riccardo, et al.
Pubblicazione: (2026)
di: Terrenzi, Riccardo, et al.
Pubblicazione: (2026)
White-Basilisk: A Hybrid Model for Code Vulnerability Detection
di: Lamprou, Ioannis, et al.
Pubblicazione: (2025)
di: Lamprou, Ioannis, et al.
Pubblicazione: (2025)
State Space Models for Bioacoustics: A Comparative Evaluation with Transformers
di: Tang, Chengyu, et al.
Pubblicazione: (2025)
di: Tang, Chengyu, et al.
Pubblicazione: (2025)
STree: Speculative Tree Decoding for Hybrid State-Space Models
di: Wu, Yangchao, et al.
Pubblicazione: (2025)
di: Wu, Yangchao, et al.
Pubblicazione: (2025)
Neural Architecture Retrieval
di: Pei, Xiaohuan, et al.
Pubblicazione: (2023)
di: Pei, Xiaohuan, et al.
Pubblicazione: (2023)
Multi-property Steering of Large Language Models with Dynamic Activation Composition
di: Scalena, Daniel, et al.
Pubblicazione: (2024)
di: Scalena, Daniel, et al.
Pubblicazione: (2024)
Pickup & Delivery with Time Windows and Transfers: combining decomposition with metaheuristics
di: Avgerinos, Ioannis, et al.
Pubblicazione: (2025)
di: Avgerinos, Ioannis, et al.
Pubblicazione: (2025)
Retrieval-Infused Reasoning Sandbox: A Benchmark for Decoupling Retrieval and Reasoning Capabilities
di: Ying, Shuangshuang, et al.
Pubblicazione: (2026)
di: Ying, Shuangshuang, et al.
Pubblicazione: (2026)
Non Verbis, Sed Rebus: Large Language Models are Weak Solvers of Italian Rebuses
di: Sarti, Gabriele, et al.
Pubblicazione: (2024)
di: Sarti, Gabriele, et al.
Pubblicazione: (2024)
Practising responsibility: Ethics in NLP as a hands-on course
di: Nissim, Malvina, et al.
Pubblicazione: (2025)
di: Nissim, Malvina, et al.
Pubblicazione: (2025)
Monte-Carlo Tree Search with Neural Network Guidance for Lane-Free Autonomous Driving
di: Peridis, Ioannis, et al.
Pubblicazione: (2026)
di: Peridis, Ioannis, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Enhancing Continual Learning in Visual Question Answering with Modality-Aware Feature Distillation
di: Nikandrou, Malvina, et al.
Pubblicazione: (2024) -
Shaking Up VLMs: Comparing Transformers and Structured State Space Models for Vision & Language Modeling
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2024) -
CROPE: Evaluating In-Context Adaptation of Vision and Language Models to Culture-Specific Concepts
di: Nikandrou, Malvina, et al.
Pubblicazione: (2024) -
Learning To See But Forgetting To Follow: Visual Instruction Tuning Makes LLMs More Prone To Jailbreak Attacks
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2024) -
Lost in Space: Probing Fine-grained Spatial Understanding in Vision and Language Resamplers
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2024)