Inner Loop Inference for Pretrained Transformers: Unlocking Latent Capabilities Without Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lys, Jonathan, Gripon, Vincent, Pasdeloup, Bastien, Marmoret, Axel, Mauch, Lukas, Cardinaux, Fabien, Hacene, Ghouthi Boukli |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Residual Connections and the Causal Shift: Uncovering a Structural Misalignment in Transformers
von: Lys, Jonathan, et al.
Veröffentlicht: (2026)
von: Lys, Jonathan, et al.
Veröffentlicht: (2026)
D5P4: Partition Determinantal Point Process for Diversity in Parallel Discrete Diffusion Decoding
von: Lys, Jonathan, et al.
Veröffentlicht: (2026)
von: Lys, Jonathan, et al.
Veröffentlicht: (2026)
LLM meets Vision-Language Models for Zero-Shot One-Class Classification
von: Bendou, Yassir, et al.
Veröffentlicht: (2024)
von: Bendou, Yassir, et al.
Veröffentlicht: (2024)
A Novel Benchmark for Few-Shot Semantic Segmentation in the Era of Foundation Models
von: Bensaid, Reda, et al.
Veröffentlicht: (2024)
von: Bensaid, Reda, et al.
Veröffentlicht: (2024)
GaLLoP: Gradient-based Sparse Learning on Low-Magnitude Parameters
von: Choudhary, Anand, et al.
Veröffentlicht: (2025)
von: Choudhary, Anand, et al.
Veröffentlicht: (2025)
TensLoRA: Tensor Alternatives for Low-Rank Adaptation
von: Marmoret, Axel, et al.
Veröffentlicht: (2025)
von: Marmoret, Axel, et al.
Veröffentlicht: (2025)
REVE: A Foundation Model for EEG -- Adapting to Any Setup with Large-Scale Pretraining on 25,000 Subjects
von: Ouahidi, Yassine El, et al.
Veröffentlicht: (2025)
von: Ouahidi, Yassine El, et al.
Veröffentlicht: (2025)
SKILL: Similarity-aware Knowledge distILLation for Speech Self-Supervised Learning
von: Zampierin, Luca, et al.
Veröffentlicht: (2024)
von: Zampierin, Luca, et al.
Veröffentlicht: (2024)
Unsupervised Adaptive Deep Learning Method For BCI Motor Imagery Decoding
von: Ouahidi, Yassine El, et al.
Veröffentlicht: (2024)
von: Ouahidi, Yassine El, et al.
Veröffentlicht: (2024)
A Strong and Simple Deep Learning Baseline for BCI MI Decoding
von: Ouahidi, Yassine El, et al.
Veröffentlicht: (2023)
von: Ouahidi, Yassine El, et al.
Veröffentlicht: (2023)
SAFT: Towards Out-of-Distribution Generalization in Fine-Tuning
von: Nguyen, Bac, et al.
Veröffentlicht: (2024)
von: Nguyen, Bac, et al.
Veröffentlicht: (2024)
Unsupervised Evaluation of Deep Audio Embeddings for Music Structure Analysis
von: Marmoret, Axel
Veröffentlicht: (2026)
von: Marmoret, Axel
Veröffentlicht: (2026)
Few and Fewer: Learning Better from Few Examples Using Fewer Base Classes
von: Lafargue, Raphael, et al.
Veröffentlicht: (2024)
von: Lafargue, Raphael, et al.
Veröffentlicht: (2024)
Towards Robust FastSpeech 2 by Modelling Residual Multimodality
von: Kögel, Fabian, et al.
Veröffentlicht: (2023)
von: Kögel, Fabian, et al.
Veröffentlicht: (2023)
AutoMashup: Automatic Music Mashups Creation
von: Delabaere, Marine, et al.
Veröffentlicht: (2025)
von: Delabaere, Marine, et al.
Veröffentlicht: (2025)
TPTT: Transforming Pretrained Transformers into Titans
von: Furfaro, Fabien
Veröffentlicht: (2025)
von: Furfaro, Fabien
Veröffentlicht: (2025)
Can pre-trained Deep Learning models predict groove ratings?
von: Marmoret, Axel, et al.
Veröffentlicht: (2026)
von: Marmoret, Axel, et al.
Veröffentlicht: (2026)
Multimodal Object Detection via Probabilistic a priori Information Integration
von: Hafyani, Hafsa El, et al.
Veröffentlicht: (2024)
von: Hafyani, Hafsa El, et al.
Veröffentlicht: (2024)
Order-Preserving GFlowNets
von: Chen, Yihang, et al.
Veröffentlicht: (2023)
von: Chen, Yihang, et al.
Veröffentlicht: (2023)
LoopUS: Recasting Pretrained LLMs into Looped Latent Refinement Models
von: Park, Taekhyun, et al.
Veröffentlicht: (2026)
von: Park, Taekhyun, et al.
Veröffentlicht: (2026)
Explicit Consumption Functions with Borrowing Constraints: a Continuous Time Approach
von: Roulleau-Pasdeloup, Jordan
Veröffentlicht: (2025)
von: Roulleau-Pasdeloup, Jordan
Veröffentlicht: (2025)
Training-Free Looped Transformers
von: Chen, Lizhang, et al.
Veröffentlicht: (2026)
von: Chen, Lizhang, et al.
Veröffentlicht: (2026)
Input Resolution Downsizing as a Compression Technique for Vision Deep Learning Systems
von: Morlier, Jeremy, et al.
Veröffentlicht: (2025)
von: Morlier, Jeremy, et al.
Veröffentlicht: (2025)
On Transfer in Classification: How Well do Subsets of Classes Generalize?
von: Baena, Raphael, et al.
Veröffentlicht: (2024)
von: Baena, Raphael, et al.
Veröffentlicht: (2024)
FFCL: Forward-Forward Net with Cortical Loops, Training and Inference on Edge Without Backpropagation
von: Karkehabadi, Ali, et al.
Veröffentlicht: (2024)
von: Karkehabadi, Ali, et al.
Veröffentlicht: (2024)
The Pathology of Plenty
von: Kulamadayil, Lys
Veröffentlicht: (2025)
von: Kulamadayil, Lys
Veröffentlicht: (2025)
Make Some Noise: Unlocking Language Model Parallel Inference Capability through Noisy Training
von: Wang, Yixuan, et al.
Veröffentlicht: (2024)
von: Wang, Yixuan, et al.
Veröffentlicht: (2024)
Event Classification of Accelerometer Data for Industrial Package Monitoring with Embedded Deep Learning
von: Renault, Manon, et al.
Veröffentlicht: (2025)
von: Renault, Manon, et al.
Veröffentlicht: (2025)
Alimentando a tradição e valorizando o conhecimento tradicional na Amazônia: o caso da castanha-da-amazônia na Terra Indígena Mãe Maria
von: Nina Lys Nunes
Veröffentlicht: (2023)
von: Nina Lys Nunes
Veröffentlicht: (2023)
Apuntes para la territorialización universitaria. Análisis de una experiencia en el sur de la CABA
von: Ivanna Lys Petz
Veröffentlicht: (2024)
von: Ivanna Lys Petz
Veröffentlicht: (2024)
Metodologia participativa para o ensino de ciências em locais megabiodiversidade
von: Nina Lys Nunes
Veröffentlicht: (2022)
von: Nina Lys Nunes
Veröffentlicht: (2022)
AH Plus extrusion into periapical tissue: literature review of main related properties and report of clinical cases
von: Karen Lys Hubbe
Veröffentlicht: (2016)
von: Karen Lys Hubbe
Veröffentlicht: (2016)
Presentación
von: Ivanna Lys Petz
Veröffentlicht: (2019)
von: Ivanna Lys Petz
Veröffentlicht: (2019)
Uso de computador e ergonomia: um estudo sobre as escolas de ensino fundamental e médio de São Paulo
von: Lys Esther Rocha
Veröffentlicht: (2003)
von: Lys Esther Rocha
Veröffentlicht: (2003)
»Vrummmummmmm FVISH!«
von: Mauch, Addrich
Veröffentlicht: (2024)
von: Mauch, Addrich
Veröffentlicht: (2024)
Paradise Blues
von: Mauch, Christof
Veröffentlicht: (2024)
von: Mauch, Christof
Veröffentlicht: (2024)
Contando policiais: os registros de pessoal como fonte
von: Cláudia Mauch
Veröffentlicht: (2012)
von: Cláudia Mauch
Veröffentlicht: (2012)
Militares, milicianos e policiais: instituições, representações e práticas
von: Cláudia Mauch
Veröffentlicht: (2012)
von: Cláudia Mauch
Veröffentlicht: (2012)
Unlocking the Working Memory of Large Language Models for Latent Reasoning
von: Aichberger, Lukas, et al.
Veröffentlicht: (2026)
von: Aichberger, Lukas, et al.
Veröffentlicht: (2026)
Object-Centric Cropping for Visual Few-Shot Classification
von: Abdali, Aymane, et al.
Veröffentlicht: (2025)
von: Abdali, Aymane, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Residual Connections and the Causal Shift: Uncovering a Structural Misalignment in Transformers
von: Lys, Jonathan, et al.
Veröffentlicht: (2026) -
D5P4: Partition Determinantal Point Process for Diversity in Parallel Discrete Diffusion Decoding
von: Lys, Jonathan, et al.
Veröffentlicht: (2026) -
LLM meets Vision-Language Models for Zero-Shot One-Class Classification
von: Bendou, Yassir, et al.
Veröffentlicht: (2024) -
A Novel Benchmark for Few-Shot Semantic Segmentation in the Era of Foundation Models
von: Bensaid, Reda, et al.
Veröffentlicht: (2024) -
GaLLoP: Gradient-based Sparse Learning on Low-Magnitude Parameters
von: Choudhary, Anand, et al.
Veröffentlicht: (2025)