Universal In-Context Approximation By Prompting Fully Recurrent Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Petrov, Aleksandar, Lamb, Tom A., Paren, Alasdair, Torr, Philip H. S., Bibi, Adel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Prompting a Pretrained Transformer Can Be a Universal Approximator
por: Petrov, Aleksandar, et al.
Publicado: (2024)
por: Petrov, Aleksandar, et al.
Publicado: (2024)
Focus On This, Not That! Steering LLMs with Adaptive Feature Specification
por: Lamb, Tom A., et al.
Publicado: (2024)
por: Lamb, Tom A., et al.
Publicado: (2024)
When Do Prompting and Prefix-Tuning Work? A Theory of Capabilities and Limitations
por: Petrov, Aleksandar, et al.
Publicado: (2023)
por: Petrov, Aleksandar, et al.
Publicado: (2023)
Shh, don't say that! Domain Certification in LLMs
por: Emde, Cornelius, et al.
Publicado: (2025)
por: Emde, Cornelius, et al.
Publicado: (2025)
Bi-Factorial Preference Optimization: Balancing Safety-Helpfulness in Language Models
por: Zhang, Wenxuan, et al.
Publicado: (2024)
por: Zhang, Wenxuan, et al.
Publicado: (2024)
Do as I do (Safely): Mitigating Task-Specific Fine-tuning Risks in Large Language Models
por: Eiras, Francisco, et al.
Publicado: (2024)
por: Eiras, Francisco, et al.
Publicado: (2024)
Model Merging and Safety Alignment: One Bad Model Spoils the Bunch
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2024)
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2024)
SynthCLIP: Are We Ready for a Fully Synthetic CLIP Training?
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2024)
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2024)
Towards Interpretable Sequence Continuation: Analyzing Shared Circuits in Large Language Models
por: Lan, Michael, et al.
Publicado: (2023)
por: Lan, Michael, et al.
Publicado: (2023)
On the Coexistence and Ensembling of Watermarks
por: Petrov, Aleksandar, et al.
Publicado: (2025)
por: Petrov, Aleksandar, et al.
Publicado: (2025)
Quantifying Feature Space Universality Across Large Language Models via Sparse Autoencoders
por: Lan, Michael, et al.
Publicado: (2024)
por: Lan, Michael, et al.
Publicado: (2024)
MIP against Agent: Malicious Image Patches Hijacking Multimodal OS Agents
por: Aichberger, Lukas, et al.
Publicado: (2025)
por: Aichberger, Lukas, et al.
Publicado: (2025)
OMNI-LEAK: Orchestrator Multi-Agent Network Induced Data Leakage
por: Naik, Akshat, et al.
Publicado: (2026)
por: Naik, Akshat, et al.
Publicado: (2026)
PromptIntern: Saving Inference Costs by Internalizing Recurrent Prompt during Large Language Model Fine-tuning
por: Zou, Jiaru, et al.
Publicado: (2024)
por: Zou, Jiaru, et al.
Publicado: (2024)
Detecting LLM Hallucination Through Layer-wise Information Deficiency: Analysis of Ambiguous Prompts and Unanswerable Questions
por: Kim, Hazel, et al.
Publicado: (2024)
por: Kim, Hazel, et al.
Publicado: (2024)
FORCE: Transferable Visual Jailbreaking Attacks via Feature Over-Reliance CorrEction
por: Lin, Runqi, et al.
Publicado: (2025)
por: Lin, Runqi, et al.
Publicado: (2025)
Rethinking Visual Prompting for Multimodal Large Language Models with External Knowledge
por: Lin, Yuanze, et al.
Publicado: (2024)
por: Lin, Yuanze, et al.
Publicado: (2024)
Revisiting Uncertainty Estimation and Calibration of Large Language Models
por: Tao, Linwei, et al.
Publicado: (2025)
por: Tao, Linwei, et al.
Publicado: (2025)
When Prompts Override Vision: Prompt-Induced Hallucinations in LVLMs
por: Khayatan, Pegah, et al.
Publicado: (2026)
por: Khayatan, Pegah, et al.
Publicado: (2026)
Teaching Pretrained Language Models to Think Deeper with Retrofitted Recurrence
por: McLeish, Sean, et al.
Publicado: (2025)
por: McLeish, Sean, et al.
Publicado: (2025)
GateLoop: Fully Data-Controlled Linear Recurrence for Sequence Modeling
por: Katsch, Tobias
Publicado: (2023)
por: Katsch, Tobias
Publicado: (2023)
Recurrent Knowledge Identification and Fusion for Language Model Continual Learning
por: Feng, Yujie, et al.
Publicado: (2025)
por: Feng, Yujie, et al.
Publicado: (2025)
Subgraph-level Universal Prompt Tuning
por: Lee, Junhyun, et al.
Publicado: (2024)
por: Lee, Junhyun, et al.
Publicado: (2024)
Unforgotten Safety: Preserving Safety Alignment of Large Language Models with Continual Learning
por: Alssum, Lama, et al.
Publicado: (2025)
por: Alssum, Lama, et al.
Publicado: (2025)
MemoryPrompt: A Light Wrapper to Improve Context Tracking in Pre-trained Language Models
por: Rakotonirina, Nathanaël Carraz, et al.
Publicado: (2024)
por: Rakotonirina, Nathanaël Carraz, et al.
Publicado: (2024)
Batch Calibration: Rethinking Calibration for In-Context Learning and Prompt Engineering
por: Zhou, Han, et al.
Publicado: (2023)
por: Zhou, Han, et al.
Publicado: (2023)
In-Context Learning Learns Label Relationships but Is Not Conventional Learning
por: Kossen, Jannik, et al.
Publicado: (2023)
por: Kossen, Jannik, et al.
Publicado: (2023)
BiasBusters: Uncovering and Mitigating Tool Selection Bias in Large Language Models
por: Blankenstein, Thierry, et al.
Publicado: (2025)
por: Blankenstein, Thierry, et al.
Publicado: (2025)
Dynamic Universal Approximation Theory: The Basic Theory for Transformer-based Large Language Models
por: Wang, Wei, et al.
Publicado: (2024)
por: Wang, Wei, et al.
Publicado: (2024)
Adapting LLMs for Efficient Context Processing through Soft Prompt Compression
por: Wang, Cangqing, et al.
Publicado: (2024)
por: Wang, Cangqing, et al.
Publicado: (2024)
Dynamic Context-oriented Decomposition for Task-aware Low-rank Adaptation with Less Forgetting and Faster Convergence
por: Yang, Yibo, et al.
Publicado: (2025)
por: Yang, Yibo, et al.
Publicado: (2025)
Instella: Fully Open Language Models with Stellar Performance
por: Liu, Jiang, et al.
Publicado: (2025)
por: Liu, Jiang, et al.
Publicado: (2025)
Efficient Joint Prediction of Multiple Future Tokens
por: Ahn, Kwangjun, et al.
Publicado: (2025)
por: Ahn, Kwangjun, et al.
Publicado: (2025)
RecurrentGemma: Moving Past Transformers for Efficient Open Language Models
por: Botev, Aleksandar, et al.
Publicado: (2024)
por: Botev, Aleksandar, et al.
Publicado: (2024)
Selective Prompting Tuning for Personalized Conversations with LLMs
por: Huang, Qiushi, et al.
Publicado: (2024)
por: Huang, Qiushi, et al.
Publicado: (2024)
Fundamental Limits of Prompt Tuning Transformers: Universality, Capacity and Efficiency
por: Hu, Jerry Yao-Chieh, et al.
Publicado: (2024)
por: Hu, Jerry Yao-Chieh, et al.
Publicado: (2024)
Prompting with Phonemes: Enhancing LLMs' Multilinguality for Non-Latin Script Languages
por: Nguyen, Hoang H, et al.
Publicado: (2024)
por: Nguyen, Hoang H, et al.
Publicado: (2024)
On Pretraining Data Diversity for Self-Supervised Learning
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2024)
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2024)
BrokenMath: A Benchmark for Sycophancy in Theorem Proving with LLMs
por: Petrov, Ivo, et al.
Publicado: (2025)
por: Petrov, Ivo, et al.
Publicado: (2025)
POSIX: A Prompt Sensitivity Index For Large Language Models
por: Chatterjee, Anwoy, et al.
Publicado: (2024)
por: Chatterjee, Anwoy, et al.
Publicado: (2024)
Ejemplares similares
-
Prompting a Pretrained Transformer Can Be a Universal Approximator
por: Petrov, Aleksandar, et al.
Publicado: (2024) -
Focus On This, Not That! Steering LLMs with Adaptive Feature Specification
por: Lamb, Tom A., et al.
Publicado: (2024) -
When Do Prompting and Prefix-Tuning Work? A Theory of Capabilities and Limitations
por: Petrov, Aleksandar, et al.
Publicado: (2023) -
Shh, don't say that! Domain Certification in LLMs
por: Emde, Cornelius, et al.
Publicado: (2025) -
Bi-Factorial Preference Optimization: Balancing Safety-Helpfulness in Language Models
por: Zhang, Wenxuan, et al.
Publicado: (2024)