Dynamic Cheatsheet: Test-Time Learning with Adaptive Memory
Fuente:
arXiv
Salvato in:
| Autori principali: | Suzgun, Mirac, Yuksekgonul, Mert, Bianchi, Federico, Jurafsky, Dan, Zou, James |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Cost-of-Pass: An Economic Framework for Evaluating Language Models
di: Erol, Mehmet Hamza, et al.
Pubblicazione: (2025)
di: Erol, Mehmet Hamza, et al.
Pubblicazione: (2025)
Safety-Tuned LLaMAs: Lessons From Improving the Safety of Large Language Models that Follow Instructions
di: Bianchi, Federico, et al.
Pubblicazione: (2023)
di: Bianchi, Federico, et al.
Pubblicazione: (2023)
Belief in the Machine: Investigating Epistemological Blind Spots of Language Models
di: Suzgun, Mirac, et al.
Pubblicazione: (2024)
di: Suzgun, Mirac, et al.
Pubblicazione: (2024)
TextGrad: Automatic "Differentiation" via Text
di: Yuksekgonul, Mert, et al.
Pubblicazione: (2024)
di: Yuksekgonul, Mert, et al.
Pubblicazione: (2024)
How Well Can LLMs Negotiate? NegotiationArena Platform and Analysis
di: Bianchi, Federico, et al.
Pubblicazione: (2024)
di: Bianchi, Federico, et al.
Pubblicazione: (2024)
Evaluating Commercial AI Chatbots as News Intermediaries
di: Suzgun, Mirac, et al.
Pubblicazione: (2026)
di: Suzgun, Mirac, et al.
Pubblicazione: (2026)
A Benchmark for Learning to Translate a New Language from One Grammar Book
di: Tanzer, Garrett, et al.
Pubblicazione: (2023)
di: Tanzer, Garrett, et al.
Pubblicazione: (2023)
The Responsible Foundation Model Development Cheatsheet: A Review of Tools & Resources
di: Longpre, Shayne, et al.
Pubblicazione: (2024)
di: Longpre, Shayne, et al.
Pubblicazione: (2024)
Sparse Reward Subsystem in Large Language Models
di: Xu, Guowei, et al.
Pubblicazione: (2026)
di: Xu, Guowei, et al.
Pubblicazione: (2026)
Meta-Prompting: Enhancing Language Models with Task-Agnostic Scaffolding
di: Suzgun, Mirac, et al.
Pubblicazione: (2024)
di: Suzgun, Mirac, et al.
Pubblicazione: (2024)
Learning to Discover at Test Time
di: Yuksekgonul, Mert, et al.
Pubblicazione: (2026)
di: Yuksekgonul, Mert, et al.
Pubblicazione: (2026)
Inefficiencies of Meta Agents for Agent Design
di: El, Batu, et al.
Pubblicazione: (2025)
di: El, Batu, et al.
Pubblicazione: (2025)
Can LLM feedback enhance review quality? A randomized study of 20K reviews at ICLR 2025
di: Thakkar, Nitya, et al.
Pubblicazione: (2025)
di: Thakkar, Nitya, et al.
Pubblicazione: (2025)
metaTextGrad: Automatically optimizing language model optimizers
di: Xu, Guowei, et al.
Pubblicazione: (2025)
di: Xu, Guowei, et al.
Pubblicazione: (2025)
Attention Satisfies: A Constraint-Satisfaction Lens on Factual Errors of Language Models
di: Yuksekgonul, Mert, et al.
Pubblicazione: (2023)
di: Yuksekgonul, Mert, et al.
Pubblicazione: (2023)
Cooking Up Creativity: Enhancing LLM Creativity through Structured Recombination
di: Mizrahi, Moran, et al.
Pubblicazione: (2025)
di: Mizrahi, Moran, et al.
Pubblicazione: (2025)
Adaptive Memory Replay for Continual Learning
di: Smith, James Seale, et al.
Pubblicazione: (2024)
di: Smith, James Seale, et al.
Pubblicazione: (2024)
Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models
di: Dahl, Matthew, et al.
Pubblicazione: (2024)
di: Dahl, Matthew, et al.
Pubblicazione: (2024)
Do Language Models Know When They're Hallucinating References?
di: Agrawal, Ayush, et al.
Pubblicazione: (2023)
di: Agrawal, Ayush, et al.
Pubblicazione: (2023)
TTRL: Test-Time Reinforcement Learning
di: Zuo, Yuxin, et al.
Pubblicazione: (2025)
di: Zuo, Yuxin, et al.
Pubblicazione: (2025)
TABED: Test-Time Adaptive Ensemble Drafting for Robust Speculative Decoding in LVLMs
di: Lee, Minjae, et al.
Pubblicazione: (2026)
di: Lee, Minjae, et al.
Pubblicazione: (2026)
Gated KalmaNet: A Fading Memory Layer Through Test-Time Ridge Regression
di: Peng, Liangzu, et al.
Pubblicazione: (2025)
di: Peng, Liangzu, et al.
Pubblicazione: (2025)
When to Ponder: Adaptive Compute Allocation for Code Generation via Test-Time Training
di: Sim, Gihyeon
Pubblicazione: (2025)
di: Sim, Gihyeon
Pubblicazione: (2025)
ATLAS: Adaptive Test-Time Latent Steering with External Verifiers for Enhancing LLMs Reasoning
di: Nguyen, Tuc, et al.
Pubblicazione: (2026)
di: Nguyen, Tuc, et al.
Pubblicazione: (2026)
TaTToo: Tool-Grounded Thinking PRM for Test-Time Scaling in Tabular Reasoning
di: Zou, Jiaru, et al.
Pubblicazione: (2025)
di: Zou, Jiaru, et al.
Pubblicazione: (2025)
AdaFRUGAL: Adaptive Memory-Efficient Training with Dynamic Control
di: Bui, Quang-Hung, et al.
Pubblicazione: (2025)
di: Bui, Quang-Hung, et al.
Pubblicazione: (2025)
Zero-Overhead Introspection for Adaptive Test-Time Compute
di: Manvi, Rohin, et al.
Pubblicazione: (2025)
di: Manvi, Rohin, et al.
Pubblicazione: (2025)
Bayesian scaling laws for in-context learning
di: Arora, Aryaman, et al.
Pubblicazione: (2024)
di: Arora, Aryaman, et al.
Pubblicazione: (2024)
$\texttt{SPECS}$: Faster Test-Time Scaling through Speculative Drafts
di: Cemri, Mert, et al.
Pubblicazione: (2025)
di: Cemri, Mert, et al.
Pubblicazione: (2025)
Bayesian Preference Learning for Test-Time Steerable Reward Models
di: Hong, Jiwoo, et al.
Pubblicazione: (2026)
di: Hong, Jiwoo, et al.
Pubblicazione: (2026)
ReFT: Representation Finetuning for Language Models
di: Wu, Zhengxuan, et al.
Pubblicazione: (2024)
di: Wu, Zhengxuan, et al.
Pubblicazione: (2024)
Test-Time Speculation
di: Kumar, Avinash, et al.
Pubblicazione: (2026)
di: Kumar, Avinash, et al.
Pubblicazione: (2026)
PERK: Long-Context Reasoning as Parameter-Efficient Test-Time Learning
di: Chen, Zeming, et al.
Pubblicazione: (2025)
di: Chen, Zeming, et al.
Pubblicazione: (2025)
Learning a Continue-Thinking Token for Enhanced Test-Time Scaling
di: Ringel, Liran, et al.
Pubblicazione: (2025)
di: Ringel, Liran, et al.
Pubblicazione: (2025)
Memory Is All You Need: Testing How Model Memory Affects LLM Performance in Annotation Tasks
di: Timoneda, Joan C., et al.
Pubblicazione: (2025)
di: Timoneda, Joan C., et al.
Pubblicazione: (2025)
Phasor Memory Networks: Stable Backpropagation Through Time for Scalable Explicit Memory
di: Goo, Sungwoo, et al.
Pubblicazione: (2026)
di: Goo, Sungwoo, et al.
Pubblicazione: (2026)
DART-ing Through the Drift: Dynamic Tracing of Knowledge Neurons for Adaptive Inference-Time Pruning
di: Tyagi, Abhishek, et al.
Pubblicazione: (2026)
di: Tyagi, Abhishek, et al.
Pubblicazione: (2026)
CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition
di: Bartelds, Martijn, et al.
Pubblicazione: (2025)
di: Bartelds, Martijn, et al.
Pubblicazione: (2025)
e3: Learning to Explore Enables Extrapolation of Test-Time Compute for LLMs
di: Setlur, Amrith, et al.
Pubblicazione: (2025)
di: Setlur, Amrith, et al.
Pubblicazione: (2025)
SPINE: Token-Selective Test-Time Reinforcement Learning with Entropy-Band Regularization
di: Wu, Jianghao, et al.
Pubblicazione: (2025)
di: Wu, Jianghao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Cost-of-Pass: An Economic Framework for Evaluating Language Models
di: Erol, Mehmet Hamza, et al.
Pubblicazione: (2025) -
Safety-Tuned LLaMAs: Lessons From Improving the Safety of Large Language Models that Follow Instructions
di: Bianchi, Federico, et al.
Pubblicazione: (2023) -
Belief in the Machine: Investigating Epistemological Blind Spots of Language Models
di: Suzgun, Mirac, et al.
Pubblicazione: (2024) -
TextGrad: Automatic "Differentiation" via Text
di: Yuksekgonul, Mert, et al.
Pubblicazione: (2024) -
How Well Can LLMs Negotiate? NegotiationArena Platform and Analysis
di: Bianchi, Federico, et al.
Pubblicazione: (2024)