Unsupervised Layer-Wise Dynamic Test Time Adaptation for LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Xu, Longhuan, Chen, Cunjian, Yin, Feng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AsyncDiff: Asynchronous Timestep Conditioning for Enhanced Text-to-Image Diffusion Inference
por: Xu, Longhuan, et al.
Publicado: (2025)
por: Xu, Longhuan, et al.
Publicado: (2025)
How to Alleviate Catastrophic Forgetting in LLMs Finetuning? Hierarchical Layer-Wise and Element-Wise Regularization
por: Song, Shezheng, et al.
Publicado: (2025)
por: Song, Shezheng, et al.
Publicado: (2025)
Test-Time Policy Adaptation for Enhanced Multi-Turn Interactions with LLMs
por: Wei, Chenxing, et al.
Publicado: (2025)
por: Wei, Chenxing, et al.
Publicado: (2025)
Less is More: Geometric Unlearning for LLMs with Minimal Data Disclosure
por: Tan, Chenchen, et al.
Publicado: (2026)
por: Tan, Chenchen, et al.
Publicado: (2026)
Evaluating LLMs Without Oracle Feedback: Agentic Annotation Evaluation Through Unsupervised Consistency Signals
por: Chen, Cheng, et al.
Publicado: (2025)
por: Chen, Cheng, et al.
Publicado: (2025)
Wisdom is Knowing What not to Say: Hallucination-Free LLMs Unlearning via Attention Shifting
por: Tan, Chenchen, et al.
Publicado: (2025)
por: Tan, Chenchen, et al.
Publicado: (2025)
Time-Reversal Provides Unsupervised Feedback to LLMs
por: Varun, Yerram, et al.
Publicado: (2024)
por: Varun, Yerram, et al.
Publicado: (2024)
Amulet: ReAlignment During Test Time for Personalized Preference Adaptation of LLMs
por: Zhang, Zhaowei, et al.
Publicado: (2025)
por: Zhang, Zhaowei, et al.
Publicado: (2025)
LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling
por: Zheng, Tong, et al.
Publicado: (2026)
por: Zheng, Tong, et al.
Publicado: (2026)
WilKE: Wise-Layer Knowledge Editor for Lifelong Knowledge Editing
por: Hu, Chenhui, et al.
Publicado: (2024)
por: Hu, Chenhui, et al.
Publicado: (2024)
IMO: Greedy Layer-Wise Sparse Representation Learning for Out-of-Distribution Text Classification with Pre-trained Models
por: Feng, Tao, et al.
Publicado: (2024)
por: Feng, Tao, et al.
Publicado: (2024)
StreamAdapter: Efficient Test Time Adaptation from Contextual Streams
por: Muhtar, Dilxat, et al.
Publicado: (2024)
por: Muhtar, Dilxat, et al.
Publicado: (2024)
Adaptive Layer Selection for Layer-Wise Token Pruning in LLM Inference
por: Taniguchi, Rei, et al.
Publicado: (2026)
por: Taniguchi, Rei, et al.
Publicado: (2026)
Examining Test-Time Adaptation for Personalized Child Speech Recognition
por: Shi, Zhonghao, et al.
Publicado: (2024)
por: Shi, Zhonghao, et al.
Publicado: (2024)
When Less Is More? Diagnosing ASR Predictions in Sardinian via Layer-Wise Decoding
por: De Cristofaro, Domenico, et al.
Publicado: (2026)
por: De Cristofaro, Domenico, et al.
Publicado: (2026)
LayAlign: Enhancing Multilingual Reasoning in Large Language Models via Layer-Wise Adaptive Fusion and Alignment Strategy
por: Ruan, Zhiwen, et al.
Publicado: (2025)
por: Ruan, Zhiwen, et al.
Publicado: (2025)
KVSharer: Efficient Inference via Layer-Wise Dissimilar KV Cache Sharing
por: Yang, Yifei, et al.
Publicado: (2024)
por: Yang, Yifei, et al.
Publicado: (2024)
OrScale: Orthogonalised Optimization with Layer-Wise Trust-Ratio Scaling
por: Lou, Yuxuan, et al.
Publicado: (2026)
por: Lou, Yuxuan, et al.
Publicado: (2026)
LLMCache: Layer-Wise Caching Strategies for Accelerated Reuse in Transformer Inference
por: Bansal, Harsh Vardhan
Publicado: (2025)
por: Bansal, Harsh Vardhan
Publicado: (2025)
Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation
por: Shu, Huizhen, et al.
Publicado: (2025)
por: Shu, Huizhen, et al.
Publicado: (2025)
How Large Language Models Encode Context Knowledge? A Layer-Wise Probing Study
por: Ju, Tianjie, et al.
Publicado: (2024)
por: Ju, Tianjie, et al.
Publicado: (2024)
Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability
por: Liang, Xiao, et al.
Publicado: (2026)
por: Liang, Xiao, et al.
Publicado: (2026)
Rewiring the Transformer with Depth-Wise LSTMs
por: Xu, Hongfei, et al.
Publicado: (2020)
por: Xu, Hongfei, et al.
Publicado: (2020)
Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning
por: Fatemi, Bahare, et al.
Publicado: (2024)
por: Fatemi, Bahare, et al.
Publicado: (2024)
Crown, Frame, Reverse: Layer-Wise Scaling Variants for LLM Pre-Training
por: Baroian, Andrei, et al.
Publicado: (2025)
por: Baroian, Andrei, et al.
Publicado: (2025)
LPCD: Unified Framework from Layer-Wise to Submodule Quantization
por: Ichikawa, Yuma, et al.
Publicado: (2025)
por: Ichikawa, Yuma, et al.
Publicado: (2025)
Unsupervised Domain Adaptation for Keyphrase Generation using Citation Contexts
por: Boudin, Florian, et al.
Publicado: (2024)
por: Boudin, Florian, et al.
Publicado: (2024)
Less Data Less Tokens: Multilingual Unification Learning for Efficient Test-Time Reasoning in LLMs
por: Chen, Kang, et al.
Publicado: (2025)
por: Chen, Kang, et al.
Publicado: (2025)
CTTA-T: Continual Test-Time Adaptation for Text Understanding via Teacher-Student with a Domain-aware and Generalized Teacher
por: Liu, Tianlun, et al.
Publicado: (2025)
por: Liu, Tianlun, et al.
Publicado: (2025)
Test-Time Adaptation via Many-Shot Prompting: Benefits, Limits, and Pitfalls
por: Upasani, Shubhangi, et al.
Publicado: (2026)
por: Upasani, Shubhangi, et al.
Publicado: (2026)
Efficient Low-Resource Language Adaptation via Multi-Source Dynamic Logit Fusion
por: Zhang, Chen, et al.
Publicado: (2026)
por: Zhang, Chen, et al.
Publicado: (2026)
ASLoRA: Adaptive Sharing Low-Rank Adaptation Across Layers
por: Hu, Junyan, et al.
Publicado: (2024)
por: Hu, Junyan, et al.
Publicado: (2024)
Dynamic Experts Search: Enhancing Reasoning in Mixture-of-Experts LLMs at Test Time
por: Han, Yixuan, et al.
Publicado: (2025)
por: Han, Yixuan, et al.
Publicado: (2025)
Tracing Representation Progression: Analyzing and Enhancing Layer-Wise Similarity
por: Jiang, Jiachen, et al.
Publicado: (2024)
por: Jiang, Jiachen, et al.
Publicado: (2024)
You only need 4 extra tokens: Synergistic Test-time Adaptation for LLMs
por: Xu, Yijie, et al.
Publicado: (2025)
por: Xu, Yijie, et al.
Publicado: (2025)
DiSCTT: Consensus-Guided Self-Curriculum for Efficient Test-Time Adaptation in Reasoning
por: Moradi, Mohammad Mahdi, et al.
Publicado: (2026)
por: Moradi, Mohammad Mahdi, et al.
Publicado: (2026)
Faster and Better LLMs via Latency-Aware Test-Time Scaling
por: Wang, Zili, et al.
Publicado: (2025)
por: Wang, Zili, et al.
Publicado: (2025)
ETT: Expanding the Long Context Understanding Capability of LLMs at Test-Time
por: Zahirnia, Kiarash, et al.
Publicado: (2025)
por: Zahirnia, Kiarash, et al.
Publicado: (2025)
Pack of LLMs: Model Fusion at Test-Time via Perplexity Optimization
por: Mavromatis, Costas, et al.
Publicado: (2024)
por: Mavromatis, Costas, et al.
Publicado: (2024)
Evaluating the Representation of Vowels in Wav2Vec Feature Extractor: A Layer-Wise Analysis Using MFCCs
por: De Cristofaro, Domenico, et al.
Publicado: (2025)
por: De Cristofaro, Domenico, et al.
Publicado: (2025)
Ejemplares similares
-
AsyncDiff: Asynchronous Timestep Conditioning for Enhanced Text-to-Image Diffusion Inference
por: Xu, Longhuan, et al.
Publicado: (2025) -
How to Alleviate Catastrophic Forgetting in LLMs Finetuning? Hierarchical Layer-Wise and Element-Wise Regularization
por: Song, Shezheng, et al.
Publicado: (2025) -
Test-Time Policy Adaptation for Enhanced Multi-Turn Interactions with LLMs
por: Wei, Chenxing, et al.
Publicado: (2025) -
Less is More: Geometric Unlearning for LLMs with Minimal Data Disclosure
por: Tan, Chenchen, et al.
Publicado: (2026) -
Evaluating LLMs Without Oracle Feedback: Agentic Annotation Evaluation Through Unsupervised Consistency Signals
por: Chen, Cheng, et al.
Publicado: (2025)