A Study on the Calibration of In-context Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Hanlin, Zhang, Yi-Fan, Yu, Yaodong, Madeka, Dhruv, Foster, Dean, Xing, Eric, Lakkaraju, Himabindu, Kakade, Sham |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Follow My Instruction and Spill the Beans: Scalable Data Extraction from Retrieval-Augmented Generation Systems
di: Qi, Zhenting, et al.
Pubblicazione: (2024)
di: Qi, Zhenting, et al.
Pubblicazione: (2024)
EvoLM: In Search of Lost Language Model Training Dynamics
di: Qi, Zhenting, et al.
Pubblicazione: (2025)
di: Qi, Zhenting, et al.
Pubblicazione: (2025)
Scaling Laws for Imitation Learning in Single-Agent Games
di: Tuyls, Jens, et al.
Pubblicazione: (2023)
di: Tuyls, Jens, et al.
Pubblicazione: (2023)
Discovering Hierarchical Latent Capabilities of Language Models via Causal Representation Learning
di: Jin, Jikai, et al.
Pubblicazione: (2025)
di: Jin, Jikai, et al.
Pubblicazione: (2025)
Prescriptive Scaling Reveals the Evolution of Language Model Capabilities
di: Zhang, Hanlin, et al.
Pubblicazione: (2026)
di: Zhang, Hanlin, et al.
Pubblicazione: (2026)
CoLoR-Filter: Conditional Loss Reduction Filtering for Targeted Language Model Pre-training
di: Brandfonbrener, David, et al.
Pubblicazione: (2024)
di: Brandfonbrener, David, et al.
Pubblicazione: (2024)
Connections between Schedule-Free Optimizers, AdEMAMix, and Accelerated SGD Variants
di: Morwani, Depen, et al.
Pubblicazione: (2025)
di: Morwani, Depen, et al.
Pubblicazione: (2025)
Manipulating Large Language Models to Increase Product Visibility
di: Kumar, Aounon, et al.
Pubblicazione: (2024)
di: Kumar, Aounon, et al.
Pubblicazione: (2024)
Learning Recourse Costs from Pairwise Feature Comparisons
di: Rawal, Kaivalya, et al.
Pubblicazione: (2024)
di: Rawal, Kaivalya, et al.
Pubblicazione: (2024)
Evaluating Adversarial Robustness of Concept Representations in Sparse Autoencoders
di: Li, Aaron J., et al.
Pubblicazione: (2025)
di: Li, Aaron J., et al.
Pubblicazione: (2025)
Neural Coordination and Capacity Control for Inventory Management
di: Eisenach, Carson, et al.
Pubblicazione: (2024)
di: Eisenach, Carson, et al.
Pubblicazione: (2024)
Peer-Predictive Self-Training for Language Model Reasoning
di: Feng, Shi, et al.
Pubblicazione: (2026)
di: Feng, Shi, et al.
Pubblicazione: (2026)
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
di: Song, Yuda, et al.
Pubblicazione: (2024)
di: Song, Yuda, et al.
Pubblicazione: (2024)
Learning an Inventory Control Policy with General Inventory Arrival Dynamics
di: Andaz, Sohrab, et al.
Pubblicazione: (2023)
di: Andaz, Sohrab, et al.
Pubblicazione: (2023)
In-Context Explainers: Harnessing LLMs for Explaining Black Box Models
di: Kroeger, Nicholas, et al.
Pubblicazione: (2023)
di: Kroeger, Nicholas, et al.
Pubblicazione: (2023)
How Does Critical Batch Size Scale in Pre-training?
di: Zhang, Hanlin, et al.
Pubblicazione: (2024)
di: Zhang, Hanlin, et al.
Pubblicazione: (2024)
Repeat After Me: Transformers are Better than State Space Models at Copying
di: Jelassi, Samy, et al.
Pubblicazione: (2024)
di: Jelassi, Samy, et al.
Pubblicazione: (2024)
Temporal Sparse Autoencoders: Leveraging the Sequential Nature of Language for Interpretability
di: Bhalla, Usha, et al.
Pubblicazione: (2025)
di: Bhalla, Usha, et al.
Pubblicazione: (2025)
More RLHF, More Trust? On The Impact of Preference Alignment On Trustworthiness
di: Li, Aaron J., et al.
Pubblicazione: (2024)
di: Li, Aaron J., et al.
Pubblicazione: (2024)
How Post-Training Reshapes LLMs: A Mechanistic View on Knowledge, Truthfulness, Refusal, and Confidence
di: Du, Hongzhe, et al.
Pubblicazione: (2025)
di: Du, Hongzhe, et al.
Pubblicazione: (2025)
Loss-to-Loss Prediction: Scaling Laws for All Datasets
di: Brandfonbrener, David, et al.
Pubblicazione: (2024)
di: Brandfonbrener, David, et al.
Pubblicazione: (2024)
Monitorability as a Free Gift: How RLVR Spontaneously Aligns Reasoning
di: Xiong, Zidi, et al.
Pubblicazione: (2026)
di: Xiong, Zidi, et al.
Pubblicazione: (2026)
Towards Unified Attribution in Explainable AI, Data-Centric AI, and Mechanistic Interpretability
di: Zhang, Shichang, et al.
Pubblicazione: (2025)
di: Zhang, Shichang, et al.
Pubblicazione: (2025)
The Role of Sparsity for Length Generalization in Transformers
di: Golowich, Noah, et al.
Pubblicazione: (2025)
di: Golowich, Noah, et al.
Pubblicazione: (2025)
Enhancing In-context Learning via Linear Probe Calibration
di: Abbas, Momin, et al.
Pubblicazione: (2024)
di: Abbas, Momin, et al.
Pubblicazione: (2024)
Who Gets Credit or Blame? Attributing Accountability in Modern AI Systems
di: Zhang, Shichang, et al.
Pubblicazione: (2025)
di: Zhang, Shichang, et al.
Pubblicazione: (2025)
In-Context Unlearning: Language Models as Few Shot Unlearners
di: Pawelczyk, Martin, et al.
Pubblicazione: (2023)
di: Pawelczyk, Martin, et al.
Pubblicazione: (2023)
Certifying LLM Safety against Adversarial Prompting
di: Kumar, Aounon, et al.
Pubblicazione: (2023)
di: Kumar, Aounon, et al.
Pubblicazione: (2023)
Learning Hidden Markov Models Using Conditional Samples
di: Kakade, Sham M., et al.
Pubblicazione: (2023)
di: Kakade, Sham M., et al.
Pubblicazione: (2023)
STeCa: Step-level Trajectory Calibration for LLM Agent Learning
di: Wang, Hanlin, et al.
Pubblicazione: (2025)
di: Wang, Hanlin, et al.
Pubblicazione: (2025)
Generalized Group Data Attribution
di: Ley, Dan, et al.
Pubblicazione: (2024)
di: Ley, Dan, et al.
Pubblicazione: (2024)
The Potential of Second-Order Optimization for LLMs: A Study with Full Gauss-Newton
di: Abreu, Natalie, et al.
Pubblicazione: (2025)
di: Abreu, Natalie, et al.
Pubblicazione: (2025)
Can large language models explore in-context?
di: Krishnamurthy, Akshay, et al.
Pubblicazione: (2024)
di: Krishnamurthy, Akshay, et al.
Pubblicazione: (2024)
Is a Good Foundation Necessary for Efficient Reinforcement Learning? The Computational Role of the Base Model in Exploration
di: Foster, Dylan J., et al.
Pubblicazione: (2025)
di: Foster, Dylan J., et al.
Pubblicazione: (2025)
Towards Interpretable Soft Prompts
di: Patel, Oam, et al.
Pubblicazione: (2025)
di: Patel, Oam, et al.
Pubblicazione: (2025)
Operationalizing the Blueprint for an AI Bill of Rights: Recommendations for Practitioners, Researchers, and Policy Makers
di: Oesterling, Alex, et al.
Pubblicazione: (2024)
di: Oesterling, Alex, et al.
Pubblicazione: (2024)
The Disagreement Problem in Explainable Machine Learning: A Practitioner's Perspective
di: Krishna, Satyapriya, et al.
Pubblicazione: (2022)
di: Krishna, Satyapriya, et al.
Pubblicazione: (2022)
Generalizing Trust: Weak-to-Strong Trustworthiness in Language Models
di: Pawelczyk, Martin, et al.
Pubblicazione: (2024)
di: Pawelczyk, Martin, et al.
Pubblicazione: (2024)
ALAS: Autonomous Learning Agent for Self-Updating Language Models
di: Atreja, Dhruv
Pubblicazione: (2025)
di: Atreja, Dhruv
Pubblicazione: (2025)
Does LLM Alignment Really Need Diversity? An Empirical Study of Adapting RLVR Methods for Moral Reasoning
di: Zhang, Zhaowei, et al.
Pubblicazione: (2026)
di: Zhang, Zhaowei, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Follow My Instruction and Spill the Beans: Scalable Data Extraction from Retrieval-Augmented Generation Systems
di: Qi, Zhenting, et al.
Pubblicazione: (2024) -
EvoLM: In Search of Lost Language Model Training Dynamics
di: Qi, Zhenting, et al.
Pubblicazione: (2025) -
Scaling Laws for Imitation Learning in Single-Agent Games
di: Tuyls, Jens, et al.
Pubblicazione: (2023) -
Discovering Hierarchical Latent Capabilities of Language Models via Causal Representation Learning
di: Jin, Jikai, et al.
Pubblicazione: (2025) -
Prescriptive Scaling Reveals the Evolution of Language Model Capabilities
di: Zhang, Hanlin, et al.
Pubblicazione: (2026)