Is In-Context Learning Sufficient for Instruction Following in LLMs?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Hao, Andriushchenko, Maksym, Croce, Francesco, Flammarion, Nicolas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Does Refusal Training in LLMs Generalize to the Past Tense?
von: Andriushchenko, Maksym, et al.
Veröffentlicht: (2024)
von: Andriushchenko, Maksym, et al.
Veröffentlicht: (2024)
Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks
von: Andriushchenko, Maksym, et al.
Veröffentlicht: (2024)
von: Andriushchenko, Maksym, et al.
Veröffentlicht: (2024)
Competition Report: Finding Universal Jailbreak Backdoors in Aligned LLMs
von: Rando, Javier, et al.
Veröffentlicht: (2024)
von: Rando, Javier, et al.
Veröffentlicht: (2024)
Long Is More for Alignment: A Simple but Tough-to-Beat Baseline for Instruction Fine-Tuning
von: Zhao, Hao, et al.
Veröffentlicht: (2024)
von: Zhao, Hao, et al.
Veröffentlicht: (2024)
HalluHard: A Hard Multi-Turn Hallucination Benchmark
von: Fan, Dongyang, et al.
Veröffentlicht: (2026)
von: Fan, Dongyang, et al.
Veröffentlicht: (2026)
Capability-Based Scaling Trends for LLM-Based Red-Teaming
von: Panfilov, Alexander, et al.
Veröffentlicht: (2025)
von: Panfilov, Alexander, et al.
Veröffentlicht: (2025)
Decomposing and Measuring Evaluation Awareness
von: Li, Changling, et al.
Veröffentlicht: (2026)
von: Li, Changling, et al.
Veröffentlicht: (2026)
From Language Modeling to Instruction Following: Understanding the Behavior Shift in LLMs after Instruction Tuning
von: Wu, Xuansheng, et al.
Veröffentlicht: (2023)
von: Wu, Xuansheng, et al.
Veröffentlicht: (2023)
FutureSim: Replaying World Events to Evaluate Adaptive Agents
von: Goel, Shashwat, et al.
Veröffentlicht: (2026)
von: Goel, Shashwat, et al.
Veröffentlicht: (2026)
Replay Failures as Successes: Sample-Efficient Reinforcement Learning for Instruction Following
von: Zhang, Kongcheng, et al.
Veröffentlicht: (2025)
von: Zhang, Kongcheng, et al.
Veröffentlicht: (2025)
Light-IF: Endowing LLMs with Generalizable Reasoning via Preview and Self-Checking for Complex Instruction Following
von: Wang, Chenyang, et al.
Veröffentlicht: (2025)
von: Wang, Chenyang, et al.
Veröffentlicht: (2025)
QuantSightBench: Evaluating LLM Quantitative Forecasting with Prediction Intervals
von: Qin, Jeremy, et al.
Veröffentlicht: (2026)
von: Qin, Jeremy, et al.
Veröffentlicht: (2026)
ReIFE: Re-evaluating Instruction-Following Evaluation
von: Liu, Yixin, et al.
Veröffentlicht: (2024)
von: Liu, Yixin, et al.
Veröffentlicht: (2024)
Enhancing and Assessing Instruction-Following with Fine-Grained Instruction Variants
von: Yang, Jiuding, et al.
Veröffentlicht: (2024)
von: Yang, Jiuding, et al.
Veröffentlicht: (2024)
Financial Instruction Following Evaluation (FIFE)
von: Matlin, Glenn, et al.
Veröffentlicht: (2025)
von: Matlin, Glenn, et al.
Veröffentlicht: (2025)
Learning Algorithms in the Limit
von: Papazov, Hristo, et al.
Veröffentlicht: (2025)
von: Papazov, Hristo, et al.
Veröffentlicht: (2025)
Why Do We Need Weight Decay in Modern Deep Learning?
von: D'Angelo, Francesco, et al.
Veröffentlicht: (2023)
von: D'Angelo, Francesco, et al.
Veröffentlicht: (2023)
OS-Harm: A Benchmark for Measuring Safety of Computer Use Agents
von: Kuntz, Thomas, et al.
Veröffentlicht: (2025)
von: Kuntz, Thomas, et al.
Veröffentlicht: (2025)
Non-instructional Fine-tuning: Enabling Instruction-Following Capabilities in Pre-trained Language Models without Instruction-Following Data
von: Xie, Juncheng, et al.
Veröffentlicht: (2024)
von: Xie, Juncheng, et al.
Veröffentlicht: (2024)
Learning and Enforcing Context-Sensitive Control for LLMs
von: Albinhassan, Mohammad, et al.
Veröffentlicht: (2026)
von: Albinhassan, Mohammad, et al.
Veröffentlicht: (2026)
SAIF: A Sparse Autoencoder Framework for Interpreting and Steering Instruction Following of Language Models
von: He, Zirui, et al.
Veröffentlicht: (2025)
von: He, Zirui, et al.
Veröffentlicht: (2025)
Improving Instruction Following in Language Models through Proxy-Based Uncertainty Estimation
von: Lee, JoonHo, et al.
Veröffentlicht: (2024)
von: Lee, JoonHo, et al.
Veröffentlicht: (2024)
SUPERNOVA: Eliciting General Reasoning in LLMs with Reinforcement Learning on Natural Instructions
von: Suvarna, Ashima, et al.
Veröffentlicht: (2026)
von: Suvarna, Ashima, et al.
Veröffentlicht: (2026)
Can LLMs Follow Simple Rules?
von: Mu, Norman, et al.
Veröffentlicht: (2023)
von: Mu, Norman, et al.
Veröffentlicht: (2023)
Boosting In-Context Learning in LLMs Through the Lens of Classical Supervised Learning
von: Gundem, Korel, et al.
Veröffentlicht: (2025)
von: Gundem, Korel, et al.
Veröffentlicht: (2025)
Improving Instruction-Following in Language Models through Activation Steering
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
Infer Human's Intentions Before Following Natural Language Instructions
von: Wan, Yanming, et al.
Veröffentlicht: (2024)
von: Wan, Yanming, et al.
Veröffentlicht: (2024)
Instruction Following by Principled Boosting Attention of Large Language Models
von: Guardieiro, Vitoria, et al.
Veröffentlicht: (2025)
von: Guardieiro, Vitoria, et al.
Veröffentlicht: (2025)
Instruction-Tuned, but Not More Verifiable Instruction-Following: A Cross-Task Diagnosis for LoRA Adapters
von: Zou, Junyi
Veröffentlicht: (2026)
von: Zou, Junyi
Veröffentlicht: (2026)
Exact Learning of Arithmetic with Differentiable Agents
von: Papazov, Hristo, et al.
Veröffentlicht: (2025)
von: Papazov, Hristo, et al.
Veröffentlicht: (2025)
Exploring Memorization and Copyright Violation in Frontier LLMs: A Study of the New York Times v. OpenAI 2023 Lawsuit
von: Freeman, Joshua, et al.
Veröffentlicht: (2024)
von: Freeman, Joshua, et al.
Veröffentlicht: (2024)
Fair In-Context Learning via Latent Concept Variables
von: Bhaila, Karuna, et al.
Veröffentlicht: (2024)
von: Bhaila, Karuna, et al.
Veröffentlicht: (2024)
Instruction Learning Paradigms: A Dual Perspective on White-box and Black-box LLMs
von: Ren, Yanwei, et al.
Veröffentlicht: (2025)
von: Ren, Yanwei, et al.
Veröffentlicht: (2025)
Conifer: Improving Complex Constrained Instruction-Following Ability of Large Language Models
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
Semantic Anchors in In-Context Learning: Why Small LLMs Cannot Flip Their Labels
von: Kumar, Anantha Padmanaban Krishna
Veröffentlicht: (2025)
von: Kumar, Anantha Padmanaban Krishna
Veröffentlicht: (2025)
LLMs Are In-Context Bandit Reinforcement Learners
von: Monea, Giovanni, et al.
Veröffentlicht: (2024)
von: Monea, Giovanni, et al.
Veröffentlicht: (2024)
Deep Language Geometry: Constructing a Metric Space from LLM Weights
von: Shamrai, Maksym, et al.
Veröffentlicht: (2025)
von: Shamrai, Maksym, et al.
Veröffentlicht: (2025)
Pragmatic Instruction Following and Goal Assistance via Cooperative Language-Guided Inverse Planning
von: Zhi-Xuan, Tan, et al.
Veröffentlicht: (2024)
von: Zhi-Xuan, Tan, et al.
Veröffentlicht: (2024)
Fragile Knowledge, Robust Instruction-Following: The Width Pruning Dichotomy in Llama-3.2
von: Martra, Pere
Veröffentlicht: (2025)
von: Martra, Pere
Veröffentlicht: (2025)
XAI4LLM. Let Machine Learning Models and LLMs Collaborate for Enhanced In-Context Learning in Healthcare
von: Nazary, Fatemeh, et al.
Veröffentlicht: (2024)
von: Nazary, Fatemeh, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Does Refusal Training in LLMs Generalize to the Past Tense?
von: Andriushchenko, Maksym, et al.
Veröffentlicht: (2024) -
Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks
von: Andriushchenko, Maksym, et al.
Veröffentlicht: (2024) -
Competition Report: Finding Universal Jailbreak Backdoors in Aligned LLMs
von: Rando, Javier, et al.
Veröffentlicht: (2024) -
Long Is More for Alignment: A Simple but Tough-to-Beat Baseline for Instruction Fine-Tuning
von: Zhao, Hao, et al.
Veröffentlicht: (2024) -
HalluHard: A Hard Multi-Turn Hallucination Benchmark
von: Fan, Dongyang, et al.
Veröffentlicht: (2026)