Towards Compute-Optimal Many-Shot In-Context Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Golchin, Shahriar, Chen, Yanfei, Han, Rujun, Gandhi, Manan, Yu, Tianli, Mishra, Swaroop, Surdeanu, Mihai, Agarwal, Rishabh, Lee, Chen-Yu, Pfister, Tomas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Time Travel in LLMs: Tracing Data Contamination in Large Language Models
by: Golchin, Shahriar, et al.
Published: (2023)
by: Golchin, Shahriar, et al.
Published: (2023)
Data Contamination Quiz: A Tool to Detect and Estimate Contamination in Large Language Models
by: Golchin, Shahriar, et al.
Published: (2023)
by: Golchin, Shahriar, et al.
Published: (2023)
Memorization in In-Context Learning
by: Golchin, Shahriar, et al.
Published: (2024)
by: Golchin, Shahriar, et al.
Published: (2024)
Structured Semantic Information Helps Retrieve Better Examples for In-Context Learning Applied to Few-Shot Relation Extraction
by: Chakma, Aunabil, et al.
Published: (2026)
by: Chakma, Aunabil, et al.
Published: (2026)
The Alchemy of Thought: Understanding In-Context Learning Through Supervised Classification
by: Narnoli, Harshita, et al.
Published: (2026)
by: Narnoli, Harshita, et al.
Published: (2026)
Speculative Knowledge Distillation: Bridging the Teacher-Student Gap Through Interleaved Sampling
by: Xu, Wenda, et al.
Published: (2024)
by: Xu, Wenda, et al.
Published: (2024)
Towards Realistic Few-Shot Relation Extraction: A New Meta Dataset and Evaluation
by: Alam, Fahmida, et al.
Published: (2024)
by: Alam, Fahmida, et al.
Published: (2024)
Reverse Thinking Makes LLMs Stronger Reasoners
by: Chen, Justin Chih-Yao, et al.
Published: (2024)
by: Chen, Justin Chih-Yao, et al.
Published: (2024)
Intent Laundering: AI Safety Datasets Are Not What They Seem
by: Golchin, Shahriar, et al.
Published: (2026)
by: Golchin, Shahriar, et al.
Published: (2026)
Many-Shot In-Context Learning
by: Agarwal, Rishabh, et al.
Published: (2024)
by: Agarwal, Rishabh, et al.
Published: (2024)
Supervised Reinforcement Learning: From Expert Trajectories to Step-wise Reasoning
by: Deng, Yihe, et al.
Published: (2025)
by: Deng, Yihe, et al.
Published: (2025)
SAGE: Steerable Agentic Data Generation for Deep Search with Execution Feedback
by: Xu, Fangyuan, et al.
Published: (2026)
by: Xu, Fangyuan, et al.
Published: (2026)
RubricEM: Meta-RL with Rubric-guided Policy Decomposition beyond Verifiable Rewards
by: Li, Gaotang, et al.
Published: (2026)
by: Li, Gaotang, et al.
Published: (2026)
Re-Invoke: Tool Invocation Rewriting for Zero-Shot Tool Retrieval
by: Chen, Yanfei, et al.
Published: (2024)
by: Chen, Yanfei, et al.
Published: (2024)
A Lightweight Explainable Guardrail for Prompt Safety
by: Islam, Md Asiful, et al.
Published: (2026)
by: Islam, Md Asiful, et al.
Published: (2026)
Fane at SemEval-2025 Task 10: Zero-Shot Entity Framing with Large Language Models
by: Fane, Enfa, et al.
Published: (2025)
by: Fane, Enfa, et al.
Published: (2025)
Chain of Agents: Large Language Models Collaborating on Long-Context Tasks
by: Zhang, Yusen, et al.
Published: (2024)
by: Zhang, Yusen, et al.
Published: (2024)
Search-Adaptor: Embedding Customization for Information Retrieval
by: Yoon, Jinsung, et al.
Published: (2023)
by: Yoon, Jinsung, et al.
Published: (2023)
From Words to Numbers: Your Large Language Model Is Secretly A Capable Regressor When Given In-Context Examples
by: Vacareanu, Robert, et al.
Published: (2024)
by: Vacareanu, Robert, et al.
Published: (2024)
SkillOS: Learning Skill Curation for Self-Evolving Agents
by: Ouyang, Siru, et al.
Published: (2026)
by: Ouyang, Siru, et al.
Published: (2026)
Bridging the Long-Tail Gap: Robust Retrieval-Augmented Relation Completion via Multi-Stage Paraphrase Infusion
by: Alam, Fahmida, et al.
Published: (2026)
by: Alam, Fahmida, et al.
Published: (2026)
Say Less, Mean More: Leveraging Pragmatics in Retrieval-Augmented Generation
by: Riaz, Haris, et al.
Published: (2025)
by: Riaz, Haris, et al.
Published: (2025)
Finding a Wolf in Sheep's Clothing: Combating Adversarial Text-To-Image Prompts with Text Summarization
by: Cooper, Portia, et al.
Published: (2024)
by: Cooper, Portia, et al.
Published: (2024)
Scaling Laws for Many-Shot In-Context Learning with Self-Generated Annotations
by: Gu, Zhengyao, et al.
Published: (2025)
by: Gu, Zhengyao, et al.
Published: (2025)
PlanGEN: A Multi-Agent Framework for Generating Planning and Reasoning Trajectories for Complex Problem Solving
by: Parmar, Mihir, et al.
Published: (2025)
by: Parmar, Mihir, et al.
Published: (2025)
Many-Shot In-Context Learning for Molecular Inverse Design
by: Moayedpour, Saeed, et al.
Published: (2024)
by: Moayedpour, Saeed, et al.
Published: (2024)
In Prospect and Retrospect: Reflective Memory Management for Long-term Personalized Dialogue Agents
by: Tan, Zhen, et al.
Published: (2025)
by: Tan, Zhen, et al.
Published: (2025)
Grading Massive Open Online Courses Using Large Language Models
by: Golchin, Shahriar, et al.
Published: (2024)
by: Golchin, Shahriar, et al.
Published: (2024)
Large Language Models As MOOCs Graders
by: Golchin, Shahriar, et al.
Published: (2024)
by: Golchin, Shahriar, et al.
Published: (2024)
ELLEN: Extremely Lightly Supervised Learning For Efficient Named Entity Recognition
by: Riaz, Haris, et al.
Published: (2024)
by: Riaz, Haris, et al.
Published: (2024)
Enhancing Transformer RNNs with Multiple Temporal Perspectives
by: Dumitru, Razvan-Gabriel, et al.
Published: (2024)
by: Dumitru, Razvan-Gabriel, et al.
Published: (2024)
How Is LLM Reasoning Distracted by Irrelevant Context? An Analysis Using a Controlled Benchmark
by: Yang, Minglai, et al.
Published: (2025)
by: Yang, Minglai, et al.
Published: (2025)
Many-Shot In-Context Learning in Multimodal Foundation Models
by: Jiang, Yixing, et al.
Published: (2024)
by: Jiang, Yixing, et al.
Published: (2024)
MAPLE: Many-Shot Adaptive Pseudo-Labeling for In-Context Learning
by: Chen, Zihan, et al.
Published: (2025)
by: Chen, Zihan, et al.
Published: (2025)
Budget-Aware Tool-Use Enables Effective Agent Scaling
by: Liu, Tengxiao, et al.
Published: (2025)
by: Liu, Tengxiao, et al.
Published: (2025)
Compressing Many-Shots in In-Context Learning
by: Khatri, Devvrit, et al.
Published: (2025)
by: Khatri, Devvrit, et al.
Published: (2025)
Layer-Wise Quantization: A Pragmatic and Effective Method for Quantizing LLMs Beyond Integer Bit-Levels
by: Dumitru, Razvan-Gabriel, et al.
Published: (2024)
by: Dumitru, Razvan-Gabriel, et al.
Published: (2024)
On Many-Shot In-Context Learning for Long-Context Evaluation
by: Zou, Kaijian, et al.
Published: (2024)
by: Zou, Kaijian, et al.
Published: (2024)
Large Language Models Can Automatically Engineer Features for Few-Shot Tabular Learning
by: Han, Sungwon, et al.
Published: (2024)
by: Han, Sungwon, et al.
Published: (2024)
Can LLMs Judge Debates? Evaluating Non-Linear Reasoning via Argumentation Theory Semantics
by: Sanayei, Reza, et al.
Published: (2025)
by: Sanayei, Reza, et al.
Published: (2025)
Similar Items
-
Time Travel in LLMs: Tracing Data Contamination in Large Language Models
by: Golchin, Shahriar, et al.
Published: (2023) -
Data Contamination Quiz: A Tool to Detect and Estimate Contamination in Large Language Models
by: Golchin, Shahriar, et al.
Published: (2023) -
Memorization in In-Context Learning
by: Golchin, Shahriar, et al.
Published: (2024) -
Structured Semantic Information Helps Retrieve Better Examples for In-Context Learning Applied to Few-Shot Relation Extraction
by: Chakma, Aunabil, et al.
Published: (2026) -
The Alchemy of Thought: Understanding In-Context Learning Through Supervised Classification
by: Narnoli, Harshita, et al.
Published: (2026)