In-Context Learning on a Budget: A Case Study in Token Classification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Berger, Uri, Baumel, Tal, Stanovsky, Gabriel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Time to Talk: LLM Agents for Asynchronous Group Communication in Mafia Games
von: Eckhaus, Niv, et al.
Veröffentlicht: (2025)
von: Eckhaus, Niv, et al.
Veröffentlicht: (2025)
Surveying the Landscape of Image Captioning Evaluation: A Comprehensive Taxonomy, Trends and Metrics Analysis
von: Berger, Uri, et al.
Veröffentlicht: (2024)
von: Berger, Uri, et al.
Veröffentlicht: (2024)
Improving Image Captioning by Mimicking Human Reformulation Feedback at Inference-time
von: Berger, Uri, et al.
Veröffentlicht: (2025)
von: Berger, Uri, et al.
Veröffentlicht: (2025)
SAUCE: Synchronous and Asynchronous User-Customizable Environment for Multi-Agent LLM Interaction
von: Neuberger, Shlomo, et al.
Veröffentlicht: (2024)
von: Neuberger, Shlomo, et al.
Veröffentlicht: (2024)
Planted in Pretraining, Swayed by Finetuning: A Case Study on the Origins of Cognitive Biases in LLMs
von: Itzhak, Itay, et al.
Veröffentlicht: (2025)
von: Itzhak, Itay, et al.
Veröffentlicht: (2025)
The State and Fate of Summarization Datasets: A Survey
von: Dahan, Noam, et al.
Veröffentlicht: (2024)
von: Dahan, Noam, et al.
Veröffentlicht: (2024)
Do Zombies Understand? A Choose-Your-Own-Adventure Exploration of Machine Cognition
von: Goldstein, Ariel, et al.
Veröffentlicht: (2024)
von: Goldstein, Ariel, et al.
Veröffentlicht: (2024)
Comparing Humans and Models on a Similar Scale: Towards Cognitive Gender Bias Evaluation in Coreference Resolution
von: Lior, Gili, et al.
Veröffentlicht: (2023)
von: Lior, Gili, et al.
Veröffentlicht: (2023)
Looking Beyond The Top-1: Transformers Determine Top Tokens In Order
von: Lioubashevski, Daria, et al.
Veröffentlicht: (2024)
von: Lioubashevski, Daria, et al.
Veröffentlicht: (2024)
Leveraging Collection-Wide Similarities for Unsupervised Document Structure Extraction
von: Lior, Gili, et al.
Veröffentlicht: (2024)
von: Lior, Gili, et al.
Veröffentlicht: (2024)
Leveraging Digitized Newspapers to Collect Summarization Data in Low-Resource Languages
von: Dahan, Noam, et al.
Veröffentlicht: (2025)
von: Dahan, Noam, et al.
Veröffentlicht: (2025)
Beyond Memorization: Distinguishing between Reductive and Epistemic Reasoning in LLMs using Classic Logic Puzzles
von: Gabay, Adi, et al.
Veröffentlicht: (2026)
von: Gabay, Adi, et al.
Veröffentlicht: (2026)
Multilingual Large Language Models and Curse of Multilinguality
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
PromptSuite: A Task-Agnostic Framework for Multi-Prompt Generation
von: Habba, Eliya, et al.
Veröffentlicht: (2025)
von: Habba, Eliya, et al.
Veröffentlicht: (2025)
Comparing the Framing Effect in Humans and LLMs on Naturally Occurring Texts
von: Lior, Gili, et al.
Veröffentlicht: (2025)
von: Lior, Gili, et al.
Veröffentlicht: (2025)
Controllable Synthetic Clinical Note Generation with Privacy Guarantees
von: Baumel, Tal, et al.
Veröffentlicht: (2024)
von: Baumel, Tal, et al.
Veröffentlicht: (2024)
PRISM: PRIor from corpus Statistics for topic Modeling
von: Ishon, Tal, et al.
Veröffentlicht: (2026)
von: Ishon, Tal, et al.
Veröffentlicht: (2026)
Applying Intrinsic Debiasing on Downstream Tasks: Challenges and Considerations for Machine Translation
von: Iluz, Bar, et al.
Veröffentlicht: (2024)
von: Iluz, Bar, et al.
Veröffentlicht: (2024)
Cross-Lingual and Cross-Cultural Variation in Image Descriptions
von: Berger, Uri, et al.
Veröffentlicht: (2024)
von: Berger, Uri, et al.
Veröffentlicht: (2024)
ReliableEval: A Recipe for Stochastic LLM Evaluation via Method of Moments
von: Lior, Gili, et al.
Veröffentlicht: (2025)
von: Lior, Gili, et al.
Veröffentlicht: (2025)
In-Context Learning with Long-Context Models: An In-Depth Exploration
von: Bertsch, Amanda, et al.
Veröffentlicht: (2024)
von: Bertsch, Amanda, et al.
Veröffentlicht: (2024)
More Documents, Same Length: Isolating the Challenge of Multiple Documents in RAG
von: Levy, Shahar, et al.
Veröffentlicht: (2025)
von: Levy, Shahar, et al.
Veröffentlicht: (2025)
Can LLMs Help Uncover Insights about LLMs? A Large-Scale, Evolving Literature Analysis of Frontier LLMs
von: Park, Jungsoo, et al.
Veröffentlicht: (2025)
von: Park, Jungsoo, et al.
Veröffentlicht: (2025)
Learning Metadata-Agnostic Representations for Text-to-SQL In-Context Example Selection
von: Mai, Chuhong, et al.
Veröffentlicht: (2024)
von: Mai, Chuhong, et al.
Veröffentlicht: (2024)
SEAM: A Stochastic Benchmark for Multi-Document Tasks
von: Lior, Gili, et al.
Veröffentlicht: (2024)
von: Lior, Gili, et al.
Veröffentlicht: (2024)
State of What Art? A Call for Multi-Prompt LLM Evaluation
von: Mizrahi, Moran, et al.
Veröffentlicht: (2023)
von: Mizrahi, Moran, et al.
Veröffentlicht: (2023)
Emotion Classification In-Context in Spanish
von: Thapa, Bipul, et al.
Veröffentlicht: (2025)
von: Thapa, Bipul, et al.
Veröffentlicht: (2025)
Anticipatory Evaluation of Language Models
von: Park, Jungsoo, et al.
Veröffentlicht: (2025)
von: Park, Jungsoo, et al.
Veröffentlicht: (2025)
Beyond Benchmarks: On The False Promise of AI Regulation
von: Stanovsky, Gabriel, et al.
Veröffentlicht: (2025)
von: Stanovsky, Gabriel, et al.
Veröffentlicht: (2025)
From Feelings to Metrics: Understanding and Formalizing How Users Vibe-Test LLMs
von: Itzhak, Itay, et al.
Veröffentlicht: (2026)
von: Itzhak, Itay, et al.
Veröffentlicht: (2026)
Token-Budget-Aware LLM Reasoning
von: Han, Tingxu, et al.
Veröffentlicht: (2024)
von: Han, Tingxu, et al.
Veröffentlicht: (2024)
In-context Learning Generalizes, But Not Always Robustly: The Case of Syntax
von: Mueller, Aaron, et al.
Veröffentlicht: (2023)
von: Mueller, Aaron, et al.
Veröffentlicht: (2023)
Reasoning in Token Economies: Budget-Aware Evaluation of LLM Reasoning Strategies
von: Wang, Junlin, et al.
Veröffentlicht: (2024)
von: Wang, Junlin, et al.
Veröffentlicht: (2024)
ScheMatiQ: From Research Question to Structured Data through Interactive Schema Discovery
von: Levy, Shahar, et al.
Veröffentlicht: (2026)
von: Levy, Shahar, et al.
Veröffentlicht: (2026)
Schema-Driven Information Extraction from Heterogeneous Tables
von: Bai, Fan, et al.
Veröffentlicht: (2023)
von: Bai, Fan, et al.
Veröffentlicht: (2023)
Evaluating In-Context Translation with Synchronous Context-Free Grammar Transduction
von: Petty, Jackson, et al.
Veröffentlicht: (2026)
von: Petty, Jackson, et al.
Veröffentlicht: (2026)
AdaGReS:Adaptive Greedy Context Selection via Redundancy-Aware Scoring for Token-Budgeted RAG
von: Peng, Chao, et al.
Veröffentlicht: (2025)
von: Peng, Chao, et al.
Veröffentlicht: (2025)
Language Models Struggle to Use Representations Learned In-Context
von: Lepori, Michael A., et al.
Veröffentlicht: (2026)
von: Lepori, Michael A., et al.
Veröffentlicht: (2026)
A Nurse is Blue and Elephant is Rugby: Cross Domain Alignment in Large Language Models Reveal Human-like Patterns
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
Dual-Pool Token-Budget Routing for Cost-Efficient and Reliable LLM Serving
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Time to Talk: LLM Agents for Asynchronous Group Communication in Mafia Games
von: Eckhaus, Niv, et al.
Veröffentlicht: (2025) -
Surveying the Landscape of Image Captioning Evaluation: A Comprehensive Taxonomy, Trends and Metrics Analysis
von: Berger, Uri, et al.
Veröffentlicht: (2024) -
Improving Image Captioning by Mimicking Human Reformulation Feedback at Inference-time
von: Berger, Uri, et al.
Veröffentlicht: (2025) -
SAUCE: Synchronous and Asynchronous User-Customizable Environment for Multi-Agent LLM Interaction
von: Neuberger, Shlomo, et al.
Veröffentlicht: (2024) -
Planted in Pretraining, Swayed by Finetuning: A Case Study on the Origins of Cognitive Biases in LLMs
von: Itzhak, Itay, et al.
Veröffentlicht: (2025)