Cost-aware LLM-based Online Dataset Annotation
Fuente:
arXiv
Saved in:
| Main Authors: | Elumar, Eray Can, Tekin, Cem, Yagan, Osman |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bandits with Anytime Knapsacks
by: Elumar, Eray Can, et al.
Published: (2025)
by: Elumar, Eray Can, et al.
Published: (2025)
On Balancing Sparsity with Reliable Connectivity in Distributed Network Design with Random K-out Graphs
by: Sood, Mansi, et al.
Published: (2025)
by: Sood, Mansi, et al.
Published: (2025)
InfAlign: Inference-aware language model alignment
by: Balashankar, Ananth, et al.
Published: (2024)
by: Balashankar, Ananth, et al.
Published: (2024)
A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability
by: Omidvar, Hamed, et al.
Published: (2026)
by: Omidvar, Hamed, et al.
Published: (2026)
Multi-turn Training with Basic Human Feedback Helps Little on LLM Reasoning
by: Liu, Qiang, et al.
Published: (2025)
by: Liu, Qiang, et al.
Published: (2025)
An Enhanced Text Compression Approach Using Transformer-based Language Models
by: Rahman, Chowdhury Mofizur, et al.
Published: (2024)
by: Rahman, Chowdhury Mofizur, et al.
Published: (2024)
Evaluation of LLM-based Strategies for the Extraction of Food Product Information from Online Shops
by: Brosch, Christoph, et al.
Published: (2025)
by: Brosch, Christoph, et al.
Published: (2025)
Learning is Forgetting: LLM Training As Lossy Compression
by: Conklin, Henry C., et al.
Published: (2026)
by: Conklin, Henry C., et al.
Published: (2026)
A Training-free Method for LLM Text Attribution
by: Radvand, Tara, et al.
Published: (2025)
by: Radvand, Tara, et al.
Published: (2025)
CAFe: Cost and Age aware Federated Learning
by: Liyanaarachchi, Sahan, et al.
Published: (2024)
by: Liyanaarachchi, Sahan, et al.
Published: (2024)
A Rate-Distortion Framework for Summarization
by: Arda, Enes, et al.
Published: (2025)
by: Arda, Enes, et al.
Published: (2025)
An Information-theoretic Multi-task Representation Learning Framework for Natural Language Understanding
by: Hu, Dou, et al.
Published: (2025)
by: Hu, Dou, et al.
Published: (2025)
What Makes the Preferred Thinking Direction for LLMs in Multiple-choice Questions?
by: Zhang, Yizhe, et al.
Published: (2025)
by: Zhang, Yizhe, et al.
Published: (2025)
Iterative Counterfactual Data Augmentation
by: Plyler, Mitchell, et al.
Published: (2025)
by: Plyler, Mitchell, et al.
Published: (2025)
Attention with Markov: A Framework for Principled Analysis of Transformers via Markov Chains
by: Makkuva, Ashok Vardhan, et al.
Published: (2024)
by: Makkuva, Ashok Vardhan, et al.
Published: (2024)
Proposal and study of statistical features for string similarity computation and classification
by: Rodrigues, E. O., et al.
Published: (2026)
by: Rodrigues, E. O., et al.
Published: (2026)
Efficient Learned Data Compression via Dual-Stream Feature Decoupling
by: Ma, Huidong, et al.
Published: (2026)
by: Ma, Huidong, et al.
Published: (2026)
Theoretical guarantees on the best-of-n alignment policy
by: Beirami, Ahmad, et al.
Published: (2024)
by: Beirami, Ahmad, et al.
Published: (2024)
Effective Context in Transformers: An Analysis of Fragmentation and Tokenization
by: Fesharaki, Amirmehdi Jafari, et al.
Published: (2026)
by: Fesharaki, Amirmehdi Jafari, et al.
Published: (2026)
Understanding Factual Recall in Transformers via Associative Memories
by: Nichani, Eshaan, et al.
Published: (2024)
by: Nichani, Eshaan, et al.
Published: (2024)
RateQuant: Optimal Mixed-Precision KV Cache Quantization via Rate-Distortion Theory
by: Zuo, Fei, et al.
Published: (2026)
by: Zuo, Fei, et al.
Published: (2026)
Filtering Beats Fine Tuning: A Bayesian Kalman View of In Context Learning in LLMs
by: Kiruluta, Andrew
Published: (2026)
by: Kiruluta, Andrew
Published: (2026)
Speculative Decoding Scaling Laws (SDSL): Throughput Optimization Made Simple
by: Bozorgkhoo, Amirhossein, et al.
Published: (2026)
by: Bozorgkhoo, Amirhossein, et al.
Published: (2026)
A Mathematical Theory for Learning Semantic Languages by Abstract Learners
by: Liao, Kuo-Yu, et al.
Published: (2024)
by: Liao, Kuo-Yu, et al.
Published: (2024)
Transformers on Markov Data: Constant Depth Suffices
by: Rajaraman, Nived, et al.
Published: (2024)
by: Rajaraman, Nived, et al.
Published: (2024)
Fundamental Limits of Prompt Compression: A Rate-Distortion Framework for Black-Box Language Models
by: Nagle, Alliot, et al.
Published: (2024)
by: Nagle, Alliot, et al.
Published: (2024)
MultiTok: Variable-Length Tokenization for Efficient LLMs Adapted from LZW Compression
by: Elias, Noel, et al.
Published: (2024)
by: Elias, Noel, et al.
Published: (2024)
Latent Space Alignment for Semantic Channel Equalization
by: Hüttebräucker, Tomás, et al.
Published: (2024)
by: Hüttebräucker, Tomás, et al.
Published: (2024)
Information-Theoretic Generative Clustering of Documents
by: Du, Xin, et al.
Published: (2024)
by: Du, Xin, et al.
Published: (2024)
Semantic Faithfulness and Entropy Production Measures to Tame Your LLM Demons and Manage Hallucinations
by: Halperin, Igor
Published: (2025)
by: Halperin, Igor
Published: (2025)
Quantifying Logical Consistency in Transformers via Query-Key Alignment
by: Tulchinskii, Eduard, et al.
Published: (2025)
by: Tulchinskii, Eduard, et al.
Published: (2025)
Theoretical Limits of Language Model Alignment
by: Paes, Lucas Monteiro, et al.
Published: (2026)
by: Paes, Lucas Monteiro, et al.
Published: (2026)
LLM vs. Lawyers: Identifying a Subset of Summary Judgments in a Large UK Case Law Dataset
by: Izzidien, Ahmed, et al.
Published: (2024)
by: Izzidien, Ahmed, et al.
Published: (2024)
Forgetting-MarI: LLM Unlearning via Marginal Information Regularization
by: Xu, Shizhou, et al.
Published: (2025)
by: Xu, Shizhou, et al.
Published: (2025)
Challenges and Considerations in Annotating Legal Data: A Comprehensive Overview
by: Darji, Harshil, et al.
Published: (2024)
by: Darji, Harshil, et al.
Published: (2024)
OD-Stega: LLM-Based Relatively Secure Steganography via Optimized Distributions
by: Huang, Yu-Shin, et al.
Published: (2024)
by: Huang, Yu-Shin, et al.
Published: (2024)
ABCD-LINK: Annotation Bootstrapping for Cross-Document Fine-Grained Links
by: Basch, Serwar, et al.
Published: (2025)
by: Basch, Serwar, et al.
Published: (2025)
Achieving Logarithmic Regret in KL-Regularized Zero-Sum Markov Games
by: Nayak, Anupam, et al.
Published: (2025)
by: Nayak, Anupam, et al.
Published: (2025)
Optimize Incompatible Parameters through Compatibility-aware Knowledge Integration
by: Lv, Zheqi, et al.
Published: (2025)
by: Lv, Zheqi, et al.
Published: (2025)
GORAG: Graph-based Online Retrieval Augmented Generation for Dynamic Few-shot Social Media Text Classification
by: Wang, Yubo, et al.
Published: (2025)
by: Wang, Yubo, et al.
Published: (2025)
Similar Items
-
Bandits with Anytime Knapsacks
by: Elumar, Eray Can, et al.
Published: (2025) -
On Balancing Sparsity with Reliable Connectivity in Distributed Network Design with Random K-out Graphs
by: Sood, Mansi, et al.
Published: (2025) -
InfAlign: Inference-aware language model alignment
by: Balashankar, Ananth, et al.
Published: (2024) -
A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability
by: Omidvar, Hamed, et al.
Published: (2026) -
Multi-turn Training with Basic Human Feedback Helps Little on LLM Reasoning
by: Liu, Qiang, et al.
Published: (2025)