Reframing Data Value for Large Language Models Through the Lens of Plausibility
Fuente:
arXiv
Salvato in:
| Autori principali: | Rammal, Mohamad Rida, Zhou, Ruida, Diggavi, Suhas |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Learning Mathematical Rules with Large Language Models
di: Gorceix, Antoine, et al.
Pubblicazione: (2024)
di: Gorceix, Antoine, et al.
Pubblicazione: (2024)
One Jump Is All You Need: Short-Cutting Transformers for Early Exit Prediction with One Jump to Fit All Exit Levels
di: Seshadri, Amrit Diggavi
Pubblicazione: (2025)
di: Seshadri, Amrit Diggavi
Pubblicazione: (2025)
DISPO: Enhancing Training Efficiency and Stability in Reinforcement Learning for Large Language Model Mathematical Reasoning
di: Karaman, Batuhan K., et al.
Pubblicazione: (2026)
di: Karaman, Batuhan K., et al.
Pubblicazione: (2026)
Regurgitative Training: The Value of Real Data in Training Large Language Models
di: Zhang, Jinghui, et al.
Pubblicazione: (2024)
di: Zhang, Jinghui, et al.
Pubblicazione: (2024)
Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer
di: Askin, Baris, et al.
Pubblicazione: (2026)
di: Askin, Baris, et al.
Pubblicazione: (2026)
SPIRE: Conditional Personalization for Federated Diffusion Generative Models
di: Ozkara, Kaan, et al.
Pubblicazione: (2025)
di: Ozkara, Kaan, et al.
Pubblicazione: (2025)
Relative Value Biases in Large Language Models
di: Hayes, William M., et al.
Pubblicazione: (2024)
di: Hayes, William M., et al.
Pubblicazione: (2024)
Rethinking Reward Model Evaluation Through the Lens of Reward Overoptimization
di: Kim, Sunghwan, et al.
Pubblicazione: (2025)
di: Kim, Sunghwan, et al.
Pubblicazione: (2025)
CodeVisionary: An Agent-based Framework for Evaluating Large Language Models in Code Generation
di: Wang, Xinchen, et al.
Pubblicazione: (2025)
di: Wang, Xinchen, et al.
Pubblicazione: (2025)
Evaluation of Large Language Models via Coupled Token Generation
di: Benz, Nina Corvelo, et al.
Pubblicazione: (2025)
di: Benz, Nina Corvelo, et al.
Pubblicazione: (2025)
CataLM: Empowering Catalyst Design Through Large Language Models
di: Wang, Ludi, et al.
Pubblicazione: (2024)
di: Wang, Ludi, et al.
Pubblicazione: (2024)
WKVQuant: Quantizing Weight and Key/Value Cache for Large Language Models Gains More
di: Yue, Yuxuan, et al.
Pubblicazione: (2024)
di: Yue, Yuxuan, et al.
Pubblicazione: (2024)
Internal Value Alignment in Large Language Models through Controlled Value Vector Activation
di: Jin, Haoran, et al.
Pubblicazione: (2025)
di: Jin, Haoran, et al.
Pubblicazione: (2025)
Explaining Large Language Models Decisions Using Shapley Values
di: Mohammadi, Behnam
Pubblicazione: (2024)
di: Mohammadi, Behnam
Pubblicazione: (2024)
DEMO: Reframing Dialogue Interaction with Fine-grained Element Modeling
di: Wang, Minzheng, et al.
Pubblicazione: (2024)
di: Wang, Minzheng, et al.
Pubblicazione: (2024)
Diversity-oriented Data Augmentation with Large Language Models
di: Wang, Zaitian, et al.
Pubblicazione: (2025)
di: Wang, Zaitian, et al.
Pubblicazione: (2025)
Lookback Lens: Detecting and Mitigating Contextual Hallucinations in Large Language Models Using Only Attention Maps
di: Chuang, Yung-Sung, et al.
Pubblicazione: (2024)
di: Chuang, Yung-Sung, et al.
Pubblicazione: (2024)
AttriLens-Mol: Attribute Guided Reinforcement Learning for Molecular Property Prediction with Large Language Models
di: Lin, Xuan, et al.
Pubblicazione: (2025)
di: Lin, Xuan, et al.
Pubblicazione: (2025)
Prediction-Powered Ranking of Large Language Models
di: Chatzi, Ivi, et al.
Pubblicazione: (2024)
di: Chatzi, Ivi, et al.
Pubblicazione: (2024)
Backward Lens: Projecting Language Model Gradients into the Vocabulary Space
di: Katz, Shahar, et al.
Pubblicazione: (2024)
di: Katz, Shahar, et al.
Pubblicazione: (2024)
GUARD: Guided Unlearning and Retention via Data Attribution for Large Language Models
di: Niu, Peizhi, et al.
Pubblicazione: (2025)
di: Niu, Peizhi, et al.
Pubblicazione: (2025)
Efficient Prompt Optimization Through the Lens of Best Arm Identification
di: Shi, Chengshuai, et al.
Pubblicazione: (2024)
di: Shi, Chengshuai, et al.
Pubblicazione: (2024)
KaSA: Knowledge-Aware Singular-Value Adaptation of Large Language Models
di: Wang, Fan, et al.
Pubblicazione: (2024)
di: Wang, Fan, et al.
Pubblicazione: (2024)
An Information-Theoretic Approach to Understanding Transformers' In-Context Learning of Variable-Order Markov Chains
di: Zhou, Ruida, et al.
Pubblicazione: (2024)
di: Zhou, Ruida, et al.
Pubblicazione: (2024)
Reasoning is Periodicity? Improving Large Language Models Through Effective Periodicity Modeling
di: Dong, Yihong, et al.
Pubblicazione: (2025)
di: Dong, Yihong, et al.
Pubblicazione: (2025)
Boosting In-Context Learning in LLMs Through the Lens of Classical Supervised Learning
di: Gundem, Korel, et al.
Pubblicazione: (2025)
di: Gundem, Korel, et al.
Pubblicazione: (2025)
How Instruction and Reasoning Data shape Post-Training: Data Quality through the Lens of Layer-wise Gradients
di: Li, Ming, et al.
Pubblicazione: (2025)
di: Li, Ming, et al.
Pubblicazione: (2025)
How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition
di: Dong, Guanting, et al.
Pubblicazione: (2023)
di: Dong, Guanting, et al.
Pubblicazione: (2023)
HARP: Hesitation-Aware Reframing in Transformer Inference Pass
di: Storaï, Romain, et al.
Pubblicazione: (2024)
di: Storaï, Romain, et al.
Pubblicazione: (2024)
Attention Satisfies: A Constraint-Satisfaction Lens on Factual Errors of Language Models
di: Yuksekgonul, Mert, et al.
Pubblicazione: (2023)
di: Yuksekgonul, Mert, et al.
Pubblicazione: (2023)
DavIR: Data Selection via Implicit Reward for Large Language Models
di: Zhou, Haotian, et al.
Pubblicazione: (2023)
di: Zhou, Haotian, et al.
Pubblicazione: (2023)
Peering Through Preferences: Unraveling Feedback Acquisition for Aligning Large Language Models
di: Bansal, Hritik, et al.
Pubblicazione: (2023)
di: Bansal, Hritik, et al.
Pubblicazione: (2023)
SePer: Measure Retrieval Utility Through The Lens Of Semantic Perplexity Reduction
di: Dai, Lu, et al.
Pubblicazione: (2025)
di: Dai, Lu, et al.
Pubblicazione: (2025)
Exploring the Plausibility of Hate and Counter Speech Detectors with Explainable AI
di: Böck, Adrian Jaques, et al.
Pubblicazione: (2024)
di: Böck, Adrian Jaques, et al.
Pubblicazione: (2024)
Beware of Calibration Data for Pruning Large Language Models
di: Ji, Yixin, et al.
Pubblicazione: (2024)
di: Ji, Yixin, et al.
Pubblicazione: (2024)
LLMSurgeon: Diagnosing Data Mixture of Large Language Models
di: Luo, Yaxin, et al.
Pubblicazione: (2026)
di: Luo, Yaxin, et al.
Pubblicazione: (2026)
DIDS: Domain Impact-aware Data Sampling for Large Language Model Training
di: Shi, Weijie, et al.
Pubblicazione: (2025)
di: Shi, Weijie, et al.
Pubblicazione: (2025)
ICQuant: Index Coding enables Low-bit LLM Quantization
di: Li, Xinlin, et al.
Pubblicazione: (2025)
di: Li, Xinlin, et al.
Pubblicazione: (2025)
Large Language Models as Optimizers
di: Yang, Chengrun, et al.
Pubblicazione: (2023)
di: Yang, Chengrun, et al.
Pubblicazione: (2023)
Data Advisor: Dynamic Data Curation for Safety Alignment of Large Language Models
di: Wang, Fei, et al.
Pubblicazione: (2024)
di: Wang, Fei, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Learning Mathematical Rules with Large Language Models
di: Gorceix, Antoine, et al.
Pubblicazione: (2024) -
One Jump Is All You Need: Short-Cutting Transformers for Early Exit Prediction with One Jump to Fit All Exit Levels
di: Seshadri, Amrit Diggavi
Pubblicazione: (2025) -
DISPO: Enhancing Training Efficiency and Stability in Reinforcement Learning for Large Language Model Mathematical Reasoning
di: Karaman, Batuhan K., et al.
Pubblicazione: (2026) -
Regurgitative Training: The Value of Real Data in Training Large Language Models
di: Zhang, Jinghui, et al.
Pubblicazione: (2024) -
Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer
di: Askin, Baris, et al.
Pubblicazione: (2026)