Salvato in:
| Autori principali: | Pang, Jinlong, Zhu, Zhaowei, Di, Na, Zhang, Yichi, Wang, Yaxuan, Qian, Chen, Liu, Yang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2602.00954 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Evaluating LLM-Contaminated Crowdsourcing Data Without Ground Truth
di: Zhang, Yichi, et al.
Pubblicazione: (2025)
di: Zhang, Yichi, et al.
Pubblicazione: (2025)
Token Cleaning: Fine-Grained Data Selection for LLM Supervised Fine-Tuning
di: Pang, Jinlong, et al.
Pubblicazione: (2025)
di: Pang, Jinlong, et al.
Pubblicazione: (2025)
Improving Data Efficiency via Curating LLM-Driven Rating Systems
di: Pang, Jinlong, et al.
Pubblicazione: (2024)
di: Pang, Jinlong, et al.
Pubblicazione: (2024)
Fairness Without Harm: An Influence-Guided Active Sampling Approach
di: Pang, Jinlong, et al.
Pubblicazione: (2024)
di: Pang, Jinlong, et al.
Pubblicazione: (2024)
Are You Sure? Rank Them Again: Repeated Ranking For Better Preference Datasets
di: Devine, Peter
Pubblicazione: (2024)
di: Devine, Peter
Pubblicazione: (2024)
SPPD: Self-training with Process Preference Learning Using Dynamic Value Margin
di: Yi, Hao, et al.
Pubblicazione: (2025)
di: Yi, Hao, et al.
Pubblicazione: (2025)
Unmasking and Improving Data Credibility: A Study with Datasets for Training Harmless Language Models
di: Zhu, Zhaowei, et al.
Pubblicazione: (2023)
di: Zhu, Zhaowei, et al.
Pubblicazione: (2023)
RLPO: Residual Listwise Preference Optimization for Long-Context Review Ranking
di: Jiang, Hao, et al.
Pubblicazione: (2026)
di: Jiang, Hao, et al.
Pubblicazione: (2026)
Incentivizing High-quality Participation From Federated Learning Agents
di: Pang, Jinlong, et al.
Pubblicazione: (2025)
di: Pang, Jinlong, et al.
Pubblicazione: (2025)
LLM Unlearning via Loss Adjustment with Only Forget Data
di: Wang, Yaxuan, et al.
Pubblicazione: (2024)
di: Wang, Yaxuan, et al.
Pubblicazione: (2024)
If You Can't Use Them, Recycle Them: Optimizing Merging at Scale Mitigates Performance Tradeoffs
di: Khalifa, Muhammad, et al.
Pubblicazione: (2024)
di: Khalifa, Muhammad, et al.
Pubblicazione: (2024)
Conditioning Matters: Training Diffusion Policies is Faster Than You Think
di: Dong, Zibin, et al.
Pubblicazione: (2025)
di: Dong, Zibin, et al.
Pubblicazione: (2025)
Knowledgeable Preference Alignment for LLMs in Domain-specific Question Answering
di: Zhang, Yichi, et al.
Pubblicazione: (2023)
di: Zhang, Yichi, et al.
Pubblicazione: (2023)
Transcendence: Generative Models Can Outperform The Experts That Train Them
di: Zhang, Edwin, et al.
Pubblicazione: (2024)
di: Zhang, Edwin, et al.
Pubblicazione: (2024)
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization
di: Dong, Zhijin
Pubblicazione: (2025)
di: Dong, Zhijin
Pubblicazione: (2025)
Larger or Smaller Reward Margins to Select Preferences for Alignment?
di: Huang, Kexin, et al.
Pubblicazione: (2025)
di: Huang, Kexin, et al.
Pubblicazione: (2025)
Towards Understanding the Influence of Reward Margin on Preference Model Performance
di: Qin, Bowen, et al.
Pubblicazione: (2024)
di: Qin, Bowen, et al.
Pubblicazione: (2024)
Sponsored Questions and How to Auction Them
di: Bhawalkar, Kshipra, et al.
Pubblicazione: (2025)
di: Bhawalkar, Kshipra, et al.
Pubblicazione: (2025)
Automatic Dataset Construction (ADC): Sample Collection, Data Curation, and Beyond
di: Liu, Minghao, et al.
Pubblicazione: (2024)
di: Liu, Minghao, et al.
Pubblicazione: (2024)
Legend: Leveraging Representation Engineering to Annotate Safety Margin for Preference Datasets
di: Feng, Duanyu, et al.
Pubblicazione: (2024)
di: Feng, Duanyu, et al.
Pubblicazione: (2024)
Adaptive Margin RLHF via Preference over Preferences
di: Chittepu, Yaswanth, et al.
Pubblicazione: (2025)
di: Chittepu, Yaswanth, et al.
Pubblicazione: (2025)
AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization
di: Wu, Junkang, et al.
Pubblicazione: (2024)
di: Wu, Junkang, et al.
Pubblicazione: (2024)
Just Say What You Want: Only-prompting Self-rewarding Online Preference Optimization
di: Xu, Ruijie, et al.
Pubblicazione: (2024)
di: Xu, Ruijie, et al.
Pubblicazione: (2024)
Training on the Benchmark Is Not All You Need
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
CHILL at SemEval-2025 Task 2: You Can't Just Throw Entities and Hope -- Make Your LLM to Get Them Right
di: Lee, Jaebok, et al.
Pubblicazione: (2025)
di: Lee, Jaebok, et al.
Pubblicazione: (2025)
Using Large Language Models to Assess Teachers' Pedagogical Content Knowledge
di: Yang, Yaxuan, et al.
Pubblicazione: (2025)
di: Yang, Yaxuan, et al.
Pubblicazione: (2025)
DRAGON: Guard LLM Unlearning in Context via Negative Detection and Reasoning
di: Wang, Yaxuan, et al.
Pubblicazione: (2025)
di: Wang, Yaxuan, et al.
Pubblicazione: (2025)
Look Before You Decide: Prompting Active Deduction of MLLMs for Assumptive Reasoning
di: Li, Yian, et al.
Pubblicazione: (2024)
di: Li, Yian, et al.
Pubblicazione: (2024)
Observations and Remedies for Large Language Model Bias in Self-Consuming Performative Loop
di: Wang, Yaxuan, et al.
Pubblicazione: (2026)
di: Wang, Yaxuan, et al.
Pubblicazione: (2026)
Preference Consistency Matters: Enhancing Preference Learning in Language Models with Automated Self-Curation of Training Corpora
di: Lee, JoonHo, et al.
Pubblicazione: (2024)
di: Lee, JoonHo, et al.
Pubblicazione: (2024)
LRR-Bench: Left, Right or Rotate? Vision-Language models Still Struggle With Spatial Understanding Tasks
di: Kong, Fei, et al.
Pubblicazione: (2025)
di: Kong, Fei, et al.
Pubblicazione: (2025)
$ξ$-DPO: Direct Preference Optimization via Ratio Reward Margin
di: Fan, Zhengyuan, et al.
Pubblicazione: (2026)
di: Fan, Zhengyuan, et al.
Pubblicazione: (2026)
STAS: Spatial-Temporal Return Decomposition for Multi-agent Reinforcement Learning
di: Chen, Sirui, et al.
Pubblicazione: (2023)
di: Chen, Sirui, et al.
Pubblicazione: (2023)
POST: Prior-Observation Adversarial Learning of Spatio-Temporal Associations for Multivariate Time Series Anomaly Detection
di: Zhang, Suofei, et al.
Pubblicazione: (2026)
di: Zhang, Suofei, et al.
Pubblicazione: (2026)
What LLMs Think When You Don't Tell Them What to Think About?
di: Kwon, Yongchan, et al.
Pubblicazione: (2026)
di: Kwon, Yongchan, et al.
Pubblicazione: (2026)
Learning to Generate Formally Verifiable Step-by-Step Logic Reasoning via Structured Formal Intermediaries
di: Chen, Luoxin, et al.
Pubblicazione: (2026)
di: Chen, Luoxin, et al.
Pubblicazione: (2026)
Learning What Matters Now: Dynamic Preference Inference under Contextual Shifts
di: Cao, Xianwei, et al.
Pubblicazione: (2026)
di: Cao, Xianwei, et al.
Pubblicazione: (2026)
You Only Forward Once: An Efficient Compositional Judging Paradigm
di: Zhang, Tianlong, et al.
Pubblicazione: (2025)
di: Zhang, Tianlong, et al.
Pubblicazione: (2025)
Large Language Model Unlearning via Embedding-Corrupted Prompts
di: Liu, Chris Yuhao, et al.
Pubblicazione: (2024)
di: Liu, Chris Yuhao, et al.
Pubblicazione: (2024)
Search Still Matters: Information Retrieval in the Era of Generative AI
di: Hersh, William R.
Pubblicazione: (2023)
di: Hersh, William R.
Pubblicazione: (2023)
Documenti analoghi
-
Evaluating LLM-Contaminated Crowdsourcing Data Without Ground Truth
di: Zhang, Yichi, et al.
Pubblicazione: (2025) -
Token Cleaning: Fine-Grained Data Selection for LLM Supervised Fine-Tuning
di: Pang, Jinlong, et al.
Pubblicazione: (2025) -
Improving Data Efficiency via Curating LLM-Driven Rating Systems
di: Pang, Jinlong, et al.
Pubblicazione: (2024) -
Fairness Without Harm: An Influence-Guided Active Sampling Approach
di: Pang, Jinlong, et al.
Pubblicazione: (2024) -
Are You Sure? Rank Them Again: Repeated Ranking For Better Preference Datasets
di: Devine, Peter
Pubblicazione: (2024)