ParetoHqD: Fast Offline Multiobjective Alignment of Large Language Models using Pareto High-quality Data
Fuente:
arXiv
Saved in:
| Main Authors: | Gu, Haoran, Wang, Handing, Mei, Yi, Zhang, Mengjie, Jin, Yaochu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
One Trigger Token Is Enough: A Defense Strategy for Balancing Safety and Usability in Large Language Models
by: Gu, Haoran, et al.
Published: (2025)
by: Gu, Haoran, et al.
Published: (2025)
Overlooked Safety Vulnerability in LLMs: Malicious Intelligent Optimization Algorithm Request and its Jailbreak
by: Gu, Haoran, et al.
Published: (2026)
by: Gu, Haoran, et al.
Published: (2026)
Pareto Multi-Objective Alignment for Language Models
by: He, Qiang, et al.
Published: (2025)
by: He, Qiang, et al.
Published: (2025)
AutoSG: LLM-Driven Solver Generation Solely from Task Prompts for Expensive Optimization
by: Gu, Haoran, et al.
Published: (2026)
by: Gu, Haoran, et al.
Published: (2026)
Solver-Independent Automated Problem Formulation via LLMs for High-Cost Simulation-Driven Design
by: Li, Yuchen, et al.
Published: (2025)
by: Li, Yuchen, et al.
Published: (2025)
Pareto Optimal Learning for Estimating Large Language Model Errors
by: Zhao, Theodore, et al.
Published: (2023)
by: Zhao, Theodore, et al.
Published: (2023)
Jailbreak-Zero: A Path to Pareto Optimal Red Teaming for Large Language Models
by: Hu, Kai, et al.
Published: (2025)
by: Hu, Kai, et al.
Published: (2025)
Implicit Jailbreak Attacks via Cross-Modal Information Concealment on Vision-Language Models
by: Wang, Zhaoxin, et al.
Published: (2025)
by: Wang, Zhaoxin, et al.
Published: (2025)
Self-Improvement Towards Pareto Optimality: Mitigating Preference Conflicts in Multi-Objective Alignment
by: Li, Moxin, et al.
Published: (2025)
by: Li, Moxin, et al.
Published: (2025)
Preventing Catastrophic Overfitting in Fast Adversarial Training: A Bi-level Optimization Perspective
by: Wang, Zhaoxin, et al.
Published: (2024)
by: Wang, Zhaoxin, et al.
Published: (2024)
Navigating the Alignment-Calibration Trade-off: A Pareto-Superior Frontier via Model Merging
by: Hu, Tiancheng, et al.
Published: (2025)
by: Hu, Tiancheng, et al.
Published: (2025)
Pareto-optimal Non-uniform Language Generation
by: Charikar, Moses, et al.
Published: (2025)
by: Charikar, Moses, et al.
Published: (2025)
Identification of Energy Management Configuration Concepts from a Set of Pareto-optimal Solutions
by: Lanfermann, Felix, et al.
Published: (2023)
by: Lanfermann, Felix, et al.
Published: (2023)
RULE: Reinforcement UnLEarning Achieves Forget-Retain Pareto Optimality
by: Zhang, Chenlong, et al.
Published: (2025)
by: Zhang, Chenlong, et al.
Published: (2025)
OCEAN: Offline Chain-of-thought Evaluation and Alignment in Large Language Models
by: Wu, Junda, et al.
Published: (2024)
by: Wu, Junda, et al.
Published: (2024)
UMOEA/D: A Multiobjective Evolutionary Algorithm for Uniform Pareto Objectives based on Decomposition
by: Zhang, Xiaoyuan, et al.
Published: (2024)
by: Zhang, Xiaoyuan, et al.
Published: (2024)
Panacea: Pareto Alignment via Preference Adaptation for LLMs
by: Zhong, Yifan, et al.
Published: (2024)
by: Zhong, Yifan, et al.
Published: (2024)
Pareto-Conditioned Diffusion Models for Offline Multi-Objective Optimization
by: Shrestha, Jatan, et al.
Published: (2026)
by: Shrestha, Jatan, et al.
Published: (2026)
OutSafe-Bench: A Benchmark for Multimodal Offensive Content Detection in Large Language Models
by: Yan, Yuping, et al.
Published: (2025)
by: Yan, Yuping, et al.
Published: (2025)
Enhancing the Effectiveness and Durability of Backdoor Attacks in Federated Learning through Maximizing Task Distinction
by: Wang, Zhaoxin, et al.
Published: (2025)
by: Wang, Zhaoxin, et al.
Published: (2025)
Compute-Accuracy Pareto Frontiers for Open-Source Reasoning Large Language Models
by: Prucs, Ákos, et al.
Published: (2025)
by: Prucs, Ákos, et al.
Published: (2025)
Activation-Informed Pareto-Guided Low-Rank Compression for Efficient LLM/VLM
by: Solgi, Ryan, et al.
Published: (2025)
by: Solgi, Ryan, et al.
Published: (2025)
LacaDM: A Latent Causal Diffusion Model for Multiobjective Reinforcement Learning
by: Yan, Xueming, et al.
Published: (2025)
by: Yan, Xueming, et al.
Published: (2025)
Towards Pareto Optimal Throughput in Small Language Model Serving
by: Recasens, Pol G., et al.
Published: (2024)
by: Recasens, Pol G., et al.
Published: (2024)
ParetoLens: A Visual Analytics Framework for Exploring Solution Sets of Multi-objective Evolutionary Algorithms
by: Ma, Yuxin, et al.
Published: (2025)
by: Ma, Yuxin, et al.
Published: (2025)
Internal Value Alignment in Large Language Models through Controlled Value Vector Activation
by: Jin, Haoran, et al.
Published: (2025)
by: Jin, Haoran, et al.
Published: (2025)
Accelerated Preference Optimization for Large Language Model Alignment
by: He, Jiafan, et al.
Published: (2024)
by: He, Jiafan, et al.
Published: (2024)
Voronoi-grid-based Pareto Front Learning and Its Application to Collaborative Federated Learning
by: Chen, Mengmeng, et al.
Published: (2025)
by: Chen, Mengmeng, et al.
Published: (2025)
Efficient Alignment of Large Language Models via Data Sampling
by: Khera, Amrit, et al.
Published: (2024)
by: Khera, Amrit, et al.
Published: (2024)
ParetoQ: Improving Scaling Laws in Extremely Low-bit LLM Quantization
by: Liu, Zechun, et al.
Published: (2025)
by: Liu, Zechun, et al.
Published: (2025)
MULTIVERSE: Exposing Large Language Model Alignment Problems in Diverse Worlds
by: Jin, Xiaolong, et al.
Published: (2024)
by: Jin, Xiaolong, et al.
Published: (2024)
Pareto Optimal Code Generation
by: Orlanski, Gabriel, et al.
Published: (2025)
by: Orlanski, Gabriel, et al.
Published: (2025)
Self-Play with Adversarial Critic: Provable and Scalable Offline Alignment for Language Models
by: Ji, Xiang, et al.
Published: (2024)
by: Ji, Xiang, et al.
Published: (2024)
Recurrent Drafter for Fast Speculative Decoding in Large Language Models
by: Cheng, Yunfei, et al.
Published: (2024)
by: Cheng, Yunfei, et al.
Published: (2024)
Offline Learning and Forgetting for Reasoning with Large Language Models
by: Ni, Tianwei, et al.
Published: (2025)
by: Ni, Tianwei, et al.
Published: (2025)
UC-MOA: Utility-Conditioned Multi-Objective Alignment for Distributional Pareto-Optimality
by: Cheng, Zelei, et al.
Published: (2025)
by: Cheng, Zelei, et al.
Published: (2025)
Pareto-Optimal Offline Reinforcement Learning via Smooth Tchebysheff Scalarization
by: Bhatnagar, Aadyot, et al.
Published: (2026)
by: Bhatnagar, Aadyot, et al.
Published: (2026)
Pareto-Optimal Energy Alignment for Designing Nature-Like Antibodies
by: Wen, Yibo, et al.
Published: (2024)
by: Wen, Yibo, et al.
Published: (2024)
Multiobjective Optimization under Uncertainties using Conditional Pareto Fronts
by: Trappler, Victor, et al.
Published: (2025)
by: Trappler, Victor, et al.
Published: (2025)
AP-BMM: Approximating Capability-Cost Pareto Sets of LLMs via Asynchronous Prior-Guided Bayesian Model Merging
by: Chen, Kesheng, et al.
Published: (2025)
by: Chen, Kesheng, et al.
Published: (2025)
Similar Items
-
One Trigger Token Is Enough: A Defense Strategy for Balancing Safety and Usability in Large Language Models
by: Gu, Haoran, et al.
Published: (2025) -
Overlooked Safety Vulnerability in LLMs: Malicious Intelligent Optimization Algorithm Request and its Jailbreak
by: Gu, Haoran, et al.
Published: (2026) -
Pareto Multi-Objective Alignment for Language Models
by: He, Qiang, et al.
Published: (2025) -
AutoSG: LLM-Driven Solver Generation Solely from Task Prompts for Expensive Optimization
by: Gu, Haoran, et al.
Published: (2026) -
Solver-Independent Automated Problem Formulation via LLMs for High-Cost Simulation-Driven Design
by: Li, Yuchen, et al.
Published: (2025)