Directional Consistency as a Complementary Optimization Signal: The GONO Framework
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Gera, Victor Daniel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Behavior-Based Knowledge Representation Improves Prediction of Players' Moves in Chess by 25%
von: Skidanov, Benny, et al.
Veröffentlicht: (2025)
von: Skidanov, Benny, et al.
Veröffentlicht: (2025)
MLSD: A Novel Few-Shot Learning Approach to Enhance Cross-Target and Cross-Domain Stance Detection
von: Gera, Parush, et al.
Veröffentlicht: (2025)
von: Gera, Parush, et al.
Veröffentlicht: (2025)
Complementary Recommendation in E-commerce: Definition, Approaches, and Future Directions
von: Li, Linyue, et al.
Veröffentlicht: (2024)
von: Li, Linyue, et al.
Veröffentlicht: (2024)
An Interpretable Framework Applying Protein Words to Predict Protein-Small Molecule Complementary Pairing Rules
von: Chen, Jingke, et al.
Veröffentlicht: (2026)
von: Chen, Jingke, et al.
Veröffentlicht: (2026)
Curriculum Direct Preference Optimization for Diffusion and Consistency Models
von: Croitoru, Florinel-Alin, et al.
Veröffentlicht: (2024)
von: Croitoru, Florinel-Alin, et al.
Veröffentlicht: (2024)
QSpec: Speculative Decoding with Complementary Quantization Schemes
von: Zhao, Juntao, et al.
Veröffentlicht: (2024)
von: Zhao, Juntao, et al.
Veröffentlicht: (2024)
Grid and Road Expressions Are Complementary for Trajectory Representation Learning
von: Zhou, Silin, et al.
Veröffentlicht: (2024)
von: Zhou, Silin, et al.
Veröffentlicht: (2024)
COSEE: Consistency-Oriented Signal-Based Early Exiting via Calibrated Sample Weighting Mechanism
von: He, Jianing, et al.
Veröffentlicht: (2024)
von: He, Jianing, et al.
Veröffentlicht: (2024)
DirectMultiStep: Direct Route Generation for Multistep Retrosynthesis
von: Shee, Yu, et al.
Veröffentlicht: (2024)
von: Shee, Yu, et al.
Veröffentlicht: (2024)
Continuous-Utility Direct Preference Optimization
von: Mohsin, Muhammad Ahmed, et al.
Veröffentlicht: (2026)
von: Mohsin, Muhammad Ahmed, et al.
Veröffentlicht: (2026)
ADPO: Anchored Direct Preference Optimization
von: Zixian, Wang
Veröffentlicht: (2025)
von: Zixian, Wang
Veröffentlicht: (2025)
LLM and GNN are Complementary: Distilling LLM for Multimodal Graph Learning
von: Xu, Junjie, et al.
Veröffentlicht: (2024)
von: Xu, Junjie, et al.
Veröffentlicht: (2024)
Same Signal, Opposite Meaning: Direction-Informed Adaptive Learning for LLM Agents
von: Li, Ziming, et al.
Veröffentlicht: (2026)
von: Li, Ziming, et al.
Veröffentlicht: (2026)
Tuning for Trustworthiness -- Balancing Performance and Explanation Consistency in Neural Network Optimization
von: Hinterleitner, Alexander, et al.
Veröffentlicht: (2025)
von: Hinterleitner, Alexander, et al.
Veröffentlicht: (2025)
Self-Consistency Preference Optimization
von: Prasad, Archiki, et al.
Veröffentlicht: (2024)
von: Prasad, Archiki, et al.
Veröffentlicht: (2024)
What If Consensus Lies? Selective-Complementary Reinforcement Learning at Test Time
von: Yan, Dong, et al.
Veröffentlicht: (2026)
von: Yan, Dong, et al.
Veröffentlicht: (2026)
Aligning CodeLLMs with Direct Preference Optimization
von: Miao, Yibo, et al.
Veröffentlicht: (2024)
von: Miao, Yibo, et al.
Veröffentlicht: (2024)
Information-Consistent Language Model Recommendations through Group Relative Policy Optimization
von: Prabhune, Sonal, et al.
Veröffentlicht: (2025)
von: Prabhune, Sonal, et al.
Veröffentlicht: (2025)
$β$-DPO: Direct Preference Optimization with Dynamic $β$
von: Wu, Junkang, et al.
Veröffentlicht: (2024)
von: Wu, Junkang, et al.
Veröffentlicht: (2024)
COPO: Consistency-Aware Policy Optimization
von: Han, Jinghang, et al.
Veröffentlicht: (2025)
von: Han, Jinghang, et al.
Veröffentlicht: (2025)
Engineering Artificial Intelligence: Framework, Challenges, and Future Direction
von: Lee, Jay, et al.
Veröffentlicht: (2025)
von: Lee, Jay, et al.
Veröffentlicht: (2025)
Rethinking Multimodality: Optimizing Multimodal Deep Learning for Biomedical Signal Classification
von: Oladunni, Timothy, et al.
Veröffentlicht: (2025)
von: Oladunni, Timothy, et al.
Veröffentlicht: (2025)
Adaptive Batch-Wise Sample Scheduling for Direct Preference Optimization
von: Huang, Zixuan, et al.
Veröffentlicht: (2025)
von: Huang, Zixuan, et al.
Veröffentlicht: (2025)
KL Penalty Control via Perturbation for Direct Preference Optimization
von: Lee, Sangkyu, et al.
Veröffentlicht: (2025)
von: Lee, Sangkyu, et al.
Veröffentlicht: (2025)
Inference-Time Alignment of Diffusion Models with Direct Noise Optimization
von: Tang, Zhiwei, et al.
Veröffentlicht: (2024)
von: Tang, Zhiwei, et al.
Veröffentlicht: (2024)
MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning
von: Lin, Yunze
Veröffentlicht: (2025)
von: Lin, Yunze
Veröffentlicht: (2025)
C2-DPO: Constrained Controlled Direct Preference Optimization
von: Asadi, Kavosh, et al.
Veröffentlicht: (2025)
von: Asadi, Kavosh, et al.
Veröffentlicht: (2025)
Federated Fine-Tuning of LLMs: Framework Comparison and Research Directions
von: Yan, Na, et al.
Veröffentlicht: (2025)
von: Yan, Na, et al.
Veröffentlicht: (2025)
Process Reward Models for LLM Agents: Practical Framework and Directions
von: Choudhury, Sanjiban
Veröffentlicht: (2025)
von: Choudhury, Sanjiban
Veröffentlicht: (2025)
Aligning Findings with Diagnosis: A Self-Consistent Reinforcement Learning Framework for Trustworthy Radiology Reporting
von: Zhao, Kun, et al.
Veröffentlicht: (2026)
von: Zhao, Kun, et al.
Veröffentlicht: (2026)
Complementary Learning System Empowers Online Continual Learning of Vehicle Motion Forecasting in Smart Cities
von: Li, Zirui, et al.
Veröffentlicht: (2025)
von: Li, Zirui, et al.
Veröffentlicht: (2025)
Modality vs. Morphology: A Framework for Time Series Classification for Biological Signals
von: Tschida, Jordan, et al.
Veröffentlicht: (2026)
von: Tschida, Jordan, et al.
Veröffentlicht: (2026)
$ξ$-DPO: Direct Preference Optimization via Ratio Reward Margin
von: Fan, Zhengyuan, et al.
Veröffentlicht: (2026)
von: Fan, Zhengyuan, et al.
Veröffentlicht: (2026)
Forward versus Backward: Comparing Reasoning Objectives in Direct Preference Optimization
von: Nikzad, Murtaza, et al.
Veröffentlicht: (2026)
von: Nikzad, Murtaza, et al.
Veröffentlicht: (2026)
ActiveDPO: Active Direct Preference Optimization for Sample-Efficient Alignment
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2025)
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2025)
Robust LLM Alignment via Distributionally Robust Direct Preference Optimization
von: Xu, Zaiyan, et al.
Veröffentlicht: (2025)
von: Xu, Zaiyan, et al.
Veröffentlicht: (2025)
Intelligently Weighting Multiple Reference Models for Direct Preference Optimization of LLMs
von: Wu, Skyler, et al.
Veröffentlicht: (2025)
von: Wu, Skyler, et al.
Veröffentlicht: (2025)
Risk-aware Direct Preference Optimization under Nested Risk Measure
von: Zhang, Lijun, et al.
Veröffentlicht: (2025)
von: Zhang, Lijun, et al.
Veröffentlicht: (2025)
EventADL: Open-Box Anomaly Detection and Localization Framework for Events in Cloud-Based Service Systems
von: Pham, Luan, et al.
Veröffentlicht: (2026)
von: Pham, Luan, et al.
Veröffentlicht: (2026)
Robustness as an Emergent Property of Task Performance
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2026)
von: Ashury-Tahan, Shir, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Behavior-Based Knowledge Representation Improves Prediction of Players' Moves in Chess by 25%
von: Skidanov, Benny, et al.
Veröffentlicht: (2025) -
MLSD: A Novel Few-Shot Learning Approach to Enhance Cross-Target and Cross-Domain Stance Detection
von: Gera, Parush, et al.
Veröffentlicht: (2025) -
Complementary Recommendation in E-commerce: Definition, Approaches, and Future Directions
von: Li, Linyue, et al.
Veröffentlicht: (2024) -
An Interpretable Framework Applying Protein Words to Predict Protein-Small Molecule Complementary Pairing Rules
von: Chen, Jingke, et al.
Veröffentlicht: (2026) -
Curriculum Direct Preference Optimization for Diffusion and Consistency Models
von: Croitoru, Florinel-Alin, et al.
Veröffentlicht: (2024)