Saved in:
| Main Author: | Nagpal, Chirag |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2508.13189 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models
by: Maura-Rivero, Roberto-Rafael, et al.
Published: (2025)
by: Maura-Rivero, Roberto-Rafael, et al.
Published: (2025)
Continuous-Utility Direct Preference Optimization
by: Mohsin, Muhammad Ahmed, et al.
Published: (2026)
by: Mohsin, Muhammad Ahmed, et al.
Published: (2026)
Predicting Deterioration in Mild Cognitive Impairment with Survival Transformers, Extreme Gradient Boosting and Cox Proportional Hazard Modelling
by: Musto, Henry, et al.
Published: (2024)
by: Musto, Henry, et al.
Published: (2024)
Rethinking Explainability in the Era of Multimodal AI
by: Agarwal, Chirag
Published: (2025)
by: Agarwal, Chirag
Published: (2025)
time2time: Causal Intervention in Hidden States to Simulate Rare Events in Time Series Foundation Models
by: Sanyal, Debdeep, et al.
Published: (2025)
by: Sanyal, Debdeep, et al.
Published: (2025)
Analyzing Memorization in Large Language Models through the Lens of Model Attribution
by: Menta, Tarun Ram, et al.
Published: (2025)
by: Menta, Tarun Ram, et al.
Published: (2025)
Holistic Utility Preference Learning for Listwise Alignment
by: Zhou, Jiacong, et al.
Published: (2024)
by: Zhou, Jiacong, et al.
Published: (2024)
Optimistic Rates for Learning from Label Proportions
by: Li, Gene, et al.
Published: (2024)
by: Li, Gene, et al.
Published: (2024)
Generative Modeling from Black-box Corruptions via Self-Consistent Stochastic Interpolants
by: Modi, Chirag, et al.
Published: (2025)
by: Modi, Chirag, et al.
Published: (2025)
Class-Proportional Coreset Selection for Difficulty-Separable Data
by: Tsai, Elisa, et al.
Published: (2025)
by: Tsai, Elisa, et al.
Published: (2025)
Transparency and Proportionality in Post-Processing Algorithmic Bias Correction
by: Ferreira, Juliett Suárez, et al.
Published: (2025)
by: Ferreira, Juliett Suárez, et al.
Published: (2025)
Quantifying the Gain in Weak-to-Strong Generalization
by: Charikar, Moses, et al.
Published: (2024)
by: Charikar, Moses, et al.
Published: (2024)
To Predict or Not To Predict? Proportionally Masked Autoencoders for Tabular Data Imputation
by: Kim, Jungkyu, et al.
Published: (2024)
by: Kim, Jungkyu, et al.
Published: (2024)
Modeling Multi-Objective Tradeoffs with Monotonic Utility Functions
by: Chen, Edward, et al.
Published: (2024)
by: Chen, Edward, et al.
Published: (2024)
Discrete-Choice Model with Generalized Additive Utility Network
by: Nishi, Tomoki, et al.
Published: (2023)
by: Nishi, Tomoki, et al.
Published: (2023)
Posts of Peril: Detecting Information About Hazards in Text
by: Burghardt, Keith, et al.
Published: (2024)
by: Burghardt, Keith, et al.
Published: (2024)
Beyond RLHF and NLHF: Population-Proportional Alignment under an Axiomatic Framework
by: Kim, Kihyun, et al.
Published: (2025)
by: Kim, Kihyun, et al.
Published: (2025)
Mixture Proportion Estimation and Weakly-supervised Kernel Test for Conditional Independence
by: Hirose, Yushi, et al.
Published: (2026)
by: Hirose, Yushi, et al.
Published: (2026)
Trajectory Modeling via Random Utility Inverse Reinforcement Learning
by: Pitombeira-Neto, Anselmo R., et al.
Published: (2021)
by: Pitombeira-Neto, Anselmo R., et al.
Published: (2021)
Clear Preferences Leave Traces: Reference Model-Guided Sampling for Preference Learning
by: Diwan, Nirav, et al.
Published: (2025)
by: Diwan, Nirav, et al.
Published: (2025)
ZeroFlood: Flood Hazard Mapping from Single-Modality SAR Using Geo-Foundation Models
by: Kim, Hyeongkyun, et al.
Published: (2025)
by: Kim, Hyeongkyun, et al.
Published: (2025)
Agnostic Language Identification and Generation
by: Høgsgaard, Mikael Møller, et al.
Published: (2026)
by: Høgsgaard, Mikael Møller, et al.
Published: (2026)
Learning from Label Proportions: Bootstrapping Supervised Learners via Belief Propagation
by: Havaldar, Shreyas, et al.
Published: (2023)
by: Havaldar, Shreyas, et al.
Published: (2023)
Evaluating Sample Utility for Efficient Data Selection by Mimicking Model Weights
by: Huang, Tzu-Heng, et al.
Published: (2025)
by: Huang, Tzu-Heng, et al.
Published: (2025)
Efficient Data Selection for Multimodal Models via Incremental Optimization Utility
by: Jing, Jinhao, et al.
Published: (2026)
by: Jing, Jinhao, et al.
Published: (2026)
From Narratives to Probabilistic Reasoning: Predicting and Interpreting Drivers' Hazardous Actions in Crashes Using Large Language Model
by: Chen, Boyou, et al.
Published: (2025)
by: Chen, Boyou, et al.
Published: (2025)
A Novel Approach to Balance Convenience and Nutrition in Meals With Long-Term Group Recommendations and Reasoning on Multimodal Recipes and its Implementation in BEACON
by: Nagpal, Vansh, et al.
Published: (2024)
by: Nagpal, Vansh, et al.
Published: (2024)
Sharper Error Bounds in Late Fusion Multi-view Clustering Using Eigenvalue Proportion
by: Du, Liang, et al.
Published: (2024)
by: Du, Liang, et al.
Published: (2024)
Scalable Utility-Aware Multiclass Calibration
by: Hegazy, Mahmoud, et al.
Published: (2025)
by: Hegazy, Mahmoud, et al.
Published: (2025)
Cumulative Hazard Function Based Efficient Multivariate Temporal Point Process Learning
by: Liu, Bingqing
Published: (2024)
by: Liu, Bingqing
Published: (2024)
Towards Operationalizing Right to Data Protection
by: Java, Abhinav, et al.
Published: (2024)
by: Java, Abhinav, et al.
Published: (2024)
Preference-Based Alignment of Discrete Diffusion Models
by: Borso, Umberto, et al.
Published: (2025)
by: Borso, Umberto, et al.
Published: (2025)
In-Context Reward Adaptation for Robust Preference Modeling
by: Sun, Zhenyu, et al.
Published: (2026)
by: Sun, Zhenyu, et al.
Published: (2026)
Forming Auxiliary High-confident Instance-level Loss to Promote Learning from Label Proportions
by: Ma, Tianhao, et al.
Published: (2024)
by: Ma, Tianhao, et al.
Published: (2024)
In-Context Explainers: Harnessing LLMs for Explaining Black Box Models
by: Kroeger, Nicholas, et al.
Published: (2023)
by: Kroeger, Nicholas, et al.
Published: (2023)
Risk Awareness Injection: Calibrating Vision-Language Models for Safety without Compromising Utility
by: Wang, Mengxuan, et al.
Published: (2026)
by: Wang, Mengxuan, et al.
Published: (2026)
Grounding and Enhancing Informativeness and Utility in Dataset Distillation
by: Wang, Shaobo, et al.
Published: (2026)
by: Wang, Shaobo, et al.
Published: (2026)
Policy-labeled Preference Learning: Is Preference Enough for RLHF?
by: Cho, Taehyun, et al.
Published: (2025)
by: Cho, Taehyun, et al.
Published: (2025)
Preference as Reward, Maximum Preference Optimization with Importance Sampling
by: Jiang, Zaifan, et al.
Published: (2023)
by: Jiang, Zaifan, et al.
Published: (2023)
IRPM: Intergroup Relative Preference Modeling for Pointwise Generative Reward Models
by: Song, Haonan, et al.
Published: (2026)
by: Song, Haonan, et al.
Published: (2026)
Similar Items
-
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models
by: Maura-Rivero, Roberto-Rafael, et al.
Published: (2025) -
Continuous-Utility Direct Preference Optimization
by: Mohsin, Muhammad Ahmed, et al.
Published: (2026) -
Predicting Deterioration in Mild Cognitive Impairment with Survival Transformers, Extreme Gradient Boosting and Cox Proportional Hazard Modelling
by: Musto, Henry, et al.
Published: (2024) -
Rethinking Explainability in the Era of Multimodal AI
by: Agarwal, Chirag
Published: (2025) -
time2time: Causal Intervention in Hidden States to Simulate Rare Events in Time Series Foundation Models
by: Sanyal, Debdeep, et al.
Published: (2025)