An Efficient Plugin Method for Metric Optimization of Black-Box Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Devic, Siddartha, Choudhary, Nurendra, Srinivasan, Anirudh, Genc, Sahika, Kveton, Branislav, Hiranandani, Gaurush |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain
von: Deb, Rohan, et al.
Veröffentlicht: (2025)
von: Deb, Rohan, et al.
Veröffentlicht: (2025)
Comparing Few to Rank Many: Active Human Preference Learning using Randomized Frank-Wolfe
von: Thekumparampil, Kiran Koshy, et al.
Veröffentlicht: (2024)
von: Thekumparampil, Kiran Koshy, et al.
Veröffentlicht: (2024)
All Against Some: Efficient Integration of Large Language Models for Message Passing in Graph Neural Networks
von: Jaiswal, Ajay, et al.
Veröffentlicht: (2024)
von: Jaiswal, Ajay, et al.
Veröffentlicht: (2024)
Context-Aware Clustering using Large Language Models
von: Tipirneni, Sindhu, et al.
Veröffentlicht: (2024)
von: Tipirneni, Sindhu, et al.
Veröffentlicht: (2024)
Logits are All We Need to Adapt Closed Models
von: Hiranandani, Gaurush, et al.
Veröffentlicht: (2025)
von: Hiranandani, Gaurush, et al.
Veröffentlicht: (2025)
Efficient and Interpretable Bandit Algorithms
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2023)
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2023)
Learning from a single labeled face and a stream of unlabeled data
von: Kveton, Branislav, et al.
Veröffentlicht: (2026)
von: Kveton, Branislav, et al.
Veröffentlicht: (2026)
Stability and Multigroup Fairness in Ranking with Uncertain Predictions
von: Devic, Siddartha, et al.
Veröffentlicht: (2024)
von: Devic, Siddartha, et al.
Veröffentlicht: (2024)
When is Multicalibration Post-Processing Necessary?
von: Hansen, Dutch, et al.
Veröffentlicht: (2024)
von: Hansen, Dutch, et al.
Veröffentlicht: (2024)
Pessimistic Off-Policy Optimization for Learning to Rank
von: Cief, Matej, et al.
Veröffentlicht: (2022)
von: Cief, Matej, et al.
Veröffentlicht: (2022)
Are LLM Decisions Faithful to Verbal Confidence?
von: Wang, Jiawei, et al.
Veröffentlicht: (2026)
von: Wang, Jiawei, et al.
Veröffentlicht: (2026)
LLM-as-Judge on a Budget
von: Saha, Aadirupa, et al.
Veröffentlicht: (2026)
von: Saha, Aadirupa, et al.
Veröffentlicht: (2026)
Cross-Validated Off-Policy Evaluation
von: Cief, Matej, et al.
Veröffentlicht: (2024)
von: Cief, Matej, et al.
Veröffentlicht: (2024)
Auditability and the Landscape of Distance to Multicalibration
von: Derhake, Nathan, et al.
Veröffentlicht: (2025)
von: Derhake, Nathan, et al.
Veröffentlicht: (2025)
Flexible Routing via Uncertainty Decomposition
von: Peale, Charlotte, et al.
Veröffentlicht: (2026)
von: Peale, Charlotte, et al.
Veröffentlicht: (2026)
Proper Learnability and the Role of Unlabeled Data
von: Asilis, Julian, et al.
Veröffentlicht: (2025)
von: Asilis, Julian, et al.
Veröffentlicht: (2025)
Regularization and Optimal Multiclass Learning
von: Asilis, Julian, et al.
Veröffentlicht: (2023)
von: Asilis, Julian, et al.
Veröffentlicht: (2023)
Semi-supervised learning with max-margin graph cuts
von: Kveton, Branislav, et al.
Veröffentlicht: (2026)
von: Kveton, Branislav, et al.
Veröffentlicht: (2026)
Spectral bandits for smooth graph functions
von: Valko, Michal, et al.
Veröffentlicht: (2026)
von: Valko, Michal, et al.
Veröffentlicht: (2026)
Online semi-supervised perception: Real-time learning without explicit feedback
von: Kveton, Branislav, et al.
Veröffentlicht: (2026)
von: Kveton, Branislav, et al.
Veröffentlicht: (2026)
Multi-Objective Alignment of Large Language Models Through Hypervolume Maximization
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2024)
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2024)
Learning to Reason in LLMs by Expectation Maximization
von: Lee, Junghyun, et al.
Veröffentlicht: (2025)
von: Lee, Junghyun, et al.
Veröffentlicht: (2025)
Spectral bandits for smooth graph functions with applications in recommender systems
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
Off-Policy Evaluation from Logged Human Feedback
von: Bhargava, Aniruddha, et al.
Veröffentlicht: (2024)
von: Bhargava, Aniruddha, et al.
Veröffentlicht: (2024)
Online Posterior Sampling with a Diffusion Prior
von: Kveton, Branislav, et al.
Veröffentlicht: (2024)
von: Kveton, Branislav, et al.
Veröffentlicht: (2024)
Evidence-based anomaly detection in clinical domains
von: Hauskrecht, Milos, et al.
Veröffentlicht: (2026)
von: Hauskrecht, Milos, et al.
Veröffentlicht: (2026)
Finite-Time Logarithmic Bayes Regret Upper Bounds
von: Atsidakou, Alexia, et al.
Veröffentlicht: (2023)
von: Atsidakou, Alexia, et al.
Veröffentlicht: (2023)
Active Learning for Direct Preference Optimization
von: Kveton, Branislav, et al.
Veröffentlicht: (2025)
von: Kveton, Branislav, et al.
Veröffentlicht: (2025)
Partial Policy Gradients for RL in LLMs
von: Mathur, Puneet, et al.
Veröffentlicht: (2026)
von: Mathur, Puneet, et al.
Veröffentlicht: (2026)
Transductive Learning Is Compact
von: Asilis, Julian, et al.
Veröffentlicht: (2024)
von: Asilis, Julian, et al.
Veröffentlicht: (2024)
Conditional anomaly detection with soft harmonic functions
von: Valko, Michal, et al.
Veröffentlicht: (2026)
von: Valko, Michal, et al.
Veröffentlicht: (2026)
Conditional anomaly detection using soft harmonic functions: An application to clinical alerting
von: Valko, Michal, et al.
Veröffentlicht: (2026)
von: Valko, Michal, et al.
Veröffentlicht: (2026)
MADA: Meta-Adaptive Optimizers through hyper-gradient Descent
von: Ozkara, Kaan, et al.
Veröffentlicht: (2024)
von: Ozkara, Kaan, et al.
Veröffentlicht: (2024)
Experimental Design for Active Transductive Inference in Large Language Models
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2024)
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2024)
ML-Tool-Bench: Tool-Augmented Planning for ML Tasks
von: Chittepu, Yaswanth, et al.
Veröffentlicht: (2025)
von: Chittepu, Yaswanth, et al.
Veröffentlicht: (2025)
Spectral bandits
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
von: Kocák, Tomáš, et al.
Veröffentlicht: (2026)
Agentic Planning with Reasoning for Image Styling via Offline RL
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2026)
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2026)
Language-Model Prior Overcomes Cold-Start Items
von: Wang, Shiyu, et al.
Veröffentlicht: (2024)
von: Wang, Shiyu, et al.
Veröffentlicht: (2024)
RADAR: Reasoning-Ability and Difficulty-Aware Routing for Reasoning LLMs
von: Fernandez, Nigel, et al.
Veröffentlicht: (2025)
von: Fernandez, Nigel, et al.
Veröffentlicht: (2025)
Surrogate-Based Black-Box Optimization Method for Costly Molecular Properties
von: Leguy, Jules, et al.
Veröffentlicht: (2021)
von: Leguy, Jules, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain
von: Deb, Rohan, et al.
Veröffentlicht: (2025) -
Comparing Few to Rank Many: Active Human Preference Learning using Randomized Frank-Wolfe
von: Thekumparampil, Kiran Koshy, et al.
Veröffentlicht: (2024) -
All Against Some: Efficient Integration of Large Language Models for Message Passing in Graph Neural Networks
von: Jaiswal, Ajay, et al.
Veröffentlicht: (2024) -
Context-Aware Clustering using Large Language Models
von: Tipirneni, Sindhu, et al.
Veröffentlicht: (2024) -
Logits are All We Need to Adapt Closed Models
von: Hiranandani, Gaurush, et al.
Veröffentlicht: (2025)