Saved in:
| Main Authors: | Salvy, Nicolas, Talbot, Hugues, Thirion, Bertrand |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2602.16449 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhanced Generative Model Evaluation with Clipped Density and Coverage
by: Salvy, Nicolas, et al.
Published: (2025)
by: Salvy, Nicolas, et al.
Published: (2025)
Riemannian Flow Matching for Brain Connectivity Matrices via Pullback Geometry
by: Collas, Antoine, et al.
Published: (2025)
by: Collas, Antoine, et al.
Published: (2025)
A foundation for exact binarized morphological neural networks
by: Aouad, Theodore, et al.
Published: (2024)
by: Aouad, Theodore, et al.
Published: (2024)
Adaptive Repetition for Mitigating Position Bias in LLM-Based Ranking
by: Vardasbi, Ali, et al.
Published: (2025)
by: Vardasbi, Ali, et al.
Published: (2025)
Performance Prediction of Hub-Based Swarms
by: Jain, Puneet, et al.
Published: (2024)
by: Jain, Puneet, et al.
Published: (2024)
Distance-Based Tree-Sliced Wasserstein Distance
by: Tran, Hoang V., et al.
Published: (2025)
by: Tran, Hoang V., et al.
Published: (2025)
Logistics Hub Location Optimization: A K-Means and P-Median Model Hybrid Approach Using Road Network Distances
by: Rahman, Muhammad Abdul, et al.
Published: (2023)
by: Rahman, Muhammad Abdul, et al.
Published: (2023)
SPD Matrix Learning for Neuroimaging Analysis: Perspectives, Methods, and Challenges
by: Ju, Ce, et al.
Published: (2025)
by: Ju, Ce, et al.
Published: (2025)
Predicting Treatment Response in Body Dysmorphic Disorder with Interpretable Machine Learning
by: Costilla-Reyes, Omar, et al.
Published: (2025)
by: Costilla-Reyes, Omar, et al.
Published: (2025)
Diffusion-Based, Data-Assimilation-Enabled Super-Resolution of Hub-height Winds
by: Ma, Xiaolong, et al.
Published: (2025)
by: Ma, Xiaolong, et al.
Published: (2025)
Generalized Tree Edit Distance (GTED): A Faithful Evaluation Metric for Statement Autoformalization
by: Liu, Yuntian, et al.
Published: (2025)
by: Liu, Yuntian, et al.
Published: (2025)
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling
by: He, Shenghong
Published: (2025)
by: He, Shenghong
Published: (2025)
HEAL: Resilient and Self-* Hub-based Learning
by: Legheraba, Mohamed Amine, et al.
Published: (2026)
by: Legheraba, Mohamed Amine, et al.
Published: (2026)
CDRRM: Contrast-Driven Rubric Generation for Reliable and Interpretable Reward Modeling
by: Liu, Dengcan, et al.
Published: (2026)
by: Liu, Dengcan, et al.
Published: (2026)
BenchHub: A Unified Benchmark Suite for Holistic and Customizable LLM Evaluation
by: Kim, Eunsu, et al.
Published: (2025)
by: Kim, Eunsu, et al.
Published: (2025)
Soft Prompts for Evaluation: Measuring Conditional Distance of Capabilities
by: Nordby, Ross
Published: (2025)
by: Nordby, Ross
Published: (2025)
OptunaHub: A Platform for Black-Box Optimization
by: Ozaki, Yoshihiko, et al.
Published: (2025)
by: Ozaki, Yoshihiko, et al.
Published: (2025)
Revisiting the robustness of post-hoc interpretability methods
by: Wei, Jiawen, et al.
Published: (2024)
by: Wei, Jiawen, et al.
Published: (2024)
Enhancing Reliability in LLM-Based Secure Code Generation
by: Kharma, Mohammed F., et al.
Published: (2026)
by: Kharma, Mohammed F., et al.
Published: (2026)
CARE: Confounder-Aware Aggregation for Reliable LLM Evaluation
by: Zhao, Jitian, et al.
Published: (2026)
by: Zhao, Jitian, et al.
Published: (2026)
Mitigating Exposure Bias in Score-Based Generation of Molecular Conformations
by: Wang, Sijia, et al.
Published: (2024)
by: Wang, Sijia, et al.
Published: (2024)
Reliable and Efficient Amortized Model-based Evaluation
by: Truong, Sang, et al.
Published: (2025)
by: Truong, Sang, et al.
Published: (2025)
Evaluation of post-hoc interpretability methods in time-series classification
by: Turbé, Hugues, et al.
Published: (2022)
by: Turbé, Hugues, et al.
Published: (2022)
Understanding and Mitigating Memorization in Generative Models via Sharpness of Probability Landscapes
by: Jeon, Dongjae, et al.
Published: (2024)
by: Jeon, Dongjae, et al.
Published: (2024)
Evaluating Large Language Models for Fair and Reliable Organ Allocation
by: Kim, Brian Hyeongseok, et al.
Published: (2025)
by: Kim, Brian Hyeongseok, et al.
Published: (2025)
Towards Reliable Evaluation of Adversarial Robustness for Spiking Neural Networks
by: Wang, Jihang, et al.
Published: (2025)
by: Wang, Jihang, et al.
Published: (2025)
A Reliable Cryptographic Framework for Empirical Machine Unlearning Evaluation
by: Tu, Yiwen, et al.
Published: (2024)
by: Tu, Yiwen, et al.
Published: (2024)
Auxiliary Reward Generation with Transition Distance Representation Learning
by: Li, Siyuan, et al.
Published: (2024)
by: Li, Siyuan, et al.
Published: (2024)
Beyond Reasoning Gains: Mitigating General Capabilities Forgetting in Large Reasoning Models
by: Phan, Hoang, et al.
Published: (2025)
by: Phan, Hoang, et al.
Published: (2025)
Quantification of Credal Uncertainty: A Distance-Based Approach
by: Gonzalez-Garcia, Xabier, et al.
Published: (2026)
by: Gonzalez-Garcia, Xabier, et al.
Published: (2026)
Learn-to-Distance: Distance Learning for Detecting LLM-Generated Text
by: Zhou, Hongyi, et al.
Published: (2026)
by: Zhou, Hongyi, et al.
Published: (2026)
Causally Reliable Concept Bottleneck Models
by: De Felice, Giovanni, et al.
Published: (2025)
by: De Felice, Giovanni, et al.
Published: (2025)
Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators
by: Roytburg, Dani, et al.
Published: (2025)
by: Roytburg, Dani, et al.
Published: (2025)
Embedding Reliability Verification Constraints into Generation Expansion Planning
by: Liu, Peng, et al.
Published: (2025)
by: Liu, Peng, et al.
Published: (2025)
Evaluating the Reliability and Fidelity of Automated Judgment Systems of Large Language Models
by: Biskupski, Tom, et al.
Published: (2026)
by: Biskupski, Tom, et al.
Published: (2026)
Reliable Evaluation and Benchmarks for Statement Autoformalization
by: Poiroux, Auguste, et al.
Published: (2024)
by: Poiroux, Auguste, et al.
Published: (2024)
How to Mitigate Overfitting in Weak-to-strong Generalization?
by: Shi, Junhao, et al.
Published: (2025)
by: Shi, Junhao, et al.
Published: (2025)
Revisiting Gradient Staleness: Evaluating Distance Metrics for Asynchronous Federated Learning Aggregation
by: Wilhelm, Patrick, et al.
Published: (2026)
by: Wilhelm, Patrick, et al.
Published: (2026)
Utilizing Class Separation Distance for the Evaluation of Corruption Robustness of Machine Learning Classifiers
by: Siedel, Georg, et al.
Published: (2022)
by: Siedel, Georg, et al.
Published: (2022)
Interactive Critique-Revision Training for Reliable Structured LLM Generation
by: Yu, Fei Xu, et al.
Published: (2026)
by: Yu, Fei Xu, et al.
Published: (2026)
Similar Items
-
Enhanced Generative Model Evaluation with Clipped Density and Coverage
by: Salvy, Nicolas, et al.
Published: (2025) -
Riemannian Flow Matching for Brain Connectivity Matrices via Pullback Geometry
by: Collas, Antoine, et al.
Published: (2025) -
A foundation for exact binarized morphological neural networks
by: Aouad, Theodore, et al.
Published: (2024) -
Adaptive Repetition for Mitigating Position Bias in LLM-Based Ranking
by: Vardasbi, Ali, et al.
Published: (2025) -
Performance Prediction of Hub-Based Swarms
by: Jain, Puneet, et al.
Published: (2024)