Significativity Indices for Agreement Values
Fuente:
arXiv
Guardado en:
| Autores principales: | Casagrande, Alberto, Fabris, Francesco, Girometti, Rossano, Pagliarini, Roberto |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Model Agreement via Anchoring
por: Eaton, Eric, et al.
Publicado: (2026)
por: Eaton, Eric, et al.
Publicado: (2026)
Modeling Teams Performance Using Deep Representational Learning on Graphs
por: Carli, Francesco, et al.
Publicado: (2022)
por: Carli, Francesco, et al.
Publicado: (2022)
Highway Value Iteration Networks
por: Wang, Yuhui, et al.
Publicado: (2024)
por: Wang, Yuhui, et al.
Publicado: (2024)
The Price of Agreement: Measuring LLM Sycophancy in Agentic Financial Applications
por: Zhao, Zhenyu, et al.
Publicado: (2026)
por: Zhao, Zhenyu, et al.
Publicado: (2026)
Test-Time Adaptation Induces Stronger Accuracy and Agreement-on-the-Line
por: Kim, Eungyeup, et al.
Publicado: (2023)
por: Kim, Eungyeup, et al.
Publicado: (2023)
Sycophantic Anchors: Localizing and Quantifying User Agreement in Reasoning Models
por: Duszenko, Jacek
Publicado: (2026)
por: Duszenko, Jacek
Publicado: (2026)
Beyond Gradient Averaging in Parallel Optimization: Improved Robustness through Gradient Agreement Filtering
por: Chaubard, Francois, et al.
Publicado: (2024)
por: Chaubard, Francois, et al.
Publicado: (2024)
Reinforcement Learning for Durable Algorithmic Recourse
por: Ceccon, Marina, et al.
Publicado: (2025)
por: Ceccon, Marina, et al.
Publicado: (2025)
DeepDFA: Injecting Temporal Logic in Deep Learning for Sequential Subsymbolic Applications
por: Umili, Elena, et al.
Publicado: (2026)
por: Umili, Elena, et al.
Publicado: (2026)
Neural Reward Machines
por: Umili, Elena, et al.
Publicado: (2024)
por: Umili, Elena, et al.
Publicado: (2024)
Don't Push the Button! Exploring Data Leakage Risks in Machine Learning and Transfer Learning
por: Apicella, Andrea, et al.
Publicado: (2024)
por: Apicella, Andrea, et al.
Publicado: (2024)
Multi-Armed Bandits With Best-Action Queries
por: Bacchiocchi, Francesco, et al.
Publicado: (2026)
por: Bacchiocchi, Francesco, et al.
Publicado: (2026)
Fully Dynamic Rebalancing in Dockless Bike-Sharing Systems via Deep Reinforcement Learning
por: Scarpel, Edoardo, et al.
Publicado: (2026)
por: Scarpel, Edoardo, et al.
Publicado: (2026)
Anomalous Agreement: How to find the Ideal Number of Anomaly Classes in Correlated, Multivariate Time Series Data
por: Rewicki, Ferdinand, et al.
Publicado: (2025)
por: Rewicki, Ferdinand, et al.
Publicado: (2025)
Agreement-Constrained Probabilistic Minimum Bayes Risk Decoding
por: Natsumi, Koki, et al.
Publicado: (2025)
por: Natsumi, Koki, et al.
Publicado: (2025)
Pretrain Value, Not Reward: Decoupled Value Policy Optimization
por: Huang, Chenghua, et al.
Publicado: (2025)
por: Huang, Chenghua, et al.
Publicado: (2025)
Instance-Adaptive Parametrization for Amortized Variational Inference
por: Pollastro, Andrea, et al.
Publicado: (2026)
por: Pollastro, Andrea, et al.
Publicado: (2026)
Don't stop me now: Rethinking Validation Criteria for Model Parameter Selection
por: Apicella, Andrea, et al.
Publicado: (2026)
por: Apicella, Andrea, et al.
Publicado: (2026)
Value Flows
por: Dong, Perry, et al.
Publicado: (2025)
por: Dong, Perry, et al.
Publicado: (2025)
Multi-Value Alignment for LLMs via Value Decorrelation and Extrapolation
por: Xu, Hefei, et al.
Publicado: (2025)
por: Xu, Hefei, et al.
Publicado: (2025)
Interpretable Discriminative Text Representations via Agreement and Label Disentanglement
por: Wang, Tong, et al.
Publicado: (2026)
por: Wang, Tong, et al.
Publicado: (2026)
An Analysis of Action-Value Temporal-Difference Methods That Learn State Values
por: Daley, Brett, et al.
Publicado: (2025)
por: Daley, Brett, et al.
Publicado: (2025)
Regression-adjusted Monte Carlo Estimators for Shapley Values and Probabilistic Values
por: Witter, R. Teal, et al.
Publicado: (2025)
por: Witter, R. Teal, et al.
Publicado: (2025)
Is there Value in Reinforcement Learning?
por: Fox, Lior, et al.
Publicado: (2025)
por: Fox, Lior, et al.
Publicado: (2025)
Value Imprint: A Technique for Auditing the Human Values Embedded in RLHF Datasets
por: Obi, Ike, et al.
Publicado: (2024)
por: Obi, Ike, et al.
Publicado: (2024)
VC-Soup: Value-Consistency Guided Multi-Value Alignment for Large Language Models
por: Xu, Hefei, et al.
Publicado: (2026)
por: Xu, Hefei, et al.
Publicado: (2026)
Toward Optimal Regret in Robust Pricing: Decoupling Corruption and Time
por: Kalupahana, Kalana, et al.
Publicado: (2026)
por: Kalupahana, Kalana, et al.
Publicado: (2026)
IMPACTX: improving model performance by appropriately constraining the training with teacher explanations
por: Apicella, Andrea, et al.
Publicado: (2025)
por: Apicella, Andrea, et al.
Publicado: (2025)
Universal Value-Function Uncertainties
por: Zanger, Moritz A., et al.
Publicado: (2025)
por: Zanger, Moritz A., et al.
Publicado: (2025)
Suboptimal Shapley Value Explanations
por: Lu, Xiaolei
Publicado: (2025)
por: Lu, Xiaolei
Publicado: (2025)
An Odd Estimator for Shapley Values
por: Fumagalli, Fabian, et al.
Publicado: (2026)
por: Fumagalli, Fabian, et al.
Publicado: (2026)
On the Inflation of KNN-Shapley Value
por: Yang, Ziao, et al.
Publicado: (2024)
por: Yang, Ziao, et al.
Publicado: (2024)
Value Improved Actor Critic Algorithms
por: Oren, Yaniv, et al.
Publicado: (2024)
por: Oren, Yaniv, et al.
Publicado: (2024)
Quasimetric Value Functions with Dense Rewards
por: Valieva, Khadichabonu, et al.
Publicado: (2024)
por: Valieva, Khadichabonu, et al.
Publicado: (2024)
Exactly Computing do-Shapley Values
por: Witter, R. Teal, et al.
Publicado: (2026)
por: Witter, R. Teal, et al.
Publicado: (2026)
Explaining Drift using Shapley Values
por: Edakunni, Narayanan U., et al.
Publicado: (2024)
por: Edakunni, Narayanan U., et al.
Publicado: (2024)
Retrosynthetic Planning with Dual Value Networks
por: Liu, Guoqing, et al.
Publicado: (2023)
por: Liu, Guoqing, et al.
Publicado: (2023)
Generalized Priority-Aware Shapley Value
por: Lee, Kiljae, et al.
Publicado: (2026)
por: Lee, Kiljae, et al.
Publicado: (2026)
No-Regret Learning Under Adversarial Resource Constraints: A Spending Plan Is All You Need!
por: Stradi, Francesco Emanuele, et al.
Publicado: (2025)
por: Stradi, Francesco Emanuele, et al.
Publicado: (2025)
Inferring Transition Dynamics from Value Functions
por: Adamczyk, Jacob
Publicado: (2025)
por: Adamczyk, Jacob
Publicado: (2025)
Ejemplares similares
-
Model Agreement via Anchoring
por: Eaton, Eric, et al.
Publicado: (2026) -
Modeling Teams Performance Using Deep Representational Learning on Graphs
por: Carli, Francesco, et al.
Publicado: (2022) -
Highway Value Iteration Networks
por: Wang, Yuhui, et al.
Publicado: (2024) -
The Price of Agreement: Measuring LLM Sycophancy in Agentic Financial Applications
por: Zhao, Zhenyu, et al.
Publicado: (2026) -
Test-Time Adaptation Induces Stronger Accuracy and Agreement-on-the-Line
por: Kim, Eungyeup, et al.
Publicado: (2023)