Granular feedback merits sophisticated aggregation
Fuente:
arXiv
Saved in:
| Main Authors: | Kagrecha, Anmol, Marklund, Henrik, Manakul, Potsawee, Zeckhauser, Richard, Van Roy, Benjamin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Crowdsourcing Via Self-Supervised Learning
by: Kagrecha, Anmol, et al.
Published: (2024)
by: Kagrecha, Anmol, et al.
Published: (2024)
SkillAggregation: Reference-free LLM-Dependent Aggregation
by: Sun, Guangzhi, et al.
Published: (2024)
by: Sun, Guangzhi, et al.
Published: (2024)
Choice Between Partial Trajectories: Disentangling Goals from Beliefs
by: Marklund, Henrik, et al.
Published: (2024)
by: Marklund, Henrik, et al.
Published: (2024)
Misalignment from Treating Means as Ends
by: Marklund, Henrik, et al.
Published: (2025)
by: Marklund, Henrik, et al.
Published: (2025)
Consequentialist Objectives and Catastrophe
by: Marklund, Henrik, et al.
Published: (2026)
by: Marklund, Henrik, et al.
Published: (2026)
Maintaining Plasticity in Continual Learning via Regenerative Regularization
by: Kumar, Saurabh, et al.
Published: (2023)
by: Kumar, Saurabh, et al.
Published: (2023)
Unlearning vs. Obfuscation: Are We Truly Removing Knowledge?
by: Sun, Guangzhi, et al.
Published: (2025)
by: Sun, Guangzhi, et al.
Published: (2025)
Continual Learning as Computationally Constrained Reinforcement Learning
by: Kumar, Saurabh, et al.
Published: (2023)
by: Kumar, Saurabh, et al.
Published: (2023)
Test-Time Alignment via Hypothesis Reweighting
by: Lee, Yoonho, et al.
Published: (2024)
by: Lee, Yoonho, et al.
Published: (2024)
Aggregation Strategies for Efficient Annotation of Bioacoustic Sound Events Using Active Learning
by: Lindholm, Richard, et al.
Published: (2025)
by: Lindholm, Richard, et al.
Published: (2025)
Risk-averse learning with delayed feedback
by: Wang, Siyi, et al.
Published: (2024)
by: Wang, Siyi, et al.
Published: (2024)
Prior Prompt Engineering for Reinforcement Fine-Tuning
by: Taveekitworachai, Pittawat, et al.
Published: (2025)
by: Taveekitworachai, Pittawat, et al.
Published: (2025)
Adapting Language-Specific LLMs to a Reasoning Model in One Day via Model Merging -- An Open Recipe
by: Pipatanakul, Kunat, et al.
Published: (2025)
by: Pipatanakul, Kunat, et al.
Published: (2025)
Typhoon T1: An Open Thai Reasoning Model
by: Taveekitworachai, Pittawat, et al.
Published: (2025)
by: Taveekitworachai, Pittawat, et al.
Published: (2025)
LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models
by: Liusie, Adian, et al.
Published: (2023)
by: Liusie, Adian, et al.
Published: (2023)
Non-Stationary Bandit Learning via Predictive Sampling
by: Liu, Yueyang, et al.
Published: (2022)
by: Liu, Yueyang, et al.
Published: (2022)
Posterior Sampling for Continuing Environments
by: Xu, Wanqiao, et al.
Published: (2022)
by: Xu, Wanqiao, et al.
Published: (2022)
Generation of Granular-Balls for Clustering Based on the Principle of Justifiable Granularity
by: Jia, Zihang, et al.
Published: (2024)
by: Jia, Zihang, et al.
Published: (2024)
Information-Theoretic Foundations for Neural Scaling Laws
by: Jeon, Hong Jun, et al.
Published: (2024)
by: Jeon, Hong Jun, et al.
Published: (2024)
Information-Theoretic Foundations for Machine Learning
by: Jeon, Hong Jun, et al.
Published: (2024)
by: Jeon, Hong Jun, et al.
Published: (2024)
Aligning AI Agents via Information-Directed Sampling
by: Jeon, Hong Jun, et al.
Published: (2024)
by: Jeon, Hong Jun, et al.
Published: (2024)
Prior Diffusiveness and Regret in the Linear-Gaussian Bandit
by: Zhu, Yifan, et al.
Published: (2026)
by: Zhu, Yifan, et al.
Published: (2026)
Learning from negative feedback, or positive feedback or both
by: Abdolmaleki, Abbas, et al.
Published: (2024)
by: Abdolmaleki, Abbas, et al.
Published: (2024)
Exploration Unbound
by: Arumugam, Dilip, et al.
Published: (2024)
by: Arumugam, Dilip, et al.
Published: (2024)
Attention Isn't All You Need for Emotion Recognition:Domain Features Outperform Transformers on the EAV Dataset
by: Guragain, Anmol
Published: (2026)
by: Guragain, Anmol
Published: (2026)
Real-Time Risky Fault-Chain Search using Time-Varying Graph RNNs
by: Dwivedi, Anmol, et al.
Published: (2025)
by: Dwivedi, Anmol, et al.
Published: (2025)
Auditing Language Model Unlearning via Information Decomposition
by: Goel, Anmol, et al.
Published: (2026)
by: Goel, Anmol, et al.
Published: (2026)
Extending Audio Context for Long-Form Understanding in Large Audio-Language Models
by: Chaichana, Yuatyong, et al.
Published: (2025)
by: Chaichana, Yuatyong, et al.
Published: (2025)
Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models
by: Manakul, Potsawee, et al.
Published: (2024)
by: Manakul, Potsawee, et al.
Published: (2024)
Operationalization of Machine Learning with Serverless Architecture: An Industrial Operationalization of Machine Learning with Serverless Architecture: An Industrial Implementation for Harmonized System Code Prediction
by: Kandappareddigari, Sai Vineeth, et al.
Published: (2026)
by: Kandappareddigari, Sai Vineeth, et al.
Published: (2026)
The Limitations of Model Retraining in the Face of Performativity
by: Kabra, Anmol, et al.
Published: (2024)
by: Kabra, Anmol, et al.
Published: (2024)
Granular Ball Twin Support Vector Machine
by: Quadir, A., et al.
Published: (2024)
by: Quadir, A., et al.
Published: (2024)
Multi-view Granular-ball Contrastive Clustering
by: Su, Peng, et al.
Published: (2024)
by: Su, Peng, et al.
Published: (2024)
Active clustering with bandit feedback
by: Thuot, Victor, et al.
Published: (2024)
by: Thuot, Victor, et al.
Published: (2024)
Satisficing Exploration for Deep Reinforcement Learning
by: Arumugam, Dilip, et al.
Published: (2024)
by: Arumugam, Dilip, et al.
Published: (2024)
Mixed-signal implementation of feedback-control optimizer for single-layer Spiking Neural Networks
by: Haag, Jonathan, et al.
Published: (2026)
by: Haag, Jonathan, et al.
Published: (2026)
Mix-of-Granularity: Optimize the Chunking Granularity for Retrieval-Augmented Generation
by: Zhong, Zijie, et al.
Published: (2024)
by: Zhong, Zijie, et al.
Published: (2024)
LGBQPC: Local Granular-Ball Quality Peaks Clustering
by: Jia, Zihang, et al.
Published: (2025)
by: Jia, Zihang, et al.
Published: (2025)
Adaptive Inverted-Index Routing for Granular Mixtures-of-Experts
by: Kladny, Klaus-Rudolf, et al.
Published: (2026)
by: Kladny, Klaus-Rudolf, et al.
Published: (2026)
Conformal online model aggregation
by: Gasparin, Matteo, et al.
Published: (2024)
by: Gasparin, Matteo, et al.
Published: (2024)
Similar Items
-
Adaptive Crowdsourcing Via Self-Supervised Learning
by: Kagrecha, Anmol, et al.
Published: (2024) -
SkillAggregation: Reference-free LLM-Dependent Aggregation
by: Sun, Guangzhi, et al.
Published: (2024) -
Choice Between Partial Trajectories: Disentangling Goals from Beliefs
by: Marklund, Henrik, et al.
Published: (2024) -
Misalignment from Treating Means as Ends
by: Marklund, Henrik, et al.
Published: (2025) -
Consequentialist Objectives and Catastrophe
by: Marklund, Henrik, et al.
Published: (2026)