AsymmetryZero: A Framework for Operationalizing Human Expert Preferences as Semantic Evals
Fuente:
arXiv
Saved in:
| Main Authors: | Looram, Tadhg, Nuzzi, Lucas, Waters, Kyle, Dillmann, Steven |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
COMPOSITE-Stem
by: Waters, Kyle, et al.
Published: (2026)
by: Waters, Kyle, et al.
Published: (2026)
AI Biases as Asymmetries: A Review to Guide Practice
by: Waters, Gabriella, et al.
Published: (2025)
by: Waters, Gabriella, et al.
Published: (2025)
Operationalizing Data Minimization for Privacy-Preserving LLM Prompting
by: Zhou, Jijie, et al.
Published: (2025)
by: Zhou, Jijie, et al.
Published: (2025)
MoE++: Accelerating Mixture-of-Experts Methods with Zero-Computation Experts
by: Jin, Peng, et al.
Published: (2024)
by: Jin, Peng, et al.
Published: (2024)
Soft Preference Optimization: Aligning Language Models to Expert Distributions
by: Sharifnassab, Arsalan, et al.
Published: (2024)
by: Sharifnassab, Arsalan, et al.
Published: (2024)
Towards Operationalizing Right to Data Protection
by: Java, Abhinav, et al.
Published: (2024)
by: Java, Abhinav, et al.
Published: (2024)
How Many Experts Are Enough? Towards Optimal Semantic Specialization for Mixture-of-Experts
by: Park, Sumin, et al.
Published: (2025)
by: Park, Sumin, et al.
Published: (2025)
OrdMoE: Preference Alignment via Hierarchical Expert Group Ranking in Multimodal Mixture-of-Experts LLMs
by: Gao, Yuting, et al.
Published: (2025)
by: Gao, Yuting, et al.
Published: (2025)
Operationalizing Fairness: Post-Hoc Threshold Optimization Under Hard Resource Limits
by: Singh, Moirangthem Tiken, et al.
Published: (2026)
by: Singh, Moirangthem Tiken, et al.
Published: (2026)
HumanEval on Latest GPT Models -- 2024
by: Li, Daniel, et al.
Published: (2024)
by: Li, Daniel, et al.
Published: (2024)
Zero-shot Generalizable Graph Anomaly Detection with Mixture of Riemannian Experts
by: Zhao, Xinyu, et al.
Published: (2026)
by: Zhao, Xinyu, et al.
Published: (2026)
Operationalizing Longitudinal Causal Discovery Under Real-World Workflow Constraints
by: Okuda, Tadahisa, et al.
Published: (2026)
by: Okuda, Tadahisa, et al.
Published: (2026)
ExpProof : Operationalizing Explanations for Confidential Models with ZKPs
by: Yadav, Chhavi, et al.
Published: (2025)
by: Yadav, Chhavi, et al.
Published: (2025)
RFBES at SemEval-2024 Task 8: Investigating Syntactic and Semantic Features for Distinguishing AI-Generated and Human-Written Texts
by: Rad, Mohammad Heydari, et al.
Published: (2024)
by: Rad, Mohammad Heydari, et al.
Published: (2024)
Capturing Individual Human Preferences with Reward Features
by: Barreto, André, et al.
Published: (2025)
by: Barreto, André, et al.
Published: (2025)
Bayesian Inference for Correlated Human Experts and Classifiers
by: Kelly, Markelle, et al.
Published: (2025)
by: Kelly, Markelle, et al.
Published: (2025)
Demystifying the Mythos or Disrupting Bugonomics? From Zero-Day Asymmetry to Defender Remediation Throughput
by: Pesoli, Alfredo, et al.
Published: (2026)
by: Pesoli, Alfredo, et al.
Published: (2026)
Learning Representations of Event Time Series with Sparse Autoencoders for Anomaly Detection, Similarity Search, and Unsupervised Classification
by: Dillmann, Steven, et al.
Published: (2025)
by: Dillmann, Steven, et al.
Published: (2025)
Semantic-Inductive Attribute Selection for Zero-Shot Learning
by: Herrera-Aranda, Juan Jose, et al.
Published: (2025)
by: Herrera-Aranda, Juan Jose, et al.
Published: (2025)
Trust, Don't Trust, or Flip: Robust Preference-Based Reinforcement Learning with Multi-Expert Feedback
by: Hosseini, Seyed Amir, et al.
Published: (2026)
by: Hosseini, Seyed Amir, et al.
Published: (2026)
Adaptive Preference Scaling for Reinforcement Learning with Human Feedback
by: Hong, Ilgee, et al.
Published: (2024)
by: Hong, Ilgee, et al.
Published: (2024)
LEAD: Minimizing Learner-Expert Asymmetry in End-to-End Driving
by: Nguyen, Long, et al.
Published: (2025)
by: Nguyen, Long, et al.
Published: (2025)
Towards Stable Preferences for Stakeholder-aligned Machine Learning
by: Sheraz, Haleema, et al.
Published: (2024)
by: Sheraz, Haleema, et al.
Published: (2024)
Operationalizing Document AI: A Microservice Architecture for OCR and LLM Pipelines in Production
by: Fehlis, Yao, et al.
Published: (2026)
by: Fehlis, Yao, et al.
Published: (2026)
Taxon: Hierarchical Tax Code Prediction with Semantically Aligned LLM Expert Guidance
by: Li, Jihang, et al.
Published: (2026)
by: Li, Jihang, et al.
Published: (2026)
From Demonstrations to Rewards: Alignment Without Explicit Human Preferences
by: Zeng, Siliang, et al.
Published: (2025)
by: Zeng, Siliang, et al.
Published: (2025)
Neural Dueling Bandits: Preference-Based Optimization with Human Feedback
by: Verma, Arun, et al.
Published: (2024)
by: Verma, Arun, et al.
Published: (2024)
LyS at SemEval 2025 Task 8: Zero-Shot Code Generation for Tabular QA
by: Gude, Adrián, et al.
Published: (2025)
by: Gude, Adrián, et al.
Published: (2025)
Multimodal Alignment and Preference Optimization for Zero-Shot Conditional RNA Generation
by: Klypa, Roman, et al.
Published: (2026)
by: Klypa, Roman, et al.
Published: (2026)
MLPMoE: Zero-Shot Architectural Metamorphosis of Dense LLM MLPs into Static Mixture-of-Experts
by: Novikov, Ivan
Published: (2025)
by: Novikov, Ivan
Published: (2025)
Cer-Eval: Certifiable and Cost-Efficient Evaluation Framework for LLMs
by: Wang, Ganghua, et al.
Published: (2025)
by: Wang, Ganghua, et al.
Published: (2025)
Predictive Preference Learning from Human Interventions
by: Cai, Haoyuan, et al.
Published: (2025)
by: Cai, Haoyuan, et al.
Published: (2025)
A Unifying Framework for Learning Argumentation Semantics
by: Mileva, Zlatina, et al.
Published: (2023)
by: Mileva, Zlatina, et al.
Published: (2023)
AlphaEval: A Comprehensive and Efficient Evaluation Framework for Formula Alpha Mining
by: Ding, Hongjun, et al.
Published: (2025)
by: Ding, Hongjun, et al.
Published: (2025)
GraphEval: A Knowledge-Graph Based LLM Hallucination Evaluation Framework
by: Sansford, Hannah, et al.
Published: (2024)
by: Sansford, Hannah, et al.
Published: (2024)
Contrastive Preference Learning: Learning from Human Feedback without RL
by: Hejna, Joey, et al.
Published: (2023)
by: Hejna, Joey, et al.
Published: (2023)
Directly Aligning the Full Diffusion Trajectory with Fine-Grained Human Preference
by: Shen, Xiangwei, et al.
Published: (2025)
by: Shen, Xiangwei, et al.
Published: (2025)
Human Alignment of Large Language Models through Online Preference Optimisation
by: Calandriello, Daniele, et al.
Published: (2024)
by: Calandriello, Daniele, et al.
Published: (2024)
COPR: Continual Human Preference Learning via Optimal Policy Regularization
by: Zhang, Han, et al.
Published: (2024)
by: Zhang, Han, et al.
Published: (2024)
UniRL-Zero: Reinforcement Learning on Unified Models with Joint Language Model and Diffusion Model Experts
by: Wang, Fu-Yun, et al.
Published: (2025)
by: Wang, Fu-Yun, et al.
Published: (2025)
Similar Items
-
COMPOSITE-Stem
by: Waters, Kyle, et al.
Published: (2026) -
AI Biases as Asymmetries: A Review to Guide Practice
by: Waters, Gabriella, et al.
Published: (2025) -
Operationalizing Data Minimization for Privacy-Preserving LLM Prompting
by: Zhou, Jijie, et al.
Published: (2025) -
MoE++: Accelerating Mixture-of-Experts Methods with Zero-Computation Experts
by: Jin, Peng, et al.
Published: (2024) -
Soft Preference Optimization: Aligning Language Models to Expert Distributions
by: Sharifnassab, Arsalan, et al.
Published: (2024)