GrandJury: A Collaborative Machine Learning Model Evaluation Protocol for Dynamic Quality Rubrics
Fuente:
arXiv
Saved in:
| Main Author: | Cho, Arthur |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Humanizing AI Grading: Student-Centered Insights on Fairness, Trust, Consistency and Transparency
by: Riahi, Bahare, et al.
Published: (2026)
by: Riahi, Bahare, et al.
Published: (2026)
Improving Knowledge Extraction from LLMs for Task Learning through Agent Analysis
by: Kirk, James R., et al.
Published: (2023)
by: Kirk, James R., et al.
Published: (2023)
Scalable Interactive Machine Learning for Future Command and Control
by: Madison, Anna, et al.
Published: (2024)
by: Madison, Anna, et al.
Published: (2024)
Relational Intervention During Functional Collapse in Large Language Models: A Lexical-Statistical Ablation and a Structure x Register Factorial
by: Santana, Franco, et al.
Published: (2026)
by: Santana, Franco, et al.
Published: (2026)
Co-Writing with AI, on Human Terms: Aligning Research with User Demands Across the Writing Process
by: Reza, Mohi, et al.
Published: (2025)
by: Reza, Mohi, et al.
Published: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
by: Fadli, Samih
Published: (2025)
by: Fadli, Samih
Published: (2025)
Emotion-Attended Stateful Memory (EASM):The Architecture for Hyper-Personalization at Scale
by: Kotecha, Vineet, et al.
Published: (2026)
by: Kotecha, Vineet, et al.
Published: (2026)
Reflective Verbal Reward Design for Pluralistic Alignment
by: Blair, Carter, et al.
Published: (2025)
by: Blair, Carter, et al.
Published: (2025)
COA-GPT: Generative Pre-trained Transformers for Accelerated Course of Action Development in Military Operations
by: Goecks, Vinicius G., et al.
Published: (2024)
by: Goecks, Vinicius G., et al.
Published: (2024)
Re-Envisioning Command and Control
by: McDowell, Kaleb, et al.
Published: (2024)
by: McDowell, Kaleb, et al.
Published: (2024)
Alternating Reinforcement Learning with Contextual Rubric Rewards: Beyond the Scalarization Strategy
by: Lan, Guangchen, et al.
Published: (2026)
by: Lan, Guangchen, et al.
Published: (2026)
Shattered Compositionality: Counterintuitive Learning Dynamics of Transformers for Arithmetic
by: Zhao, Xingyu, et al.
Published: (2026)
by: Zhao, Xingyu, et al.
Published: (2026)
CoE: Collaborative Entropy for Uncertainty Quantification in Agentic Multi-LLM Systems
by: Sun, Kangkang, et al.
Published: (2026)
by: Sun, Kangkang, et al.
Published: (2026)
OFMU: Optimization-Driven Framework for Machine Unlearning
by: Asif, Sadia, et al.
Published: (2025)
by: Asif, Sadia, et al.
Published: (2025)
Conversation Tree Architecture: A Structured Framework for Context-Aware Multi-Branch LLM Conversations
by: Hemanth, Pranav, et al.
Published: (2026)
by: Hemanth, Pranav, et al.
Published: (2026)
Automated CAD Modeling Sequence Generation from Text Descriptions via Transformer-Based Large Language Models
by: Liao, Jianxing, et al.
Published: (2025)
by: Liao, Jianxing, et al.
Published: (2025)
Towards More Human-like AI Communication: A Review of Emergent Communication Research
by: Brandizzi, Nicolo'
Published: (2023)
by: Brandizzi, Nicolo'
Published: (2023)
FastForward Pruning: Efficient LLM Pruning via Single-Step Reinforcement Learning
by: Yuan, Xin, et al.
Published: (2025)
by: Yuan, Xin, et al.
Published: (2025)
FastGRPO: Accelerating Policy Optimization via Concurrency-aware Speculative Decoding and Online Draft Learning
by: Zhang, Yizhou, et al.
Published: (2025)
by: Zhang, Yizhou, et al.
Published: (2025)
BitCal-TTS: Bit-Calibrated Test-Time Scaling for Quantized Reasoning Models
by: Patarlapalli, Sai Babu, et al.
Published: (2026)
by: Patarlapalli, Sai Babu, et al.
Published: (2026)
The Geometry of Thought: How Scale Restructures Reasoning In Large Language Models
by: Anderson, Samuel Cyrenius
Published: (2026)
by: Anderson, Samuel Cyrenius
Published: (2026)
On Semantic Loss Fine-Tuning Approach for Preventing Model Collapse in Causal Reasoning
by: Deshmukh, Pratik, et al.
Published: (2026)
by: Deshmukh, Pratik, et al.
Published: (2026)
Quantization Undoes Alignment: Bias Emergence in Compressed LLMs Across Models and Precision Levels
by: Rath, Plawan Kumar, et al.
Published: (2026)
by: Rath, Plawan Kumar, et al.
Published: (2026)
Reasoning Large Language Model Errors Arise from Hallucinating Critical Problem Features
by: Heyman, Alex, et al.
Published: (2025)
by: Heyman, Alex, et al.
Published: (2025)
The Anti-Ouroboros Effect: Emergent Resilience in Large Language Models from Recursive Selective Feedback
by: Adapala, Sai Teja Reddy
Published: (2025)
by: Adapala, Sai Teja Reddy
Published: (2025)
Neurocognitive Modeling for Text Generation: Deep Learning Architecture for EEG Data
by: Khushiyant
Published: (2025)
by: Khushiyant
Published: (2025)
Control Reinforcement Learning: Interpretable Token-Level Steering of LLMs via Sparse Autoencoder Features
by: Cho, Seonglae, et al.
Published: (2026)
by: Cho, Seonglae, et al.
Published: (2026)
No Free Swap: Protocol-Dependent Layer Redundancy in Transformers
by: Garcia, Gabriel
Published: (2026)
by: Garcia, Gabriel
Published: (2026)
JURY-RL: Votes Propose, Proofs Dispose for Label-Free RLVR
by: Chen, Xinjie, et al.
Published: (2026)
by: Chen, Xinjie, et al.
Published: (2026)
OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind
by: Srishty, Sharmin Sultana, et al.
Published: (2026)
by: Srishty, Sharmin Sultana, et al.
Published: (2026)
Planning vs Reasoning: Ablations to Test Capabilities of LoRA layers
by: Redkar, Neel
Published: (2024)
by: Redkar, Neel
Published: (2024)
Curveball Steering: The Right Direction To Steer Isn't Always Linear
by: Raval, Shivam, et al.
Published: (2026)
by: Raval, Shivam, et al.
Published: (2026)
Generalizing Numerical Reasoning in Table Data through Operation Sketches and Self-Supervised Learning
by: Cho, Hanjun, et al.
Published: (2026)
by: Cho, Hanjun, et al.
Published: (2026)
Social Cooperation in Conversational AI Agents
by: Çelikok, Mustafa Mert, et al.
Published: (2025)
by: Çelikok, Mustafa Mert, et al.
Published: (2025)
How Does Unfaithful Reasoning Emerge from Autoregressive Training? A Study of Synthetic Experiments
by: Wang, Fuxin, et al.
Published: (2026)
by: Wang, Fuxin, et al.
Published: (2026)
Extreme AutoML: Analysis of Classification, Regression, and NLP Performance
by: Ratner, Edward, et al.
Published: (2024)
by: Ratner, Edward, et al.
Published: (2024)
CircuitProbe: Predicting Reasoning Circuits in Transformers via Stability Zone Detection
by: Panuganti, Rajkiran
Published: (2026)
by: Panuganti, Rajkiran
Published: (2026)
PRPO: Aligning Process Reward with Outcome Reward in Policy Optimization
by: Ding, Ruiyi, et al.
Published: (2026)
by: Ding, Ruiyi, et al.
Published: (2026)
BLOCK-EM: Preventing Emergent Misalignment via Latent Blocking
by: Ustaomeroglu, Muhammed, et al.
Published: (2026)
by: Ustaomeroglu, Muhammed, et al.
Published: (2026)
Persona Features Control Emergent Misalignment
by: Wang, Miles, et al.
Published: (2025)
by: Wang, Miles, et al.
Published: (2025)
Similar Items
-
Humanizing AI Grading: Student-Centered Insights on Fairness, Trust, Consistency and Transparency
by: Riahi, Bahare, et al.
Published: (2026) -
Improving Knowledge Extraction from LLMs for Task Learning through Agent Analysis
by: Kirk, James R., et al.
Published: (2023) -
Scalable Interactive Machine Learning for Future Command and Control
by: Madison, Anna, et al.
Published: (2024) -
Relational Intervention During Functional Collapse in Large Language Models: A Lexical-Statistical Ablation and a Structure x Register Factorial
by: Santana, Franco, et al.
Published: (2026) -
Co-Writing with AI, on Human Terms: Aligning Research with User Demands Across the Writing Process
by: Reza, Mohi, et al.
Published: (2025)