Saved in:
| Main Authors: | Kumar, Bhavesh, Feng, Dylan, Tang, Leonard |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2603.07990 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CurvaDion: Curvature-Adaptive Distributed Orthonormalization
by: Kumar, Bhavesh, et al.
Published: (2025)
by: Kumar, Bhavesh, et al.
Published: (2025)
Verification of Visual Controllers via Compositional Geometric Transformations
by: Estornell, Alexander, et al.
Published: (2025)
by: Estornell, Alexander, et al.
Published: (2025)
Synthpop++: A Hybrid Framework for Generating A Country-scale Synthetic Population
by: Neekhra, Bhavesh, et al.
Published: (2023)
by: Neekhra, Bhavesh, et al.
Published: (2023)
On the (In)Significance of Feature Selection in High-Dimensional Datasets
by: Neekhra, Bhavesh, et al.
Published: (2025)
by: Neekhra, Bhavesh, et al.
Published: (2025)
Optimization for Neural Operators can Benefit from Width
by: Cisneros-Velarde, Pedro, et al.
Published: (2025)
by: Cisneros-Velarde, Pedro, et al.
Published: (2025)
Hybrid CNN with Chebyshev Polynomial Expansion for Medical Image Analysis
by: Roy, Abhinav, et al.
Published: (2025)
by: Roy, Abhinav, et al.
Published: (2025)
MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?
by: Chen, Zhaorun, et al.
Published: (2024)
by: Chen, Zhaorun, et al.
Published: (2024)
TRACE: Grounding Time Series in Context for Multimodal Embedding and Retrieval
by: Chen, Jialin, et al.
Published: (2025)
by: Chen, Jialin, et al.
Published: (2025)
Curiosity-Driven LLM-as-a-judge for Personalized Creative Judgment
by: Kumar, Vanya Bannihatti, et al.
Published: (2025)
by: Kumar, Vanya Bannihatti, et al.
Published: (2025)
Advancing Parkinson's Disease Progression Prediction: Comparing Long Short-Term Memory Networks and Kolmogorov-Arnold Networks
by: Roy, Abhinav, et al.
Published: (2024)
by: Roy, Abhinav, et al.
Published: (2024)
ProKAN: Progressive Stacking of Kolmogorov-Arnold Networks for Efficient Liver Segmentation
by: Gyanchandani, Bhavesh, et al.
Published: (2024)
by: Gyanchandani, Bhavesh, et al.
Published: (2024)
Ask Again, Then Fail: Large Language Models' Vacillations in Judgment
by: Xie, Qiming, et al.
Published: (2023)
by: Xie, Qiming, et al.
Published: (2023)
Grounding Multimodal Large Language Models in Actions
by: Szot, Andrew, et al.
Published: (2024)
by: Szot, Andrew, et al.
Published: (2024)
When Are Multimodal Predictions Biologically Supported? A Diagnostic Evaluation Framework
by: Steiner, Dylan, et al.
Published: (2026)
by: Steiner, Dylan, et al.
Published: (2026)
Multi-Stage Prototype Learning for Interpretable Time Series Classification
by: Kalisetti, Bhavesh, et al.
Published: (2021)
by: Kalisetti, Bhavesh, et al.
Published: (2021)
Benchmarking LLMs' Judgments with No Gold Standard
by: Xu, Shengwei, et al.
Published: (2024)
by: Xu, Shengwei, et al.
Published: (2024)
FactReview: Evidence-Grounded Peer Review with Execution-Based Claim Verification
by: Yue, Ling, et al.
Published: (2026)
by: Yue, Ling, et al.
Published: (2026)
TriSpec: Ternary Speculative Decoding via Lightweight Proxy Verification
by: Jiang, Haoyun, et al.
Published: (2026)
by: Jiang, Haoyun, et al.
Published: (2026)
Execution-Grounded Credit Assignment for GRPO in Code Generation
by: Kumar, Abhijit, et al.
Published: (2026)
by: Kumar, Abhijit, et al.
Published: (2026)
Multimodal Reference Visual Grounding
by: Lu, Yangxiao, et al.
Published: (2025)
by: Lu, Yangxiao, et al.
Published: (2025)
Rating Quality of Diverse Time Series Data by Meta-learning from LLM Judgment
by: Wu, Shunyu, et al.
Published: (2025)
by: Wu, Shunyu, et al.
Published: (2025)
QEDCartographer: Automating Formal Verification Using Reward-Free Reinforcement Learning
by: Sanchez-Stern, Alex, et al.
Published: (2024)
by: Sanchez-Stern, Alex, et al.
Published: (2024)
Perception-Driven Bias Detection in Machine Learning via Crowdsourced Visual Judgment
by: Tupakula, Chirudeep, et al.
Published: (2025)
by: Tupakula, Chirudeep, et al.
Published: (2025)
When to Accept Automated Predictions and When to Defer to Human Judgment?
by: Sikar, Daniel, et al.
Published: (2024)
by: Sikar, Daniel, et al.
Published: (2024)
Efficient MAP Estimation of LLM Judgment Performance with Prior Transfer
by: Qu, Huaizhi, et al.
Published: (2025)
by: Qu, Huaizhi, et al.
Published: (2025)
Grounding Multilingual Multimodal LLMs With Cultural Knowledge
by: Nyandwi, Jean de Dieu, et al.
Published: (2025)
by: Nyandwi, Jean de Dieu, et al.
Published: (2025)
Auto-GDA: Automatic Domain Adaptation for Efficient Grounding Verification in Retrieval-Augmented Generation
by: Leemann, Tobias, et al.
Published: (2024)
by: Leemann, Tobias, et al.
Published: (2024)
Data-Driven Graph Filters via Adaptive Spectral Shaping
by: Sandfelder, Dylan, et al.
Published: (2026)
by: Sandfelder, Dylan, et al.
Published: (2026)
Tree-of-Evidence: Efficient "System 2" Search for Faithful Multimodal Grounding
by: Nnamdi, Micky C., et al.
Published: (2026)
by: Nnamdi, Micky C., et al.
Published: (2026)
Adjusted Overfitting Regression
by: Wilson, Dylan
Published: (2024)
by: Wilson, Dylan
Published: (2024)
Legal Judgment Reimagined: PredEx and the Rise of Intelligent AI Interpretation in Indian Courts
by: Nigam, Shubham Kumar, et al.
Published: (2024)
by: Nigam, Shubham Kumar, et al.
Published: (2024)
Internalizing Curriculum Judgment for LLM Reinforcement Fine-Tuning
by: Zheng, Han, et al.
Published: (2026)
by: Zheng, Han, et al.
Published: (2026)
Statistical Runtime Verification for LLMs via Robustness Estimation
by: Levy, Natan, et al.
Published: (2025)
by: Levy, Natan, et al.
Published: (2025)
Degraded Polygons Raise Fundamental Questions of Neural Network Perception
by: Tang, Leonard, et al.
Published: (2023)
by: Tang, Leonard, et al.
Published: (2023)
ReGal: A First Look at PPO-based Legal AI for Judgment Prediction and Summarization in India
by: Nigam, Shubham Kumar, et al.
Published: (2025)
by: Nigam, Shubham Kumar, et al.
Published: (2025)
An Iterative Utility Judgment Framework Inspired by Philosophical Relevance via LLMs
by: Zhang, Hengran, et al.
Published: (2024)
by: Zhang, Hengran, et al.
Published: (2024)
Spectral Clustering of Categorical and Mixed-type Data via Extra Graph Nodes
by: Soemitro, Dylan, et al.
Published: (2024)
by: Soemitro, Dylan, et al.
Published: (2024)
Verde: Verification via Refereed Delegation for Machine Learning Programs
by: Arun, Arasu, et al.
Published: (2025)
by: Arun, Arasu, et al.
Published: (2025)
Verification of Unknown Dynamical Systems via Autoencoder Latent Space
by: Reed, Robert, et al.
Published: (2025)
by: Reed, Robert, et al.
Published: (2025)
Beyond Binary Moral Judgment: Modeling Ethical Pluralism in AI
by: Aijaz, Aisha, et al.
Published: (2026)
by: Aijaz, Aisha, et al.
Published: (2026)
Similar Items
-
CurvaDion: Curvature-Adaptive Distributed Orthonormalization
by: Kumar, Bhavesh, et al.
Published: (2025) -
Verification of Visual Controllers via Compositional Geometric Transformations
by: Estornell, Alexander, et al.
Published: (2025) -
Synthpop++: A Hybrid Framework for Generating A Country-scale Synthetic Population
by: Neekhra, Bhavesh, et al.
Published: (2023) -
On the (In)Significance of Feature Selection in High-Dimensional Datasets
by: Neekhra, Bhavesh, et al.
Published: (2025) -
Optimization for Neural Operators can Benefit from Width
by: Cisneros-Velarde, Pedro, et al.
Published: (2025)