Towards a Scalable Reference-Free Evaluation of Generative Models
Fuente:
arXiv
Saved in:
| Main Authors: | Ospanov, Azim, Zhang, Jingwei, Jalali, Mohammad, Cao, Xuenan, Bogdanov, Andrej, Farnia, Farzan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exposing Diversity Bias in Deep Generative Models: Statistical Origins and Correction of Diversity Error
by: Farnia, Farzan, et al.
Published: (2026)
by: Farnia, Farzan, et al.
Published: (2026)
Do Vendi Scores Converge with Finite Samples? Truncated Vendi Score for Finite-Sample Convergence Guarantees
by: Ospanov, Azim, et al.
Published: (2024)
by: Ospanov, Azim, et al.
Published: (2024)
Conditional Vendi Score: An Information-Theoretic Approach to Diversity Evaluation of Prompt-based Generative Models
by: Jalali, Mohammad, et al.
Published: (2024)
by: Jalali, Mohammad, et al.
Published: (2024)
miniF2F-Lean Revisited: Reviewing Limitations and Charting a Path Forward
by: Ospanov, Azim, et al.
Published: (2025)
by: Ospanov, Azim, et al.
Published: (2025)
PromptSplit: Revealing Prompt-Level Disagreement in Generative Models
by: Lotfian, Mehdi, et al.
Published: (2026)
by: Lotfian, Mehdi, et al.
Published: (2026)
Scendi Score: Prompt-Aware Diversity Evaluation via Schur Complement of CLIP Embeddings
by: Ospanov, Azim, et al.
Published: (2024)
by: Ospanov, Azim, et al.
Published: (2024)
SPARKE: Scalable Prompt-Aware Diversity and Novelty Guidance in Diffusion Models via RKE Score
by: Jalali, Mohammad, et al.
Published: (2025)
by: Jalali, Mohammad, et al.
Published: (2025)
APOLLO: Automated LLM and Lean Collaboration for Advanced Formal Reasoning
by: Ospanov, Azim, et al.
Published: (2025)
by: Ospanov, Azim, et al.
Published: (2025)
Towards an Explainable Comparison and Alignment of Feature Embeddings
by: Jalali, Mohammad, et al.
Published: (2025)
by: Jalali, Mohammad, et al.
Published: (2025)
Training-Free Distribution Adaptation for Diffusion Models via Maximum Mean Discrepancy Guidance
by: Sani, Matina Mahdizadeh, et al.
Published: (2026)
by: Sani, Matina Mahdizadeh, et al.
Published: (2026)
Unveiling Differences in Generative Models: A Scalable Differential Clustering Approach
by: Zhang, Jingwei, et al.
Published: (2024)
by: Zhang, Jingwei, et al.
Published: (2024)
On the Fragility of AI-Based Channel Decoders under Small Channel Perturbations
by: Lei, Haoyu, et al.
Published: (2026)
by: Lei, Haoyu, et al.
Published: (2026)
A Lean Dataset for International Math Olympiad: Small Steps towards Writing Math Proofs for Hard Problems
by: Yousefzadeh, Roozbeh, et al.
Published: (2024)
by: Yousefzadeh, Roozbeh, et al.
Published: (2024)
Certified Adversarial Robustness via Partition-based Randomized Smoothing
by: Goli, Hossein, et al.
Published: (2024)
by: Goli, Hossein, et al.
Published: (2024)
HERMES: Towards Efficient and Verifiable Mathematical Reasoning in LLMs
by: Ospanov, Azim, et al.
Published: (2025)
by: Ospanov, Azim, et al.
Published: (2025)
When Exploration Comes for Free with Mixture-Greedy: Do we need UCB in Diversity-Aware Multi-Armed Bandits?
by: Nia, Bahar Dibaei, et al.
Published: (2026)
by: Nia, Bahar Dibaei, et al.
Published: (2026)
PromptWise: Online Learning for Cost-Aware Prompt Assignment in Generative Models
by: Hu, Xiaoyan, et al.
Published: (2025)
by: Hu, Xiaoyan, et al.
Published: (2025)
On the Inductive Biases of Demographic Parity-based Fair Learning Algorithms
by: Lei, Haoyu, et al.
Published: (2024)
by: Lei, Haoyu, et al.
Published: (2024)
An Interpretable Evaluation of Entropy-based Novelty of Generative Models
by: Zhang, Jingwei, et al.
Published: (2024)
by: Zhang, Jingwei, et al.
Published: (2024)
Boosting Cross-problem Generalization in Diffusion-Based Neural Combinatorial Solver via Inference Time Adaptation
by: Lei, Haoyu, et al.
Published: (2025)
by: Lei, Haoyu, et al.
Published: (2025)
Sparse Domain Transfer via Elastic Net Regularization
by: Zhang, Jingwei, et al.
Published: (2024)
by: Zhang, Jingwei, et al.
Published: (2024)
Syndrome-Flow Consistency Model Achieves One-step Denoising Error Correction Codes
by: Lei, Haoyu, et al.
Published: (2025)
by: Lei, Haoyu, et al.
Published: (2025)
Interpretation of Neural Networks is Susceptible to Universal Adversarial Perturbations
by: Oskouie, Haniyeh Ehsani, et al.
Published: (2022)
by: Oskouie, Haniyeh Ehsani, et al.
Published: (2022)
PermLLM: Learnable Channel Permutation for N:M Sparse Large Language Models
by: Zou, Lancheng, et al.
Published: (2025)
by: Zou, Lancheng, et al.
Published: (2025)
When Kernels Multiply, Clusters Unify: Fusing Embeddings with the Kronecker Product
by: Wu, Youqi, et al.
Published: (2025)
by: Wu, Youqi, et al.
Published: (2025)
Stability and Generalization in Free Adversarial Training
by: Cheng, Xiwei, et al.
Published: (2024)
by: Cheng, Xiwei, et al.
Published: (2024)
DAK-UCB: Diversity-Aware Prompt Routing for LLMs and Generative Models
by: Jafari, Donya, et al.
Published: (2026)
by: Jafari, Donya, et al.
Published: (2026)
Advocate for Complete Benchmarks for Formal Reasoning with Formal/Informal Statements and Formal/Informal Proofs
by: Yousefzadeh, Roozbeh, et al.
Published: (2025)
by: Yousefzadeh, Roozbeh, et al.
Published: (2025)
A Multi-Armed Bandit Approach to Online Selection and Evaluation of Generative Models
by: Hu, Xiaoyan, et al.
Published: (2024)
by: Hu, Xiaoyan, et al.
Published: (2024)
Scalable Evaluation and Neural Models for Compositional Generalization
by: Camposampiero, Giacomo, et al.
Published: (2025)
by: Camposampiero, Giacomo, et al.
Published: (2025)
Towards General-Purpose Model-Free Reinforcement Learning
by: Fujimoto, Scott, et al.
Published: (2025)
by: Fujimoto, Scott, et al.
Published: (2025)
On the Distributed Evaluation of Generative Models
by: Wang, Zixiao, et al.
Published: (2023)
by: Wang, Zixiao, et al.
Published: (2023)
SORA: Free Second-Order Attacks in Fast Adversarial Training
by: Teymourian, Mazdak, et al.
Published: (2026)
by: Teymourian, Mazdak, et al.
Published: (2026)
MoreauPruner: Robust Pruning of Large Language Models against Weight Perturbations
by: Wang, Zixiao, et al.
Published: (2024)
by: Wang, Zixiao, et al.
Published: (2024)
Certifiably Robust Model Evaluation in Federated Learning under Meta-Distributional Shifts
by: Najafi, Amir, et al.
Published: (2024)
by: Najafi, Amir, et al.
Published: (2024)
Learning to Play Air Hockey with Model-Based Deep Reinforcement Learning
by: Orsula, Andrej
Published: (2024)
by: Orsula, Andrej
Published: (2024)
A Data-Centric Perspective on Evaluating Machine Learning Models for Tabular Data
by: Tschalzev, Andrej, et al.
Published: (2024)
by: Tschalzev, Andrej, et al.
Published: (2024)
Loquetier: A Virtualized Multi-LoRA Framework for Unified LLM Fine-tuning and Serving
by: Zhang, Yuchen, et al.
Published: (2025)
by: Zhang, Yuchen, et al.
Published: (2025)
PAK-UCB Contextual Bandit: An Online Learning Approach to Prompt-Aware Selection of Generative Models and LLMs
by: Hu, Xiaoyan, et al.
Published: (2024)
by: Hu, Xiaoyan, et al.
Published: (2024)
The Maximum von Neumann Entropy Principle: Theory and Applications in Machine Learning
by: Wu, Youqi, et al.
Published: (2026)
by: Wu, Youqi, et al.
Published: (2026)
Similar Items
-
Exposing Diversity Bias in Deep Generative Models: Statistical Origins and Correction of Diversity Error
by: Farnia, Farzan, et al.
Published: (2026) -
Do Vendi Scores Converge with Finite Samples? Truncated Vendi Score for Finite-Sample Convergence Guarantees
by: Ospanov, Azim, et al.
Published: (2024) -
Conditional Vendi Score: An Information-Theoretic Approach to Diversity Evaluation of Prompt-based Generative Models
by: Jalali, Mohammad, et al.
Published: (2024) -
miniF2F-Lean Revisited: Reviewing Limitations and Charting a Path Forward
by: Ospanov, Azim, et al.
Published: (2025) -
PromptSplit: Revealing Prompt-Level Disagreement in Generative Models
by: Lotfian, Mehdi, et al.
Published: (2026)