What's Producible May Not Be Reachable: Measuring the Steerability of Generative Models
Fuente:
arXiv
Saved in:
| Main Authors: | Vafa, Keyon, Bentley, Sarah, Kleinberg, Jon, Mullainathan, Sendhil |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Do Large Language Models Perform the Way People Expect? Measuring the Human Generalization Function
by: Vafa, Keyon, et al.
Published: (2024)
by: Vafa, Keyon, et al.
Published: (2024)
Evaluating the World Model Implicit in a Generative Model
by: Vafa, Keyon, et al.
Published: (2024)
by: Vafa, Keyon, et al.
Published: (2024)
Language Generation in the Limit
by: Kleinberg, Jon, et al.
Published: (2024)
by: Kleinberg, Jon, et al.
Published: (2024)
Potemkin Understanding in Large Language Models
by: Mancoridis, Marina, et al.
Published: (2025)
by: Mancoridis, Marina, et al.
Published: (2025)
It's a Feature, Not a Bug: Measuring Creative Fluidity in Image Generators
by: Ramaswamy, Aditi, et al.
Published: (2024)
by: Ramaswamy, Aditi, et al.
Published: (2024)
A Foundational Generative Model for Breast Ultrasound Image Analysis
by: Yu, Haojun, et al.
Published: (2025)
by: Yu, Haojun, et al.
Published: (2025)
Semantic Approach to Quantifying the Consistency of Diffusion Model Image Generation
by: Bent, Brinnae
Published: (2024)
by: Bent, Brinnae
Published: (2024)
Generating Synthetic Satellite Imagery for Rare Objects: An Empirical Comparison of Models and Metrics
by: Nguyen, Tuong Vy, et al.
Published: (2024)
by: Nguyen, Tuong Vy, et al.
Published: (2024)
What Has a Foundation Model Found? Using Inductive Bias to Probe for World Models
by: Vafa, Keyon, et al.
Published: (2025)
by: Vafa, Keyon, et al.
Published: (2025)
Generating Synthetic Satellite Imagery With Deep-Learning Text-to-Image Models -- Technical Challenges and Implications for Monitoring and Verification
by: Nguyen, Tuong Vy, et al.
Published: (2024)
by: Nguyen, Tuong Vy, et al.
Published: (2024)
DebiasPI: Inference-time Debiasing by Prompt Iteration of a Text-to-Image Generative Model
by: Bonna, Sarah, et al.
Published: (2025)
by: Bonna, Sarah, et al.
Published: (2025)
Yume: An Interactive World Generation Model
by: Mao, Xiaofeng, et al.
Published: (2025)
by: Mao, Xiaofeng, et al.
Published: (2025)
Efficient Personalization of Generative User Interfaces
by: Peng, Yi-Hao, et al.
Published: (2026)
by: Peng, Yi-Hao, et al.
Published: (2026)
Generalization of CNNs on Relational Reasoning with Bar Charts
by: Cui, Zhenxing, et al.
Published: (2025)
by: Cui, Zhenxing, et al.
Published: (2025)
Revision Matters: Generative Design Guided by Revision Edits
by: Li, Tao, et al.
Published: (2024)
by: Li, Tao, et al.
Published: (2024)
Characterizing Photorealism and Artifacts in Diffusion Model-Generated Images
by: Kamali, Negar, et al.
Published: (2025)
by: Kamali, Negar, et al.
Published: (2025)
Deep Generative Domain Adaptation with Temporal Attention for Cross-User Activity Recognition
by: Ye, Xiaozhou, et al.
Published: (2024)
by: Ye, Xiaozhou, et al.
Published: (2024)
EmoGene: Audio-Driven Emotional 3D Talking-Head Generation
by: Wang, Wenqing, et al.
Published: (2024)
by: Wang, Wenqing, et al.
Published: (2024)
Screen2AX: Vision-Based Approach for Automatic macOS Accessibility Generation
by: Muryn, Viktor, et al.
Published: (2025)
by: Muryn, Viktor, et al.
Published: (2025)
Deep Generative Domain Adaptation with Temporal Relation Knowledge for Cross-User Activity Recognition
by: Ye, Xiaozhou, et al.
Published: (2024)
by: Ye, Xiaozhou, et al.
Published: (2024)
FERGI: Automatic Scoring of User Preferences for Text-to-Image Generation from Spontaneous Facial Expression Reaction
by: Feng, Shuangquan, et al.
Published: (2023)
by: Feng, Shuangquan, et al.
Published: (2023)
See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding
by: Sun, Boyuan, et al.
Published: (2026)
by: Sun, Boyuan, et al.
Published: (2026)
ShowUI-$π$: Flow-based Generative Models as GUI Dexterous Hands
by: Hu, Siyuan, et al.
Published: (2025)
by: Hu, Siyuan, et al.
Published: (2025)
Efficient Retail Video Annotation: A Robust Key Frame Generation Approach for Product and Customer Interaction Analysis
by: Mannam, Varun, et al.
Published: (2025)
by: Mannam, Varun, et al.
Published: (2025)
Measuring Agreeableness Bias in Multimodal Models
by: Lim, Jaehyuk, et al.
Published: (2024)
by: Lim, Jaehyuk, et al.
Published: (2024)
MNIST-Gen: A Modular MNIST-Style Dataset Generation Using Hierarchical Semantics, Reinforcement Learning, and Category Theory
by: Shaeri, Pouya, et al.
Published: (2025)
by: Shaeri, Pouya, et al.
Published: (2025)
Training a Vision Language Model as Smartphone Assistant
by: Dorka, Nicolai, et al.
Published: (2024)
by: Dorka, Nicolai, et al.
Published: (2024)
LEDITS++: Limitless Image Editing using Text-to-Image Models
by: Brack, Manuel, et al.
Published: (2023)
by: Brack, Manuel, et al.
Published: (2023)
InterVLS: Interactive Model Understanding and Improvement with Vision-Language Surrogates
by: Huang, Jinbin, et al.
Published: (2023)
by: Huang, Jinbin, et al.
Published: (2023)
Predicting and Explaining Mobile UI Tappability with Vision Modeling and Saliency Analysis
by: Schoop, Eldon, et al.
Published: (2022)
by: Schoop, Eldon, et al.
Published: (2022)
I-CEE: Tailoring Explanations of Image Classification Models to User Expertise
by: Rong, Yao, et al.
Published: (2023)
by: Rong, Yao, et al.
Published: (2023)
Advancing the Understanding and Evaluation of AR-Generated Scenes: When Vision-Language Models Shine and Stumble
by: Duan, Lin, et al.
Published: (2025)
by: Duan, Lin, et al.
Published: (2025)
Quantitative Movement Testing: Measuring Patient Movements from a Single Smartphone Video
by: Mahajan, Pranav, et al.
Published: (2026)
by: Mahajan, Pranav, et al.
Published: (2026)
SpiderNets: Vision Models Predict Human Fear From Aversive Images
by: Pegler, Dominik, et al.
Published: (2025)
by: Pegler, Dominik, et al.
Published: (2025)
HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning
by: Hiranaka, Ayano, et al.
Published: (2024)
by: Hiranaka, Ayano, et al.
Published: (2024)
ExpressEdit: Fast Editing of Stylized Facial Expressions with Diffusion Models in Photoshop
by: Tang, Kenan, et al.
Published: (2026)
by: Tang, Kenan, et al.
Published: (2026)
LLAniMAtion: LLAMA Driven Gesture Animation
by: Windle, Jonathan, et al.
Published: (2024)
by: Windle, Jonathan, et al.
Published: (2024)
Interaction as Explanation: A User Interaction-based Method for Explaining Image Classification Models
by: Yun, Hyeonggeun
Published: (2024)
by: Yun, Hyeonggeun
Published: (2024)
K-Sort Arena: Efficient and Reliable Benchmarking for Generative Models via K-wise Human Preferences
by: Li, Zhikai, et al.
Published: (2024)
by: Li, Zhikai, et al.
Published: (2024)
Vision-Language Models for Ergonomic Assessment of Manual Lifting Tasks: Estimating Horizontal and Vertical Hand Distances from RGB Video
by: Rajabi, Mohammad Sadra, et al.
Published: (2026)
by: Rajabi, Mohammad Sadra, et al.
Published: (2026)
Similar Items
-
Do Large Language Models Perform the Way People Expect? Measuring the Human Generalization Function
by: Vafa, Keyon, et al.
Published: (2024) -
Evaluating the World Model Implicit in a Generative Model
by: Vafa, Keyon, et al.
Published: (2024) -
Language Generation in the Limit
by: Kleinberg, Jon, et al.
Published: (2024) -
Potemkin Understanding in Large Language Models
by: Mancoridis, Marina, et al.
Published: (2025) -
It's a Feature, Not a Bug: Measuring Creative Fluidity in Image Generators
by: Ramaswamy, Aditi, et al.
Published: (2024)