What's Producible May Not Be Reachable: Measuring the Steerability of Generative Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vafa, Keyon, Bentley, Sarah, Kleinberg, Jon, Mullainathan, Sendhil |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Do Large Language Models Perform the Way People Expect? Measuring the Human Generalization Function
von: Vafa, Keyon, et al.
Veröffentlicht: (2024)
von: Vafa, Keyon, et al.
Veröffentlicht: (2024)
Evaluating the World Model Implicit in a Generative Model
von: Vafa, Keyon, et al.
Veröffentlicht: (2024)
von: Vafa, Keyon, et al.
Veröffentlicht: (2024)
Language Generation in the Limit
von: Kleinberg, Jon, et al.
Veröffentlicht: (2024)
von: Kleinberg, Jon, et al.
Veröffentlicht: (2024)
Potemkin Understanding in Large Language Models
von: Mancoridis, Marina, et al.
Veröffentlicht: (2025)
von: Mancoridis, Marina, et al.
Veröffentlicht: (2025)
It's a Feature, Not a Bug: Measuring Creative Fluidity in Image Generators
von: Ramaswamy, Aditi, et al.
Veröffentlicht: (2024)
von: Ramaswamy, Aditi, et al.
Veröffentlicht: (2024)
A Foundational Generative Model for Breast Ultrasound Image Analysis
von: Yu, Haojun, et al.
Veröffentlicht: (2025)
von: Yu, Haojun, et al.
Veröffentlicht: (2025)
Semantic Approach to Quantifying the Consistency of Diffusion Model Image Generation
von: Bent, Brinnae
Veröffentlicht: (2024)
von: Bent, Brinnae
Veröffentlicht: (2024)
Generating Synthetic Satellite Imagery for Rare Objects: An Empirical Comparison of Models and Metrics
von: Nguyen, Tuong Vy, et al.
Veröffentlicht: (2024)
von: Nguyen, Tuong Vy, et al.
Veröffentlicht: (2024)
What Has a Foundation Model Found? Using Inductive Bias to Probe for World Models
von: Vafa, Keyon, et al.
Veröffentlicht: (2025)
von: Vafa, Keyon, et al.
Veröffentlicht: (2025)
Generating Synthetic Satellite Imagery With Deep-Learning Text-to-Image Models -- Technical Challenges and Implications for Monitoring and Verification
von: Nguyen, Tuong Vy, et al.
Veröffentlicht: (2024)
von: Nguyen, Tuong Vy, et al.
Veröffentlicht: (2024)
DebiasPI: Inference-time Debiasing by Prompt Iteration of a Text-to-Image Generative Model
von: Bonna, Sarah, et al.
Veröffentlicht: (2025)
von: Bonna, Sarah, et al.
Veröffentlicht: (2025)
Yume: An Interactive World Generation Model
von: Mao, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Mao, Xiaofeng, et al.
Veröffentlicht: (2025)
Efficient Personalization of Generative User Interfaces
von: Peng, Yi-Hao, et al.
Veröffentlicht: (2026)
von: Peng, Yi-Hao, et al.
Veröffentlicht: (2026)
Generalization of CNNs on Relational Reasoning with Bar Charts
von: Cui, Zhenxing, et al.
Veröffentlicht: (2025)
von: Cui, Zhenxing, et al.
Veröffentlicht: (2025)
Revision Matters: Generative Design Guided by Revision Edits
von: Li, Tao, et al.
Veröffentlicht: (2024)
von: Li, Tao, et al.
Veröffentlicht: (2024)
Characterizing Photorealism and Artifacts in Diffusion Model-Generated Images
von: Kamali, Negar, et al.
Veröffentlicht: (2025)
von: Kamali, Negar, et al.
Veröffentlicht: (2025)
Deep Generative Domain Adaptation with Temporal Attention for Cross-User Activity Recognition
von: Ye, Xiaozhou, et al.
Veröffentlicht: (2024)
von: Ye, Xiaozhou, et al.
Veröffentlicht: (2024)
EmoGene: Audio-Driven Emotional 3D Talking-Head Generation
von: Wang, Wenqing, et al.
Veröffentlicht: (2024)
von: Wang, Wenqing, et al.
Veröffentlicht: (2024)
Screen2AX: Vision-Based Approach for Automatic macOS Accessibility Generation
von: Muryn, Viktor, et al.
Veröffentlicht: (2025)
von: Muryn, Viktor, et al.
Veröffentlicht: (2025)
Deep Generative Domain Adaptation with Temporal Relation Knowledge for Cross-User Activity Recognition
von: Ye, Xiaozhou, et al.
Veröffentlicht: (2024)
von: Ye, Xiaozhou, et al.
Veröffentlicht: (2024)
FERGI: Automatic Scoring of User Preferences for Text-to-Image Generation from Spontaneous Facial Expression Reaction
von: Feng, Shuangquan, et al.
Veröffentlicht: (2023)
von: Feng, Shuangquan, et al.
Veröffentlicht: (2023)
See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding
von: Sun, Boyuan, et al.
Veröffentlicht: (2026)
von: Sun, Boyuan, et al.
Veröffentlicht: (2026)
ShowUI-$π$: Flow-based Generative Models as GUI Dexterous Hands
von: Hu, Siyuan, et al.
Veröffentlicht: (2025)
von: Hu, Siyuan, et al.
Veröffentlicht: (2025)
Efficient Retail Video Annotation: A Robust Key Frame Generation Approach for Product and Customer Interaction Analysis
von: Mannam, Varun, et al.
Veröffentlicht: (2025)
von: Mannam, Varun, et al.
Veröffentlicht: (2025)
Measuring Agreeableness Bias in Multimodal Models
von: Lim, Jaehyuk, et al.
Veröffentlicht: (2024)
von: Lim, Jaehyuk, et al.
Veröffentlicht: (2024)
MNIST-Gen: A Modular MNIST-Style Dataset Generation Using Hierarchical Semantics, Reinforcement Learning, and Category Theory
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025)
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025)
Training a Vision Language Model as Smartphone Assistant
von: Dorka, Nicolai, et al.
Veröffentlicht: (2024)
von: Dorka, Nicolai, et al.
Veröffentlicht: (2024)
LEDITS++: Limitless Image Editing using Text-to-Image Models
von: Brack, Manuel, et al.
Veröffentlicht: (2023)
von: Brack, Manuel, et al.
Veröffentlicht: (2023)
InterVLS: Interactive Model Understanding and Improvement with Vision-Language Surrogates
von: Huang, Jinbin, et al.
Veröffentlicht: (2023)
von: Huang, Jinbin, et al.
Veröffentlicht: (2023)
Predicting and Explaining Mobile UI Tappability with Vision Modeling and Saliency Analysis
von: Schoop, Eldon, et al.
Veröffentlicht: (2022)
von: Schoop, Eldon, et al.
Veröffentlicht: (2022)
I-CEE: Tailoring Explanations of Image Classification Models to User Expertise
von: Rong, Yao, et al.
Veröffentlicht: (2023)
von: Rong, Yao, et al.
Veröffentlicht: (2023)
Advancing the Understanding and Evaluation of AR-Generated Scenes: When Vision-Language Models Shine and Stumble
von: Duan, Lin, et al.
Veröffentlicht: (2025)
von: Duan, Lin, et al.
Veröffentlicht: (2025)
Quantitative Movement Testing: Measuring Patient Movements from a Single Smartphone Video
von: Mahajan, Pranav, et al.
Veröffentlicht: (2026)
von: Mahajan, Pranav, et al.
Veröffentlicht: (2026)
SpiderNets: Vision Models Predict Human Fear From Aversive Images
von: Pegler, Dominik, et al.
Veröffentlicht: (2025)
von: Pegler, Dominik, et al.
Veröffentlicht: (2025)
HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning
von: Hiranaka, Ayano, et al.
Veröffentlicht: (2024)
von: Hiranaka, Ayano, et al.
Veröffentlicht: (2024)
ExpressEdit: Fast Editing of Stylized Facial Expressions with Diffusion Models in Photoshop
von: Tang, Kenan, et al.
Veröffentlicht: (2026)
von: Tang, Kenan, et al.
Veröffentlicht: (2026)
LLAniMAtion: LLAMA Driven Gesture Animation
von: Windle, Jonathan, et al.
Veröffentlicht: (2024)
von: Windle, Jonathan, et al.
Veröffentlicht: (2024)
Interaction as Explanation: A User Interaction-based Method for Explaining Image Classification Models
von: Yun, Hyeonggeun
Veröffentlicht: (2024)
von: Yun, Hyeonggeun
Veröffentlicht: (2024)
K-Sort Arena: Efficient and Reliable Benchmarking for Generative Models via K-wise Human Preferences
von: Li, Zhikai, et al.
Veröffentlicht: (2024)
von: Li, Zhikai, et al.
Veröffentlicht: (2024)
Vision-Language Models for Ergonomic Assessment of Manual Lifting Tasks: Estimating Horizontal and Vertical Hand Distances from RGB Video
von: Rajabi, Mohammad Sadra, et al.
Veröffentlicht: (2026)
von: Rajabi, Mohammad Sadra, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Do Large Language Models Perform the Way People Expect? Measuring the Human Generalization Function
von: Vafa, Keyon, et al.
Veröffentlicht: (2024) -
Evaluating the World Model Implicit in a Generative Model
von: Vafa, Keyon, et al.
Veröffentlicht: (2024) -
Language Generation in the Limit
von: Kleinberg, Jon, et al.
Veröffentlicht: (2024) -
Potemkin Understanding in Large Language Models
von: Mancoridis, Marina, et al.
Veröffentlicht: (2025) -
It's a Feature, Not a Bug: Measuring Creative Fluidity in Image Generators
von: Ramaswamy, Aditi, et al.
Veröffentlicht: (2024)