Stylus: Automatic Adapter Selection for Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Michael, Wong, Justin, Trabucco, Brandon, Huang, Yanping, Gonzalez, Joseph E., Chen, Zhifeng, Salakhutdinov, Ruslan, Stoica, Ion |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Effective Data Augmentation With Diffusion Models
by: Trabucco, Brandon, et al.
Published: (2023)
by: Trabucco, Brandon, et al.
Published: (2023)
A Study of the Framework and Real-World Applications of Language Embedding for 3D Scene Understanding
by: Zaouali, Mahmoud Chick, et al.
Published: (2025)
by: Zaouali, Mahmoud Chick, et al.
Published: (2025)
Understanding Visual Concepts Across Models
by: Trabucco, Brandon, et al.
Published: (2024)
by: Trabucco, Brandon, et al.
Published: (2024)
Generic 3D Diffusion Adapter Using Controlled Multi-View Editing
by: Chen, Hansheng, et al.
Published: (2024)
by: Chen, Hansheng, et al.
Published: (2024)
Haptic Stylus vs. Handheld Controllers: A Comparative Study for Surface Visualization Interactions
by: Afzaal, Hamza, et al.
Published: (2024)
by: Afzaal, Hamza, et al.
Published: (2024)
FIFA: Unified Faithfulness Evaluation Framework for Text-to-Video and Video-to-Text Generation
by: Jing, Liqiang, et al.
Published: (2025)
by: Jing, Liqiang, et al.
Published: (2025)
Is this chart lying to me? Automating the detection of misleading visualizations
by: Tonglet, Jonathan, et al.
Published: (2025)
by: Tonglet, Jonathan, et al.
Published: (2025)
ComfyGen: Prompt-Adaptive Workflows for Text-to-Image Generation
by: Gal, Rinon, et al.
Published: (2024)
by: Gal, Rinon, et al.
Published: (2024)
TexGS-VolVis: Expressive Scene Editing for Volume Visualization via Textured Gaussian Splatting
by: Tang, Kaiyuan, et al.
Published: (2025)
by: Tang, Kaiyuan, et al.
Published: (2025)
Is Your World Simulator a Good Story Presenter? A Consecutive Events-Based Benchmark for Future Long Video Generation
by: Wang, Yiping, et al.
Published: (2024)
by: Wang, Yiping, et al.
Published: (2024)
FlairGPT: Repurposing LLMs for Interior Designs
by: Littlefair, Gabrielle, et al.
Published: (2025)
by: Littlefair, Gabrielle, et al.
Published: (2025)
Co-Layout: LLM-driven Co-optimization for Interior Layout
by: Xiang, Chucheng, et al.
Published: (2025)
by: Xiang, Chucheng, et al.
Published: (2025)
T$^3$-S2S: Training-free Triplet Tuning for Sketch to Scene Synthesis in Controllable Concept Art Generation
by: Sun, Zhenhong, et al.
Published: (2024)
by: Sun, Zhenhong, et al.
Published: (2024)
Global Position Aware Group Choreography using Large Language Model
by: Pang, Haozhou, et al.
Published: (2025)
by: Pang, Haozhou, et al.
Published: (2025)
CAP: Evaluation of Persuasive and Creative Image Generation
by: Aghazadeh, Aysan, et al.
Published: (2024)
by: Aghazadeh, Aysan, et al.
Published: (2024)
Token Perturbation Guidance for Diffusion Models
by: Rajabi, Javad, et al.
Published: (2025)
by: Rajabi, Javad, et al.
Published: (2025)
How to Train Your Dragon: Automatic Diffusion-Based Rigging for Characters with Diverse Topologies
by: Gu, Zeqi, et al.
Published: (2025)
by: Gu, Zeqi, et al.
Published: (2025)
VisionGPT-3D: A Generalized Multimodal Agent for Enhanced 3D Vision Understanding
by: Kelly, Chris, et al.
Published: (2024)
by: Kelly, Chris, et al.
Published: (2024)
STGA: Selective-Training Gaussian Head Avatars
by: Guo, Hanzhi, et al.
Published: (2025)
by: Guo, Hanzhi, et al.
Published: (2025)
ReverBERT: A State Space Model for Efficient Text-Driven Speech Style Transfer
by: Brown, Michael, et al.
Published: (2025)
by: Brown, Michael, et al.
Published: (2025)
Towards Interactive Intelligence for Digital Humans
by: Cai, Yiyi, et al.
Published: (2025)
by: Cai, Yiyi, et al.
Published: (2025)
Back to Basics: Motion Representation Matters for Human Motion Generation Using Diffusion Model
by: Jin, Yuduo, et al.
Published: (2025)
by: Jin, Yuduo, et al.
Published: (2025)
LGTM: Local-to-Global Text-Driven Human Motion Diffusion Model
by: Sun, Haowen, et al.
Published: (2024)
by: Sun, Haowen, et al.
Published: (2024)
Physics-based Scene Layout Generation from Human Motion
by: Li, Jianan, et al.
Published: (2024)
by: Li, Jianan, et al.
Published: (2024)
Learning to Control Physically-simulated 3D Characters via Generating and Mimicking 2D Motions
by: Li, Jianan, et al.
Published: (2025)
by: Li, Jianan, et al.
Published: (2025)
OpenCOLE: Towards Reproducible Automatic Graphic Design Generation
by: Inoue, Naoto, et al.
Published: (2024)
by: Inoue, Naoto, et al.
Published: (2024)
AniGaussian: Animatable Gaussian Avatar with Pose-guided Deformation
by: Li, Mengtian, et al.
Published: (2025)
by: Li, Mengtian, et al.
Published: (2025)
DynamicGTR: Leveraging Graph Topology Representation Preferences to Boost VLM Capabilities on Graph QAs
by: Wei, Yanbin, et al.
Published: (2026)
by: Wei, Yanbin, et al.
Published: (2026)
ORACLE: Orchestrate NPC Daily Activities using Contrastive Learning with Transformer-CVAE
by: Hong, Seong-Eun, et al.
Published: (2026)
by: Hong, Seong-Eun, et al.
Published: (2026)
PALP: Prompt Aligned Personalization of Text-to-Image Models
by: Arar, Moab, et al.
Published: (2024)
by: Arar, Moab, et al.
Published: (2024)
Grounding Language in Multi-Perspective Referential Communication
by: Tang, Zineng, et al.
Published: (2024)
by: Tang, Zineng, et al.
Published: (2024)
Towards Understanding Graphical Perception in Large Multimodal Models
by: Zhang, Kai, et al.
Published: (2025)
by: Zhang, Kai, et al.
Published: (2025)
SIMS: Simulating Stylized Human-Scene Interactions with Retrieval-Augmented Script Generation
by: Wang, Wenjia, et al.
Published: (2024)
by: Wang, Wenjia, et al.
Published: (2024)
Generative Powers of Ten
by: Wang, Xiaojuan, et al.
Published: (2023)
by: Wang, Xiaojuan, et al.
Published: (2023)
Image Generation Models: A Technical History
by: Shirvani, Rouzbeh
Published: (2026)
by: Shirvani, Rouzbeh
Published: (2026)
MM-Conv: A Multi-modal Conversational Dataset for Virtual Humans
by: Deichler, Anna, et al.
Published: (2024)
by: Deichler, Anna, et al.
Published: (2024)
Advancements and limitations of LLMs in replicating human color-word associations
by: Fukushima, Makoto, et al.
Published: (2024)
by: Fukushima, Makoto, et al.
Published: (2024)
StyleMotif: Multi-Modal Motion Stylization using Style-Content Cross Fusion
by: Guo, Ziyu, et al.
Published: (2025)
by: Guo, Ziyu, et al.
Published: (2025)
In-Context LoRA for Diffusion Transformers
by: Huang, Lianghua, et al.
Published: (2024)
by: Huang, Lianghua, et al.
Published: (2024)
SAMa: Material-aware 3D Selection and Segmentation
by: Fischer, Michael, et al.
Published: (2024)
by: Fischer, Michael, et al.
Published: (2024)
Similar Items
-
Effective Data Augmentation With Diffusion Models
by: Trabucco, Brandon, et al.
Published: (2023) -
A Study of the Framework and Real-World Applications of Language Embedding for 3D Scene Understanding
by: Zaouali, Mahmoud Chick, et al.
Published: (2025) -
Understanding Visual Concepts Across Models
by: Trabucco, Brandon, et al.
Published: (2024) -
Generic 3D Diffusion Adapter Using Controlled Multi-View Editing
by: Chen, Hansheng, et al.
Published: (2024) -
Haptic Stylus vs. Handheld Controllers: A Comparative Study for Surface Visualization Interactions
by: Afzaal, Hamza, et al.
Published: (2024)