Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Bingchen, Akhgari, Ehsan, Visheratin, Alexander, Kamko, Aleks, Xu, Linmiao, Shrirao, Shivam, Lambert, Chase, Souza, Joao, Doshi, Suhail, Li, Daiqing |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation
by: Li, Daiqing, et al.
Published: (2024)
by: Li, Daiqing, et al.
Published: (2024)
Quantum Machine Learning Playground
by: Debus, Pascal, et al.
Published: (2025)
by: Debus, Pascal, et al.
Published: (2025)
GraphShaper: Geometry-aware Alignment for Improving Transfer Learning in Text-Attributed Graphs
by: Zhang, Heng, et al.
Published: (2025)
by: Zhang, Heng, et al.
Published: (2025)
From Delays to Densities: Exploring Data Uncertainty through Speech, Text, and Visualization
by: Chase Stokes, et al.
Published: (2024)
by: Chase Stokes, et al.
Published: (2024)
Can Representation Gaps Be the Key to Enhancing Robustness in Graph-Text Alignment?
by: Zhang, Heng, et al.
Published: (2025)
by: Zhang, Heng, et al.
Published: (2025)
Deep SVBRDF Acquisition and Modelling: A Survey
by: Behnaz Kavoosighafi, et al.
Published: (2024)
by: Behnaz Kavoosighafi, et al.
Published: (2024)
ETBHD‐HMF: A Hierarchical Multimodal Fusion Architecture for Enhanced Text‐Based Hair Design
by: Rong He, et al.
Published: (2024)
by: Rong He, et al.
Published: (2024)
Bilateral Guided Radiance Field Processing
by: Wang, Yuehao, et al.
Published: (2024)
by: Wang, Yuehao, et al.
Published: (2024)
Harness Local Rewards for Global Benefits: Effective Text-to-Video Generation Alignment with Patch-level Reward Models
by: Wang, Shuting, et al.
Published: (2025)
by: Wang, Shuting, et al.
Published: (2025)
DLSF: Dual-Layer Synergistic Fusion for High-Fidelity Image Syn-thesis
by: Chen, Zhen-Qi, et al.
Published: (2025)
by: Chen, Zhen-Qi, et al.
Published: (2025)
WaFusion: A Wavelet-Enhanced Diffusion Framework for Face Morph Generation
by: Hosseini, Seyed Rasoul, et al.
Published: (2025)
by: Hosseini, Seyed Rasoul, et al.
Published: (2025)
FlowMotion: Target-Predictive Conditional Flow Matching for Jitter-Reduced Text-Driven Human Motion Generation
by: Cuba, Manolo Canales, et al.
Published: (2025)
by: Cuba, Manolo Canales, et al.
Published: (2025)
Fuse3D: Generating 3D Assets Controlled by Multi-Image Fusion
by: Jin, Xuancheng, et al.
Published: (2025)
by: Jin, Xuancheng, et al.
Published: (2025)
Simplicial Approximation of Deforming 3D Spaces for Visualizing Fusion Plasma Simulation Data
by: Ren, Congrong, et al.
Published: (2023)
by: Ren, Congrong, et al.
Published: (2023)
Controllable Segmentation-Based Text-Guided Style Editing
by: Li, Jingwen, et al.
Published: (2025)
by: Li, Jingwen, et al.
Published: (2025)
FatigueFusion: Latent Space Fusion for Fatigue-Driven Motion Synthesis
by: Loi, Iliana, et al.
Published: (2026)
by: Loi, Iliana, et al.
Published: (2026)
Text-based Transfer Function Design for Semantic Volume Rendering
by: Jeong, Sangwon, et al.
Published: (2024)
by: Jeong, Sangwon, et al.
Published: (2024)
T2Bs: Text-to-Character Blendshapes via Video Generation
by: Luo, Jiahao, et al.
Published: (2025)
by: Luo, Jiahao, et al.
Published: (2025)
CAD-Coder: Text-to-CAD Generation with Chain-of-Thought and Geometric Reward
by: Guan, Yandong, et al.
Published: (2025)
by: Guan, Yandong, et al.
Published: (2025)
SketchDream: Sketch-based Text-to-3D Generation and Editing
by: Liu, Feng-Lin, et al.
Published: (2024)
by: Liu, Feng-Lin, et al.
Published: (2024)
Adaptive Sampling for BRDF Acquisition
by: Behnaz Kavoosighafi, et al.
Published: (2025)
by: Behnaz Kavoosighafi, et al.
Published: (2025)
SigStyle: Signature Style Transfer via Personalized Text-to-Image Models
by: Wang, Ye, et al.
Published: (2025)
by: Wang, Ye, et al.
Published: (2025)
Event-T2M: Event-level Conditioning for Complex Text-to-Motion Synthesis
by: Hong, Seong-Eun, et al.
Published: (2026)
by: Hong, Seong-Eun, et al.
Published: (2026)
Holographic Parallax Improves 3D Perceptual Realism
by: Kim, Dongyeon, et al.
Published: (2024)
by: Kim, Dongyeon, et al.
Published: (2024)
Building An Efficient Grid On GPU
by: Costa, Vasco, et al.
Published: (2024)
by: Costa, Vasco, et al.
Published: (2024)
Selfi: Self Improving Reconstruction Engine via 3D Geometric Feature Alignment
by: Deng, Youming, et al.
Published: (2025)
by: Deng, Youming, et al.
Published: (2025)
G2rammar: Bilingual Grammar Modeling for Enhanced Text-attributed Graph Learning
by: Zheng, Heng, et al.
Published: (2025)
by: Zheng, Heng, et al.
Published: (2025)
Data Visualization for Improving Financial Literacy: A Systematic Review
by: Du, Meng, et al.
Published: (2025)
by: Du, Meng, et al.
Published: (2025)
DeepMill: Neural Accessibility Learning for Subtractive Manufacturing
by: Zhong, Fanchao, et al.
Published: (2025)
by: Zhong, Fanchao, et al.
Published: (2025)
The Rise of AI-Generated Anime Avatars: Trends, Challenges, and Opportunities
by: Yamada, Fernanda Miyuki, et al.
Published: (2026)
by: Yamada, Fernanda Miyuki, et al.
Published: (2026)
Improving Sparse IMU-based Motion Capture with Motion Label Smoothing
by: Meng, Zhaorui, et al.
Published: (2025)
by: Meng, Zhaorui, et al.
Published: (2025)
MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning
by: Zhang, Yi-Yang, et al.
Published: (2025)
by: Zhang, Yi-Yang, et al.
Published: (2025)
Text-Driven Video Style Transfer with State-Space Models: Extending StyleMamba for Temporal Coherence
by: Li, Chao, et al.
Published: (2025)
by: Li, Chao, et al.
Published: (2025)
DP-Adapter: Dual-Pathway Adapter for Boosting Fidelity and Text Consistency in Customizable Human Image Generation
by: Wang, Ye, et al.
Published: (2025)
by: Wang, Ye, et al.
Published: (2025)
Improving Global Motion Estimation in Sparse IMU-based Motion Capture with Physics
by: Yi, Xinyu, et al.
Published: (2025)
by: Yi, Xinyu, et al.
Published: (2025)
Utilizing Motion Matching with Deep Reinforcement Learning for Target Location Tasks
by: Lee, Jeongmin, et al.
Published: (2024)
by: Lee, Jeongmin, et al.
Published: (2024)
Parallelobox: Improved Decomposition for Optimized Parallel Printing using Axis-Aligned Bounding Boxes
by: Hatton, Hayley, et al.
Published: (2026)
by: Hatton, Hayley, et al.
Published: (2026)
Parametric Integration with Neural Integral Operators
by: Schied, Christoph, et al.
Published: (2025)
by: Schied, Christoph, et al.
Published: (2025)
BoxFusion: Reconstruction‐Free Open‐Vocabulary 3D Object Detection via Real‐Time Multi‐View Box Fusion
by: Yuqing Lan, et al.
Published: (2025)
by: Yuqing Lan, et al.
Published: (2025)
PartMotionEdit: Fine-Grained Text-Driven 3D Human Motion Editing via Part-Level Modulation
by: Yang, Yujie, et al.
Published: (2025)
by: Yang, Yujie, et al.
Published: (2025)
Similar Items
-
Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation
by: Li, Daiqing, et al.
Published: (2024) -
Quantum Machine Learning Playground
by: Debus, Pascal, et al.
Published: (2025) -
GraphShaper: Geometry-aware Alignment for Improving Transfer Learning in Text-Attributed Graphs
by: Zhang, Heng, et al.
Published: (2025) -
From Delays to Densities: Exploring Data Uncertainty through Speech, Text, and Visualization
by: Chase Stokes, et al.
Published: (2024) -
Can Representation Gaps Be the Key to Enhancing Robustness in Graph-Text Alignment?
by: Zhang, Heng, et al.
Published: (2025)