simpleposter: a simple baseline for product poster generation
Fuente:
arXiv
Saved in:
| Main Authors: | Cui, Benlei, Zeng, Fangao, Jiang, Weitao, Zhai, Yuwen, Hong, Haiwen, Huang, Longtao, Xue, Hui, Shang, Wenxiang, Huang, Pipei |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Diffusion Probe: Generated Image Result Prediction Using CNN Probes
by: Cui, Benlei, et al.
Published: (2026)
by: Cui, Benlei, et al.
Published: (2026)
TC-Padé: Trajectory-Consistent Padé Approximation for Diffusion Acceleration
by: Cui, Benlei, et al.
Published: (2026)
by: Cui, Benlei, et al.
Published: (2026)
Erase Diffusion: Empowering Object Removal Through Calibrating Diffusion Pathways
by: Liu, Yi, et al.
Published: (2025)
by: Liu, Yi, et al.
Published: (2025)
How LoRA Remembers? A Parametric Memory Law for LLM Finetuning
by: Xu, Ziwen, et al.
Published: (2026)
by: Xu, Ziwen, et al.
Published: (2026)
Attentive Eraser: Unleashing Diffusion Model's Object Removal Potential via Self-Attention Redirection Guidance
by: Sun, Wenhao, et al.
Published: (2024)
by: Sun, Wenhao, et al.
Published: (2024)
Boundary feature fusion network for tooth image segmentation
by: Zhang, Dongping, et al.
Published: (2024)
by: Zhang, Dongping, et al.
Published: (2024)
CME-CAD: Heterogeneous Collaborative Multi-Expert Reinforcement Learning for CAD Code Generation
by: Niu, Ke, et al.
Published: (2025)
by: Niu, Ke, et al.
Published: (2025)
Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers
by: Zhu, Jingyuan, et al.
Published: (2026)
by: Zhu, Jingyuan, et al.
Published: (2026)
Seeing but Not Thinking: Routing Distraction in Multimodal Mixture-of-Experts
by: Xu, Haolei, et al.
Published: (2026)
by: Xu, Haolei, et al.
Published: (2026)
Think When Needed: Adaptive Reasoning-Driven Multimodal Embeddings with a Dual-LoRA Architecture
by: Zhang, Longxiang, et al.
Published: (2026)
by: Zhang, Longxiang, et al.
Published: (2026)
Category Query Learning for Human-Object Interaction Classification
by: Xie, Chi, et al.
Published: (2023)
by: Xie, Chi, et al.
Published: (2023)
PromptEcho: Annotation-Free Reward from Vision-Language Models for Text-to-Image Reinforcement Learning
by: Liu, Jinlong, et al.
Published: (2026)
by: Liu, Jinlong, et al.
Published: (2026)
Denoising as Path Planning: Training-Free Acceleration of Diffusion Models with DPCache
by: Cui, Bowen, et al.
Published: (2026)
by: Cui, Bowen, et al.
Published: (2026)
Dynamic Mixture of Curriculum LoRA Experts for Continual Multimodal Instruction Tuning
by: Ge, Chendi, et al.
Published: (2025)
by: Ge, Chendi, et al.
Published: (2025)
A simple, strong baseline for building damage detection on the xBD dataset
by: Gerard, Sebastian, et al.
Published: (2024)
by: Gerard, Sebastian, et al.
Published: (2024)
3D StreetUnveiler with Semantic-aware 2DGS -- a simple baseline
by: Xu, Jingwei, et al.
Published: (2024)
by: Xu, Jingwei, et al.
Published: (2024)
Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics
by: Xu, Ziwen, et al.
Published: (2026)
by: Xu, Ziwen, et al.
Published: (2026)
Generative Video Matting
by: Ge, Yongtao, et al.
Published: (2025)
by: Ge, Yongtao, et al.
Published: (2025)
Making Training-Free Diffusion Segmentors Scale with the Generative Power
by: Meng, Benyuan, et al.
Published: (2026)
by: Meng, Benyuan, et al.
Published: (2026)
Renovating Names in Open-Vocabulary Segmentation Benchmarks
by: Huang, Haiwen, et al.
Published: (2024)
by: Huang, Haiwen, et al.
Published: (2024)
Alleviating Sparse Rewards by Modeling Step-Wise and Long-Term Sampling Effects in Flow-Based GRPO
by: Tong, Yunze, et al.
Published: (2026)
by: Tong, Yunze, et al.
Published: (2026)
Self-Improving 4D Perception via Self-Distillation
by: Huang, Nan, et al.
Published: (2026)
by: Huang, Nan, et al.
Published: (2026)
Deep Boosting Learning: A Brand-new Cooperative Approach for Image-Text Matching
by: Diao, Haiwen, et al.
Published: (2024)
by: Diao, Haiwen, et al.
Published: (2024)
Token Painter: Training-Free Text-Guided Image Inpainting via Mask Autoregressive Models
by: Jiang, Longtao, et al.
Published: (2025)
by: Jiang, Longtao, et al.
Published: (2025)
KARST: Multi-Kernel Kronecker Adaptation with Re-Scaling Transmission for Visual Classification
by: Zhu, Yue, et al.
Published: (2025)
by: Zhu, Yue, et al.
Published: (2025)
Dense Geometry Supervision for Underwater Depth Estimation
by: Gua, Wenxiang, et al.
Published: (2025)
by: Gua, Wenxiang, et al.
Published: (2025)
Benchmarking Feature Upsampling Methods for Vision Foundation Models using Interactive Segmentation
by: Havrylov, Volodymyr, et al.
Published: (2025)
by: Havrylov, Volodymyr, et al.
Published: (2025)
One-Dimensional Adapter to Rule Them All: Concepts, Diffusion Models and Erasing Applications
by: Lyu, Mengyao, et al.
Published: (2023)
by: Lyu, Mengyao, et al.
Published: (2023)
TextMaster: A Unified Framework for Realistic Text Editing via Glyph-Style Dual-Control
by: Yan, Zhenyu, et al.
Published: (2024)
by: Yan, Zhenyu, et al.
Published: (2024)
GSSF: Generalized Structural Sparse Function for Deep Cross-modal Metric Learning
by: Diao, Haiwen, et al.
Published: (2024)
by: Diao, Haiwen, et al.
Published: (2024)
Disrupting Hierarchical Reasoning: Adversarial Protection for Geographic Privacy in Multimodal Reasoning Models
by: Zhang, Jiaming, et al.
Published: (2025)
by: Zhang, Jiaming, et al.
Published: (2025)
SEDS: Semantically Enhanced Dual-Stream Encoder for Sign Language Retrieval
by: Jiang, Longtao, et al.
Published: (2024)
by: Jiang, Longtao, et al.
Published: (2024)
RoCo-Sim: Enhancing Roadside Collaborative Perception through Foreground Simulation
by: Du, Yuwen, et al.
Published: (2025)
by: Du, Yuwen, et al.
Published: (2025)
3DIS-FLUX: simple and efficient multi-instance generation with DiT rendering
by: Zhou, Dewei, et al.
Published: (2025)
by: Zhou, Dewei, et al.
Published: (2025)
Disentangle Nighttime Lens Flares: Self-supervised Generation-based Lens Flare Removal
by: He, Yuwen, et al.
Published: (2025)
by: He, Yuwen, et al.
Published: (2025)
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding
by: Zhang, Peng, et al.
Published: (2026)
by: Zhang, Peng, et al.
Published: (2026)
Depth as Points: Center Point-based Depth Estimation
by: Tu, Zhiheng, et al.
Published: (2025)
by: Tu, Zhiheng, et al.
Published: (2025)
Unveiling Encoder-Free Vision-Language Models
by: Diao, Haiwen, et al.
Published: (2024)
by: Diao, Haiwen, et al.
Published: (2024)
MindShot: A Few-Shot Brain Decoding Framework via Transferring Cross-Subject Prior and Distilling Frequency Domain Knowledge
by: Jiang, Shuai, et al.
Published: (2024)
by: Jiang, Shuai, et al.
Published: (2024)
Structure-guided Diffusion Transformer for Low-Light Image Enhancement
by: Yin, Xiangchen, et al.
Published: (2025)
by: Yin, Xiangchen, et al.
Published: (2025)
Similar Items
-
Diffusion Probe: Generated Image Result Prediction Using CNN Probes
by: Cui, Benlei, et al.
Published: (2026) -
TC-Padé: Trajectory-Consistent Padé Approximation for Diffusion Acceleration
by: Cui, Benlei, et al.
Published: (2026) -
Erase Diffusion: Empowering Object Removal Through Calibrating Diffusion Pathways
by: Liu, Yi, et al.
Published: (2025) -
How LoRA Remembers? A Parametric Memory Law for LLM Finetuning
by: Xu, Ziwen, et al.
Published: (2026) -
Attentive Eraser: Unleashing Diffusion Model's Object Removal Potential via Self-Attention Redirection Guidance
by: Sun, Wenhao, et al.
Published: (2024)