ChArtist: Generating Pictorial Charts with Unified Spatial and Subject Control
Fuente:
arXiv
Guardado en:
| Autores principales: | Xiao, Shishi, Zhou, Tongyu, Laidlaw, David, Chan, Gromit Yeuk-Yin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Omni-Effects: Unified and Spatially-Controllable Visual Effects Generation
por: Mao, Fangyuan, et al.
Publicado: (2025)
por: Mao, Fangyuan, et al.
Publicado: (2025)
START: Spatial and Textual Learning for Chart Understanding
por: Liu, Zhuoming, et al.
Publicado: (2025)
por: Liu, Zhuoming, et al.
Publicado: (2025)
MeshAnything: Artist-Created Mesh Generation with Autoregressive Transformers
por: Chen, Yiwen, et al.
Publicado: (2024)
por: Chen, Yiwen, et al.
Publicado: (2024)
Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling
por: Wang, Yuran, et al.
Publicado: (2025)
por: Wang, Yuran, et al.
Publicado: (2025)
mChartQA: A universal benchmark for multimodal Chart Question Answer based on Vision-Language Alignment and Reasoning
por: Wei, Jingxuan, et al.
Publicado: (2024)
por: Wei, Jingxuan, et al.
Publicado: (2024)
ChartM$^3$: Benchmarking Chart Editing with Multimodal Instructions
por: Yang, Donglu, et al.
Publicado: (2025)
por: Yang, Donglu, et al.
Publicado: (2025)
AskChart: Universal Chart Understanding through Textual Enhancement
por: Yang, Xudong, et al.
Publicado: (2024)
por: Yang, Xudong, et al.
Publicado: (2024)
OmniGen: Unified Image Generation
por: Xiao, Shitao, et al.
Publicado: (2024)
por: Xiao, Shitao, et al.
Publicado: (2024)
MeshAnything V2: Artist-Created Mesh Generation With Adjacent Mesh Tokenization
por: Chen, Yiwen, et al.
Publicado: (2024)
por: Chen, Yiwen, et al.
Publicado: (2024)
Video-As-Prompt: Unified Semantic Control for Video Generation
por: Bian, Yuxuan, et al.
Publicado: (2025)
por: Bian, Yuxuan, et al.
Publicado: (2025)
InfoChartQA: A Benchmark for Multimodal Question Answering on Infographic Charts
por: Xie, Tianchi, et al.
Publicado: (2025)
por: Xie, Tianchi, et al.
Publicado: (2025)
Chart-R1: Chain-of-Thought Supervision and Reinforcement for Advanced Chart Reasoner
por: Chen, Lei, et al.
Publicado: (2025)
por: Chen, Lei, et al.
Publicado: (2025)
Uni-RS: A Spatially Faithful Unified Understanding and Generation Model for Remote Sensing
por: Zhang, Weiyu, et al.
Publicado: (2026)
por: Zhang, Weiyu, et al.
Publicado: (2026)
UniEval: Unified Holistic Evaluation for Unified Multimodal Understanding and Generation
por: Li, Yi, et al.
Publicado: (2025)
por: Li, Yi, et al.
Publicado: (2025)
Improved Iterative Refinement for Chart-to-Code Generation via Structured Instruction
por: Xu, Chengzhi, et al.
Publicado: (2025)
por: Xu, Chengzhi, et al.
Publicado: (2025)
Spatial-Aware Latent Initialization for Controllable Image Generation
por: Sun, Wenqiang, et al.
Publicado: (2024)
por: Sun, Wenqiang, et al.
Publicado: (2024)
SpatialReward: Verifiable Spatial Reward Modeling for Fine-Grained Spatial Consistency in Text-to-Image Generation
por: Zhou, Sashuai, et al.
Publicado: (2026)
por: Zhou, Sashuai, et al.
Publicado: (2026)
Latent Action Control for Reasoning-Guided Unified Image Generation
por: Zhai, Fuxiang, et al.
Publicado: (2026)
por: Zhai, Fuxiang, et al.
Publicado: (2026)
Negative-Guided Subject Fidelity Optimization for Zero-Shot Subject-Driven Generation
por: Shin, Chaehun, et al.
Publicado: (2025)
por: Shin, Chaehun, et al.
Publicado: (2025)
MAGREF: Masked Guidance for Any-Reference Video Generation with Subject Disentanglement
por: Deng, Yufan, et al.
Publicado: (2025)
por: Deng, Yufan, et al.
Publicado: (2025)
ChartGen: Scaling Chart Understanding Via Code-Guided Synthetic Chart Generation
por: Kondic, Jovana, et al.
Publicado: (2025)
por: Kondic, Jovana, et al.
Publicado: (2025)
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement
por: Hu, Teng, et al.
Publicado: (2025)
por: Hu, Teng, et al.
Publicado: (2025)
DiMSUM: Diffusion Mamba -- A Scalable and Unified Spatial-Frequency Method for Image Generation
por: Phung, Hao, et al.
Publicado: (2024)
por: Phung, Hao, et al.
Publicado: (2024)
Breaking the SFT Plateau: Multimodal Structured Reinforcement Learning for Chart-to-Code Generation
por: Chen, Lei, et al.
Publicado: (2025)
por: Chen, Lei, et al.
Publicado: (2025)
Unified Thinker: A General Reasoning Modular Core for Image Generation
por: Zhou, Sashuai, et al.
Publicado: (2026)
por: Zhou, Sashuai, et al.
Publicado: (2026)
DocSAM: Unified Document Image Segmentation via Query Decomposition and Heterogeneous Mixed Learning
por: Li, Xiao-Hui, et al.
Publicado: (2025)
por: Li, Xiao-Hui, et al.
Publicado: (2025)
SpaceControl: Introducing Test-Time Spatial Control to 3D Generative Modeling
por: Fedele, Elisabetta, et al.
Publicado: (2025)
por: Fedele, Elisabetta, et al.
Publicado: (2025)
OmniCam: Unified Multimodal Video Generation via Camera Control
por: Yang, Xiaoda, et al.
Publicado: (2025)
por: Yang, Xiaoda, et al.
Publicado: (2025)
ModalChorus: Visual Probing and Alignment of Multi-modal Embeddings via Modal Fusion Map
por: Ye, Yilin, et al.
Publicado: (2024)
por: Ye, Yilin, et al.
Publicado: (2024)
Pictorial and apictorial polygonal jigsaw puzzles from arbitrary number of crossing cuts
por: Shahar, Peleg Harel Ofir Itzhak, et al.
Publicado: (2020)
por: Shahar, Peleg Harel Ofir Itzhak, et al.
Publicado: (2020)
IdGlow: Dynamic Identity Modulation for Multi-Subject Generation
por: Cai, Honghao, et al.
Publicado: (2026)
por: Cai, Honghao, et al.
Publicado: (2026)
SmartSpatial: Enhancing the 3D Spatial Arrangement Capabilities of Stable Diffusion Models and Introducing a Novel 3D Spatial Evaluation Framework
por: Huang, Mao Xun, et al.
Publicado: (2025)
por: Huang, Mao Xun, et al.
Publicado: (2025)
DICE: Disentangling Artist Style from Content via Contrastive Subspace Decomposition in Diffusion Models
por: Zhang, Tong, et al.
Publicado: (2026)
por: Zhang, Tong, et al.
Publicado: (2026)
Enhancing Multimodal Unified Representations for Cross Modal Generalization
por: Huang, Hai, et al.
Publicado: (2024)
por: Huang, Hai, et al.
Publicado: (2024)
Guess the Unified Model: How Much Can We Recover from Generated Images?
por: Cekinmez, Jasin, et al.
Publicado: (2026)
por: Cekinmez, Jasin, et al.
Publicado: (2026)
7DGS: Unified Spatial-Temporal-Angular Gaussian Splatting
por: Gao, Zhongpai, et al.
Publicado: (2025)
por: Gao, Zhongpai, et al.
Publicado: (2025)
Benchmarking Multimodal RAG through a Chart-based Document Question-Answering Generation Framework
por: Yang, Yuming, et al.
Publicado: (2025)
por: Yang, Yuming, et al.
Publicado: (2025)
RetriBooru: Leakage-Free Retrieval of Conditions from Reference Images for Subject-Driven Generation
por: Tang, Haoran, et al.
Publicado: (2023)
por: Tang, Haoran, et al.
Publicado: (2023)
SSG-Dit: A Spatial Signal Guided Framework for Controllable Video Generation
por: Hu, Peng, et al.
Publicado: (2025)
por: Hu, Peng, et al.
Publicado: (2025)
Single Image Iterative Subject-driven Generation and Editing
por: Shpitzer, Yair, et al.
Publicado: (2025)
por: Shpitzer, Yair, et al.
Publicado: (2025)
Ejemplares similares
-
Omni-Effects: Unified and Spatially-Controllable Visual Effects Generation
por: Mao, Fangyuan, et al.
Publicado: (2025) -
START: Spatial and Textual Learning for Chart Understanding
por: Liu, Zhuoming, et al.
Publicado: (2025) -
MeshAnything: Artist-Created Mesh Generation with Autoregressive Transformers
por: Chen, Yiwen, et al.
Publicado: (2024) -
Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling
por: Wang, Yuran, et al.
Publicado: (2025) -
mChartQA: A universal benchmark for multimodal Chart Question Answer based on Vision-Language Alignment and Reasoning
por: Wei, Jingxuan, et al.
Publicado: (2024)