Skywork UniPic 3.0: Unified Multi-Image Composition via Sequence Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wei, Hongyang, Liu, Hongbo, Wang, Zidong, Peng, Yi, Xu, Baixin, Wu, Size, Zhang, Xuying, He, Xianglong, Liu, Zexiang, Wang, Peiyu, Song, Xuchen, Li, Yangguang, Liu, Yang, Zhou, Yahui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Skywork UniPic 2.0: Building Kontext Model with Online RL for Unified Multimodal Model
von: Wei, Hongyang, et al.
Veröffentlicht: (2025)
von: Wei, Hongyang, et al.
Veröffentlicht: (2025)
Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation
von: Wang, Peiyu, et al.
Veröffentlicht: (2025)
von: Wang, Peiyu, et al.
Veröffentlicht: (2025)
Skywork-R1V3 Technical Report
von: Shen, Wei, et al.
Veröffentlicht: (2025)
von: Shen, Wei, et al.
Veröffentlicht: (2025)
Advances in GRPO for Generation Models: A Survey
von: Liu, Zexiang, et al.
Veröffentlicht: (2026)
von: Liu, Zexiang, et al.
Veröffentlicht: (2026)
Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning
von: Wang, Xiaokun, et al.
Veröffentlicht: (2025)
von: Wang, Xiaokun, et al.
Veröffentlicht: (2025)
Matrix-game 2.0: An open-source real-time and streaming interactive world model
von: He, Xianglong, et al.
Veröffentlicht: (2025)
von: He, Xianglong, et al.
Veröffentlicht: (2025)
Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning
von: Wang, Peiyu, et al.
Veröffentlicht: (2025)
von: Wang, Peiyu, et al.
Veröffentlicht: (2025)
Matrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory
von: Wang, Zile, et al.
Veröffentlicht: (2026)
von: Wang, Zile, et al.
Veröffentlicht: (2026)
Skywork R1V: Pioneering Multimodal Reasoning with Chain-of-Thought
von: Peng, Yi, et al.
Veröffentlicht: (2025)
von: Peng, Yi, et al.
Veröffentlicht: (2025)
Skywork-R1V4: Toward Agentic Multimodal Intelligence through Interleaved Thinking with Images and DeepResearch
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
Skywork-SWE: Unveiling Data Scaling Laws for Software Engineering in LLMs
von: Zeng, Liang, et al.
Veröffentlicht: (2025)
von: Zeng, Liang, et al.
Veröffentlicht: (2025)
Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs
von: Liu, Chris Yuhao, et al.
Veröffentlicht: (2024)
von: Liu, Chris Yuhao, et al.
Veröffentlicht: (2024)
UniDream: Unifying Diffusion Priors for Relightable Text-to-3D Generation
von: Liu, Zexiang, et al.
Veröffentlicht: (2023)
von: Liu, Zexiang, et al.
Veröffentlicht: (2023)
Skywork Open Reasoner 1 Technical Report
von: He, Jujie, et al.
Veröffentlicht: (2025)
von: He, Jujie, et al.
Veröffentlicht: (2025)
UniReason 1.0: A Unified Reasoning Framework for World Knowledge Aligned Image Generation and Editing
von: Wang, Dianyi, et al.
Veröffentlicht: (2026)
von: Wang, Dianyi, et al.
Veröffentlicht: (2026)
Skywork-Reward-V2: Scaling Preference Data Curation via Human-AI Synergy
von: Liu, Chris Yuhao, et al.
Veröffentlicht: (2025)
von: Liu, Chris Yuhao, et al.
Veröffentlicht: (2025)
ShapeGen: Towards High-Quality 3D Shape Synthesis
von: Li, Yangguang, et al.
Veröffentlicht: (2025)
von: Li, Yangguang, et al.
Veröffentlicht: (2025)
MeshCraft: Exploring Efficient and Controllable Mesh Generation with Flow-based DiTs
von: He, Xianglong, et al.
Veröffentlicht: (2025)
von: He, Xianglong, et al.
Veröffentlicht: (2025)
UniEP: Unified Expert-Parallel MoE MegaKernel for LLM Training
von: Zheng, Size, et al.
Veröffentlicht: (2026)
von: Zheng, Size, et al.
Veröffentlicht: (2026)
Skywork-Math: Data Scaling Laws for Mathematical Reasoning in Large Language Models -- The Story Goes On
von: Zeng, Liang, et al.
Veröffentlicht: (2024)
von: Zeng, Liang, et al.
Veröffentlicht: (2024)
LongSkywork: A Training Recipe for Efficiently Extending Context Length in Large Language Models
von: Zhao, Liang, et al.
Veröffentlicht: (2024)
von: Zhao, Liang, et al.
Veröffentlicht: (2024)
Uni-MMMU: A Massive Multi-discipline Multimodal Unified Benchmark
von: Zou, Kai, et al.
Veröffentlicht: (2025)
von: Zou, Kai, et al.
Veröffentlicht: (2025)
UniFormer: Unifying Convolution and Self-attention for Visual Recognition
von: Li, Kunchang, et al.
Veröffentlicht: (2022)
von: Li, Kunchang, et al.
Veröffentlicht: (2022)
UniMo: Unified Motion Generation and Understanding with Chain of Thought
von: Wang, Guocun, et al.
Veröffentlicht: (2026)
von: Wang, Guocun, et al.
Veröffentlicht: (2026)
OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation
von: Wu, Size, et al.
Veröffentlicht: (2025)
von: Wu, Size, et al.
Veröffentlicht: (2025)
Skywork-MoE: A Deep Dive into Training Techniques for Mixture-of-Experts Language Models
von: Wei, Tianwen, et al.
Veröffentlicht: (2024)
von: Wei, Tianwen, et al.
Veröffentlicht: (2024)
Particle manipulation by hydrodynamic effects in vortical Stokes flow
von: Liu, Xuchen
Veröffentlicht: (2025)
von: Liu, Xuchen
Veröffentlicht: (2025)
Uni-Edit: Intelligent Editing Is A General Task For Unified Model Tuning
von: Zheng, Dian, et al.
Veröffentlicht: (2026)
von: Zheng, Dian, et al.
Veröffentlicht: (2026)
Cryptocephalus inhumeralis Pic 1922
von: Duan, Wen-Yuan, et al.
Veröffentlicht: (2025)
von: Duan, Wen-Yuan, et al.
Veröffentlicht: (2025)
UniAlignment: Semantic Alignment for Unified Image Generation, Understanding, Manipulation and Perception
von: Song, Xinyang, et al.
Veröffentlicht: (2025)
von: Song, Xinyang, et al.
Veröffentlicht: (2025)
UniLS: End-to-End Audio-Driven Avatars for Unified Listening and Speaking
von: Chu, Xuangeng, et al.
Veröffentlicht: (2025)
von: Chu, Xuangeng, et al.
Veröffentlicht: (2025)
ResFormer: All-Time Reservoir Memory for Long Sequence Classification
von: Liu, Hongbo, et al.
Veröffentlicht: (2025)
von: Liu, Hongbo, et al.
Veröffentlicht: (2025)
UniHM: Unified Dexterous Hand Manipulation with Vision Language Model
von: Zhang, Zhenhao, et al.
Veröffentlicht: (2026)
von: Zhang, Zhenhao, et al.
Veröffentlicht: (2026)
UniVideo: Unified Understanding, Generation, and Editing for Videos
von: Wei, Cong, et al.
Veröffentlicht: (2025)
von: Wei, Cong, et al.
Veröffentlicht: (2025)
UniShield: Unified Face Attack Detection via KG-Informed Multimodal Reasoning
von: Li, Hongrui, et al.
Veröffentlicht: (2026)
von: Li, Hongrui, et al.
Veröffentlicht: (2026)
UniMotion: A Unified Framework for Motion-Text-Vision Understanding and Generation
von: Wang, Ziyi, et al.
Veröffentlicht: (2026)
von: Wang, Ziyi, et al.
Veröffentlicht: (2026)
UniAudio 2.0: A Unified Audio Language Model with Text-Aligned Factorized Audio Tokenization
von: Yang, Dongchao, et al.
Veröffentlicht: (2026)
von: Yang, Dongchao, et al.
Veröffentlicht: (2026)
Uni-Animator: Towards Unified Visual Colorization
von: Chen, Xinyuan, et al.
Veröffentlicht: (2026)
von: Chen, Xinyuan, et al.
Veröffentlicht: (2026)
Squrve: A Unified and Modular Framework for Complex Real-World Text-to-SQL Tasks
von: Wang, Yihan, et al.
Veröffentlicht: (2025)
von: Wang, Yihan, et al.
Veröffentlicht: (2025)
Does Unification Come at a Cost? Uni-SafeBench: A Safety Benchmark for Unified Multimodal Large Models
von: Peng, Zixiang, et al.
Veröffentlicht: (2026)
von: Peng, Zixiang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Skywork UniPic 2.0: Building Kontext Model with Online RL for Unified Multimodal Model
von: Wei, Hongyang, et al.
Veröffentlicht: (2025) -
Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation
von: Wang, Peiyu, et al.
Veröffentlicht: (2025) -
Skywork-R1V3 Technical Report
von: Shen, Wei, et al.
Veröffentlicht: (2025) -
Advances in GRPO for Generation Models: A Survey
von: Liu, Zexiang, et al.
Veröffentlicht: (2026) -
Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning
von: Wang, Xiaokun, et al.
Veröffentlicht: (2025)