Skywork UniPic 2.0: Building Kontext Model with Online RL for Unified Multimodal Model
Fuente:
arXiv
Salvato in:
| Autori principali: | Wei, Hongyang, Xu, Baixin, Liu, Hongbo, Wu, Size, Liu, Jie, Peng, Yi, Wang, Peiyu, Liu, Zexiang, He, Jingwen, Xietian, Yidan, Tang, Chuanxin, Wang, Zidong, Wei, Yichen, Hu, Liang, Jiang, Boyi, Li, Wei, He, Ying, Liu, Yang, Song, Xuchen, Li, Yangguang, Zhou, Yahui |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Skywork UniPic 3.0: Unified Multi-Image Composition via Sequence Modeling
di: Wei, Hongyang, et al.
Pubblicazione: (2026)
di: Wei, Hongyang, et al.
Pubblicazione: (2026)
Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation
di: Wang, Peiyu, et al.
Pubblicazione: (2025)
di: Wang, Peiyu, et al.
Pubblicazione: (2025)
Matrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory
di: Wang, Zile, et al.
Pubblicazione: (2026)
di: Wang, Zile, et al.
Pubblicazione: (2026)
Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning
di: Wang, Peiyu, et al.
Pubblicazione: (2025)
di: Wang, Peiyu, et al.
Pubblicazione: (2025)
Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning
di: Wang, Xiaokun, et al.
Pubblicazione: (2025)
di: Wang, Xiaokun, et al.
Pubblicazione: (2025)
Skywork-R1V3 Technical Report
di: Shen, Wei, et al.
Pubblicazione: (2025)
di: Shen, Wei, et al.
Pubblicazione: (2025)
Advances in GRPO for Generation Models: A Survey
di: Liu, Zexiang, et al.
Pubblicazione: (2026)
di: Liu, Zexiang, et al.
Pubblicazione: (2026)
Skywork-SWE: Unveiling Data Scaling Laws for Software Engineering in LLMs
di: Zeng, Liang, et al.
Pubblicazione: (2025)
di: Zeng, Liang, et al.
Pubblicazione: (2025)
Skywork R1V: Pioneering Multimodal Reasoning with Chain-of-Thought
di: Peng, Yi, et al.
Pubblicazione: (2025)
di: Peng, Yi, et al.
Pubblicazione: (2025)
Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs
di: Liu, Chris Yuhao, et al.
Pubblicazione: (2024)
di: Liu, Chris Yuhao, et al.
Pubblicazione: (2024)
Skywork-R1V4: Toward Agentic Multimodal Intelligence through Interleaved Thinking with Images and DeepResearch
di: Zhang, Yifan, et al.
Pubblicazione: (2025)
di: Zhang, Yifan, et al.
Pubblicazione: (2025)
Skywork Open Reasoner 1 Technical Report
di: He, Jujie, et al.
Pubblicazione: (2025)
di: He, Jujie, et al.
Pubblicazione: (2025)
Skywork-Math: Data Scaling Laws for Mathematical Reasoning in Large Language Models -- The Story Goes On
di: Zeng, Liang, et al.
Pubblicazione: (2024)
di: Zeng, Liang, et al.
Pubblicazione: (2024)
Matrix-game 2.0: An open-source real-time and streaming interactive world model
di: He, Xianglong, et al.
Pubblicazione: (2025)
di: He, Xianglong, et al.
Pubblicazione: (2025)
Skywork-Reward-V2: Scaling Preference Data Curation via Human-AI Synergy
di: Liu, Chris Yuhao, et al.
Pubblicazione: (2025)
di: Liu, Chris Yuhao, et al.
Pubblicazione: (2025)
LLM-Slice: Dedicated Wireless Network Slicing for Large Language Models
di: Liu, Boyi, et al.
Pubblicazione: (2024)
di: Liu, Boyi, et al.
Pubblicazione: (2024)
LongSkywork: A Training Recipe for Efficiently Extending Context Length in Large Language Models
di: Zhao, Liang, et al.
Pubblicazione: (2024)
di: Zhao, Liang, et al.
Pubblicazione: (2024)
UniVBench: Towards Unified Evaluation for Video Foundation Models
di: Wei, Jianhui, et al.
Pubblicazione: (2026)
di: Wei, Jianhui, et al.
Pubblicazione: (2026)
Uni-MMMU: A Massive Multi-discipline Multimodal Unified Benchmark
di: Zou, Kai, et al.
Pubblicazione: (2025)
di: Zou, Kai, et al.
Pubblicazione: (2025)
About Optimal Prefix Codes over Countably Infinite Alphabets: Probabilistic Intervals for the Codeword Lengths Assignment
di: Liu, Hongyang, et al.
Pubblicazione: (2026)
di: Liu, Hongyang, et al.
Pubblicazione: (2026)
Skywork-MoE: A Deep Dive into Training Techniques for Mixture-of-Experts Language Models
di: Wei, Tianwen, et al.
Pubblicazione: (2024)
di: Wei, Tianwen, et al.
Pubblicazione: (2024)
EdgeLoc: A Communication-Adaptive Parallel System for Real-Time Localization in Infrastructure-Assisted Autonomous Driving
di: Liu, Boyi, et al.
Pubblicazione: (2024)
di: Liu, Boyi, et al.
Pubblicazione: (2024)
UniDream: Unifying Diffusion Priors for Relightable Text-to-3D Generation
di: Liu, Zexiang, et al.
Pubblicazione: (2023)
di: Liu, Zexiang, et al.
Pubblicazione: (2023)
ShotBench: Expert-Level Cinematic Understanding in Vision-Language Models
di: Liu, Hongbo, et al.
Pubblicazione: (2025)
di: Liu, Hongbo, et al.
Pubblicazione: (2025)
Modular MeanFlow: Towards Stable and Scalable One-Step Generative Modeling
di: You, Haochen, et al.
Pubblicazione: (2025)
di: You, Haochen, et al.
Pubblicazione: (2025)
Particle manipulation by hydrodynamic effects in vortical Stokes flow
di: Liu, Xuchen
Pubblicazione: (2025)
di: Liu, Xuchen
Pubblicazione: (2025)
ShapeGen: Towards High-Quality 3D Shape Synthesis
di: Li, Yangguang, et al.
Pubblicazione: (2025)
di: Li, Yangguang, et al.
Pubblicazione: (2025)
MeshCraft: Exploring Efficient and Controllable Mesh Generation with Flow-based DiTs
di: He, Xianglong, et al.
Pubblicazione: (2025)
di: He, Xianglong, et al.
Pubblicazione: (2025)
DoTA: Weight-Decomposed Tensor Adaptation for Large Language Models
di: Hu, Xiaolin, et al.
Pubblicazione: (2024)
di: Hu, Xiaolin, et al.
Pubblicazione: (2024)
Perceive, Understand and Restore: Real-World Image Super-Resolution with Autoregressive Multimodal Generative Models
di: Wei, Hongyang, et al.
Pubblicazione: (2025)
di: Wei, Hongyang, et al.
Pubblicazione: (2025)
Regioselective [3 + 2] Cycloaddition Between Ynamides and Pyridine‐N‐Aminides Catalyzed by PicAuCl2 Anchored Onto SBA‐15
di: Boling Song, et al.
Pubblicazione: (2024)
di: Boling Song, et al.
Pubblicazione: (2024)
TreeRare: Syntax Tree-Guided Retrieval and Reasoning for Knowledge-Intensive Question Answering
di: Zhang, Boyi, et al.
Pubblicazione: (2025)
di: Zhang, Boyi, et al.
Pubblicazione: (2025)
Letter: Incremental Value and Outcome Modelling in Frailty Assessment for Older Patients With Inflammatory Bowel Disease
di: Chang Liu, et al.
Pubblicazione: (2026)
di: Chang Liu, et al.
Pubblicazione: (2026)
Hydrochromic Nanocapsule with Real Time Visual In Situ Water Level Sensing Function and Toughening Effect for Plastic Materials
di: Qin Wang, et al.
Pubblicazione: (2024)
di: Qin Wang, et al.
Pubblicazione: (2024)
Escher-Loop: Mutual Evolution by Closed-Loop Self-Referential Optimization
di: Liu, Ziyang, et al.
Pubblicazione: (2026)
di: Liu, Ziyang, et al.
Pubblicazione: (2026)
Simplifying CLIP: Unleashing the Power of Large-Scale Models on Consumer-level Computers
di: Liu, Hongbo
Pubblicazione: (2024)
di: Liu, Hongbo
Pubblicazione: (2024)
F-LMM: Grounding Frozen Large Multimodal Models
di: Wu, Size, et al.
Pubblicazione: (2024)
di: Wu, Size, et al.
Pubblicazione: (2024)
Enhancing Generalization in Medical Visual Question Answering Tasks via Gradient-Guided Model Perturbation
di: Liu, Gang, et al.
Pubblicazione: (2024)
di: Liu, Gang, et al.
Pubblicazione: (2024)
Beyond Static Pipelines: Learning Dynamic Workflows for Text-to-SQL
di: Wang, Yihan, et al.
Pubblicazione: (2026)
di: Wang, Yihan, et al.
Pubblicazione: (2026)
StreamUni: Achieving Streaming Speech Translation with a Unified Large Speech-Language Model
di: Guo, Shoutao, et al.
Pubblicazione: (2025)
di: Guo, Shoutao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Skywork UniPic 3.0: Unified Multi-Image Composition via Sequence Modeling
di: Wei, Hongyang, et al.
Pubblicazione: (2026) -
Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation
di: Wang, Peiyu, et al.
Pubblicazione: (2025) -
Matrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory
di: Wang, Zile, et al.
Pubblicazione: (2026) -
Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning
di: Wang, Peiyu, et al.
Pubblicazione: (2025) -
Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning
di: Wang, Xiaokun, et al.
Pubblicazione: (2025)