Context Unrolling in Omni Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Ceyuan, Lin, Zhijie, Zhao, Yang, Xiao, Fei, He, Hao, Zhao, Qi, Deng, Chaorui, Li, Kunchang, Ding, Zihan, Guo, Yuwei, Wang, Fuyun, Zhu, Fangqi, Nie, Xiaonan, Zhu, Shenhan, Lin, Shanchuan, Li, Hongsheng, Huang, Weilin, Shi, Guang, Fan, Haoqi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Continuous Adversarial Flow Models
von: Lin, Shanchuan, et al.
Veröffentlicht: (2026)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2026)
Adversarial Flow Models
von: Lin, Shanchuan, et al.
Veröffentlicht: (2025)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2025)
Causal Diffusion Transformers for Generative Modeling
von: Deng, Chaorui, et al.
Veröffentlicht: (2024)
von: Deng, Chaorui, et al.
Veröffentlicht: (2024)
Representation Forcing for Bottleneck-Free Unified Multimodal Models
von: Wang, Yuqing, et al.
Veröffentlicht: (2026)
von: Wang, Yuqing, et al.
Veröffentlicht: (2026)
Emerging Properties in Unified Multimodal Pretraining
von: Deng, Chaorui, et al.
Veröffentlicht: (2025)
von: Deng, Chaorui, et al.
Veröffentlicht: (2025)
Long Context Tuning for Video Generation
von: Guo, Yuwei, et al.
Veröffentlicht: (2025)
von: Guo, Yuwei, et al.
Veröffentlicht: (2025)
UniGRPO: Unified Policy Optimization for Reasoning-Driven Visual Generation
von: Liu, Jie, et al.
Veröffentlicht: (2026)
von: Liu, Jie, et al.
Veröffentlicht: (2026)
End-to-End Training for Autoregressive Video Diffusion via Self-Resampling
von: Guo, Yuwei, et al.
Veröffentlicht: (2025)
von: Guo, Yuwei, et al.
Veröffentlicht: (2025)
CameraCtrl II: Dynamic Scene Exploration via Camera-controlled Video Diffusion Models
von: He, Hao, et al.
Veröffentlicht: (2025)
von: He, Hao, et al.
Veröffentlicht: (2025)
Diffusion Adversarial Post-Training for One-Step Video Generation
von: Lin, Shanchuan, et al.
Veröffentlicht: (2025)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2025)
AnimateDiff-Lightning: Cross-Model Diffusion Distillation
von: Lin, Shanchuan, et al.
Veröffentlicht: (2024)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2024)
Diffusion Model with Perceptual Loss
von: Lin, Shanchuan, et al.
Veröffentlicht: (2023)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2023)
Autoregressive Adversarial Post-Training for Real-Time Interactive Video Generation
von: Lin, Shanchuan, et al.
Veröffentlicht: (2025)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2025)
SeedVR: Seeding Infinity in Diffusion Transformer Towards Generic Video Restoration
von: Wang, Jianyi, et al.
Veröffentlicht: (2025)
von: Wang, Jianyi, et al.
Veröffentlicht: (2025)
LightFusion: A Light-weighted, Double Fusion Framework for Unified Multimodal Understanding and Generation
von: Wang, Zeyu, et al.
Veröffentlicht: (2025)
von: Wang, Zeyu, et al.
Veröffentlicht: (2025)
CameraCtrl: Enabling Camera Control for Text-to-Video Generation
von: He, Hao, et al.
Veröffentlicht: (2024)
von: He, Hao, et al.
Veröffentlicht: (2024)
SeedVR2: One-Step Video Restoration via Diffusion Adversarial Post-Training
von: Wang, Jianyi, et al.
Veröffentlicht: (2025)
von: Wang, Jianyi, et al.
Veröffentlicht: (2025)
LSH-MoE: Communication-efficient MoE Training via Locality-Sensitive Hashing
von: Nie, Xiaonan, et al.
Veröffentlicht: (2024)
von: Nie, Xiaonan, et al.
Veröffentlicht: (2024)
SDXL-Lightning: Progressive Adversarial Diffusion Distillation
von: Lin, Shanchuan, et al.
Veröffentlicht: (2024)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2024)
Common Diffusion Noise Schedules and Sample Steps are Flawed
von: Lin, Shanchuan, et al.
Veröffentlicht: (2023)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2023)
Improving Automatic Parallel Training via Balanced Memory Workload Optimization
von: Wang, Yujie, et al.
Veröffentlicht: (2023)
von: Wang, Yujie, et al.
Veröffentlicht: (2023)
VQ-VA World: Towards High-Quality Visual Question-Visual Answering
von: Gou, Chenhui, et al.
Veröffentlicht: (2025)
von: Gou, Chenhui, et al.
Veröffentlicht: (2025)
Self-Supervised Cross-Encoder for Neurodegenerative Disease Diagnosis
von: Cheng, Fangqi, et al.
Veröffentlicht: (2025)
von: Cheng, Fangqi, et al.
Veröffentlicht: (2025)
Unleashing Scalable Context Parallelism for Foundation Models Pre-Training via FCP
von: Zhao, Yilong, et al.
Veröffentlicht: (2026)
von: Zhao, Yilong, et al.
Veröffentlicht: (2026)
Dynamic Spectral Denoising with Global-Context Attention for Multi-Behavior Recommendation
von: Cai, Miaomiao, et al.
Veröffentlicht: (2026)
von: Cai, Miaomiao, et al.
Veröffentlicht: (2026)
Memory-Efficient Gradient Unrolling for Large-Scale Bi-level Optimization
von: Shen, Qianli, et al.
Veröffentlicht: (2024)
von: Shen, Qianli, et al.
Veröffentlicht: (2024)
Unrolled Decomposed Unpaired Learning for Controllable Low-Light Video Enhancement
von: Zhu, Lingyu, et al.
Veröffentlicht: (2024)
von: Zhu, Lingyu, et al.
Veröffentlicht: (2024)
HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context
von: Yang, Qize, et al.
Veröffentlicht: (2025)
von: Yang, Qize, et al.
Veröffentlicht: (2025)
HIFICL: High-Fidelity In-Context Learning for Multimodal Tasks
von: Li, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Li, Xiaoyu, et al.
Veröffentlicht: (2026)
SAIL-Embedding Technical Report: Omni-modal Embedding Foundation Model
von: Lin, Lin, et al.
Veröffentlicht: (2025)
von: Lin, Lin, et al.
Veröffentlicht: (2025)
From Cd(SCN)2(CH4N2S)2 to Cd(SCN)2(C4H6N2)2: Controlling Sulfur Content in Thiocyanate Systems Significantly Improves the Overall Performance of UV Nonlinear Optical Materials
von: Yuwei Kang, et al.
Veröffentlicht: (2024)
von: Yuwei Kang, et al.
Veröffentlicht: (2024)
Seedance 1.0: Exploring the Boundaries of Video Generation Models
von: Gao, Yu, et al.
Veröffentlicht: (2025)
von: Gao, Yu, et al.
Veröffentlicht: (2025)
Mixture of Contexts for Long Video Generation
von: Cai, Shengqu, et al.
Veröffentlicht: (2025)
von: Cai, Shengqu, et al.
Veröffentlicht: (2025)
Revisiting 360 Depth Estimation with PanoGabor: A New Fusion Perspective
von: Shen, Zhijie, et al.
Veröffentlicht: (2024)
von: Shen, Zhijie, et al.
Veröffentlicht: (2024)
Revisiting Monocular 3D Object Detection with Depth Thickness Field
von: Zhang, Qiude, et al.
Veröffentlicht: (2024)
von: Zhang, Qiude, et al.
Veröffentlicht: (2024)
OmniDiagram: Advancing Unified Diagram Code Generation via Visual Interrogation Reward
von: Yang, Haoyue, et al.
Veröffentlicht: (2026)
von: Yang, Haoyue, et al.
Veröffentlicht: (2026)
A Canonical Transformation for the Anderson Lattice Hamiltonian with f–f Electron Coupling
von: Guang-Lin Zhao
Veröffentlicht: (2024)
von: Guang-Lin Zhao
Veröffentlicht: (2024)
Omni-Reward: Towards Generalist Omni-Modal Reward Modeling with Free-Form Preferences
von: Jin, Zhuoran, et al.
Veröffentlicht: (2025)
von: Jin, Zhuoran, et al.
Veröffentlicht: (2025)
OmniGUI: Benchmarking GUI Agents in Omni-Modal Smartphone Environments
von: Henry, Felix, et al.
Veröffentlicht: (2026)
von: Henry, Felix, et al.
Veröffentlicht: (2026)
Unrolling Plug-and-Play Network for Hyperspectral Unmixing
von: Zhao, Min, et al.
Veröffentlicht: (2024)
von: Zhao, Min, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Continuous Adversarial Flow Models
von: Lin, Shanchuan, et al.
Veröffentlicht: (2026) -
Adversarial Flow Models
von: Lin, Shanchuan, et al.
Veröffentlicht: (2025) -
Causal Diffusion Transformers for Generative Modeling
von: Deng, Chaorui, et al.
Veröffentlicht: (2024) -
Representation Forcing for Bottleneck-Free Unified Multimodal Models
von: Wang, Yuqing, et al.
Veröffentlicht: (2026) -
Emerging Properties in Unified Multimodal Pretraining
von: Deng, Chaorui, et al.
Veröffentlicht: (2025)