Tele-Omni: a Unified Multimodal Framework for Video Generation and Editing
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Jialun, Li, Tian, Cao, Xiao, Ma, Yukuo, Shang, Gonghu, Huang, Haibin, Zhang, Chi, Chang, Xiangzhen, Huang, Zhiyong, Hu, Jiakui, Li, Zuoxin, Liang, Yuanzhi, Liu, Cong, Liu, Junqi, Tan, Robby T., Tang, Haitong, Weng, Qizhen, Xu, Yifan, Yang, Liying, Yang, Xiaoyan, Yu, Peng, Zhang, Shiwen, Li, Xuelong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
TeleStyle: Content-Preserving Style Transfer in Images and Videos
di: Zhang, Shiwen, et al.
Pubblicazione: (2026)
di: Zhang, Shiwen, et al.
Pubblicazione: (2026)
TeleWorld: Towards Dynamic Multimodal Synthesis with a 4D World Model
di: Chen, Yabo, et al.
Pubblicazione: (2025)
di: Chen, Yabo, et al.
Pubblicazione: (2025)
UniModel: A Visual-Only Framework for Unified Multimodal Understanding and Generation
di: Zhang, Chi, et al.
Pubblicazione: (2025)
di: Zhang, Chi, et al.
Pubblicazione: (2025)
TempoMaster: Efficient Long Video Generation via Next-Frame-Rate Prediction
di: Ma, Yukuo, et al.
Pubblicazione: (2025)
di: Ma, Yukuo, et al.
Pubblicazione: (2025)
Spatial-Temporal State Propagation Autoregressive Model for 4D Object Generation
di: Yang, Liying, et al.
Pubblicazione: (2026)
di: Yang, Liying, et al.
Pubblicazione: (2026)
TeleBoost: A Systematic Alignment Framework for High-Fidelity, Controllable, and Robust Video Generation
di: Liang, Yuanzhi, et al.
Pubblicazione: (2026)
di: Liang, Yuanzhi, et al.
Pubblicazione: (2026)
QwenStyle: Content-Preserving Style Transfer with Qwen-Image-Edit
di: Zhang, Shiwen, et al.
Pubblicazione: (2026)
di: Zhang, Shiwen, et al.
Pubblicazione: (2026)
Computation-Bandwidth-Memory Trade-offs: A Unified Paradigm for AI Infrastructure
di: Fan, Yuankai, et al.
Pubblicazione: (2025)
di: Fan, Yuankai, et al.
Pubblicazione: (2025)
Product market competition and disclosure content differentiation: A topic modeling analysis
di: Yongqiang Chu, et al.
Pubblicazione: (2024)
di: Yongqiang Chu, et al.
Pubblicazione: (2024)
Macro-from-Micro Planning for High-Quality and Parallelized Autoregressive Long Video Generation
di: Xiang, Xunzhi, et al.
Pubblicazione: (2025)
di: Xiang, Xunzhi, et al.
Pubblicazione: (2025)
Seeing What Matters: Visual Preference Policy Optimization for Visual Generation
di: Ni, Ziqi, et al.
Pubblicazione: (2025)
di: Ni, Ziqi, et al.
Pubblicazione: (2025)
OmniVDiff: Omni Controllable Video Diffusion for Generation and Understanding
di: Xi, Dianbing, et al.
Pubblicazione: (2025)
di: Xi, Dianbing, et al.
Pubblicazione: (2025)
CtrlVDiff: Controllable Video Generation via Unified Multimodal Video Diffusion
di: Xi, Dianbing, et al.
Pubblicazione: (2025)
di: Xi, Dianbing, et al.
Pubblicazione: (2025)
TelePhysics: Physics-Grounded Multi-Object Scene Generation from a Single Image with Real-Time Interaction
di: Zhang, Xin, et al.
Pubblicazione: (2026)
di: Zhang, Xin, et al.
Pubblicazione: (2026)
Geometry-as-context: Modulating Explicit 3D in Scene-consistent Video Generation to Geometry Context
di: Hu, JiaKui, et al.
Pubblicazione: (2026)
di: Hu, JiaKui, et al.
Pubblicazione: (2026)
FSSD: Feature Fusion Single Shot Multibox Detector
di: Li, Zuoxin, et al.
Pubblicazione: (2017)
di: Li, Zuoxin, et al.
Pubblicazione: (2017)
Learning What to Trust: Bayesian Prior-Guided Optimization for Visual Generation
di: Liu, Ruiying, et al.
Pubblicazione: (2025)
di: Liu, Ruiying, et al.
Pubblicazione: (2025)
Tele-FLM Technical Report
di: Li, Xiang, et al.
Pubblicazione: (2024)
di: Li, Xiang, et al.
Pubblicazione: (2024)
PRTS: A Primitive Reasoning and Tasking System via Contrastive Representations
di: Zhang, Yang, et al.
Pubblicazione: (2026)
di: Zhang, Yang, et al.
Pubblicazione: (2026)
Rethinking Reward Signals in Video GRPO: When Scores Become Targets
di: Li, Rui, et al.
Pubblicazione: (2025)
di: Li, Rui, et al.
Pubblicazione: (2025)
TeleChat Technical Report
di: He, Zhongjiang, et al.
Pubblicazione: (2024)
di: He, Zhongjiang, et al.
Pubblicazione: (2024)
Point2Insert: Video Object Insertion via Sparse Point Guidance
di: Zhou, Yu, et al.
Pubblicazione: (2026)
di: Zhou, Yu, et al.
Pubblicazione: (2026)
NFIG: Multi-Scale Autoregressive Image Generation via Frequency Ordering
di: Huang, Zhihao, et al.
Pubblicazione: (2025)
di: Huang, Zhihao, et al.
Pubblicazione: (2025)
Reward-Aware Trajectory Shaping for Few-step Visual Generation
di: Li, Rui, et al.
Pubblicazione: (2026)
di: Li, Rui, et al.
Pubblicazione: (2026)
Uni-Inter: Unifying 3D Human Motion Synthesis Across Diverse Interaction Contexts
di: Liu, Sheng, et al.
Pubblicazione: (2025)
di: Liu, Sheng, et al.
Pubblicazione: (2025)
InterSyn: Interleaved Learning for Dynamic Motion Synthesis in the Wild
di: Ma, Yiyi, et al.
Pubblicazione: (2025)
di: Ma, Yiyi, et al.
Pubblicazione: (2025)
52B to 1T: Lessons Learned via Tele-FLM Series
di: Li, Xiang, et al.
Pubblicazione: (2024)
di: Li, Xiang, et al.
Pubblicazione: (2024)
Technical Report of TeleChat2, TeleChat2.5 and T1
di: Wang, Zihan, et al.
Pubblicazione: (2025)
di: Wang, Zihan, et al.
Pubblicazione: (2025)
VAST 1.0: A Unified Framework for Controllable and Consistent Video Generation
di: Zhang, Chi, et al.
Pubblicazione: (2024)
di: Zhang, Chi, et al.
Pubblicazione: (2024)
Two-timescale Derivative Free Optimization for Performative Prediction with Markovian Data
di: Liu, Haitong, et al.
Pubblicazione: (2023)
di: Liu, Haitong, et al.
Pubblicazione: (2023)
Training Report of TeleChat3-MoE
di: Liu, Xinzhang, et al.
Pubblicazione: (2025)
di: Liu, Xinzhang, et al.
Pubblicazione: (2025)
Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale
di: Zhang, Liujie, et al.
Pubblicazione: (2026)
di: Zhang, Liujie, et al.
Pubblicazione: (2026)
Securing High-Concurrency Ticket Sales: A Framework Based on Microservice
di: Zhang, Zhiyong, et al.
Pubblicazione: (2025)
di: Zhang, Zhiyong, et al.
Pubblicazione: (2025)
TeleEgo: Benchmarking Egocentric AI Assistants in the Wild
di: Yan, Jiaqi, et al.
Pubblicazione: (2025)
di: Yan, Jiaqi, et al.
Pubblicazione: (2025)
TeleAI-Safety: A comprehensive LLM jailbreaking benchmark towards attacks, defenses, and evaluations
di: Chen, Xiuyuan, et al.
Pubblicazione: (2025)
di: Chen, Xiuyuan, et al.
Pubblicazione: (2025)
OmniBench: Towards The Future of Universal Omni-Language Models
di: Li, Yizhi, et al.
Pubblicazione: (2024)
di: Li, Yizhi, et al.
Pubblicazione: (2024)
Learning to Credit the Right Steps: Objective-aware Process Optimization for Visual Generation
di: Li, Rui, et al.
Pubblicazione: (2026)
di: Li, Rui, et al.
Pubblicazione: (2026)
Haptic-Based User Authentication for Tele-robotic System
di: Yu, Rongyu, et al.
Pubblicazione: (2025)
di: Yu, Rongyu, et al.
Pubblicazione: (2025)
Enhance Vision-Language Alignment with Noise
di: Huang, Sida, et al.
Pubblicazione: (2024)
di: Huang, Sida, et al.
Pubblicazione: (2024)
Tele-Correlation: Calibrating Shear-Shear Correlation with Real Data
di: Shen, Zhi, et al.
Pubblicazione: (2024)
di: Shen, Zhi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
TeleStyle: Content-Preserving Style Transfer in Images and Videos
di: Zhang, Shiwen, et al.
Pubblicazione: (2026) -
TeleWorld: Towards Dynamic Multimodal Synthesis with a 4D World Model
di: Chen, Yabo, et al.
Pubblicazione: (2025) -
UniModel: A Visual-Only Framework for Unified Multimodal Understanding and Generation
di: Zhang, Chi, et al.
Pubblicazione: (2025) -
TempoMaster: Efficient Long Video Generation via Next-Frame-Rate Prediction
di: Ma, Yukuo, et al.
Pubblicazione: (2025) -
Spatial-Temporal State Propagation Autoregressive Model for 4D Object Generation
di: Yang, Liying, et al.
Pubblicazione: (2026)