EVA01: Unified Native 3D Understanding and Generation via Mixture-of-Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Zongyuan, Yi, Mingjing, Ma, Wanli, Fan, Chenzhuo, Li, Bocheng, Liu, Baolin, Lou, Yuke, Song, Yingde, Xiong, Yongping, Guo, Zhengdong, Wang, Shimu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DirectL: Efficient Radiance Fields Rendering for 3D Light Field Displays
von: Yang, Zongyuan, et al.
Veröffentlicht: (2024)
von: Yang, Zongyuan, et al.
Veröffentlicht: (2024)
TextDiff: Mask-Guided Residual Diffusion Models for Scene Text Image Super-Resolution
von: Liu, Baolin, et al.
Veröffentlicht: (2023)
von: Liu, Baolin, et al.
Veröffentlicht: (2023)
Towards Scalable Training for Handwritten Mathematical Expression Recognition
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
EVA: Mixture-of-Experts Semantic Variant Alignment for Compositional Zero-Shot Learning
von: Zhang, Xiao, et al.
Veröffentlicht: (2025)
von: Zhang, Xiao, et al.
Veröffentlicht: (2025)
UniViTAR: Unified Vision Transformer with Native Resolution
von: Qiao, Limeng, et al.
Veröffentlicht: (2025)
von: Qiao, Limeng, et al.
Veröffentlicht: (2025)
Efficient Video to Audio Mapper with Visual Scene Detection
von: Yi, Mingjing, et al.
Veröffentlicht: (2024)
von: Yi, Mingjing, et al.
Veröffentlicht: (2024)
Thermal characteristics and microstructural evolution mechanism of coal oxidation under methane atmospheres
von: Rongkun Pan, et al.
Veröffentlicht: (2026)
von: Rongkun Pan, et al.
Veröffentlicht: (2026)
CDiT: Conditional Diffusion Transformer for Geometry-Aware Terahertz Cross Far- and Near-Field Channel Generation
von: Hu, Zhengdong, et al.
Veröffentlicht: (2026)
von: Hu, Zhengdong, et al.
Veröffentlicht: (2026)
EVA: Zero-shot Accurate Attributes and Multi-Object Video Editing
von: Yang, Xiangpeng, et al.
Veröffentlicht: (2024)
von: Yang, Xiangpeng, et al.
Veröffentlicht: (2024)
EM Approaches to Nonparametric Estimation for Mixture of Linear Regressions
von: Welbaum, Andrew, et al.
Veröffentlicht: (2025)
von: Welbaum, Andrew, et al.
Veröffentlicht: (2025)
(Sketch) Custrock: A Planetary Custody Architecture for Near-Earth Object Defense
von: Yuan, Mingjing
Veröffentlicht: (2026)
von: Yuan, Mingjing
Veröffentlicht: (2026)
(Sketch) Custrock: A Planetary Custody Architecture for Near-Earth Object Defense
von: Yuan, Mingjing
Veröffentlicht: (2026)
von: Yuan, Mingjing
Veröffentlicht: (2026)
Understanding Catastrophic Interference: On the Identifibility of Latent Representations
von: Li, Yuke, et al.
Veröffentlicht: (2025)
von: Li, Yuke, et al.
Veröffentlicht: (2025)
Unifying Continuous and Discrete Text Diffusion with Non-simultaneous Diffusion Processes
von: Li, Bocheng, et al.
Veröffentlicht: (2025)
von: Li, Bocheng, et al.
Veröffentlicht: (2025)
TUMLU: A Unified and Native Language Understanding Benchmark for Turkic Languages
von: Isbarov, Jafar, et al.
Veröffentlicht: (2025)
von: Isbarov, Jafar, et al.
Veröffentlicht: (2025)
Understanding Auditory Evoked Brain Signal via Physics-informed Embedding Network with Multi-Task Transformer
von: Ma, Wanli, et al.
Veröffentlicht: (2024)
von: Ma, Wanli, et al.
Veröffentlicht: (2024)
LAMP: Lift Image-Editing as General 3D Priors for Open-world Manipulation
von: Wang, Jingjing, et al.
Veröffentlicht: (2026)
von: Wang, Jingjing, et al.
Veröffentlicht: (2026)
QLIP: Text-Aligned Visual Tokenization Unifies Auto-Regressive Multimodal Understanding and Generation
von: Zhao, Yue, et al.
Veröffentlicht: (2025)
von: Zhao, Yue, et al.
Veröffentlicht: (2025)
The midgut of Aedes albopictus shapes its bacteriome but not its mycobiome
von: Baolin Song, et al.
Veröffentlicht: (2026)
von: Baolin Song, et al.
Veröffentlicht: (2026)
The mosquito midgut harbors stable bacteria that enhance host hemolymph immunity
von: Baolin Song, et al.
Veröffentlicht: (2026)
von: Baolin Song, et al.
Veröffentlicht: (2026)
Native-Resolution Image Synthesis
von: Wang, Zidong, et al.
Veröffentlicht: (2025)
von: Wang, Zidong, et al.
Veröffentlicht: (2025)
MoSA: Mixture of Sparse Adapters for Visual Efficient Tuning
von: Zhang, Qizhe, et al.
Veröffentlicht: (2023)
von: Zhang, Qizhe, et al.
Veröffentlicht: (2023)
EVA02-AT: Egocentric Video-Language Understanding with Spatial-Temporal Rotary Positional Embeddings and Symmetric Optimization
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2025)
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2025)
Generative Adversarial Network on Motion-Blur Image Restoration
von: Li, Zhengdong
Veröffentlicht: (2024)
von: Li, Zhengdong
Veröffentlicht: (2024)
A Survey on Blockchain-based Supply Chain Finance with Progress and Future directions
von: Luo, Zhengdong
Veröffentlicht: (2024)
von: Luo, Zhengdong
Veröffentlicht: (2024)
An Overview of Machine Learning-Driven Resource Allocation in IoT Networks
von: Li, Zhengdong
Veröffentlicht: (2024)
von: Li, Zhengdong
Veröffentlicht: (2024)
HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer
von: Cai, Qi, et al.
Veröffentlicht: (2026)
von: Cai, Qi, et al.
Veröffentlicht: (2026)
Condensation of Composite Bosonic Trions in Interacting Bose-Fermi Mixtures
von: Song, Qi, et al.
Veröffentlicht: (2024)
von: Song, Qi, et al.
Veröffentlicht: (2024)
UniCTokens: Boosting Personalized Understanding and Generation via Unified Concept Tokens
von: An, Ruichuan, et al.
Veröffentlicht: (2025)
von: An, Ruichuan, et al.
Veröffentlicht: (2025)
Aria: An Open Multimodal Native Mixture-of-Experts Model
von: Li, Dongxu, et al.
Veröffentlicht: (2024)
von: Li, Dongxu, et al.
Veröffentlicht: (2024)
Generalized factorials characterized by Dirichlet convolution
von: Ma, Wanli
Veröffentlicht: (2025)
von: Ma, Wanli
Veröffentlicht: (2025)
A Generalized Digit Map: Periodicity, Prouhet-Tarry-Escott Solutions, and Summation Identities
von: Ma, Wanli
Veröffentlicht: (2025)
von: Ma, Wanli
Veröffentlicht: (2025)
Transfer Learning Enabled Transformer based Generative Adversarial Networks (TT-GAN) for Terahertz Channel Modeling and Generating
von: Hu, Zhengdong, et al.
Veröffentlicht: (2024)
von: Hu, Zhengdong, et al.
Veröffentlicht: (2024)
Median of Means Sampling for the Keister Function
von: Zhang, Bocheng
Veröffentlicht: (2025)
von: Zhang, Bocheng
Veröffentlicht: (2025)
EVA NICK
von:
Veröffentlicht: (2007)
von:
Veröffentlicht: (2007)
Securing Transformer-based AI Execution via Unified TEEs and Crypto-protected Accelerators
von: Xue, Jiaqi, et al.
Veröffentlicht: (2025)
von: Xue, Jiaqi, et al.
Veröffentlicht: (2025)
Stable higher-order vortex quantum droplets in an annular potential
von: Dong, Liangwei, et al.
Veröffentlicht: (2024)
von: Dong, Liangwei, et al.
Veröffentlicht: (2024)
EVA: An Embodied World Model for Future Video Anticipation
von: Chi, Xiaowei, et al.
Veröffentlicht: (2024)
von: Chi, Xiaowei, et al.
Veröffentlicht: (2024)
Visible‐Light‐Promoted Phosphorylation of Quinoxalin‐2(1 H )‐One With Diarylphosphine Oxide
von: Zhengdong Xu, et al.
Veröffentlicht: (2026)
von: Zhengdong Xu, et al.
Veröffentlicht: (2026)
A 0.875‐ppm/°C Bandgap Reference With Piecewise Compensation: β ‐Correction and Source/Sink Current Balancing for Concave Curvature
von: Zhengdong Tang, et al.
Veröffentlicht: (2025)
von: Zhengdong Tang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DirectL: Efficient Radiance Fields Rendering for 3D Light Field Displays
von: Yang, Zongyuan, et al.
Veröffentlicht: (2024) -
TextDiff: Mask-Guided Residual Diffusion Models for Scene Text Image Super-Resolution
von: Liu, Baolin, et al.
Veröffentlicht: (2023) -
Towards Scalable Training for Handwritten Mathematical Expression Recognition
von: Li, Haoyang, et al.
Veröffentlicht: (2025) -
EVA: Mixture-of-Experts Semantic Variant Alignment for Compositional Zero-Shot Learning
von: Zhang, Xiao, et al.
Veröffentlicht: (2025) -
UniViTAR: Unified Vision Transformer with Native Resolution
von: Qiao, Limeng, et al.
Veröffentlicht: (2025)