Gespeichert in:
| Hauptverfasser: | Wang, Weijie, He, Xiaoxuan, Gu, Youping, Yang, Yifan, Zhang, Zeyu, He, Yefei, Ding, Yanbo, Hu, Xirui, Chen, Donny Y., He, Zhiyuan, Yang, Yuqing, Zhuang, Bohan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2604.24764 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Less Detail, Better Answers: Degradation-Driven Prompting for VQA
von: Han, Haoxuan, et al.
Veröffentlicht: (2026)
von: Han, Haoxuan, et al.
Veröffentlicht: (2026)
PSA: Pyramid Sparse Attention for Efficient Video Understanding and Generation
von: Li, Xiaolong, et al.
Veröffentlicht: (2025)
von: Li, Xiaolong, et al.
Veröffentlicht: (2025)
FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation
von: Zhou, Junkang, et al.
Veröffentlicht: (2026)
von: Zhou, Junkang, et al.
Veröffentlicht: (2026)
ZPressor: Bottleneck-Aware Compression for Scalable Feed-Forward 3DGS
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
BLADE: Block-Sparse Attention Meets Step Distillation for Efficient Video Generation
von: Gu, Youping, et al.
Veröffentlicht: (2025)
von: Gu, Youping, et al.
Veröffentlicht: (2025)
ReCA: Multi-Shot Long Video Extrapolation via Recursive Context Allocation
von: Liu, Akide, et al.
Veröffentlicht: (2026)
von: Liu, Akide, et al.
Veröffentlicht: (2026)
Revisiting Depth Representations for Feed-Forward 3D Gaussian Splatting
von: Shi, Duochao, et al.
Veröffentlicht: (2025)
von: Shi, Duochao, et al.
Veröffentlicht: (2025)
EfficientDM: Efficient Quantization-Aware Fine-Tuning of Low-Bit Diffusion Models
von: He, Yefei, et al.
Veröffentlicht: (2023)
von: He, Yefei, et al.
Veröffentlicht: (2023)
Neighboring Autoregressive Modeling for Efficient Visual Generation
von: He, Yefei, et al.
Veröffentlicht: (2025)
von: He, Yefei, et al.
Veröffentlicht: (2025)
ZipAR: Parallel Auto-regressive Image Generation through Spatial Locality
von: He, Yefei, et al.
Veröffentlicht: (2024)
von: He, Yefei, et al.
Veröffentlicht: (2024)
Sparsity Forcing: Reinforcing Token Sparsity of MLLMs
von: Chen, Feng, et al.
Veröffentlicht: (2025)
von: Chen, Feng, et al.
Veröffentlicht: (2025)
TriSplat: Simulation-Ready Feed-Forward 3D Scene Reconstruction
von: Wang, Weijie, et al.
Veröffentlicht: (2026)
von: Wang, Weijie, et al.
Veröffentlicht: (2026)
ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification
von: He, Yefei, et al.
Veröffentlicht: (2024)
von: He, Yefei, et al.
Veröffentlicht: (2024)
ME-Switch: A Memory-Efficient Expert Switching Framework for Large Language Models
von: Liu, Jing, et al.
Veröffentlicht: (2024)
von: Liu, Jing, et al.
Veröffentlicht: (2024)
MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
von: Liu, Akide, et al.
Veröffentlicht: (2024)
von: Liu, Akide, et al.
Veröffentlicht: (2024)
VolSplat: Rethinking Feed-Forward 3D Gaussian Splatting with Voxel-Aligned Prediction
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
Flash-GRPO: Efficient Alignment for Video Diffusion via One-Step Policy Optimization
von: He, Xiaoxuan, et al.
Veröffentlicht: (2026)
von: He, Xiaoxuan, et al.
Veröffentlicht: (2026)
OmniSparse: Training-Aware Fine-Grained Sparse Attention for Long-Video MLLMs
von: Chen, Feng, et al.
Veröffentlicht: (2025)
von: Chen, Feng, et al.
Veröffentlicht: (2025)
ZipVL: Efficient Large Vision-Language Models with Dynamic Token Sparsification
von: He, Yefei, et al.
Veröffentlicht: (2024)
von: He, Yefei, et al.
Veröffentlicht: (2024)
LiveWorld: Simulating Out-of-Sight Dynamics in Generative Video World Models
von: Duan, Zicheng, et al.
Veröffentlicht: (2026)
von: Duan, Zicheng, et al.
Veröffentlicht: (2026)
BlockVid: Block Diffusion for High-Quality and Consistent Minute-Long Video Generation
von: Zhang, Zeyu, et al.
Veröffentlicht: (2025)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2025)
Unified Medical Image Pre-training in Language-Guided Common Semantic Space
von: He, Xiaoxuan, et al.
Veröffentlicht: (2023)
von: He, Xiaoxuan, et al.
Veröffentlicht: (2023)
Towards Eliminating Hard Label Constraints in Gradient Inversion Attacks
von: Wang, Yanbo, et al.
Veröffentlicht: (2024)
von: Wang, Yanbo, et al.
Veröffentlicht: (2024)
A People's History of the Second World War
von: Gluckstein, Donny
Veröffentlicht: (2018)
von: Gluckstein, Donny
Veröffentlicht: (2018)
Congestion Control System Optimization with Large Language Models
von: He, Zhiyuan, et al.
Veröffentlicht: (2025)
von: He, Zhiyuan, et al.
Veröffentlicht: (2025)
3D Printing of Bioceramic Multifunctional Scaffolds for Bone Tissue Engineering
von: Huifeng Shao, et al.
Veröffentlicht: (2025)
von: Huifeng Shao, et al.
Veröffentlicht: (2025)
Epilepsy and systemic autoimmune diseases: A bidirectional two‐sample Mendelian randomization study
von: Ruoshi Tian, et al.
Veröffentlicht: (2025)
von: Ruoshi Tian, et al.
Veröffentlicht: (2025)
HiTVideo: Hierarchical Tokenizers for Enhancing Text-to-Video Generation with Autoregressive Large Language Models
von: Zhou, Ziqin, et al.
Veröffentlicht: (2025)
von: Zhou, Ziqin, et al.
Veröffentlicht: (2025)
MTVCraft: Tokenizing 4D Motion for Arbitrary Character Animation
von: Ding, Yanbo, et al.
Veröffentlicht: (2025)
von: Ding, Yanbo, et al.
Veröffentlicht: (2025)
Inferix: A Block-Diffusion based Next-Generation Inference Engine for World Simulation
von: Inferix Team, et al.
Veröffentlicht: (2025)
von: Inferix Team, et al.
Veröffentlicht: (2025)
LongVLM: Efficient Long Video Understanding via Large Language Models
von: Weng, Yuetian, et al.
Veröffentlicht: (2024)
von: Weng, Yuetian, et al.
Veröffentlicht: (2024)
Act While Thinking: Accelerating LLM Agents via Pattern-Aware Speculative Tool Execution
von: Sui, Yifan, et al.
Veröffentlicht: (2026)
von: Sui, Yifan, et al.
Veröffentlicht: (2026)
VL Norm: Rethink Loss Aggregation in RLVR
von: He, Zhiyuan, et al.
Veröffentlicht: (2025)
von: He, Zhiyuan, et al.
Veröffentlicht: (2025)
Optimize Replica Server Placement in a Satellite Network
von: He, Zhiyuan, et al.
Veröffentlicht: (2025)
von: He, Zhiyuan, et al.
Veröffentlicht: (2025)
Streaming Video Diffusion: Online Video Editing with Diffusion Models
von: Chen, Feng, et al.
Veröffentlicht: (2024)
von: Chen, Feng, et al.
Veröffentlicht: (2024)
Constraints on the Hot Circumgalactic Medium around Nearby L* Galaxies from SRG/eROSITA All Sky Survey
von: He, Lin, et al.
Veröffentlicht: (2026)
von: He, Lin, et al.
Veröffentlicht: (2026)
Reinforcement Learning with Inverse Rewards for World Model Post-training
von: Ye, Yang, et al.
Veröffentlicht: (2025)
von: Ye, Yang, et al.
Veröffentlicht: (2025)
DyST-XL: Dynamic Layout Planning and Content Control for Compositional Text-to-Video Generation
von: He, Weijie, et al.
Veröffentlicht: (2025)
von: He, Weijie, et al.
Veröffentlicht: (2025)
Agent Lightning: Train ANY AI Agents with Reinforcement Learning
von: Luo, Xufang, et al.
Veröffentlicht: (2025)
von: Luo, Xufang, et al.
Veröffentlicht: (2025)
Warp-as-History: Generalizable Camera-Controlled Video Generation from One Training Video
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Less Detail, Better Answers: Degradation-Driven Prompting for VQA
von: Han, Haoxuan, et al.
Veröffentlicht: (2026) -
PSA: Pyramid Sparse Attention for Efficient Video Understanding and Generation
von: Li, Xiaolong, et al.
Veröffentlicht: (2025) -
FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation
von: Zhou, Junkang, et al.
Veröffentlicht: (2026) -
ZPressor: Bottleneck-Aware Compression for Scalable Feed-Forward 3DGS
von: Wang, Weijie, et al.
Veröffentlicht: (2025) -
BLADE: Block-Sparse Attention Meets Step Distillation for Efficient Video Generation
von: Gu, Youping, et al.
Veröffentlicht: (2025)