LaTtE-Flow: Layerwise Timestep-Expert Flow-based Transformer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shen, Ying, Xu, Zhiyang, Chen, Jiuhai, Diao, Shizhe, Zhang, Jiaxin, Yao, Yuguang, Rimchala, Joy, Lourentzou, Ismini, Huang, Lifu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
R2I-Bench: Benchmarking Reasoning-Driven Text-to-Image Generation
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
SuperFlow: Training Flow Matching Models with RL on the Fly
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
Modality-Specialized Synergizers for Interleaved Vision-Language Generalists
von: Xu, Zhiyang, et al.
Veröffentlicht: (2024)
von: Xu, Zhiyang, et al.
Veröffentlicht: (2024)
Holistic Evaluation for Interleaved Text-and-Image Generation
von: Liu, Minqian, et al.
Veröffentlicht: (2024)
von: Liu, Minqian, et al.
Veröffentlicht: (2024)
Commonsense for Zero-Shot Natural Language Video Localization
von: Holla, Meghana, et al.
Veröffentlicht: (2023)
von: Holla, Meghana, et al.
Veröffentlicht: (2023)
RE$^2$: Region-Aware Relation Extraction from Visually Rich Documents
von: Ramu, Pritika, et al.
Veröffentlicht: (2023)
von: Ramu, Pritika, et al.
Veröffentlicht: (2023)
ELBA: Learning by Asking for Embodied Visual Navigation and Task Completion
von: Shen, Ying, et al.
Veröffentlicht: (2023)
von: Shen, Ying, et al.
Veröffentlicht: (2023)
Phantom: Physics-Infused Video Generation via Joint Modeling of Visual and Latent Physical Dynamics
von: Shen, Ying, et al.
Veröffentlicht: (2026)
von: Shen, Ying, et al.
Veröffentlicht: (2026)
3D-VCD: Hallucination Mitigation in 3D-LLM Embodied Agents through Visual Contrastive Decoding
von: Ogunleye, Makanjuola, et al.
Veröffentlicht: (2026)
von: Ogunleye, Makanjuola, et al.
Veröffentlicht: (2026)
mTSBench: Benchmarking Multivariate Time Series Anomaly Detection and Model Selection at Scale
von: Zhou, Xiaona, et al.
Veröffentlicht: (2025)
von: Zhou, Xiaona, et al.
Veröffentlicht: (2025)
uaMix-MAE: Efficient Tuning of Pretrained Audio Transformers with Unsupervised Audio Mixtures
von: Tabassum, Afrina, et al.
Veröffentlicht: (2024)
von: Tabassum, Afrina, et al.
Veröffentlicht: (2024)
CoRe3D: Collaborative Reasoning as a Foundation for 3D Intelligence
von: Yu, Tianjiao, et al.
Veröffentlicht: (2025)
von: Yu, Tianjiao, et al.
Veröffentlicht: (2025)
Grounded-VideoLLM: Sharpening Fine-grained Temporal Grounding in Video Large Language Models
von: Wang, Haibo, et al.
Veröffentlicht: (2024)
von: Wang, Haibo, et al.
Veröffentlicht: (2024)
Hierarchical Dataset Selection for High-Quality Data Sharing
von: Zhou, Xiaona, et al.
Veröffentlicht: (2025)
von: Zhou, Xiaona, et al.
Veröffentlicht: (2025)
FAIR: Facilitating Artificial Intelligence Resilience in Manufacturing Industrial Internet
von: Zeng, Yingyan, et al.
Veröffentlicht: (2025)
von: Zeng, Yingyan, et al.
Veröffentlicht: (2025)
RAVEL: Rare Concept Generation and Editing via Graph-driven Relational Guidance
von: Venkatesh, Kavana, et al.
Veröffentlicht: (2024)
von: Venkatesh, Kavana, et al.
Veröffentlicht: (2024)
MULTISCRIPT: Multimodal Script Learning for Supporting Open Domain Everyday Tasks
von: Qi, Jingyuan, et al.
Veröffentlicht: (2023)
von: Qi, Jingyuan, et al.
Veröffentlicht: (2023)
SPARTUN3D: Situated Spatial Understanding of 3D World in Large Language Models
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
LISA: Layerwise Importance Sampling for Memory-Efficient Large Language Model Fine-Tuning
von: Pan, Rui, et al.
Veröffentlicht: (2024)
von: Pan, Rui, et al.
Veröffentlicht: (2024)
Evaluating Cognitive Age Alignment in Interactive AI Agents
von: Shen, Yifan, et al.
Veröffentlicht: (2026)
von: Shen, Yifan, et al.
Veröffentlicht: (2026)
DreamPartGen: Semantically Grounded Part-Level 3D Generation via Collaborative Latent Denoising
von: Yu, Tianjiao, et al.
Veröffentlicht: (2026)
von: Yu, Tianjiao, et al.
Veröffentlicht: (2026)
Data for TtAgo
von: Liu, Jinchu, et al.
Veröffentlicht: (2025)
von: Liu, Jinchu, et al.
Veröffentlicht: (2025)
Part$^{2}$GS: Part-aware Modeling of Articulated Objects using 3D Gaussian Splatting
von: Yu, Tianjiao, et al.
Veröffentlicht: (2025)
von: Yu, Tianjiao, et al.
Veröffentlicht: (2025)
Layerwise Recurrent Router for Mixture-of-Experts
von: Qiu, Zihan, et al.
Veröffentlicht: (2024)
von: Qiu, Zihan, et al.
Veröffentlicht: (2024)
RewardFlow: Generate Images by Optimizing What You Reward
von: Susladkar, Onkar, et al.
Veröffentlicht: (2026)
von: Susladkar, Onkar, et al.
Veröffentlicht: (2026)
LLM Braces: Straightening Out LLM Predictions with Relevant Sub-Updates
von: Shen, Ying, et al.
Veröffentlicht: (2025)
von: Shen, Ying, et al.
Veröffentlicht: (2025)
Multimodal Instruction Tuning with Conditional Mixture of LoRA
von: Shen, Ying, et al.
Veröffentlicht: (2024)
von: Shen, Ying, et al.
Veröffentlicht: (2024)
MMPlanner: Zero-Shot Multimodal Procedural Planning with Chain-of-Thought Object State Reasoning
von: Tabassum, Afrina, et al.
Veröffentlicht: (2025)
von: Tabassum, Afrina, et al.
Veröffentlicht: (2025)
Tiny but Trusted: Efficient Vision-Language Reasoning for Time-Series Anomaly Detection
von: Zhou, Xiaona, et al.
Veröffentlicht: (2026)
von: Zhou, Xiaona, et al.
Veröffentlicht: (2026)
ChunkFlow: Communication-Aware Chunked Prefetching for Layerwise Offloading in Distributed Diffusion Transformer Inference
von: Meng, Han, et al.
Veröffentlicht: (2026)
von: Meng, Han, et al.
Veröffentlicht: (2026)
TheoremLlama: Transforming General-Purpose LLMs into Lean4 Experts
von: Wang, Ruida, et al.
Veröffentlicht: (2024)
von: Wang, Ruida, et al.
Veröffentlicht: (2024)
Learning Hamiltonian Flow Maps: Mean Flow Consistency for Large-Timestep Molecular Dynamics
von: Ripken, Winfried, et al.
Veröffentlicht: (2026)
von: Ripken, Winfried, et al.
Veröffentlicht: (2026)
CALICO: Part-Focused Semantic Co-Segmentation with Large Vision-Language Models
von: Nguyen, Kiet A., et al.
Veröffentlicht: (2024)
von: Nguyen, Kiet A., et al.
Veröffentlicht: (2024)
AR-RAG: Autoregressive Retrieval Augmentation for Image Generation
von: Qi, Jingyuan, et al.
Veröffentlicht: (2025)
von: Qi, Jingyuan, et al.
Veröffentlicht: (2025)
Think, Act, Build: An Agentic Framework with Vision Language Models for Zero-Shot 3D Visual Grounding
von: Wang, Haibo, et al.
Veröffentlicht: (2026)
von: Wang, Haibo, et al.
Veröffentlicht: (2026)
Best of Both Worlds: Multimodal Reasoning and Generation via Unified Discrete Flow Matching
von: Susladkar, Onkar, et al.
Veröffentlicht: (2026)
von: Susladkar, Onkar, et al.
Veröffentlicht: (2026)
Minimum-Cost Network Flow with Dual Predictions
von: Chen, Zhiyang, et al.
Veröffentlicht: (2026)
von: Chen, Zhiyang, et al.
Veröffentlicht: (2026)
TCAQ-DM: Timestep-Channel Adaptive Quantization for Diffusion Models
von: Huang, Haocheng, et al.
Veröffentlicht: (2024)
von: Huang, Haocheng, et al.
Veröffentlicht: (2024)
UniHGKR: Unified Instruction-aware Heterogeneous Knowledge Retrievers
von: Min, Dehai, et al.
Veröffentlicht: (2024)
von: Min, Dehai, et al.
Veröffentlicht: (2024)
OpenCAEPoro: A Parallel Simulation Framework for Multiphase and Multicomponent Porous Media Flows
von: Li, Shizhe, et al.
Veröffentlicht: (2024)
von: Li, Shizhe, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
R2I-Bench: Benchmarking Reasoning-Driven Text-to-Image Generation
von: Chen, Kaijie, et al.
Veröffentlicht: (2025) -
SuperFlow: Training Flow Matching Models with RL on the Fly
von: Chen, Kaijie, et al.
Veröffentlicht: (2025) -
Modality-Specialized Synergizers for Interleaved Vision-Language Generalists
von: Xu, Zhiyang, et al.
Veröffentlicht: (2024) -
Holistic Evaluation for Interleaved Text-and-Image Generation
von: Liu, Minqian, et al.
Veröffentlicht: (2024) -
Commonsense for Zero-Shot Natural Language Video Localization
von: Holla, Meghana, et al.
Veröffentlicht: (2023)