Guardado en:
| Autores principales: | Luo, Yinyi, Wang, Wenwen, Bai, Hayes, Zhu, Hongyu, Chen, Hao, He, Pan, Savvides, Marios, Li, Sharon, Wang, Jindong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2604.10784 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LatentUMM: Dual Latent Alignment for Unified Multimodal Models
por: Luo, Yinyi, et al.
Publicado: (2026)
por: Luo, Yinyi, et al.
Publicado: (2026)
UniPath: Adaptive Coordination of Understanding and Generation for Unified Multimodal Reasoning
por: Bai, Hayes, et al.
Publicado: (2026)
por: Bai, Hayes, et al.
Publicado: (2026)
Self-Corrected Image Generation with Explainable Latent Rewards
por: Luo, Yinyi, et al.
Publicado: (2026)
por: Luo, Yinyi, et al.
Publicado: (2026)
KnowledgeSmith: Uncovering Knowledge Updating in LLMs with Model Editing and Unlearning
por: Luo, Yinyi, et al.
Publicado: (2025)
por: Luo, Yinyi, et al.
Publicado: (2025)
FedUMM: A General Framework for Federated Learning with Unified Multimodal Models
por: Su, Zhaolong, et al.
Publicado: (2026)
por: Su, Zhaolong, et al.
Publicado: (2026)
Image Tokenizer Needs Post-Training
por: Qiu, Kai, et al.
Publicado: (2025)
por: Qiu, Kai, et al.
Publicado: (2025)
MetaVLA: Unified Meta Co-training For Efficient Embodied Adaption
por: Li, Chen, et al.
Publicado: (2025)
por: Li, Chen, et al.
Publicado: (2025)
SOLAR: Scalable Optimization of Large-scale Architecture for Reasoning
por: Li, Chen, et al.
Publicado: (2025)
por: Li, Chen, et al.
Publicado: (2025)
UniGame: Turning a Unified Multimodal Model Into Its Own Adversary
por: Su, Zhaolong, et al.
Publicado: (2025)
por: Su, Zhaolong, et al.
Publicado: (2025)
An Embarrassingly Simple Baseline for Imbalanced Semi-Supervised Learning
por: Chen, Hao, et al.
Publicado: (2022)
por: Chen, Hao, et al.
Publicado: (2022)
A Unified Study of LoRA Variants: Taxonomy, Review, Codebase, and Empirical Evaluation
por: He, Haonan, et al.
Publicado: (2026)
por: He, Haonan, et al.
Publicado: (2026)
Robust Latent Matters: Boosting Image Generation with Sampling Error Synthesis
por: Qiu, Kai, et al.
Publicado: (2025)
por: Qiu, Kai, et al.
Publicado: (2025)
Conv-Adapter: Exploring Parameter Efficient Transfer Learning for ConvNets
por: Chen, Hao, et al.
Publicado: (2022)
por: Chen, Hao, et al.
Publicado: (2022)
SciPost Physics Codebases
Publicado: (2026)
Publicado: (2026)
Reward Evolution with Graph-of-Thoughts: A Bi-Level Language Model Framework for Reinforcement Learning
por: Yao, Changwei, et al.
Publicado: (2025)
por: Yao, Changwei, et al.
Publicado: (2025)
ChatUMM: Robust Context Tracking for Conversational Interleaved Generation
por: Dai, Wenxun, et al.
Publicado: (2026)
por: Dai, Wenxun, et al.
Publicado: (2026)
PromptBench: A Unified Library for Evaluation of Large Language Models
por: Zhu, Kaijie, et al.
Publicado: (2023)
por: Zhu, Kaijie, et al.
Publicado: (2023)
Rethinking UMM Visual Generation: Masked Modeling for Efficient Image-Only Pre-training
por: Sun, Peng, et al.
Publicado: (2026)
por: Sun, Peng, et al.
Publicado: (2026)
OpenWorldLib: A Unified Codebase and Definition of Advanced World Models
por: DataFlow Team, et al.
Publicado: (2026)
por: DataFlow Team, et al.
Publicado: (2026)
UniSD: Towards a Unified Self-Distillation Framework for Large Language Models
por: Jin, Yiqiao, et al.
Publicado: (2026)
por: Jin, Yiqiao, et al.
Publicado: (2026)
RTGen: Generating Region-Text Pairs for Open-Vocabulary Object Detection
por: Chen, Fangyi, et al.
Publicado: (2024)
por: Chen, Fangyi, et al.
Publicado: (2024)
RPG: A Repository Planning Graph for Unified and Scalable Codebase Generation
por: Luo, Jane, et al.
Publicado: (2025)
por: Luo, Jane, et al.
Publicado: (2025)
On Fairness of Unified Multimodal Large Language Model for Image Generation
por: Liu, Ming, et al.
Publicado: (2025)
por: Liu, Ming, et al.
Publicado: (2025)
Hierarchical Knowledge Graph Construction from Images for Scalable E-Commerce
por: Yang, Zhantao, et al.
Publicado: (2024)
por: Yang, Zhantao, et al.
Publicado: (2024)
Self-Ensemble Post Learning for Noisy Domain Generalization
por: Lu, Wang, et al.
Publicado: (2025)
por: Lu, Wang, et al.
Publicado: (2025)
Efficient Autoregressive Audio Modeling via Next-Scale Prediction
por: Qiu, Kai, et al.
Publicado: (2024)
por: Qiu, Kai, et al.
Publicado: (2024)
MotionVerse: A Unified Multimodal Framework for Motion Comprehension, Generation and Editing
por: Hou, Ruibing, et al.
Publicado: (2025)
por: Hou, Ruibing, et al.
Publicado: (2025)
STELAR-VISION: Self-Topology-Aware Efficient Learning for Aligned Reasoning in Vision
por: Li, Chen, et al.
Publicado: (2025)
por: Li, Chen, et al.
Publicado: (2025)
TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training
por: Liang, Wanchao, et al.
Publicado: (2024)
por: Liang, Wanchao, et al.
Publicado: (2024)
MASLab: A Unified and Comprehensive Codebase for LLM-based Multi-Agent Systems
por: Ye, Rui, et al.
Publicado: (2025)
por: Ye, Rui, et al.
Publicado: (2025)
When to Re-Commit: Temporal Abstraction Discovery for Long-Horizon Vision-Language Reasoning
por: Li, Chen, et al.
Publicado: (2026)
por: Li, Chen, et al.
Publicado: (2026)
A CLIP-based Uncertainty Modal Modeling (UMM) Framework for Pedestrian Re-Identification in Autonomous Driving
por: Li, Jialin, et al.
Publicado: (2025)
por: Li, Jialin, et al.
Publicado: (2025)
ExecuTorch -- A Unified PyTorch Solution to Run AI Models On-Device
por: Nachin, Mergen, et al.
Publicado: (2026)
por: Nachin, Mergen, et al.
Publicado: (2026)
Neural Radiance Fields with Torch Units
por: Ni, Bingnan, et al.
Publicado: (2024)
por: Ni, Bingnan, et al.
Publicado: (2024)
Can Vision Replace Text in Working Memory? Evidence from Spatial n-Back in Vision-Language Models
por: Liang, Sichu, et al.
Publicado: (2026)
por: Liang, Sichu, et al.
Publicado: (2026)
FormulaCode: Evaluating Agentic Optimization on Large Codebases
por: Sehgal, Atharva, et al.
Publicado: (2026)
por: Sehgal, Atharva, et al.
Publicado: (2026)
AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent
por: Luo, Yinyi, et al.
Publicado: (2026)
por: Luo, Yinyi, et al.
Publicado: (2026)
ForensicHub: A Unified Benchmark & Codebase for All-Domain Fake Image Detection and Localization
por: Du, Bo, et al.
Publicado: (2025)
por: Du, Bo, et al.
Publicado: (2025)
torchtune: PyTorch native post-training library
por: Obozov, Mark, et al.
Publicado: (2026)
por: Obozov, Mark, et al.
Publicado: (2026)
SWE-Adept: An LLM-Based Agentic Framework for Deep Codebase Analysis and Structured Issue Resolution
por: He, Kang, et al.
Publicado: (2026)
por: He, Kang, et al.
Publicado: (2026)
Ejemplares similares
-
LatentUMM: Dual Latent Alignment for Unified Multimodal Models
por: Luo, Yinyi, et al.
Publicado: (2026) -
UniPath: Adaptive Coordination of Understanding and Generation for Unified Multimodal Reasoning
por: Bai, Hayes, et al.
Publicado: (2026) -
Self-Corrected Image Generation with Explainable Latent Rewards
por: Luo, Yinyi, et al.
Publicado: (2026) -
KnowledgeSmith: Uncovering Knowledge Updating in LLMs with Model Editing and Unlearning
por: Luo, Yinyi, et al.
Publicado: (2025) -
FedUMM: A General Framework for Federated Learning with Unified Multimodal Models
por: Su, Zhaolong, et al.
Publicado: (2026)