Saved in:
| Main Authors: | Di, Chong, Liu, Li, Zhang, Jinglin, Li, Zhenjiang, Chen, Da, Cohen, Laurent D. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2606.00139 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dynamic Exploration on Segment-Proposal Graphs for Tubular Centerline Tracking
by: Di, Chong, et al.
Published: (2025)
by: Di, Chong, et al.
Published: (2025)
Curvature-Aware Captioning:Leveraging Geodesic Attention for 3D Scene Understanding
by: He, Ziyao, et al.
Published: (2026)
by: He, Ziyao, et al.
Published: (2026)
OmniAlpha: Aligning Transparency-Aware Generation via Multi-Task Unified Reinforcement Learning
by: Yu, Hao, et al.
Published: (2025)
by: Yu, Hao, et al.
Published: (2025)
Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising
by: Guo, Jun, et al.
Published: (2026)
by: Guo, Jun, et al.
Published: (2026)
Distilled Large Language Model-Driven Dynamic Sparse Expert Activation Mechanism
by: Chen, Qinghui, et al.
Published: (2026)
by: Chen, Qinghui, et al.
Published: (2026)
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision
by: Liu, Zeyu, et al.
Published: (2026)
by: Liu, Zeyu, et al.
Published: (2026)
Evolving High-Quality Rendering and Reconstruction in a Unified Framework with Contribution-Adaptive Regularization
by: Shen, You, et al.
Published: (2025)
by: Shen, You, et al.
Published: (2025)
Predicting and Enhancing the Fairness of DNNs with the Curvature of Perceptual Manifolds
by: Ma, Yanbiao, et al.
Published: (2023)
by: Ma, Yanbiao, et al.
Published: (2023)
Multi-modal Test-time Adaptation via Adaptive Probabilistic Gaussian Calibration
by: Xu, Jinglin, et al.
Published: (2026)
by: Xu, Jinglin, et al.
Published: (2026)
Align Your Tangent: Training Better Consistency Models via Manifold-Aligned Tangents
by: Kim, Beomsu, et al.
Published: (2025)
by: Kim, Beomsu, et al.
Published: (2025)
Deep Image Prior with L0 Gradient Regularizer for Image Smoothing
by: Tran, Nhat Thanh, et al.
Published: (2026)
by: Tran, Nhat Thanh, et al.
Published: (2026)
Exploring Semantic-constrained Adversarial Example with Instruction Uncertainty Reduction
by: Hu, Jin, et al.
Published: (2025)
by: Hu, Jin, et al.
Published: (2025)
TokenMotion: Decoupled Motion Control via Token Disentanglement for Human-centric Video Generation
by: Li, Ruineng, et al.
Published: (2025)
by: Li, Ruineng, et al.
Published: (2025)
EAvatar: Expression-Aware Head Avatar Reconstruction with Generative Geometry Priors
by: Zhang, Shikun, et al.
Published: (2025)
by: Zhang, Shikun, et al.
Published: (2025)
Neural Tangent Knowledge Distillation for Optical Convolutional Networks
by: Xiang, Jinlin, et al.
Published: (2025)
by: Xiang, Jinlin, et al.
Published: (2025)
Multi-modal Relation Distillation for Unified 3D Representation Learning
by: Wang, Huiqun, et al.
Published: (2024)
by: Wang, Huiqun, et al.
Published: (2024)
Beyond BEV: Optimizing Point-Level Tokens for Collaborative Perception
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
PriorRG: Prior-Guided Contrastive Pre-training and Coarse-to-Fine Decoding for Chest X-ray Report Generation
by: Liu, Kang, et al.
Published: (2025)
by: Liu, Kang, et al.
Published: (2025)
InfoTok: Information-Theoretic Regularization for Capacity-Constrained Shared Visual Tokenization in Unified MLLMs
by: Tang, Lv, et al.
Published: (2026)
by: Tang, Lv, et al.
Published: (2026)
UniEval: Unified Holistic Evaluation for Unified Multimodal Understanding and Generation
by: Li, Yi, et al.
Published: (2025)
by: Li, Yi, et al.
Published: (2025)
From 2D Alignment to 3D Plausibility: Unifying Heterogeneous 2D Priors and Penetration-Free Diffusion for Occlusion-Robust Two-Hand Reconstruction
by: Han, Gaoge, et al.
Published: (2025)
by: Han, Gaoge, et al.
Published: (2025)
Overcoming the Curvature Bottleneck in MeanFlow
by: Zhang, Xinxi, et al.
Published: (2025)
by: Zhang, Xinxi, et al.
Published: (2025)
UNIAA: A Unified Multi-modal Image Aesthetic Assessment Baseline and Benchmark
by: Zhou, Zhaokun, et al.
Published: (2024)
by: Zhou, Zhaokun, et al.
Published: (2024)
See Before You Code: Learning Visual Priors for Spatially Aware Educational Animation Generation
by: Li, Yuejia, et al.
Published: (2026)
by: Li, Yuejia, et al.
Published: (2026)
Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling
by: Chen, Xiaokang, et al.
Published: (2025)
by: Chen, Xiaokang, et al.
Published: (2025)
Exploiting Diffusion Prior for Out-of-Distribution Detection
by: Zhu, Armando, et al.
Published: (2024)
by: Zhu, Armando, et al.
Published: (2024)
LaneDiffusion: Improving Centerline Graph Learning via Prior Injected BEV Feature Generation
by: Wang, Zijie, et al.
Published: (2025)
by: Wang, Zijie, et al.
Published: (2025)
SurfSplat: Conquering Feedforward 2D Gaussian Splatting with Surface Continuity Priors
by: He, Bing, et al.
Published: (2026)
by: He, Bing, et al.
Published: (2026)
NeuroBridge: Bio-Inspired Self-Supervised EEG-to-Image Decoding via Cognitive Priors and Bidirectional Semantic Alignment
by: Zhang, Wenjiang, et al.
Published: (2025)
by: Zhang, Wenjiang, et al.
Published: (2025)
Archon: A Unified Multimodal Model for Holistic Digital Human Generation
by: Bao, Chong, et al.
Published: (2026)
by: Bao, Chong, et al.
Published: (2026)
Empowering Bridge Digital Twins by Bridging the Data Gap with a Unified Synthesis Framework
by: Wang, Wang, et al.
Published: (2025)
by: Wang, Wang, et al.
Published: (2025)
Uni-RS: A Spatially Faithful Unified Understanding and Generation Model for Remote Sensing
by: Zhang, Weiyu, et al.
Published: (2026)
by: Zhang, Weiyu, et al.
Published: (2026)
DeepSVU: Towards In-depth Security-oriented Video Understanding via Unified Physical-world Regularized MoE
by: Jin, Yujie, et al.
Published: (2026)
by: Jin, Yujie, et al.
Published: (2026)
Neural Distribution Prior for LiDAR Out-of-Distribution Detection
by: Li, Zizhao, et al.
Published: (2026)
by: Li, Zizhao, et al.
Published: (2026)
One Model to Translate Them All: Universal Any-to-Any Translation for Heterogeneous Collaborative Perception
by: Li, Yang, et al.
Published: (2026)
by: Li, Yang, et al.
Published: (2026)
Unlocking Visual Secrets: Inverting Features with Diffusion Priors for Image Reconstruction
by: Zhang, Sai Qian, et al.
Published: (2024)
by: Zhang, Sai Qian, et al.
Published: (2024)
Crab$^{+}$: A Scalable and Unified Audio-Visual Scene Understanding Model with Explicit Cooperation
by: Cai, Dongnuan, et al.
Published: (2026)
by: Cai, Dongnuan, et al.
Published: (2026)
Semantic Generative Tuning for Unified Multimodal Models
by: Yu, Songsong, et al.
Published: (2026)
by: Yu, Songsong, et al.
Published: (2026)
OmniCorpus: A Unified Multimodal Corpus of 10 Billion-Level Images Interleaved with Text
by: Li, Qingyun, et al.
Published: (2024)
by: Li, Qingyun, et al.
Published: (2024)
Single Image Reflection Separation via Dual Prior Interaction Transformer
by: Huang, Yue, et al.
Published: (2025)
by: Huang, Yue, et al.
Published: (2025)
Similar Items
-
Dynamic Exploration on Segment-Proposal Graphs for Tubular Centerline Tracking
by: Di, Chong, et al.
Published: (2025) -
Curvature-Aware Captioning:Leveraging Geodesic Attention for 3D Scene Understanding
by: He, Ziyao, et al.
Published: (2026) -
OmniAlpha: Aligning Transparency-Aware Generation via Multi-Task Unified Reinforcement Learning
by: Yu, Hao, et al.
Published: (2025) -
Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising
by: Guo, Jun, et al.
Published: (2026) -
Distilled Large Language Model-Driven Dynamic Sparse Expert Activation Mechanism
by: Chen, Qinghui, et al.
Published: (2026)