Stochastic Interpolants via Conditional Dependent Coupling
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Chenrui, Xiao, Xi, Wang, Tianyang, Wang, Xiao, Shen, Yanning |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Editing Pairs: Fine-Grained Instructional Image Editing via Multi-Scale Learnable Regions
by: Ma, Chenrui, et al.
Published: (2025)
by: Ma, Chenrui, et al.
Published: (2025)
Learning Straight Flows: Variational Flow Matching for Efficient Generation
by: Ma, Chenrui, et al.
Published: (2025)
by: Ma, Chenrui, et al.
Published: (2025)
CAD-VAE: Leveraging Correlation-Aware Latents for Comprehensive Fair Disentanglement
by: Ma, Chenrui, et al.
Published: (2025)
by: Ma, Chenrui, et al.
Published: (2025)
Self-Supervised Visual Prompting for Cross-Domain Road Damage Detection
by: Xiao, Xi, et al.
Published: (2025)
by: Xiao, Xi, et al.
Published: (2025)
Not All Directions Matter: Towards Structured and Task-Aware Low-Rank Model Adaptation
by: Xiao, Xi, et al.
Published: (2026)
by: Xiao, Xi, et al.
Published: (2026)
Learning Structure-Supporting Dependencies via Keypoint Interactive Transformer for General Mammal Pose Estimation
by: Xu, Tianyang, et al.
Published: (2025)
by: Xu, Tianyang, et al.
Published: (2025)
Velocity Disambiguation for Video Frame Interpolation
by: Zhong, Zhihang, et al.
Published: (2023)
by: Zhong, Zhihang, et al.
Published: (2023)
Visual Variational Autoencoder Prompt Tuning
by: Xiao, Xi, et al.
Published: (2025)
by: Xiao, Xi, et al.
Published: (2025)
ACDiT: Interpolating Autoregressive Conditional Modeling and Diffusion Transformer
by: Hu, Jinyi, et al.
Published: (2024)
by: Hu, Jinyi, et al.
Published: (2024)
Boosting Active Learning with Knowledge Transfer
by: Wang, Tianyang, et al.
Published: (2025)
by: Wang, Tianyang, et al.
Published: (2025)
C3L: Content Correlated Vision-Language Instruction Tuning Data Generation via Contrastive Learning
by: Ma, Ji, et al.
Published: (2024)
by: Ma, Ji, et al.
Published: (2024)
CIBR: Cross-modal Information Bottleneck Regularization for Robust CLIP Generalization
by: Ji, Yingrui, et al.
Published: (2025)
by: Ji, Yingrui, et al.
Published: (2025)
Framer: Interactive Frame Interpolation
by: Wang, Wen, et al.
Published: (2024)
by: Wang, Wen, et al.
Published: (2024)
4D Multimodal Co-attention Fusion Network with Latent Contrastive Alignment for Alzheimer's Diagnosis
by: Wei, Yuxiang, et al.
Published: (2025)
by: Wei, Yuxiang, et al.
Published: (2025)
Prompt-Free Conditional Diffusion for Multi-object Image Augmentation
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
FairViT: Fair Vision Transformer via Adaptive Masking
by: Tian, Bowei, et al.
Published: (2024)
by: Tian, Bowei, et al.
Published: (2024)
FOCUS: Fused Observation of Channels for Unveiling Spectra
by: Xiao, Xi, et al.
Published: (2025)
by: Xiao, Xi, et al.
Published: (2025)
TASAM: Terrain-and-Aware Segment Anything Model for Temporal-Scale Remote Sensing Segmentation
by: Wang, Tianyang, et al.
Published: (2025)
by: Wang, Tianyang, et al.
Published: (2025)
MagicID: Hybrid Preference Optimization for ID-Consistent and Dynamic-Preserved Video Customization
by: Li, Hengjia, et al.
Published: (2025)
by: Li, Hengjia, et al.
Published: (2025)
Semantic Frame Interpolation
by: Hong, Yijia, et al.
Published: (2025)
by: Hong, Yijia, et al.
Published: (2025)
FlexIP: Dynamic Control of Preservation and Personality for Customized Image Generation
by: Huang, Linyan, et al.
Published: (2025)
by: Huang, Linyan, et al.
Published: (2025)
CryoCCD: Conditional Cycle-consistent Diffusion with Biophysical Modeling for Cryo-EM Synthesis
by: Jiang, Runmin, et al.
Published: (2025)
by: Jiang, Runmin, et al.
Published: (2025)
Visual Instance-aware Prompt Tuning
by: Xiao, Xi, et al.
Published: (2025)
by: Xiao, Xi, et al.
Published: (2025)
Transition Flow Matching
by: Ma, Chenrui
Published: (2026)
by: Ma, Chenrui
Published: (2026)
TD-RD: A Top-Down Benchmark with Real-Time Framework for Road Damage Detection
by: Xiao, Xi, et al.
Published: (2025)
by: Xiao, Xi, et al.
Published: (2025)
Multi-dimension Transformer with Attention-based Filtering for Medical Image Segmentation
by: Wang, Wentao, et al.
Published: (2024)
by: Wang, Wentao, et al.
Published: (2024)
Short-LVLM: Compressing and Accelerating Large Vision-Language Models by Pruning Redundant Layers
by: Ma, Ji, et al.
Published: (2025)
by: Ma, Ji, et al.
Published: (2025)
Understanding and Mitigating Hallucinations in Multimodal Chain-of-Thought Models
by: Ma, Ji, et al.
Published: (2026)
by: Ma, Ji, et al.
Published: (2026)
DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding
by: Ran, Lingyan, et al.
Published: (2025)
by: Ran, Lingyan, et al.
Published: (2025)
Detail Consistent Stage-Wise Distillation for Efficient 3D MRI Segmentation
by: Fan, Mengchen, et al.
Published: (2026)
by: Fan, Mengchen, et al.
Published: (2026)
FreeBlend: Advancing Concept Blending with Staged Feedback-Driven Interpolation Diffusion
by: Zhou, Yufan, et al.
Published: (2025)
by: Zhou, Yufan, et al.
Published: (2025)
M$^2$IV: Towards Efficient and Fine-grained Multimodal In-Context Learning via Representation Engineering
by: Li, Yanshu, et al.
Published: (2025)
by: Li, Yanshu, et al.
Published: (2025)
StdGEN: Semantic-Decomposed 3D Character Generation from Single Images
by: He, Yuze, et al.
Published: (2024)
by: He, Yuze, et al.
Published: (2024)
Motion-Aware Generative Frame Interpolation
by: Zhang, Guozhen, et al.
Published: (2025)
by: Zhang, Guozhen, et al.
Published: (2025)
Visual Object Tracking on Multi-modal RGB-D Videos: A Review
by: Zhu, Xue-Feng, et al.
Published: (2022)
by: Zhu, Xue-Feng, et al.
Published: (2022)
Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion Models
by: Shen, Fei, et al.
Published: (2023)
by: Shen, Fei, et al.
Published: (2023)
Arbitrary Generative Video Interpolation
by: Zhang, Guozhen, et al.
Published: (2025)
by: Zhang, Guozhen, et al.
Published: (2025)
MoRE-Brain: Routed Mixture of Experts for Interpretable and Generalizable Cross-Subject fMRI Visual Decoding
by: Wei, Yuxiang, et al.
Published: (2025)
by: Wei, Yuxiang, et al.
Published: (2025)
MS-DETR: Multispectral Pedestrian Detection Transformer with Loosely Coupled Fusion and Modality-Balanced Optimization
by: Xing, Yinghui, et al.
Published: (2023)
by: Xing, Yinghui, et al.
Published: (2023)
Encoding Structural Constraints into Segment Anything Models via Probabilistic Graphical Models
by: Li, Yu, et al.
Published: (2025)
by: Li, Yu, et al.
Published: (2025)
Similar Items
-
Beyond Editing Pairs: Fine-Grained Instructional Image Editing via Multi-Scale Learnable Regions
by: Ma, Chenrui, et al.
Published: (2025) -
Learning Straight Flows: Variational Flow Matching for Efficient Generation
by: Ma, Chenrui, et al.
Published: (2025) -
CAD-VAE: Leveraging Correlation-Aware Latents for Comprehensive Fair Disentanglement
by: Ma, Chenrui, et al.
Published: (2025) -
Self-Supervised Visual Prompting for Cross-Domain Road Damage Detection
by: Xiao, Xi, et al.
Published: (2025) -
Not All Directions Matter: Towards Structured and Task-Aware Low-Rank Model Adaptation
by: Xiao, Xi, et al.
Published: (2026)