DeepInv: A Novel Self-supervised Learning Approach for Fast and Accurate Diffusion Inversion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Ziyue, Lin, Luxi, Hu, Xiaolin, Chang, Chao, Wang, HuaiXi, Zhou, Yiyi, Ji, Rongrong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EasyInv: Toward Fast and Better DDIM Inversion
von: Zhang, Ziyue, et al.
Veröffentlicht: (2024)
von: Zhang, Ziyue, et al.
Veröffentlicht: (2024)
ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling
von: Ju, Shaobo, et al.
Veröffentlicht: (2026)
von: Ju, Shaobo, et al.
Veröffentlicht: (2026)
InvSeg: Test-Time Prompt Inversion for Semantic Segmentation
von: Lin, Jiayi, et al.
Veröffentlicht: (2024)
von: Lin, Jiayi, et al.
Veröffentlicht: (2024)
InvCoSS: Inversion-driven Continual Self-supervised Learning in Medical Multi-modal Image Pre-training
von: Luo, Zihao, et al.
Veröffentlicht: (2025)
von: Luo, Zihao, et al.
Veröffentlicht: (2025)
Boosting Multimodal Large Language Models with Visual Tokens Withdrawal for Rapid Inference
von: Lin, Zhihang, et al.
Veröffentlicht: (2024)
von: Lin, Zhihang, et al.
Veröffentlicht: (2024)
Speculative Decoding Reimagined for Multimodal Large Language Models
von: Lin, Luxi, et al.
Veröffentlicht: (2025)
von: Lin, Luxi, et al.
Veröffentlicht: (2025)
Parallel Vision Token Scheduling for Fast and Accurate Multimodal LMMs Inference
von: Zhan, Wengyi, et al.
Veröffentlicht: (2025)
von: Zhan, Wengyi, et al.
Veröffentlicht: (2025)
DiffusionInv: Prior-enhanced Bayesian Full Waveform Inversion using Diffusion models
von: Li, Yuanyuan, et al.
Veröffentlicht: (2025)
von: Li, Yuanyuan, et al.
Veröffentlicht: (2025)
AccDiffusion v2: Towards More Accurate Higher-Resolution Diffusion Extrapolation
von: Lin, Zhihang, et al.
Veröffentlicht: (2024)
von: Lin, Zhihang, et al.
Veröffentlicht: (2024)
AccDiffusion: An Accurate Method for Higher-Resolution Image Generation
von: Lin, Zhihang, et al.
Veröffentlicht: (2024)
von: Lin, Zhihang, et al.
Veröffentlicht: (2024)
InvAD: Inversion-based Reconstruction-Free Anomaly Detection with Diffusion Models
von: Sakai, Shunsuke, et al.
Veröffentlicht: (2025)
von: Sakai, Shunsuke, et al.
Veröffentlicht: (2025)
FreeInv: Free Lunch for Improving DDIM Inversion
von: Bao, Yuxiang, et al.
Veröffentlicht: (2025)
von: Bao, Yuxiang, et al.
Veröffentlicht: (2025)
CutDiffusion: A Simple, Fast, Cheap, and Strong Diffusion Extrapolation Method
von: Lin, Mingbao, et al.
Veröffentlicht: (2024)
von: Lin, Mingbao, et al.
Veröffentlicht: (2024)
ObjectAdd: Adding Objects into Image via a Training-Free Diffusion Modification Fashion
von: Zhang, Ziyue, et al.
Veröffentlicht: (2024)
von: Zhang, Ziyue, et al.
Veröffentlicht: (2024)
IterInv: Iterative Inversion for Pixel-Level T2I Models
von: Tang, Chuanming, et al.
Veröffentlicht: (2023)
von: Tang, Chuanming, et al.
Veröffentlicht: (2023)
LocInv: Localization-aware Inversion for Text-Guided Image Editing
von: Tang, Chuanming, et al.
Veröffentlicht: (2024)
von: Tang, Chuanming, et al.
Veröffentlicht: (2024)
InvDiff: Invariant Guidance for Bias Mitigation in Diffusion Models
von: Hou, Min, et al.
Veröffentlicht: (2024)
von: Hou, Min, et al.
Veröffentlicht: (2024)
Inv-Adapter: ID Customization Generation via Image Inversion and Lightweight Adapter
von: Xing, Peng, et al.
Veröffentlicht: (2024)
von: Xing, Peng, et al.
Veröffentlicht: (2024)
Deep Instruction Tuning for Segment Anything Model
von: Huang, Xiaorui, et al.
Veröffentlicht: (2024)
von: Huang, Xiaorui, et al.
Veröffentlicht: (2024)
DiffusionTrend: A Minimalist Approach to Virtual Fashion Try-On
von: Zhan, Wengyi, et al.
Veröffentlicht: (2024)
von: Zhan, Wengyi, et al.
Veröffentlicht: (2024)
FA-Depth: Toward Fast and Accurate Self-supervised Monocular Depth Estimation
von: Wang, Fei, et al.
Veröffentlicht: (2024)
von: Wang, Fei, et al.
Veröffentlicht: (2024)
MBQuant: A Novel Multi-Branch Topology Method for Arbitrary Bit-width Network Quantization
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
Scaling the Long Video Understanding of Multimodal Large Language Models via Visual Memory Mechanism
von: Chen, Tao, et al.
Veröffentlicht: (2026)
von: Chen, Tao, et al.
Veröffentlicht: (2026)
Sketch and Refine: Towards Fast and Accurate Lane Detection
von: Chen, Chao, et al.
Veröffentlicht: (2024)
von: Chen, Chao, et al.
Veröffentlicht: (2024)
Routing Experts: Learning to Route Dynamic Experts in Multi-modal Large Language Models
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
Not All Attention is Needed: Parameter and Computation Efficient Transfer Learning for Multi-modal Large Language Models
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
Textual Inversion and Self-supervised Refinement for Radiology Report Generation
von: Luo, Yuanjiang, et al.
Veröffentlicht: (2024)
von: Luo, Yuanjiang, et al.
Veröffentlicht: (2024)
Self-supervised Representations and Node Embedding Graph Neural Networks for Accurate and Multi-scale Analysis of Materials
von: Kong, Jian-Gang, et al.
Veröffentlicht: (2022)
von: Kong, Jian-Gang, et al.
Veröffentlicht: (2022)
MMAPS: End-to-End Multi-Grained Multi-Modal Attribute-Aware Product Summarization
von: Chen, Tao, et al.
Veröffentlicht: (2023)
von: Chen, Tao, et al.
Veröffentlicht: (2023)
TAD: Temporal-Aware Trajectory Self-Distillation for Fast and Accurate Diffusion LLM
von: Zhou, Haoyang, et al.
Veröffentlicht: (2026)
von: Zhou, Haoyang, et al.
Veröffentlicht: (2026)
AFGI: Towards Accurate and Fast-convergent Gradient Inversion Attack in Federated Learning
von: Liu, Can, et al.
Veröffentlicht: (2024)
von: Liu, Can, et al.
Veröffentlicht: (2024)
Fast Text-to-3D-Aware Face Generation and Manipulation via Direct Cross-modal Mapping and Geometric Regularization
von: Zhang, Jinlu, et al.
Veröffentlicht: (2024)
von: Zhang, Jinlu, et al.
Veröffentlicht: (2024)
Scan-specific Self-supervised Bayesian Deep Non-linear Inversion for Undersampled MRI Reconstruction
von: Leynes, Andrew P., et al.
Veröffentlicht: (2022)
von: Leynes, Andrew P., et al.
Veröffentlicht: (2022)
InvFusion: Bridging Supervised and Zero-shot Diffusion for Inverse Problems
von: Elata, Noam, et al.
Veröffentlicht: (2025)
von: Elata, Noam, et al.
Veröffentlicht: (2025)
Semi-supervised Counting via Pixel-by-pixel Density Distribution Modelling
von: Lin, Hui, et al.
Veröffentlicht: (2024)
von: Lin, Hui, et al.
Veröffentlicht: (2024)
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
Towards Accurate Post-Training Quantization of Vision Transformers via Error Reduction
von: Zhong, Yunshan, et al.
Veröffentlicht: (2024)
von: Zhong, Yunshan, et al.
Veröffentlicht: (2024)
AnchorInv: Few-Shot Class-Incremental Learning of Physiological Signals via Representation Space Guided Inversion
von: Li, Chenqi, et al.
Veröffentlicht: (2024)
von: Li, Chenqi, et al.
Veröffentlicht: (2024)
Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models
von: Luo, Gen, et al.
Veröffentlicht: (2024)
von: Luo, Gen, et al.
Veröffentlicht: (2024)
SPLIT: Self-supervised Partitioning for Learned Inversion in Nonlinear Tomography
von: Haltmeier, Markus, et al.
Veröffentlicht: (2026)
von: Haltmeier, Markus, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
EasyInv: Toward Fast and Better DDIM Inversion
von: Zhang, Ziyue, et al.
Veröffentlicht: (2024) -
ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling
von: Ju, Shaobo, et al.
Veröffentlicht: (2026) -
InvSeg: Test-Time Prompt Inversion for Semantic Segmentation
von: Lin, Jiayi, et al.
Veröffentlicht: (2024) -
InvCoSS: Inversion-driven Continual Self-supervised Learning in Medical Multi-modal Image Pre-training
von: Luo, Zihao, et al.
Veröffentlicht: (2025) -
Boosting Multimodal Large Language Models with Visual Tokens Withdrawal for Rapid Inference
von: Lin, Zhihang, et al.
Veröffentlicht: (2024)