DICE: Discrete Inversion Enabling Controllable Editing for Multinomial Diffusion and Masked Generative Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | He, Xiaoxiao, Dao, Quan, Han, Ligong, Wen, Song, Bai, Minhao, Liu, Di, Zhang, Han, Min, Martin Renqiang, Juefei-Xu, Felix, Tan, Chaowei, Liu, Bo, Li, Kang, Li, Hongdong, Huang, Junzhou, Ahmed, Faez, Srivastava, Akash, Metaxas, Dimitris |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Discrete Noise Inversion for Next-scale Autoregressive Text-based Image Editing
von: Dao, Quan, et al.
Veröffentlicht: (2025)
von: Dao, Quan, et al.
Veröffentlicht: (2025)
Spectrum-Aware Parameter Efficient Fine-Tuning for Diffusion Models
von: Zhang, Xinxi, et al.
Veröffentlicht: (2024)
von: Zhang, Xinxi, et al.
Veröffentlicht: (2024)
MPDiT: Multi-Patch Global-to-Local Transformer Architecture For Efficient Flow Matching and Diffusion Model
von: Dao, Quan, et al.
Veröffentlicht: (2026)
von: Dao, Quan, et al.
Veröffentlicht: (2026)
Score-Guided Diffusion for 3D Human Recovery
von: Stathopoulos, Anastasis, et al.
Veröffentlicht: (2024)
von: Stathopoulos, Anastasis, et al.
Veröffentlicht: (2024)
SINE: SINgle Image Editing with Text-to-Image Diffusion Models
von: Zhang, Zhixing, et al.
Veröffentlicht: (2022)
von: Zhang, Zhixing, et al.
Veröffentlicht: (2022)
Rate-My-LoRA: Efficient and Adaptive Federated Model Tuning for Cardiac MRI Segmentation
von: He, Xiaoxiao, et al.
Veröffentlicht: (2025)
von: He, Xiaoxiao, et al.
Veröffentlicht: (2025)
Overcoming the Curvature Bottleneck in MeanFlow
von: Zhang, Xinxi, et al.
Veröffentlicht: (2025)
von: Zhang, Xinxi, et al.
Veröffentlicht: (2025)
AutoEdit: Automatic Hyperparameter Tuning for Image Editing
von: Pham, Chau, et al.
Veröffentlicht: (2025)
von: Pham, Chau, et al.
Veröffentlicht: (2025)
BLoB: Bayesian Low-Rank Adaptation by Backpropagation for Large Language Models
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
PrefGen: Multimodal Preference Learning for Preference-Conditioned Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
Self-Corrected Flow Distillation for Consistent One-Step and Few-Step Text-to-Image Generation
von: Dao, Quan, et al.
Veröffentlicht: (2024)
von: Dao, Quan, et al.
Veröffentlicht: (2024)
Improved Training Technique for Latent Consistency Models
von: Dao, Quan, et al.
Veröffentlicht: (2025)
von: Dao, Quan, et al.
Veröffentlicht: (2025)
SNLP: Layer-Parallel Inference via Structured Newton Corrections
von: Han, Ligong, et al.
Veröffentlicht: (2026)
von: Han, Ligong, et al.
Veröffentlicht: (2026)
SQuat: Subspace-orthogonal KV Cache Quantization
von: Wang, Hao, et al.
Veröffentlicht: (2025)
von: Wang, Hao, et al.
Veröffentlicht: (2025)
S2D2: Fast Decoding for Diffusion LLMs via Training-Free Self-Speculation
von: Han, Ligong, et al.
Veröffentlicht: (2026)
von: Han, Ligong, et al.
Veröffentlicht: (2026)
Few-Step Diffusion Language Models via Trajectory Self-Distillation
von: Zhang, Tunyu, et al.
Veröffentlicht: (2026)
von: Zhang, Tunyu, et al.
Veröffentlicht: (2026)
LoR-VP: Low-Rank Visual Prompting for Efficient Vision Model Adaptation
von: Jin, Can, et al.
Veröffentlicht: (2025)
von: Jin, Can, et al.
Veröffentlicht: (2025)
DiMSUM: Diffusion Mamba -- A Scalable and Unified Spatial-Frequency Method for Image Generation
von: Phung, Hao, et al.
Veröffentlicht: (2024)
von: Phung, Hao, et al.
Veröffentlicht: (2024)
Implicit In-context Learning
von: Li, Zhuowei, et al.
Veröffentlicht: (2024)
von: Li, Zhuowei, et al.
Veröffentlicht: (2024)
Hopscotch: Discovering and Skipping Redundancies in Language Models
von: Eyceoz, Mustafa, et al.
Veröffentlicht: (2025)
von: Eyceoz, Mustafa, et al.
Veröffentlicht: (2025)
Large Sign Language Models: Toward 3D American Sign Language Translation
von: Zhang, Sen, et al.
Veröffentlicht: (2025)
von: Zhang, Sen, et al.
Veröffentlicht: (2025)
Beyond Statistical Similarity: Rethinking Metrics for Deep Generative Models in Engineering Design
von: Regenwetter, Lyle, et al.
Veröffentlicht: (2023)
von: Regenwetter, Lyle, et al.
Veröffentlicht: (2023)
Optimize Any Topology: A Foundation Model for Shape- and Resolution-Free Structural Topology Optimization
von: Nobari, Amin Heyrani, et al.
Veröffentlicht: (2025)
von: Nobari, Amin Heyrani, et al.
Veröffentlicht: (2025)
TokUR: Token-Level Uncertainty Estimation for Large Language Model Reasoning
von: Zhang, Tunyu, et al.
Veröffentlicht: (2025)
von: Zhang, Tunyu, et al.
Veröffentlicht: (2025)
Multinomial mixture for spatial data
von: Nalpantidi, Anna, et al.
Veröffentlicht: (2024)
von: Nalpantidi, Anna, et al.
Veröffentlicht: (2024)
Input Snapshots Fusion for Scalable Discrete-Time Dynamic Graph Neural Networks
von: Qi, QingGuo, et al.
Veröffentlicht: (2024)
von: Qi, QingGuo, et al.
Veröffentlicht: (2024)
DICE: Discrete Interpretable Comparative Evaluation with Probabilistic Scoring for Retrieval-Augmented Generation
von: Liu, Shiyan, et al.
Veröffentlicht: (2025)
von: Liu, Shiyan, et al.
Veröffentlicht: (2025)
Embedding Inversion via Conditional Masked Diffusion Language Models
von: Xiao, Han
Veröffentlicht: (2026)
von: Xiao, Han
Veröffentlicht: (2026)
Learning to Localize Actions in Instructional Videos with LLM-Based Multi-Pathway Text-Video Alignment
von: Chen, Yuxiao, et al.
Veröffentlicht: (2024)
von: Chen, Yuxiao, et al.
Veröffentlicht: (2024)
MLLM-as-a-Judge for Image Safety without Human Labeling
von: Wang, Zhenting, et al.
Veröffentlicht: (2024)
von: Wang, Zhenting, et al.
Veröffentlicht: (2024)
Constraining Generative Models for Engineering Design with Negative Data
von: Regenwetter, Lyle, et al.
Veröffentlicht: (2023)
von: Regenwetter, Lyle, et al.
Veröffentlicht: (2023)
M^3-Bench: Multi-Modal, Multi-Hop, Multi-Threaded Tool-Using MLLM Agent Benchmark
von: Zhou, Yang, et al.
Veröffentlicht: (2025)
von: Zhou, Yang, et al.
Veröffentlicht: (2025)
Can Large Vision-Language Models Detect Images Copyright Infringement from GenAI?
von: Xu, Qipan, et al.
Veröffentlicht: (2025)
von: Xu, Qipan, et al.
Veröffentlicht: (2025)
Beyond Explicit Edges: Robust Reasoning over Noisy and Sparse Knowledge Graphs
von: Gao, Hang, et al.
Veröffentlicht: (2026)
von: Gao, Hang, et al.
Veröffentlicht: (2026)
Generative Optimization: A Perspective on AI-Enhanced Problem Solving in Engineering
von: Regenwetter, Lyle, et al.
Veröffentlicht: (2024)
von: Regenwetter, Lyle, et al.
Veröffentlicht: (2024)
LInK: Learning Joint Representations of Design and Performance Spaces through Contrastive Learning for Mechanism Synthesis
von: Nobari, Amin Heyrani, et al.
Veröffentlicht: (2024)
von: Nobari, Amin Heyrani, et al.
Veröffentlicht: (2024)
DICE: Discrete inverse continuity equation for learning population dynamics
von: Blickhan, Tobias, et al.
Veröffentlicht: (2025)
von: Blickhan, Tobias, et al.
Veröffentlicht: (2025)
Let Watermarks Speak: A Robust and Unforgeable Watermark for Language Models
von: Bai, Minhao
Veröffentlicht: (2024)
von: Bai, Minhao
Veröffentlicht: (2024)
SteerFlow: Steering Rectified Flows for Faithful Inversion-Based Image Editing
von: Dao, Thinh, et al.
Veröffentlicht: (2026)
von: Dao, Thinh, et al.
Veröffentlicht: (2026)
APEER: Automatic Prompt Engineering Enhances Large Language Model Reranking
von: Jin, Can, et al.
Veröffentlicht: (2024)
von: Jin, Can, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Discrete Noise Inversion for Next-scale Autoregressive Text-based Image Editing
von: Dao, Quan, et al.
Veröffentlicht: (2025) -
Spectrum-Aware Parameter Efficient Fine-Tuning for Diffusion Models
von: Zhang, Xinxi, et al.
Veröffentlicht: (2024) -
MPDiT: Multi-Patch Global-to-Local Transformer Architecture For Efficient Flow Matching and Diffusion Model
von: Dao, Quan, et al.
Veröffentlicht: (2026) -
Score-Guided Diffusion for 3D Human Recovery
von: Stathopoulos, Anastasis, et al.
Veröffentlicht: (2024) -
SINE: SINgle Image Editing with Text-to-Image Diffusion Models
von: Zhang, Zhixing, et al.
Veröffentlicht: (2022)