Align Your Query: Representation Alignment for Multimodality Medical Object Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Seo, Ara, Kim, Bryan Sangwoo, Chung, Hyungjin, Ye, Jong Chul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Chain-of-Zoom: Extreme Super-Resolution via Scale Autoregression and Preference Alignment
von: Kim, Bryan Sangwoo, et al.
Veröffentlicht: (2025)
von: Kim, Bryan Sangwoo, et al.
Veröffentlicht: (2025)
Free$^2$Guide: Training-Free Text-to-Video Alignment using Image LVLM
von: Kim, Jaemin, et al.
Veröffentlicht: (2024)
von: Kim, Jaemin, et al.
Veröffentlicht: (2024)
Extreme Blind Image Restoration via Prompt-Conditioned Information Bottleneck
von: Kim, Hongeun, et al.
Veröffentlicht: (2025)
von: Kim, Hongeun, et al.
Veröffentlicht: (2025)
FlowDPS: Flow-Driven Posterior Sampling for Inverse Problems
von: Kim, Jeongsol, et al.
Veröffentlicht: (2025)
von: Kim, Jeongsol, et al.
Veröffentlicht: (2025)
Decomposed Diffusion Sampler for Accelerating Large-Scale Inverse Problems
von: Chung, Hyungjin, et al.
Veröffentlicht: (2023)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2023)
Contrastive CFG: Improving CFG in Diffusion Models by Contrasting Positive and Negative Concepts
von: Chang, Jinho, et al.
Veröffentlicht: (2024)
von: Chang, Jinho, et al.
Veröffentlicht: (2024)
Tiled Prompts: Overcoming Prompt Misguidance in Image and Video Super-Resolution
von: Kim, Bryan Sangwoo, et al.
Veröffentlicht: (2026)
von: Kim, Bryan Sangwoo, et al.
Veröffentlicht: (2026)
Align Your Tangent: Training Better Consistency Models via Manifold-Aligned Tangents
von: Kim, Beomsu, et al.
Veröffentlicht: (2025)
von: Kim, Beomsu, et al.
Veröffentlicht: (2025)
Regularization by Texts for Latent Diffusion Inverse Solvers
von: Kim, Jeongsol, et al.
Veröffentlicht: (2023)
von: Kim, Jeongsol, et al.
Veröffentlicht: (2023)
Accelerating Video Inverse Problem Solvers with Autoregressive Diffusion Models
von: Kwon, Taesung, et al.
Veröffentlicht: (2026)
von: Kwon, Taesung, et al.
Veröffentlicht: (2026)
InverseCrafter: Efficient Video ReCapture as a Latent Domain Inverse Problem
von: Hong, Yeobin, et al.
Veröffentlicht: (2025)
von: Hong, Yeobin, et al.
Veröffentlicht: (2025)
Improving Diffusion Models for Inverse Problems using Manifold Constraints
von: Chung, Hyungjin, et al.
Veröffentlicht: (2022)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2022)
Diffusion Posterior Sampling for General Noisy Inverse Problems
von: Chung, Hyungjin, et al.
Veröffentlicht: (2022)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2022)
CFG++: Manifold-constrained Classifier Free Guidance for Diffusion Models
von: Chung, Hyungjin, et al.
Veröffentlicht: (2024)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2024)
ACDC: Autoregressive Coherent Multimodal Generation using Diffusion Correction
von: Chung, Hyungjin, et al.
Veröffentlicht: (2024)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2024)
Deep Diffusion Image Prior for Efficient OOD Adaptation in 3D Inverse Problems
von: Chung, Hyungjin, et al.
Veröffentlicht: (2024)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2024)
Solving 3D Inverse Problems using Pre-trained 2D Diffusion Models
von: Chung, Hyungjin, et al.
Veröffentlicht: (2022)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2022)
PCPO: Proportionate Credit Policy Optimization for Aligning Image Generation Models
von: Lee, Jeongjae, et al.
Veröffentlicht: (2025)
von: Lee, Jeongjae, et al.
Veröffentlicht: (2025)
Gradient-Free Noise Optimization for Reward Alignment in Generative Models
von: Kim, Jeongsol, et al.
Veröffentlicht: (2026)
von: Kim, Jeongsol, et al.
Veröffentlicht: (2026)
FlowAlign: Trajectory-Regularized, Inversion-Free Flow-based Image Editing
von: Kim, Jeongsol, et al.
Veröffentlicht: (2025)
von: Kim, Jeongsol, et al.
Veröffentlicht: (2025)
PromptLoop: Plug-and-Play Prompt Refinement via Latent Feedback for Diffusion Model Alignment
von: Lee, Suhyeon, et al.
Veröffentlicht: (2025)
von: Lee, Suhyeon, et al.
Veröffentlicht: (2025)
Aligning Text to Image in Diffusion Models is Easier Than You Think
von: Lee, Jaa-Yeon, et al.
Veröffentlicht: (2025)
von: Lee, Jaa-Yeon, et al.
Veröffentlicht: (2025)
MindFormer: Semantic Alignment of Multi-Subject fMRI for Brain Decoding
von: Han, Inhwa, et al.
Veröffentlicht: (2024)
von: Han, Inhwa, et al.
Veröffentlicht: (2024)
VideoGuide: Improving Video Diffusion Models without Training Through a Teacher's Guide
von: Lee, Dohun, et al.
Veröffentlicht: (2024)
von: Lee, Dohun, et al.
Veröffentlicht: (2024)
Alignment-Guided Score Matching for Text-to-Image Alignment in Diffusion Models
von: Lee, Jaa-Yeon, et al.
Veröffentlicht: (2026)
von: Lee, Jaa-Yeon, et al.
Veröffentlicht: (2026)
Amortized Posterior Sampling with Diffusion Prior Distillation
von: Mammadov, Abbas, et al.
Veröffentlicht: (2024)
von: Mammadov, Abbas, et al.
Veröffentlicht: (2024)
Latent Schrodinger Bridge: Prompting Latent Diffusion for Fast Unpaired Image-to-Image Translation
von: Kim, Jeongsol, et al.
Veröffentlicht: (2024)
von: Kim, Jeongsol, et al.
Veröffentlicht: (2024)
A Survey on Diffusion Models for Inverse Problems
von: Daras, Giannis, et al.
Veröffentlicht: (2024)
von: Daras, Giannis, et al.
Veröffentlicht: (2024)
Boost-and-Skip: A Simple Guidance-Free Diffusion for Minority Generation
von: Um, Soobin, et al.
Veröffentlicht: (2025)
von: Um, Soobin, et al.
Veröffentlicht: (2025)
OTSeg: Multi-prompt Sinkhorn Attention for Zero-Shot Semantic Segmentation
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2024)
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2024)
MotionCFG: Boosting Motion Dynamics via Stochastic Concept Perturbation
von: Kim, Byungjun, et al.
Veröffentlicht: (2026)
von: Kim, Byungjun, et al.
Veröffentlicht: (2026)
Draw Your Mind: Personalized Generation via Condition-Level Modeling in Text-to-Image Diffusion Models
von: Kim, Hyungjin, et al.
Veröffentlicht: (2025)
von: Kim, Hyungjin, et al.
Veröffentlicht: (2025)
Generalized Consistency Trajectory Models for Image Manipulation
von: Kim, Beomsu, et al.
Veröffentlicht: (2024)
von: Kim, Beomsu, et al.
Veröffentlicht: (2024)
CompoDistill: Attention Distillation for Compositional Reasoning in Multimodal LLMs
von: Kim, Jiwan, et al.
Veröffentlicht: (2025)
von: Kim, Jiwan, et al.
Veröffentlicht: (2025)
Ground-A-Video: Zero-shot Grounded Video Editing using Text-to-image Diffusion Models
von: Jeong, Hyeonho, et al.
Veröffentlicht: (2023)
von: Jeong, Hyeonho, et al.
Veröffentlicht: (2023)
FILT3R: Latent State Adaptive Kalman Filter for Streaming 3D Reconstruction
von: Jin, Seonghyun, et al.
Veröffentlicht: (2026)
von: Jin, Seonghyun, et al.
Veröffentlicht: (2026)
Solving Video Inverse Problems Using Image Diffusion Models
von: Kwon, Taesung, et al.
Veröffentlicht: (2024)
von: Kwon, Taesung, et al.
Veröffentlicht: (2024)
FlowLPS: Langevin-Proximal Sampling for Flow-based Inverse Problem Solvers
von: Park, Jonghyun, et al.
Veröffentlicht: (2025)
von: Park, Jonghyun, et al.
Veröffentlicht: (2025)
Minority-Focused Text-to-Image Generation via Prompt Optimization
von: Um, Soobin, et al.
Veröffentlicht: (2024)
von: Um, Soobin, et al.
Veröffentlicht: (2024)
VISION-XL: High Definition Video Inverse Problem Solver using Latent Image Diffusion Models
von: Kwon, Taesung, et al.
Veröffentlicht: (2024)
von: Kwon, Taesung, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Chain-of-Zoom: Extreme Super-Resolution via Scale Autoregression and Preference Alignment
von: Kim, Bryan Sangwoo, et al.
Veröffentlicht: (2025) -
Free$^2$Guide: Training-Free Text-to-Video Alignment using Image LVLM
von: Kim, Jaemin, et al.
Veröffentlicht: (2024) -
Extreme Blind Image Restoration via Prompt-Conditioned Information Bottleneck
von: Kim, Hongeun, et al.
Veröffentlicht: (2025) -
FlowDPS: Flow-Driven Posterior Sampling for Inverse Problems
von: Kim, Jeongsol, et al.
Veröffentlicht: (2025) -
Decomposed Diffusion Sampler for Accelerating Large-Scale Inverse Problems
von: Chung, Hyungjin, et al.
Veröffentlicht: (2023)