Stitched Value Model for Diffusion Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Go, Hyojun, Chung, Hyungjin, Truong, Prune, Bhat, Goutam, Mi, Li, An, Zhaochong, Zhao, Zixiang, Narnhofer, Dominik, Belongie, Serge, Tombari, Federico, Schindler, Konrad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Text-to-3D by Stitching a Multi-view Reconstruction Network to a Video Generator
von: Go, Hyojun, et al.
Veröffentlicht: (2025)
von: Go, Hyojun, et al.
Veröffentlicht: (2025)
Elastic3D: Controllable Stereo Video Conversion with Guided Latent Decoding
von: Metzger, Nando, et al.
Veröffentlicht: (2025)
von: Metzger, Nando, et al.
Veröffentlicht: (2025)
M2SVid: End-to-End Inpainting and Refinement for Monocular-to-Stereo Video Conversion
von: Shvetsova, Nina, et al.
Veröffentlicht: (2025)
von: Shvetsova, Nina, et al.
Veröffentlicht: (2025)
Understanding, Accelerating, and Improving MeanFlow Training
von: Kim, Jin-Young, et al.
Veröffentlicht: (2025)
von: Kim, Jin-Young, et al.
Veröffentlicht: (2025)
Continuous Space-Time Video Super-Resolution with 3D Fourier Fields
von: Becker, Alexander, et al.
Veröffentlicht: (2025)
von: Becker, Alexander, et al.
Veröffentlicht: (2025)
Stitch: Training-Free Position Control in Multimodal Diffusion Transformers
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
Generating Human Motion Videos using a Cascaded Text-to-Video Framework
von: Nam, Hyelin, et al.
Veröffentlicht: (2025)
von: Nam, Hyelin, et al.
Veröffentlicht: (2025)
FlowSDF: Flow Matching for Medical Image Segmentation Using Distance Transforms
von: Bogensperger, Lea, et al.
Veröffentlicht: (2024)
von: Bogensperger, Lea, et al.
Veröffentlicht: (2024)
SteerX: Creating Any Camera-Free 3D and 4D Scenes with Geometric Steering
von: Park, Byeongjun, et al.
Veröffentlicht: (2025)
von: Park, Byeongjun, et al.
Veröffentlicht: (2025)
One2Any: One-Reference 6D Pose Estimation for Any Object
von: Liu, Mengya, et al.
Veröffentlicht: (2025)
von: Liu, Mengya, et al.
Veröffentlicht: (2025)
VideoRFSplat: Direct Scene-Level Text-to-3D Gaussian Splatting Generation with Flexible Pose and Multi-View Joint Modeling
von: Go, Hyojun, et al.
Veröffentlicht: (2025)
von: Go, Hyojun, et al.
Veröffentlicht: (2025)
Generalized Few-shot 3D Point Cloud Segmentation with Vision-Language Model
von: An, Zhaochong, et al.
Veröffentlicht: (2025)
von: An, Zhaochong, et al.
Veröffentlicht: (2025)
CubeDiff: Repurposing Diffusion-Based Image Models for Panorama Generation
von: Kalischek, Nikolai, et al.
Veröffentlicht: (2025)
von: Kalischek, Nikolai, et al.
Veröffentlicht: (2025)
Revisiting the Perception-Distortion Trade-off with Spatial-Semantic Guided Super-Resolution
von: Wang, Dan, et al.
Veröffentlicht: (2026)
von: Wang, Dan, et al.
Veröffentlicht: (2026)
Solving Inverse Problems with FLAIR
von: Erbach, Julius, et al.
Veröffentlicht: (2025)
von: Erbach, Julius, et al.
Veröffentlicht: (2025)
AnyUp: Universal Feature Upsampling
von: Wimmer, Thomas, et al.
Veröffentlicht: (2025)
von: Wimmer, Thomas, et al.
Veröffentlicht: (2025)
Unified Panoramic Geometry Estimation via Multi-View Foundation Models
von: Bozic, Vukasin, et al.
Veröffentlicht: (2026)
von: Bozic, Vukasin, et al.
Veröffentlicht: (2026)
Multimodality Helps Few-shot 3D Point Cloud Semantic Segmentation
von: An, Zhaochong, et al.
Veröffentlicht: (2024)
von: An, Zhaochong, et al.
Veröffentlicht: (2024)
Rethinking Few-shot 3D Point Cloud Semantic Segmentation
von: An, Zhaochong, et al.
Veröffentlicht: (2024)
von: An, Zhaochong, et al.
Veröffentlicht: (2024)
Video Depth without Video Models
von: Ke, Bingxin, et al.
Veröffentlicht: (2024)
von: Ke, Bingxin, et al.
Veröffentlicht: (2024)
Thera: Aliasing-Free Arbitrary-Scale Super-Resolution with Neural Heat Fields
von: Becker, Alexander, et al.
Veröffentlicht: (2023)
von: Becker, Alexander, et al.
Veröffentlicht: (2023)
VGGRPO: Towards World-Consistent Video Generation with 4D Latent Reward
von: An, Zhaochong, et al.
Veröffentlicht: (2026)
von: An, Zhaochong, et al.
Veröffentlicht: (2026)
HiddenObjects: Scalable Diffusion-Distilled Spatial Priors for Object Placement
von: Schouten, Marco, et al.
Veröffentlicht: (2026)
von: Schouten, Marco, et al.
Veröffentlicht: (2026)
Bridging Implicit and Explicit Geometric Transformation for Single-Image View Synthesis
von: Park, Byeongjun, et al.
Veröffentlicht: (2022)
von: Park, Byeongjun, et al.
Veröffentlicht: (2022)
Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs
von: Chung, Hyungjin, et al.
Veröffentlicht: (2025)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2025)
Better Language Models Exhibit Higher Visual Alignment
von: Ruthardt, Jona, et al.
Veröffentlicht: (2024)
von: Ruthardt, Jona, et al.
Veröffentlicht: (2024)
Deep Diffusion Image Prior for Efficient OOD Adaptation in 3D Inverse Problems
von: Chung, Hyungjin, et al.
Veröffentlicht: (2024)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2024)
Denoising Task Difficulty-based Curriculum for Training Diffusion Models
von: Kim, Jin-Young, et al.
Veröffentlicht: (2024)
von: Kim, Jin-Young, et al.
Veröffentlicht: (2024)
Denoising Task Routing for Diffusion Models
von: Park, Byeongjun, et al.
Veröffentlicht: (2023)
von: Park, Byeongjun, et al.
Veröffentlicht: (2023)
Labeled Data Selection for Category Discovery
von: Zhao, Bingchen, et al.
Veröffentlicht: (2024)
von: Zhao, Bingchen, et al.
Veröffentlicht: (2024)
Amortized Posterior Sampling with Diffusion Prior Distillation
von: Mammadov, Abbas, et al.
Veröffentlicht: (2024)
von: Mammadov, Abbas, et al.
Veröffentlicht: (2024)
A Variational Perspective on Generative Protein Fitness Optimization
von: Bogensperger, Lea, et al.
Veröffentlicht: (2025)
von: Bogensperger, Lea, et al.
Veröffentlicht: (2025)
ACDC: Autoregressive Coherent Multimodal Generation using Diffusion Correction
von: Chung, Hyungjin, et al.
Veröffentlicht: (2024)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2024)
Switch Diffusion Transformer: Synergizing Denoising Tasks with Sparse Mixture-of-Experts
von: Park, Byeongjun, et al.
Veröffentlicht: (2024)
von: Park, Byeongjun, et al.
Veröffentlicht: (2024)
Diffusion Model Patching via Mixture-of-Prompts
von: Ham, Seokil, et al.
Veröffentlicht: (2024)
von: Ham, Seokil, et al.
Veröffentlicht: (2024)
MMEarth-Bench: Global Model Adaptation via Multimodal Test-Time Training
von: Gordon, Lucia, et al.
Veröffentlicht: (2026)
von: Gordon, Lucia, et al.
Veröffentlicht: (2026)
Unlearning-based Neural Interpretations
von: Choi, Ching Lam, et al.
Veröffentlicht: (2024)
von: Choi, Ching Lam, et al.
Veröffentlicht: (2024)
Decomposed Diffusion Sampler for Accelerating Large-Scale Inverse Problems
von: Chung, Hyungjin, et al.
Veröffentlicht: (2023)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2023)
PhysConvex: Physics-Informed 3D Dynamic Convex Radiance Fields for Reconstruction and Simulation
von: Wang, Dan, et al.
Veröffentlicht: (2026)
von: Wang, Dan, et al.
Veröffentlicht: (2026)
EditCrafter: Tuning-free High-Resolution Image Editing via Pretrained Diffusion Model
von: Kim, Kunho, et al.
Veröffentlicht: (2026)
von: Kim, Kunho, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Text-to-3D by Stitching a Multi-view Reconstruction Network to a Video Generator
von: Go, Hyojun, et al.
Veröffentlicht: (2025) -
Elastic3D: Controllable Stereo Video Conversion with Guided Latent Decoding
von: Metzger, Nando, et al.
Veröffentlicht: (2025) -
M2SVid: End-to-End Inpainting and Refinement for Monocular-to-Stereo Video Conversion
von: Shvetsova, Nina, et al.
Veröffentlicht: (2025) -
Understanding, Accelerating, and Improving MeanFlow Training
von: Kim, Jin-Young, et al.
Veröffentlicht: (2025) -
Continuous Space-Time Video Super-Resolution with 3D Fourier Fields
von: Becker, Alexander, et al.
Veröffentlicht: (2025)