End-to-End Fine-Tuning of 3D Texture Generation using Differentiable Rewards
Fuente:
arXiv
Saved in:
| Main Authors: | Zamani, AmirHossein, Xie, Tianhao, Aghdam, Amir G., Popa, Tiberiu, Belilovsky, Eugene |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Temporally Consistent Object Editing in Videos using Extended Attention
by: Zamani, AmirHossein, et al.
Published: (2024)
by: Zamani, AmirHossein, et al.
Published: (2024)
Sketch-guided Cage-based 3D Gaussian Splatting Deformation
by: Xie, Tianhao, et al.
Published: (2024)
by: Xie, Tianhao, et al.
Published: (2024)
Unsupervised Representation Learning for 3D Mesh Parameterization with Semantic and Visibility Objectives
by: Zamani, AmirHossein, et al.
Published: (2025)
by: Zamani, AmirHossein, et al.
Published: (2025)
Towards motion from video diffusion models
by: Janson, Paul, et al.
Published: (2024)
by: Janson, Paul, et al.
Published: (2024)
A Scalable Attention-Based Approach for Image-to-3D Texture Mapping
by: Rampini, Arianna, et al.
Published: (2025)
by: Rampini, Arianna, et al.
Published: (2025)
FACT-GS: Frequency-Aligned Complexity-Aware Texture Reparameterization for 2D Gaussian Splatting
by: Xie, Tianhao, et al.
Published: (2025)
by: Xie, Tianhao, et al.
Published: (2025)
Confident Splatting: Confidence-Based Compression of 3D Gaussian Splatting via Learnable Beta Distributions
by: Razlighi, AmirHossein Naghi, et al.
Published: (2025)
by: Razlighi, AmirHossein Naghi, et al.
Published: (2025)
Deep Nonlinear Hyperspectral Unmixing Using Multi-task Learning
by: Mehrdad, Saeid, et al.
Published: (2024)
by: Mehrdad, Saeid, et al.
Published: (2024)
DragD3D: Realistic Mesh Editing with Rigidity Control Driven by 2D Diffusion Priors
by: Xie, Tianhao, et al.
Published: (2023)
by: Xie, Tianhao, et al.
Published: (2023)
Sound Sparks Motion: Audio and Text Tuning for Video Editing
by: Razlighi, AmirHossein Naghi, et al.
Published: (2026)
by: Razlighi, AmirHossein Naghi, et al.
Published: (2026)
End-to-End Vision Tokenizer Tuning
by: Wang, Wenxuan, et al.
Published: (2025)
by: Wang, Wenxuan, et al.
Published: (2025)
Optimizing Data Augmentation for Real-Time Small UAV Detection: A Lightweight Context-Aware Approach
by: Zamani, Amir, et al.
Published: (2026)
by: Zamani, Amir, et al.
Published: (2026)
Neural 4D Evolution under Large Topological Changes from 2D Images
by: Razlighi, AmirHossein Naghi, et al.
Published: (2024)
by: Razlighi, AmirHossein Naghi, et al.
Published: (2024)
Hunyuan3D Studio: End-to-End AI Pipeline for Game-Ready 3D Asset Generation
by: Lei, Biwen, et al.
Published: (2025)
by: Lei, Biwen, et al.
Published: (2025)
BEEP3D: Box-Supervised End-to-End Pseudo-Mask Generation for 3D Instance Segmentation
by: Yoo, Youngju, et al.
Published: (2025)
by: Yoo, Youngju, et al.
Published: (2025)
AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning
by: Zhou, Zewei, et al.
Published: (2025)
by: Zhou, Zewei, et al.
Published: (2025)
End to End Face Reconstruction via Differentiable PnP
by: Lu, Yiren, et al.
Published: (2024)
by: Lu, Yiren, et al.
Published: (2024)
Directly Fine-Tuning Diffusion Models on Differentiable Rewards
by: Clark, Kevin, et al.
Published: (2023)
by: Clark, Kevin, et al.
Published: (2023)
REPA-E: Unlocking VAE for End-to-End Tuning with Latent Diffusion Transformers
by: Leng, Xingjian, et al.
Published: (2025)
by: Leng, Xingjian, et al.
Published: (2025)
Generative Planning with 3D-vision Language Pre-training for End-to-End Autonomous Driving
by: Li, Tengpeng, et al.
Published: (2025)
by: Li, Tengpeng, et al.
Published: (2025)
An End-to-End Decision-Aware Multi-Scale Attention-Based Model for Explainable Autonomous Driving
by: Azad, Maryam Sadat Hosseini, et al.
Published: (2026)
by: Azad, Maryam Sadat Hosseini, et al.
Published: (2026)
MangaFlow: An End-to-End Agentic Framework for Controllable Story to Manga Generation
by: Wang, Muyao, et al.
Published: (2026)
by: Wang, Muyao, et al.
Published: (2026)
Rethinking the Spatio-Temporal Alignment of End-to-End 3D Perception
by: Li, Xiaoyu, et al.
Published: (2025)
by: Li, Xiaoyu, et al.
Published: (2025)
Diffusion-DRF: Free, Rich, and Differentiable Reward for Video Diffusion Fine-Tuning
by: Wang, Yifan, et al.
Published: (2026)
by: Wang, Yifan, et al.
Published: (2026)
Rethinking End-to-End 2D to 3D Scene Segmentation in Gaussian Splatting
by: Zhu, Runsong, et al.
Published: (2025)
by: Zhu, Runsong, et al.
Published: (2025)
Differentiable NMS via Sinkhorn Matching for End-to-End Fabric Defect Detection
by: Lu, Zhengyang, et al.
Published: (2025)
by: Lu, Zhengyang, et al.
Published: (2025)
RAP: 3D Rasterization Augmented End-to-End Planning
by: Feng, Lan, et al.
Published: (2025)
by: Feng, Lan, et al.
Published: (2025)
ActAlign: Zero-Shot Fine-Grained Video Classification via Language-Guided Sequence Alignment
by: Aghdam, Amir, et al.
Published: (2025)
by: Aghdam, Amir, et al.
Published: (2025)
MALLVI: A Multi-Agent Framework for Integrated Generalized Robotics Manipulation
by: Taji, Mehrshad, et al.
Published: (2026)
by: Taji, Mehrshad, et al.
Published: (2026)
SS3D: End2End Self-Supervised 3D from Web Videos
by: Hariat, Marwane, et al.
Published: (2026)
by: Hariat, Marwane, et al.
Published: (2026)
E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models
by: Cong, Wenyan, et al.
Published: (2025)
by: Cong, Wenyan, et al.
Published: (2025)
VLM-3D:End-to-End Vision-Language Models for Open-World 3D Perception
by: Chang, Fuhao, et al.
Published: (2025)
by: Chang, Fuhao, et al.
Published: (2025)
A Differentiable Wave Optics Model for End-to-End Computational Imaging System Optimization
by: Ho, Chi-Jui, et al.
Published: (2024)
by: Ho, Chi-Jui, et al.
Published: (2024)
GenAD: Generative End-to-End Autonomous Driving
by: Zheng, Wenzhao, et al.
Published: (2024)
by: Zheng, Wenzhao, et al.
Published: (2024)
Generative Scenario Rollouts for End-to-End Autonomous Driving
by: Yasarla, Rajeev, et al.
Published: (2026)
by: Yasarla, Rajeev, et al.
Published: (2026)
End-to-End Rate-Distortion Optimized 3D Gaussian Representation
by: Wang, Henan, et al.
Published: (2024)
by: Wang, Henan, et al.
Published: (2024)
Direct Reward Fine-Tuning on Poses for Single Image to 3D Human in the Wild
by: Do, Seunguk, et al.
Published: (2026)
by: Do, Seunguk, et al.
Published: (2026)
End-to-End Autonomous Driving without Costly Modularization and 3D Manual Annotation
by: Guo, Mingzhe, et al.
Published: (2024)
by: Guo, Mingzhe, et al.
Published: (2024)
Puzzles: Unbounded Video-Depth Augmentation for Scalable End-to-End 3D Reconstruction
by: Ma, Jiahao, et al.
Published: (2025)
by: Ma, Jiahao, et al.
Published: (2025)
SafeDrive: Fine-Grained Safety Reasoning for End-to-End Driving in a Sparse World
by: Kim, Jungho, et al.
Published: (2026)
by: Kim, Jungho, et al.
Published: (2026)
Similar Items
-
Temporally Consistent Object Editing in Videos using Extended Attention
by: Zamani, AmirHossein, et al.
Published: (2024) -
Sketch-guided Cage-based 3D Gaussian Splatting Deformation
by: Xie, Tianhao, et al.
Published: (2024) -
Unsupervised Representation Learning for 3D Mesh Parameterization with Semantic and Visibility Objectives
by: Zamani, AmirHossein, et al.
Published: (2025) -
Towards motion from video diffusion models
by: Janson, Paul, et al.
Published: (2024) -
A Scalable Attention-Based Approach for Image-to-3D Texture Mapping
by: Rampini, Arianna, et al.
Published: (2025)