Direct Reward Fine-Tuning on Poses for Single Image to 3D Human in the Wild
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Do, Seunguk, Huh, Minwoo, Shin, Joonghyuk, Park, Jaesik |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
InstantDrag: Improving Interactivity in Drag-based Image Editing
von: Shin, Joonghyuk, et al.
Veröffentlicht: (2024)
von: Shin, Joonghyuk, et al.
Veröffentlicht: (2024)
SEAL-pose: Enhancing 3D Human Pose Estimation via a Learned Loss for Structural Consistency
von: Kim, Yeonsung, et al.
Veröffentlicht: (2026)
von: Kim, Yeonsung, et al.
Veröffentlicht: (2026)
Exploring Multimodal Diffusion Transformers for Enhanced Prompt-based Image Editing
von: Shin, Joonghyuk, et al.
Veröffentlicht: (2025)
von: Shin, Joonghyuk, et al.
Veröffentlicht: (2025)
Leveraging Learned Image Prior for 3D Gaussian Compression
von: Shin, Seungjoo, et al.
Veröffentlicht: (2025)
von: Shin, Seungjoo, et al.
Veröffentlicht: (2025)
JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching
von: Kwon, Mingi, et al.
Veröffentlicht: (2025)
von: Kwon, Mingi, et al.
Veröffentlicht: (2025)
MotionStream: Real-Time Video Generation with Interactive Motion Controls
von: Shin, Joonghyuk, et al.
Veröffentlicht: (2025)
von: Shin, Joonghyuk, et al.
Veröffentlicht: (2025)
360 in the Wild: Dataset for Depth Prediction and View Synthesis
von: Park, Kibaek, et al.
Veröffentlicht: (2024)
von: Park, Kibaek, et al.
Veröffentlicht: (2024)
PoseSyn: Synthesizing Diverse 3D Pose Data from In-the-Wild 2D Data
von: Yang, ChangHee, et al.
Veröffentlicht: (2025)
von: Yang, ChangHee, et al.
Veröffentlicht: (2025)
Locality-aware Gaussian Compression for Fast and High-quality Rendering
von: Shin, Seungjoo, et al.
Veröffentlicht: (2025)
von: Shin, Seungjoo, et al.
Veröffentlicht: (2025)
Metropolis-Hastings Sampling for 3D Gaussian Reconstruction
von: Kim, Hyunjin, et al.
Veröffentlicht: (2025)
von: Kim, Hyunjin, et al.
Veröffentlicht: (2025)
PandaPose: 3D Human Pose Lifting from a Single Image via Propagating 2D Pose Prior to 3D Anchor Space
von: Zheng, Jinghong, et al.
Veröffentlicht: (2026)
von: Zheng, Jinghong, et al.
Veröffentlicht: (2026)
L3D-Pose: Lifting Pose for 3D Avatars from a Single Camera in the Wild
von: Debnath, Soumyaratna, et al.
Veröffentlicht: (2025)
von: Debnath, Soumyaratna, et al.
Veröffentlicht: (2025)
Directly Fine-Tuning Diffusion Models on Differentiable Rewards
von: Clark, Kevin, et al.
Veröffentlicht: (2023)
von: Clark, Kevin, et al.
Veröffentlicht: (2023)
Extend3D: Town-Scale 3D Generation
von: Yoon, Seungwoo, et al.
Veröffentlicht: (2026)
von: Yoon, Seungwoo, et al.
Veröffentlicht: (2026)
Video Consistency Distance: Enhancing Temporal Consistency for Image-to-Video Generation via Reward-Based Fine-Tuning
von: Aoshima, Takehiro, et al.
Veröffentlicht: (2025)
von: Aoshima, Takehiro, et al.
Veröffentlicht: (2025)
Pose-RFT: Enhancing MLLMs for 3D Pose Generation via Hybrid Action Reinforcement Fine-Tuning
von: Li, Bao, et al.
Veröffentlicht: (2025)
von: Li, Bao, et al.
Veröffentlicht: (2025)
Coherent Human-Scene Reconstruction from Multi-Person Multi-View Video in a Single Pass
von: Kim, Sangmin, et al.
Veröffentlicht: (2026)
von: Kim, Sangmin, et al.
Veröffentlicht: (2026)
SnapPose3D: Diffusion-Based Single-Frame 2D-to-3D Lifting of Human Poses
von: Simoni, Alessandro, et al.
Veröffentlicht: (2026)
von: Simoni, Alessandro, et al.
Veröffentlicht: (2026)
Human Video Generation from a Single Image with 3D Pose and View Control
von: Wang, Tiantian, et al.
Veröffentlicht: (2026)
von: Wang, Tiantian, et al.
Veröffentlicht: (2026)
CF3: Compact and Fast 3D Feature Fields
von: Lee, Hyunjoon, et al.
Veröffentlicht: (2025)
von: Lee, Hyunjoon, et al.
Veröffentlicht: (2025)
EgoCast: Forecasting Egocentric Human Pose in the Wild
von: Escobar, Maria, et al.
Veröffentlicht: (2024)
von: Escobar, Maria, et al.
Veröffentlicht: (2024)
End-to-End Fine-Tuning of 3D Texture Generation using Differentiable Rewards
von: Zamani, AmirHossein, et al.
Veröffentlicht: (2025)
von: Zamani, AmirHossein, et al.
Veröffentlicht: (2025)
3Doodle: Compact Abstraction of Objects with 3D Strokes
von: Choi, Changwoon, et al.
Veröffentlicht: (2024)
von: Choi, Changwoon, et al.
Veröffentlicht: (2024)
WildPose: A Unified Framework for Robust Pose Estimation in the Wild
von: Zheng, Jianhao, et al.
Veröffentlicht: (2026)
von: Zheng, Jianhao, et al.
Veröffentlicht: (2026)
Recovering Dynamic 3D Sketches from Videos
von: Lee, Jaeah, et al.
Veröffentlicht: (2025)
von: Lee, Jaeah, et al.
Veröffentlicht: (2025)
Designing Concise ConvNets with Columnar Stages
von: Kumar, Ashish, et al.
Veröffentlicht: (2024)
von: Kumar, Ashish, et al.
Veröffentlicht: (2024)
Cross Resolution Encoding-Decoding For Detection Transformers
von: Kumar, Ashish, et al.
Veröffentlicht: (2024)
von: Kumar, Ashish, et al.
Veröffentlicht: (2024)
FinePOSE: Fine-Grained Prompt-Driven 3D Human Pose Estimation via Diffusion Models
von: Xu, Jinglin, et al.
Veröffentlicht: (2024)
von: Xu, Jinglin, et al.
Veröffentlicht: (2024)
SelfSplat: Pose-Free and 3D Prior-Free Generalizable 3D Gaussian Splatting
von: Kang, Gyeongjin, et al.
Veröffentlicht: (2024)
von: Kang, Gyeongjin, et al.
Veröffentlicht: (2024)
Referring Human Pose and Mask Estimation in the Wild
von: Miao, Bo, et al.
Veröffentlicht: (2024)
von: Miao, Bo, et al.
Veröffentlicht: (2024)
Pre-Training for 3D Hand Pose Estimation with Contrastive Learning on Large-Scale Hand Images in the Wild
von: Lin, Nie, et al.
Veröffentlicht: (2024)
von: Lin, Nie, et al.
Veröffentlicht: (2024)
OpenBox: Annotate Any Bounding Boxes in 3D
von: Lee, In-Jae, et al.
Veröffentlicht: (2025)
von: Lee, In-Jae, et al.
Veröffentlicht: (2025)
Learning SO(3)-Invariant Semantic Correspondence via Local Shape Transform
von: Park, Chunghyun, et al.
Veröffentlicht: (2024)
von: Park, Chunghyun, et al.
Veröffentlicht: (2024)
ChatPose: Chatting about 3D Human Pose
von: Feng, Yao, et al.
Veröffentlicht: (2023)
von: Feng, Yao, et al.
Veröffentlicht: (2023)
Improving the Robustness of 3D Human Pose Estimation: A Benchmark and Learning from Noisy Input
von: Hoang, Trung-Hieu, et al.
Veröffentlicht: (2023)
von: Hoang, Trung-Hieu, et al.
Veröffentlicht: (2023)
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
DressWild: Feed-Forward Pose-Agnostic Garment Sewing Pattern Generation from In-the-Wild Images
von: Tao, Zeng, et al.
Veröffentlicht: (2026)
von: Tao, Zeng, et al.
Veröffentlicht: (2026)
Canonical Pose Reconstruction from Single Depth Image for 3D Non-rigid Pose Recovery on Limited Datasets
von: Alhamazani, Fahd, et al.
Veröffentlicht: (2025)
von: Alhamazani, Fahd, et al.
Veröffentlicht: (2025)
PoseFix: Correcting 3D Human Poses with Natural Language
von: Delmas, Ginger, et al.
Veröffentlicht: (2023)
von: Delmas, Ginger, et al.
Veröffentlicht: (2023)
PoseScript: Linking 3D Human Poses and Natural Language
von: Delmas, Ginger, et al.
Veröffentlicht: (2022)
von: Delmas, Ginger, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
InstantDrag: Improving Interactivity in Drag-based Image Editing
von: Shin, Joonghyuk, et al.
Veröffentlicht: (2024) -
SEAL-pose: Enhancing 3D Human Pose Estimation via a Learned Loss for Structural Consistency
von: Kim, Yeonsung, et al.
Veröffentlicht: (2026) -
Exploring Multimodal Diffusion Transformers for Enhanced Prompt-based Image Editing
von: Shin, Joonghyuk, et al.
Veröffentlicht: (2025) -
Leveraging Learned Image Prior for 3D Gaussian Compression
von: Shin, Seungjoo, et al.
Veröffentlicht: (2025) -
JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching
von: Kwon, Mingi, et al.
Veröffentlicht: (2025)