SK-Adapter: Skeleton-Based Structural Control for Native 3D Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Anbang, Ao, Yuzhuo, Wu, Shangzhe, Tang, Chi-Keung |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ReasonNavi: Human-Inspired Global Map Reasoning for Zero-Shot Embodied Navigation
por: Ao, Yuzhuo, et al.
Publicado: (2026)
por: Ao, Yuzhuo, et al.
Publicado: (2026)
Multimodal Generation of Animatable 3D Human Models with AvatarForge
por: Liu, Xinhang, et al.
Publicado: (2025)
por: Liu, Xinhang, et al.
Publicado: (2025)
InceptionHuman: Controllable Prompt-to-NeRF for Photorealistic 3D Human Generation
por: Kao, Shiu-hong, et al.
Publicado: (2023)
por: Kao, Shiu-hong, et al.
Publicado: (2023)
Agentic 3D Scene Generation with Spatially Contextualized VLMs
por: Liu, Xinhang, et al.
Publicado: (2025)
por: Liu, Xinhang, et al.
Publicado: (2025)
FED-NeRF: Achieve High 3D Consistency and Temporal Coherence for Face Video Editing on Dynamic NeRF
por: Zhang, Hao, et al.
Publicado: (2024)
por: Zhang, Hao, et al.
Publicado: (2024)
ChatCam: Empowering Camera Control through Conversational AI
por: Liu, Xinhang, et al.
Publicado: (2024)
por: Liu, Xinhang, et al.
Publicado: (2024)
ReelWave: Multi-Agentic Movie Sound Generation through Multimodal LLM Conversation
por: Wang, Zixuan, et al.
Publicado: (2025)
por: Wang, Zixuan, et al.
Publicado: (2025)
Farm3D: Learning Articulated 3D Animals by Distilling 2D Diffusion
por: Jakab, Tomas, et al.
Publicado: (2023)
por: Jakab, Tomas, et al.
Publicado: (2023)
iDAT: inverse Distillation Adapter-Tuning
por: Ruan, Jiacheng, et al.
Publicado: (2024)
por: Ruan, Jiacheng, et al.
Publicado: (2024)
Motion-Agent: A Conversational Framework for Human Motion Generation with LLMs
por: Wu, Qi, et al.
Publicado: (2024)
por: Wu, Qi, et al.
Publicado: (2024)
WorldCraft: Photo-Realistic 3D World Creation and Customization via LLM Agents
por: Liu, Xinhang, et al.
Publicado: (2025)
por: Liu, Xinhang, et al.
Publicado: (2025)
SS4D: Native 4D Generative Model via Structured Spacetime Latents
por: Li, Zhibing, et al.
Publicado: (2025)
por: Li, Zhibing, et al.
Publicado: (2025)
Native and Compact Structured Latents for 3D Generation
por: Xiang, Jianfeng, et al.
Publicado: (2025)
por: Xiang, Jianfeng, et al.
Publicado: (2025)
Anymate: A Dataset and Baselines for Learning 3D Object Rigging
por: Deng, Yufan, et al.
Publicado: (2025)
por: Deng, Yufan, et al.
Publicado: (2025)
DualPM: Dual Posed-Canonical Point Maps for 3D Shape and Pose Reconstruction
por: Kaye, Ben, et al.
Publicado: (2024)
por: Kaye, Ben, et al.
Publicado: (2024)
Deceptive-NeRF/3DGS: Diffusion-Generated Pseudo-Observations for High-Quality Sparse-View Reconstruction
por: Liu, Xinhang, et al.
Publicado: (2023)
por: Liu, Xinhang, et al.
Publicado: (2023)
Flow4R: Unifying 4D Reconstruction and Tracking with Scene Flow
por: Qian, Shenhan, et al.
Publicado: (2026)
por: Qian, Shenhan, et al.
Publicado: (2026)
Ponymation: Learning Articulated 3D Animal Motions from Unlabeled Online Videos
por: Sun, Keqiang, et al.
Publicado: (2023)
por: Sun, Keqiang, et al.
Publicado: (2023)
Hitem3D 2.0: Multi-View Guided Native 3D Texture Generation
por: He, Huiang, et al.
Publicado: (2026)
por: He, Huiang, et al.
Publicado: (2026)
Generic 3D Diffusion Adapter Using Controlled Multi-View Editing
por: Chen, Hansheng, et al.
Publicado: (2024)
por: Chen, Hansheng, et al.
Publicado: (2024)
NeuROK: Generative 4D Neural Object Kinematics
por: Geng, Chen, et al.
Publicado: (2026)
por: Geng, Chen, et al.
Publicado: (2026)
UVRM: A Scalable 3D Reconstruction Model from Unposed Videos
por: Kao, Shiu-hong, et al.
Publicado: (2025)
por: Kao, Shiu-hong, et al.
Publicado: (2025)
3D Skeleton-Based Action Recognition: A Review
por: Liu, Mengyuan, et al.
Publicado: (2025)
por: Liu, Mengyuan, et al.
Publicado: (2025)
LACONIC: A 3D Layout Adapter for Controllable Image Creation
por: Maillard, Léopold, et al.
Publicado: (2025)
por: Maillard, Léopold, et al.
Publicado: (2025)
Think Before You Segment: High-Quality Reasoning Segmentation with GPT Chain of Thoughts
por: Kao, Shiu-hong, et al.
Publicado: (2025)
por: Kao, Shiu-hong, et al.
Publicado: (2025)
CoT-RVS: Zero-Shot Chain-of-Thought Reasoning Segmentation for Videos
por: Kao, Shiu-hong, et al.
Publicado: (2025)
por: Kao, Shiu-hong, et al.
Publicado: (2025)
Inpaint4DNeRF: Promptable Spatio-Temporal NeRF Inpainting with Generative Diffusion Models
por: Jiang, Han, et al.
Publicado: (2023)
por: Jiang, Han, et al.
Publicado: (2023)
EVCtrl: Efficient Control Adapter for Visual Generation
por: Yang, Zixiang, et al.
Publicado: (2025)
por: Yang, Zixiang, et al.
Publicado: (2025)
Grounded Question-Answering in Long Egocentric Videos
por: Di, Shangzhe, et al.
Publicado: (2023)
por: Di, Shangzhe, et al.
Publicado: (2023)
SkeletonGaussian: Editable 4D Generation through Gaussian Skeletonization
por: Wu, Lifan, et al.
Publicado: (2026)
por: Wu, Lifan, et al.
Publicado: (2026)
Marrying Text-to-Motion Generation with Skeleton-Based Action Recognition
por: Kuang, Jidong, et al.
Publicado: (2026)
por: Kuang, Jidong, et al.
Publicado: (2026)
Towards Native Generative Model for 3D Head Avatar
por: Zhuang, Yiyu, et al.
Publicado: (2024)
por: Zhuang, Yiyu, et al.
Publicado: (2024)
Learning the 3D Fauna of the Web
por: Li, Zizhang, et al.
Publicado: (2024)
por: Li, Zizhang, et al.
Publicado: (2024)
SANeRF-HQ: Segment Anything for NeRF in High Quality
por: Liu, Yichen, et al.
Publicado: (2023)
por: Liu, Yichen, et al.
Publicado: (2023)
Language-Informed Visual Concept Learning
por: Lee, Sharon, et al.
Publicado: (2023)
por: Lee, Sharon, et al.
Publicado: (2023)
Seeing a Rose in Five Thousand Ways
por: Zhang, Yunzhi, et al.
Publicado: (2022)
por: Zhang, Yunzhi, et al.
Publicado: (2022)
IPVTON: Image-based 3D Virtual Try-on with Image Prompt Adapter
por: Zhong, Xiaojing, et al.
Publicado: (2025)
por: Zhong, Xiaojing, et al.
Publicado: (2025)
StereoAdapter-2: Globally Structure-Consistent Underwater Stereo Depth Estimation
por: Ren, Zeyu, et al.
Publicado: (2026)
por: Ren, Zeyu, et al.
Publicado: (2026)
Stanford-ORB: A Real-World 3D Object Inverse Rendering Benchmark
por: Kuang, Zhengfei, et al.
Publicado: (2023)
por: Kuang, Zhengfei, et al.
Publicado: (2023)
SkeletonContext: Skeleton-side Context Prompt Learning for Zero-Shot Skeleton-based Action Recognition
por: Wang, Ning, et al.
Publicado: (2026)
por: Wang, Ning, et al.
Publicado: (2026)
Ejemplares similares
-
ReasonNavi: Human-Inspired Global Map Reasoning for Zero-Shot Embodied Navigation
por: Ao, Yuzhuo, et al.
Publicado: (2026) -
Multimodal Generation of Animatable 3D Human Models with AvatarForge
por: Liu, Xinhang, et al.
Publicado: (2025) -
InceptionHuman: Controllable Prompt-to-NeRF for Photorealistic 3D Human Generation
por: Kao, Shiu-hong, et al.
Publicado: (2023) -
Agentic 3D Scene Generation with Spatially Contextualized VLMs
por: Liu, Xinhang, et al.
Publicado: (2025) -
FED-NeRF: Achieve High 3D Consistency and Temporal Coherence for Face Video Editing on Dynamic NeRF
por: Zhang, Hao, et al.
Publicado: (2024)