Improving generalization by mimicking the human visual diet
Fuente:
arXiv
Salvato in:
| Autori principali: | Madan, Spandan, Li, You, Zhang, Mengmi, Pfister, Hanspeter, Kreiman, Gabriel |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Tuned Compositional Feature Replays for Efficient Stream Learning
di: Talbot, Morgan B., et al.
Pubblicazione: (2021)
di: Talbot, Morgan B., et al.
Pubblicazione: (2021)
LangSplatV2: High-dimensional 3D Language Gaussian Splatting with 450+ FPS
di: Li, Wanhua, et al.
Pubblicazione: (2025)
di: Li, Wanhua, et al.
Pubblicazione: (2025)
Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
di: Liu, Bingchen, et al.
Pubblicazione: (2024)
di: Liu, Bingchen, et al.
Pubblicazione: (2024)
In-Context Brush: Zero-shot Customized Subject Insertion with Context-Aware Latent Space Manipulation
di: Xu, Yu, et al.
Pubblicazione: (2025)
di: Xu, Yu, et al.
Pubblicazione: (2025)
DreamDrive: Generative 4D Scene Modeling from Street View Images
di: Mao, Jiageng, et al.
Pubblicazione: (2024)
di: Mao, Jiageng, et al.
Pubblicazione: (2024)
SparseOIT: Improving Order-Independent Transparency 3DGS via Active Set Method
di: Yang, Wentao, et al.
Pubblicazione: (2026)
di: Yang, Wentao, et al.
Pubblicazione: (2026)
Lodge++: High-quality and Long Dance Generation with Vivid Choreography Patterns
di: Li, Ronghui, et al.
Pubblicazione: (2024)
di: Li, Ronghui, et al.
Pubblicazione: (2024)
SRDiffusion: Accelerate Video Diffusion Inference via Sketching-Rendering Cooperation
di: Cheng, Shenggan, et al.
Pubblicazione: (2025)
di: Cheng, Shenggan, et al.
Pubblicazione: (2025)
Template-Guided Reconstruction of Pulmonary Segments with Neural Implicit Functions
di: Xie, Kangxian, et al.
Pubblicazione: (2025)
di: Xie, Kangxian, et al.
Pubblicazione: (2025)
Taming Video Models for 3D and 4D Generation via Zero-Shot Camera Control
di: Song, Chenxi, et al.
Pubblicazione: (2025)
di: Song, Chenxi, et al.
Pubblicazione: (2025)
ProReflow: Progressive Reflow with Decomposed Velocity
di: Ke, Lei, et al.
Pubblicazione: (2025)
di: Ke, Lei, et al.
Pubblicazione: (2025)
OmniHands: Towards Robust 4D Hand Mesh Recovery via A Versatile Transformer
di: Lin, Dixuan, et al.
Pubblicazione: (2024)
di: Lin, Dixuan, et al.
Pubblicazione: (2024)
GeometryCrafter: Consistent Geometry Estimation for Open-world Videos with Diffusion Priors
di: Xu, Tian-Xing, et al.
Pubblicazione: (2025)
di: Xu, Tian-Xing, et al.
Pubblicazione: (2025)
BAG: Body-Aligned 3D Wearable Asset Generation
di: Luo, Zhongjin, et al.
Pubblicazione: (2025)
di: Luo, Zhongjin, et al.
Pubblicazione: (2025)
HY-Motion 1.0: Scaling Flow Matching Models for Text-To-Motion Generation
di: Wen, Yuxin, et al.
Pubblicazione: (2025)
di: Wen, Yuxin, et al.
Pubblicazione: (2025)
DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos
di: Hu, Wenbo, et al.
Pubblicazione: (2024)
di: Hu, Wenbo, et al.
Pubblicazione: (2024)
FreeOrbit4D: Training-Free Arbitrary Camera Redirection for Monocular Videos via Foreground-Complete 4D Reconstruction
di: Cao, Wei, et al.
Pubblicazione: (2026)
di: Cao, Wei, et al.
Pubblicazione: (2026)
Visionary: The World Model Carrier Built on WebGPU-Powered Gaussian Splatting Platform
di: Gong, Yuning, et al.
Pubblicazione: (2025)
di: Gong, Yuning, et al.
Pubblicazione: (2025)
Neural Gaffer: Relighting Any Object via Diffusion
di: Jin, Haian, et al.
Pubblicazione: (2024)
di: Jin, Haian, et al.
Pubblicazione: (2024)
Learning Disentangled Speech- and Expression-Driven Blendshapes for 3D Talking Face Animation
di: Mao, Yuxiang, et al.
Pubblicazione: (2025)
di: Mao, Yuxiang, et al.
Pubblicazione: (2025)
Improved Baselines with Representation Autoencoders
di: Singh, Jaskirat, et al.
Pubblicazione: (2026)
di: Singh, Jaskirat, et al.
Pubblicazione: (2026)
Learning to Synthesize Graphics Programs for Geometric Artworks
di: Bing, Qi, et al.
Pubblicazione: (2024)
di: Bing, Qi, et al.
Pubblicazione: (2024)
Morse: Dual-Sampling for Lossless Acceleration of Diffusion Models
di: Li, Chao, et al.
Pubblicazione: (2025)
di: Li, Chao, et al.
Pubblicazione: (2025)
Mojito: LLM-Aided Motion Instructor with Jitter-Reduced Inertial Tokens
di: Shan, Ziwei, et al.
Pubblicazione: (2025)
di: Shan, Ziwei, et al.
Pubblicazione: (2025)
CoARF: Controllable 3D Artistic Style Transfer for Radiance Fields
di: Zhang, Deheng, et al.
Pubblicazione: (2024)
di: Zhang, Deheng, et al.
Pubblicazione: (2024)
Generative Object Insertion in Gaussian Splatting with a Multi-View Diffusion Model
di: Zhong, Hongliang, et al.
Pubblicazione: (2024)
di: Zhong, Hongliang, et al.
Pubblicazione: (2024)
PersonaTalk: Bring Attention to Your Persona in Visual Dubbing
di: Zhang, Longhao, et al.
Pubblicazione: (2024)
di: Zhang, Longhao, et al.
Pubblicazione: (2024)
Exploration and Improvement of Nerf-based 3D Scene Editing Techniques
di: Fang, Shun, et al.
Pubblicazione: (2024)
di: Fang, Shun, et al.
Pubblicazione: (2024)
SkeletonGaussian: Editable 4D Generation through Gaussian Skeletonization
di: Wu, Lifan, et al.
Pubblicazione: (2026)
di: Wu, Lifan, et al.
Pubblicazione: (2026)
BOOTPLACE: Bootstrapped Object Placement with Detection Transformers
di: Zhou, Hang, et al.
Pubblicazione: (2025)
di: Zhou, Hang, et al.
Pubblicazione: (2025)
LAYOUTDREAMER: Physics-guided Layout for Text-to-3D Compositional Scene Generation
di: Zhou, Yang, et al.
Pubblicazione: (2025)
di: Zhou, Yang, et al.
Pubblicazione: (2025)
UniLat3D: Geometry-Appearance Unified Latents for Single-Stage 3D Generation
di: Wu, Guanjun, et al.
Pubblicazione: (2025)
di: Wu, Guanjun, et al.
Pubblicazione: (2025)
Uncertainty Matters in Dynamic Gaussian Splatting for Monocular 4D Reconstruction
di: Guo, Fengzhi, et al.
Pubblicazione: (2025)
di: Guo, Fengzhi, et al.
Pubblicazione: (2025)
Image-Conditioned 3D Gaussian Splat Quantization
di: Liu, Xinshuang, et al.
Pubblicazione: (2025)
di: Liu, Xinshuang, et al.
Pubblicazione: (2025)
Light-SQ: Structure-aware Shape Abstraction with Superquadrics for Generated Meshes
di: Wang, Yuhan, et al.
Pubblicazione: (2025)
di: Wang, Yuhan, et al.
Pubblicazione: (2025)
RMD: A Simple Baseline for More General Human Motion Generation via Training-free Retrieval-Augmented Motion Diffuse
di: Liao, Zhouyingcheng, et al.
Pubblicazione: (2024)
di: Liao, Zhouyingcheng, et al.
Pubblicazione: (2024)
KinMo: Kinematic-aware Human Motion Understanding and Generation
di: Zhang, Pengfei, et al.
Pubblicazione: (2024)
di: Zhang, Pengfei, et al.
Pubblicazione: (2024)
FlexCAD: Unified and Versatile Controllable CAD Generation with Fine-tuned Large Language Models
di: Zhang, Zhanwei, et al.
Pubblicazione: (2024)
di: Zhang, Zhanwei, et al.
Pubblicazione: (2024)
BlockFusion: Expandable 3D Scene Generation using Latent Tri-plane Extrapolation
di: Wu, Zhennan, et al.
Pubblicazione: (2024)
di: Wu, Zhennan, et al.
Pubblicazione: (2024)
GazeFusion: Saliency-Guided Image Generation
di: Zhang, Yunxiang, et al.
Pubblicazione: (2024)
di: Zhang, Yunxiang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Tuned Compositional Feature Replays for Efficient Stream Learning
di: Talbot, Morgan B., et al.
Pubblicazione: (2021) -
LangSplatV2: High-dimensional 3D Language Gaussian Splatting with 450+ FPS
di: Li, Wanhua, et al.
Pubblicazione: (2025) -
Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
di: Liu, Bingchen, et al.
Pubblicazione: (2024) -
In-Context Brush: Zero-shot Customized Subject Insertion with Context-Aware Latent Space Manipulation
di: Xu, Yu, et al.
Pubblicazione: (2025) -
DreamDrive: Generative 4D Scene Modeling from Street View Images
di: Mao, Jiageng, et al.
Pubblicazione: (2024)