Masks make discriminative models great again!
Fuente:
arXiv
Salvato in:
| Autori principali: | Cao, Tianshi, Rakotosaona, Marie-Julie, Poole, Ben, Tombari, Federico, Niemeyer, Michael |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
UniSDF: Unifying Neural Representations for High-Fidelity 3D Reconstruction of Complex Scenes with Reflections
di: Wang, Fangjinhua, et al.
Pubblicazione: (2023)
di: Wang, Fangjinhua, et al.
Pubblicazione: (2023)
LODGE: Level-of-Detail Large-Scale Gaussian Splatting with Efficient Rendering
di: Kulhanek, Jonas, et al.
Pubblicazione: (2025)
di: Kulhanek, Jonas, et al.
Pubblicazione: (2025)
Learning Neural Exposure Fields for View Synthesis
di: Niemeyer, Michael, et al.
Pubblicazione: (2025)
di: Niemeyer, Michael, et al.
Pubblicazione: (2025)
P2P-Bridge: Diffusion Bridges for 3D Point Cloud Denoising
di: Vogel, Mathias, et al.
Pubblicazione: (2024)
di: Vogel, Mathias, et al.
Pubblicazione: (2024)
RadSplat: Radiance Field-Informed Gaussian Splatting for Robust Real-Time Rendering with 900+ FPS
di: Niemeyer, Michael, et al.
Pubblicazione: (2024)
di: Niemeyer, Michael, et al.
Pubblicazione: (2024)
Gaussians-to-Life: Text-Driven Animation of 3D Gaussian Splatting Scenes
di: Wimmer, Thomas, et al.
Pubblicazione: (2024)
di: Wimmer, Thomas, et al.
Pubblicazione: (2024)
AnyUp: Universal Feature Upsampling
di: Wimmer, Thomas, et al.
Pubblicazione: (2025)
di: Wimmer, Thomas, et al.
Pubblicazione: (2025)
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
di: Zheng, Shuhong, et al.
Pubblicazione: (2026)
di: Zheng, Shuhong, et al.
Pubblicazione: (2026)
CausNVS: Autoregressive Multi-view Diffusion for Flexible 3D Novel View Synthesis
di: Kong, Xin, et al.
Pubblicazione: (2025)
di: Kong, Xin, et al.
Pubblicazione: (2025)
MonoGSDF: Exploring Monocular Geometric Cues for Gaussian Splatting-Guided Implicit Surface Reconstruction
di: Li, Kunyi, et al.
Pubblicazione: (2024)
di: Li, Kunyi, et al.
Pubblicazione: (2024)
3D-LATTE: Latent Space 3D Editing from Textual Instructions
di: Parelli, Maria, et al.
Pubblicazione: (2025)
di: Parelli, Maria, et al.
Pubblicazione: (2025)
OpenGaFF: Open-Vocabulary Gaussian Feature Field with Codebook Attention
di: Li, Kunyi, et al.
Pubblicazione: (2026)
di: Li, Kunyi, et al.
Pubblicazione: (2026)
OpenNeRF: Open Set 3D Neural Scene Segmentation with Pixel-Wise Features and Rendered Novel Views
di: Engelmann, Francis, et al.
Pubblicazione: (2024)
di: Engelmann, Francis, et al.
Pubblicazione: (2024)
OracleGS: Grounding Generative Priors for Sparse-View Gaussian Splatting
di: Topaloglu, Atakan, et al.
Pubblicazione: (2025)
di: Topaloglu, Atakan, et al.
Pubblicazione: (2025)
A new baseline for edge detection: Make Encoder-Decoder great again
di: Li, Yachuan, et al.
Pubblicazione: (2024)
di: Li, Yachuan, et al.
Pubblicazione: (2024)
SING3R-SLAM: Submap-based Indoor Monocular Gaussian SLAM with 3D Reconstruction Priors
di: Li, Kunyi, et al.
Pubblicazione: (2025)
di: Li, Kunyi, et al.
Pubblicazione: (2025)
Splat-SLAM: Globally Optimized RGB-only SLAM with 3D Gaussians
di: Sandström, Erik, et al.
Pubblicazione: (2024)
di: Sandström, Erik, et al.
Pubblicazione: (2024)
Few-shot point cloud reconstruction and denoising via learned Guassian splats renderings and fine-tuned diffusion features
di: Bonazzi, Pietro, et al.
Pubblicazione: (2024)
di: Bonazzi, Pietro, et al.
Pubblicazione: (2024)
HouseLayout3D: A Benchmark and Training-Free Baseline for 3D Layout Estimation in the Wild
di: Bieri, Valentin, et al.
Pubblicazione: (2025)
di: Bieri, Valentin, et al.
Pubblicazione: (2025)
GALA: Guided Attention with Language Alignment for Open Vocabulary Gaussian Splatting
di: Alegret, Elena, et al.
Pubblicazione: (2025)
di: Alegret, Elena, et al.
Pubblicazione: (2025)
Prior2Former -- Evidential Modeling of Mask Transformers for Assumption-Free Open-World Panoptic Segmentation
di: Schmidt, Sebastian, et al.
Pubblicazione: (2025)
di: Schmidt, Sebastian, et al.
Pubblicazione: (2025)
SuperGSeg: Open-Vocabulary 3D Segmentation with Structured Super-Gaussians
di: Liang, Siyun, et al.
Pubblicazione: (2024)
di: Liang, Siyun, et al.
Pubblicazione: (2024)
InseRF: Text-Driven Generative Object Insertion in Neural 3D Scenes
di: Shahbazi, Mohamad, et al.
Pubblicazione: (2024)
di: Shahbazi, Mohamad, et al.
Pubblicazione: (2024)
Text To 3D Object Generation For Scalable Room Assembly
di: Laguna, Sonia, et al.
Pubblicazione: (2025)
di: Laguna, Sonia, et al.
Pubblicazione: (2025)
Neural Semantic Map-Learning for Autonomous Vehicles
di: Herb, Markus, et al.
Pubblicazione: (2024)
di: Herb, Markus, et al.
Pubblicazione: (2024)
Language-Guided Open-World Anomaly Segmentation
di: Reichard, Klara, et al.
Pubblicazione: (2025)
di: Reichard, Klara, et al.
Pubblicazione: (2025)
Epipolar Geometry Improves Video Generation Models
di: Kupyn, Orest, et al.
Pubblicazione: (2025)
di: Kupyn, Orest, et al.
Pubblicazione: (2025)
SegSplat: Feed-forward Gaussian Splatting and Open-Set Semantic Segmentation
di: Siegel, Peter, et al.
Pubblicazione: (2025)
di: Siegel, Peter, et al.
Pubblicazione: (2025)
LoRACLR: Contrastive Adaptation for Customization of Diffusion Models
di: Simsar, Enis, et al.
Pubblicazione: (2024)
di: Simsar, Enis, et al.
Pubblicazione: (2024)
A Taxonomy and Library for Visualizing Learned Features in Convolutional Neural Networks
di: Grün, Felix, et al.
Pubblicazione: (2016)
di: Grün, Felix, et al.
Pubblicazione: (2016)
HyperSDFusion: Bridging Hierarchical Structures in Language and Geometry for Enhanced 3D Text2Shape Generation
di: Leng, Zhiying, et al.
Pubblicazione: (2024)
di: Leng, Zhiying, et al.
Pubblicazione: (2024)
ESCAPE: Equivariant Shape Completion via Anchor Point Encoding
di: Bekci, Burak, et al.
Pubblicazione: (2024)
di: Bekci, Burak, et al.
Pubblicazione: (2024)
Physics-Encoded Graph Neural Networks for Deformation Prediction under Contact
di: Saleh, Mahdi, et al.
Pubblicazione: (2024)
di: Saleh, Mahdi, et al.
Pubblicazione: (2024)
Object-X: Learning to Reconstruct Multi-Modal 3D Object Representations
di: Di Lorenzo, Gaia, et al.
Pubblicazione: (2025)
di: Di Lorenzo, Gaia, et al.
Pubblicazione: (2025)
SmileSplat: Generalizable Gaussian Splats for Unconstrained Sparse Images
di: Li, Yanyan, et al.
Pubblicazione: (2024)
di: Li, Yanyan, et al.
Pubblicazione: (2024)
RaNeuS: Ray-adaptive Neural Surface Reconstruction
di: Wang, Yida, et al.
Pubblicazione: (2024)
di: Wang, Yida, et al.
Pubblicazione: (2024)
LaRI: Layered Ray Intersections for Single-view 3D Geometric Reasoning
di: Li, Rui, et al.
Pubblicazione: (2025)
di: Li, Rui, et al.
Pubblicazione: (2025)
Adversarial Appearance Learning in Augmented Cityscapes for Pedestrian Recognition in Autonomous Driving
di: Savkin, Artem, et al.
Pubblicazione: (2025)
di: Savkin, Artem, et al.
Pubblicazione: (2025)
Omnia de EgoTempo: Benchmarking Temporal Understanding of Multi-Modal LLMs in Egocentric Videos
di: Plizzari, Chiara, et al.
Pubblicazione: (2025)
di: Plizzari, Chiara, et al.
Pubblicazione: (2025)
Elastic3D: Controllable Stereo Video Conversion with Guided Latent Decoding
di: Metzger, Nando, et al.
Pubblicazione: (2025)
di: Metzger, Nando, et al.
Pubblicazione: (2025)
Documenti analoghi
-
UniSDF: Unifying Neural Representations for High-Fidelity 3D Reconstruction of Complex Scenes with Reflections
di: Wang, Fangjinhua, et al.
Pubblicazione: (2023) -
LODGE: Level-of-Detail Large-Scale Gaussian Splatting with Efficient Rendering
di: Kulhanek, Jonas, et al.
Pubblicazione: (2025) -
Learning Neural Exposure Fields for View Synthesis
di: Niemeyer, Michael, et al.
Pubblicazione: (2025) -
P2P-Bridge: Diffusion Bridges for 3D Point Cloud Denoising
di: Vogel, Mathias, et al.
Pubblicazione: (2024) -
RadSplat: Radiance Field-Informed Gaussian Splatting for Robust Real-Time Rendering with 900+ FPS
di: Niemeyer, Michael, et al.
Pubblicazione: (2024)