AnyUp: Universal Feature Upsampling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wimmer, Thomas, Truong, Prune, Rakotosaona, Marie-Julie, Oechsle, Michael, Tombari, Federico, Schiele, Bernt, Lenssen, Jan Eric |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Spatial Reasoners for Continuous Variables in Any Domain
von: Pogodzinski, Bart, et al.
Veröffentlicht: (2025)
von: Pogodzinski, Bart, et al.
Veröffentlicht: (2025)
Gaussians-to-Life: Text-Driven Animation of 3D Gaussian Splatting Scenes
von: Wimmer, Thomas, et al.
Veröffentlicht: (2024)
von: Wimmer, Thomas, et al.
Veröffentlicht: (2024)
MEt3R: Measuring Multi-View Consistency in Generated Images
von: Asim, Mohammad, et al.
Veröffentlicht: (2025)
von: Asim, Mohammad, et al.
Veröffentlicht: (2025)
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
von: Zheng, Shuhong, et al.
Veröffentlicht: (2026)
von: Zheng, Shuhong, et al.
Veröffentlicht: (2026)
One2Any: One-Reference 6D Pose Estimation for Any Object
von: Liu, Mengya, et al.
Veröffentlicht: (2025)
von: Liu, Mengya, et al.
Veröffentlicht: (2025)
RefAM: Attention Magnets for Zero-Shot Referral Segmentation
von: Kukleva, Anna, et al.
Veröffentlicht: (2025)
von: Kukleva, Anna, et al.
Veröffentlicht: (2025)
SimNP: Learning Self-Similarity Priors Between Neural Points
von: Wewer, Christopher, et al.
Veröffentlicht: (2023)
von: Wewer, Christopher, et al.
Veröffentlicht: (2023)
Sp2360: Sparse-view 360 Scene Reconstruction using Cascaded 2D Diffusion Priors
von: Paul, Soumava, et al.
Veröffentlicht: (2024)
von: Paul, Soumava, et al.
Veröffentlicht: (2024)
PersonaHOI: Effortlessly Improving Personalized Face with Human-Object Interaction Generation
von: Hu, Xinting, et al.
Veröffentlicht: (2025)
von: Hu, Xinting, et al.
Veröffentlicht: (2025)
Learning Neural Exposure Fields for View Synthesis
von: Niemeyer, Michael, et al.
Veröffentlicht: (2025)
von: Niemeyer, Michael, et al.
Veröffentlicht: (2025)
Masks make discriminative models great again!
von: Cao, Tianshi, et al.
Veröffentlicht: (2025)
von: Cao, Tianshi, et al.
Veröffentlicht: (2025)
Spatial Reasoning with Denoising Models
von: Wewer, Christopher, et al.
Veröffentlicht: (2025)
von: Wewer, Christopher, et al.
Veröffentlicht: (2025)
Rewis3d: Reconstruction Improves Weakly-Supervised Semantic Segmentation
von: Ernst, Jonas, et al.
Veröffentlicht: (2026)
von: Ernst, Jonas, et al.
Veröffentlicht: (2026)
Scribbles for All: Benchmarking Scribble Supervised Segmentation Across Datasets
von: Boettcher, Wolfgang, et al.
Veröffentlicht: (2024)
von: Boettcher, Wolfgang, et al.
Veröffentlicht: (2024)
latentSplat: Autoencoding Variational Gaussians for Fast Generalizable 3D Reconstruction
von: Wewer, Christopher, et al.
Veröffentlicht: (2024)
von: Wewer, Christopher, et al.
Veröffentlicht: (2024)
RadSplat: Radiance Field-Informed Gaussian Splatting for Robust Real-Time Rendering with 900+ FPS
von: Niemeyer, Michael, et al.
Veröffentlicht: (2024)
von: Niemeyer, Michael, et al.
Veröffentlicht: (2024)
Elastic3D: Controllable Stereo Video Conversion with Guided Latent Decoding
von: Metzger, Nando, et al.
Veröffentlicht: (2025)
von: Metzger, Nando, et al.
Veröffentlicht: (2025)
M2SVid: End-to-End Inpainting and Refinement for Monocular-to-Stereo Video Conversion
von: Shvetsova, Nina, et al.
Veröffentlicht: (2025)
von: Shvetsova, Nina, et al.
Veröffentlicht: (2025)
UniSDF: Unifying Neural Representations for High-Fidelity 3D Reconstruction of Complex Scenes with Reflections
von: Wang, Fangjinhua, et al.
Veröffentlicht: (2023)
von: Wang, Fangjinhua, et al.
Veröffentlicht: (2023)
R-CoV: Region-Aware Chain-of-Verification for Alleviating Object Hallucinations in LVLMs
von: Xie, Jiahao, et al.
Veröffentlicht: (2026)
von: Xie, Jiahao, et al.
Veröffentlicht: (2026)
SSL-R1: Self-Supervised Visual Reinforcement Post-Training for Multimodal Large Language Models
von: Xie, Jiahao, et al.
Veröffentlicht: (2026)
von: Xie, Jiahao, et al.
Veröffentlicht: (2026)
P2P-Bridge: Diffusion Bridges for 3D Point Cloud Denoising
von: Vogel, Mathias, et al.
Veröffentlicht: (2024)
von: Vogel, Mathias, et al.
Veröffentlicht: (2024)
Language-Unlocked ViT (LUViT): Empowering Self-Supervised Vision Transformers with LLMs
von: Kuzucu, Selim, et al.
Veröffentlicht: (2025)
von: Kuzucu, Selim, et al.
Veröffentlicht: (2025)
Text-to-3D by Stitching a Multi-view Reconstruction Network to a Video Generator
von: Go, Hyojun, et al.
Veröffentlicht: (2025)
von: Go, Hyojun, et al.
Veröffentlicht: (2025)
Test-Time Visual In-Context Tuning
von: Xie, Jiahao, et al.
Veröffentlicht: (2025)
von: Xie, Jiahao, et al.
Veröffentlicht: (2025)
Solving Inverse Problems with FLAIR
von: Erbach, Julius, et al.
Veröffentlicht: (2025)
von: Erbach, Julius, et al.
Veröffentlicht: (2025)
3D-LATTE: Latent Space 3D Editing from Textual Instructions
von: Parelli, Maria, et al.
Veröffentlicht: (2025)
von: Parelli, Maria, et al.
Veröffentlicht: (2025)
LODGE: Level-of-Detail Large-Scale Gaussian Splatting with Efficient Rendering
von: Kulhanek, Jonas, et al.
Veröffentlicht: (2025)
von: Kulhanek, Jonas, et al.
Veröffentlicht: (2025)
VITAL: More Understandable Feature Visualization through Distribution Alignment and Relevant Information Flow
von: Gorgun, Ada, et al.
Veröffentlicht: (2025)
von: Gorgun, Ada, et al.
Veröffentlicht: (2025)
Toward a Diffusion-Based Generalist for Dense Vision Tasks
von: Fan, Yue, et al.
Veröffentlicht: (2024)
von: Fan, Yue, et al.
Veröffentlicht: (2024)
CubeDiff: Repurposing Diffusion-Based Image Models for Panorama Generation
von: Kalischek, Nikolai, et al.
Veröffentlicht: (2025)
von: Kalischek, Nikolai, et al.
Veröffentlicht: (2025)
Stepper: Stepwise Immersive Scene Generation with Multiview Panoramas
von: Wimbauer, Felix, et al.
Veröffentlicht: (2026)
von: Wimbauer, Felix, et al.
Veröffentlicht: (2026)
Splat-SLAM: Globally Optimized RGB-only SLAM with 3D Gaussians
von: Sandström, Erik, et al.
Veröffentlicht: (2024)
von: Sandström, Erik, et al.
Veröffentlicht: (2024)
DiveUp: Learning Feature Upsampling from Diverse Vision Foundation Models
von: Liu, Xiaoqiong, et al.
Veröffentlicht: (2026)
von: Liu, Xiaoqiong, et al.
Veröffentlicht: (2026)
OrCo: Towards Better Generalization via Orthogonality and Contrast for Few-Shot Class-Incremental Learning
von: Ahmed, Noor, et al.
Veröffentlicht: (2024)
von: Ahmed, Noor, et al.
Veröffentlicht: (2024)
DWDN: Deep Wiener Deconvolution Network for Non-Blind Image Deblurring
von: Dong, Jiangxin, et al.
Veröffentlicht: (2021)
von: Dong, Jiangxin, et al.
Veröffentlicht: (2021)
PARCEL: Pool-Anchored Resampling with Conditioned Elastic Queries for Efficient Vision-Language Understanding
von: Kuzucu, Selim, et al.
Veröffentlicht: (2026)
von: Kuzucu, Selim, et al.
Veröffentlicht: (2026)
Upsample Anything: A Simple and Hard to Beat Baseline for Feature Upsampling
von: Seo, Minseok, et al.
Veröffentlicht: (2025)
von: Seo, Minseok, et al.
Veröffentlicht: (2025)
Improving 2D Feature Representations by 3D-Aware Fine-Tuning
von: Yue, Yuanwen, et al.
Veröffentlicht: (2024)
von: Yue, Yuanwen, et al.
Veröffentlicht: (2024)
Optimising for Interpretability: Convolutional Dynamic Alignment Networks
von: Böhle, Moritz, et al.
Veröffentlicht: (2021)
von: Böhle, Moritz, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
Spatial Reasoners for Continuous Variables in Any Domain
von: Pogodzinski, Bart, et al.
Veröffentlicht: (2025) -
Gaussians-to-Life: Text-Driven Animation of 3D Gaussian Splatting Scenes
von: Wimmer, Thomas, et al.
Veröffentlicht: (2024) -
MEt3R: Measuring Multi-View Consistency in Generated Images
von: Asim, Mohammad, et al.
Veröffentlicht: (2025) -
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
von: Zheng, Shuhong, et al.
Veröffentlicht: (2026) -
One2Any: One-Reference 6D Pose Estimation for Any Object
von: Liu, Mengya, et al.
Veröffentlicht: (2025)