NAF: Zero-Shot Feature Upsampling via Neighborhood Attention Filtering
Fuente:
arXiv
Saved in:
| Main Authors: | Chambon, Loick, Couairon, Paul, Zablocki, Eloi, Boulch, Alexandre, Thome, Nicolas, Cord, Matthieu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GaussRender: Learning 3D Occupancy with Gaussian Rendering
by: Chambon, Loïck, et al.
Published: (2025)
by: Chambon, Loïck, et al.
Published: (2025)
JAFAR: Jack up Any Feature at Any Resolution
by: Couairon, Paul, et al.
Published: (2025)
by: Couairon, Paul, et al.
Published: (2025)
PointBeV: A Sparse Approach to BeV Predictions
by: Chambon, Loick, et al.
Published: (2023)
by: Chambon, Loick, et al.
Published: (2023)
DiffCut: Catalyzing Zero-Shot Semantic Segmentation with Diffusion Features and Recursive Normalized Cut
by: Couairon, Paul, et al.
Published: (2024)
by: Couairon, Paul, et al.
Published: (2024)
Towards Motion Forecasting with Real-World Perception Inputs: Are End-to-End Approaches Competitive?
by: Xu, Yihong, et al.
Published: (2023)
by: Xu, Yihong, et al.
Published: (2023)
ReGentS: Real-World Safety-Critical Driving Scenario Generation Made Stable
by: Yin, Yuan, et al.
Published: (2024)
by: Yin, Yuan, et al.
Published: (2024)
PPT: Pretraining with Pseudo-Labeled Trajectories for Motion Forecasting
by: Xu, Yihong, et al.
Published: (2024)
by: Xu, Yihong, et al.
Published: (2024)
VidEdit: Zero-Shot and Spatially Aware Text-Driven Video Editing
by: Couairon, Paul, et al.
Published: (2023)
by: Couairon, Paul, et al.
Published: (2023)
MAD: Motion Appearance Decoupling for efficient Driving World Models
by: Rahimi, Ahmad, et al.
Published: (2026)
by: Rahimi, Ahmad, et al.
Published: (2026)
LLM-wrapper: Black-Box Semantic-Aware Adaptation of Vision-Language Models for Referring Expression Comprehension
by: Cardiel, Amaia, et al.
Published: (2024)
by: Cardiel, Amaia, et al.
Published: (2024)
Annealed Winner-Takes-All for Motion Forecasting
by: Xu, Yihong, et al.
Published: (2024)
by: Xu, Yihong, et al.
Published: (2024)
UniTraj: A Unified Framework for Scalable Vehicle Trajectory Prediction
by: Feng, Lan, et al.
Published: (2024)
by: Feng, Lan, et al.
Published: (2024)
GIFT: A Framework Towards Global Interpretable Faithful Textual Explanations of Vision Classifiers
by: Zablocki, Éloi, et al.
Published: (2024)
by: Zablocki, Éloi, et al.
Published: (2024)
FreeSeg-Diff: Training-Free Open-Vocabulary Segmentation with Diffusion Models
by: Corradini, Barbara Toniella, et al.
Published: (2024)
by: Corradini, Barbara Toniella, et al.
Published: (2024)
VaViM and VaVAM: Autonomous Driving through Video Generative Modeling
by: Bartoccioni, Florent, et al.
Published: (2025)
by: Bartoccioni, Florent, et al.
Published: (2025)
RAP: 3D Rasterization Augmented End-to-End Planning
by: Feng, Lan, et al.
Published: (2025)
by: Feng, Lan, et al.
Published: (2025)
Valeo4Cast: A Modular Approach to End-to-End Forecasting
by: Xu, Yihong, et al.
Published: (2024)
by: Xu, Yihong, et al.
Published: (2024)
Zero-to-Hero: Enhancing Zero-Shot Novel View Synthesis via Attention Map Filtering
by: Sobol, Ido, et al.
Published: (2024)
by: Sobol, Ido, et al.
Published: (2024)
Towards Generalizable Trajectory Prediction Using Dual-Level Representation Learning And Adaptive Prompting
by: Messaoud, Kaouther, et al.
Published: (2025)
by: Messaoud, Kaouther, et al.
Published: (2025)
Driving on Registers
by: Kirby, Ellington, et al.
Published: (2026)
by: Kirby, Ellington, et al.
Published: (2026)
Unsupervised Object Localization in the Era of Self-Supervised ViTs: A Survey
by: Siméoni, Oriane, et al.
Published: (2023)
by: Siméoni, Oriane, et al.
Published: (2023)
Upsample Anything: A Simple and Hard to Beat Baseline for Feature Upsampling
by: Seo, Minseok, et al.
Published: (2025)
by: Seo, Minseok, et al.
Published: (2025)
Slot Attention-based Feature Filtering for Few-Shot Learning
by: Rodenas, Javier, et al.
Published: (2025)
by: Rodenas, Javier, et al.
Published: (2025)
Weighted Reverse Convolution for Feature Upsampling
by: Li, Wentong, et al.
Published: (2026)
by: Li, Wentong, et al.
Published: (2026)
Global-Regularized Neighborhood Regression for Efficient Zero-Shot Texture Anomaly Detection
by: Yao, Haiming, et al.
Published: (2024)
by: Yao, Haiming, et al.
Published: (2024)
Beyond Task Performance: Evaluating and Reducing the Flaws of Large Multimodal Models with In-Context Learning
by: Shukor, Mustafa, et al.
Published: (2023)
by: Shukor, Mustafa, et al.
Published: (2023)
ViLU: Learning Vision-Language Uncertainties for Failure Prediction
by: Lafon, Marc, et al.
Published: (2025)
by: Lafon, Marc, et al.
Published: (2025)
Local Attention Transformers for High-Detail Optical Flow Upsampling
by: Gielisse, Alexander, et al.
Published: (2024)
by: Gielisse, Alexander, et al.
Published: (2024)
Geo6DPose: Fast Zero-Shot 6D Object Pose Estimation via Geometry-Filtered Feature Matching
by: Toro, Javier Villena, et al.
Published: (2025)
by: Toro, Javier Villena, et al.
Published: (2025)
AnyUp: Universal Feature Upsampling
by: Wimmer, Thomas, et al.
Published: (2025)
by: Wimmer, Thomas, et al.
Published: (2025)
Skipping Computations in Multimodal LLMs
by: Shukor, Mustafa, et al.
Published: (2024)
by: Shukor, Mustafa, et al.
Published: (2024)
SSDD: Single-Step Diffusion Decoder for Efficient Image Tokenization
by: Vallaeys, Théophane, et al.
Published: (2025)
by: Vallaeys, Théophane, et al.
Published: (2025)
Nearest Neighbor Classification for Classical Image Upsampling
by: Matthews, Evan, et al.
Published: (2024)
by: Matthews, Evan, et al.
Published: (2024)
Implicit Multimodal Alignment: On the Generalization of Frozen LLMs to Multimodal Inputs
by: Shukor, Mustafa, et al.
Published: (2024)
by: Shukor, Mustafa, et al.
Published: (2024)
Zero-Shot Textual Explanations via Translating Decision-Critical Features
by: Yamauchi, Toshinori, et al.
Published: (2025)
by: Yamauchi, Toshinori, et al.
Published: (2025)
NAF-DPM: A Nonlinear Activation-Free Diffusion Probabilistic Model for Document Enhancement
by: Cicchetti, Giordano, et al.
Published: (2024)
by: Cicchetti, Giordano, et al.
Published: (2024)
UPLiFT: Efficient Pixel-Dense Feature Upsampling with Local Attenders
by: Walmer, Matthew, et al.
Published: (2026)
by: Walmer, Matthew, et al.
Published: (2026)
Cross-Layer Attentive Feature Upsampling for Low-latency Semantic Segmentation
by: Cheng, Tianheng, et al.
Published: (2026)
by: Cheng, Tianheng, et al.
Published: (2026)
Soft Masked Transformer for Point Cloud Processing with Skip Attention-Based Upsampling
by: He, Yong, et al.
Published: (2024)
by: He, Yong, et al.
Published: (2024)
LDA-AQU: Adaptive Query-guided Upsampling via Local Deformable Attention
by: Du, Zewen, et al.
Published: (2024)
by: Du, Zewen, et al.
Published: (2024)
Similar Items
-
GaussRender: Learning 3D Occupancy with Gaussian Rendering
by: Chambon, Loïck, et al.
Published: (2025) -
JAFAR: Jack up Any Feature at Any Resolution
by: Couairon, Paul, et al.
Published: (2025) -
PointBeV: A Sparse Approach to BeV Predictions
by: Chambon, Loick, et al.
Published: (2023) -
DiffCut: Catalyzing Zero-Shot Semantic Segmentation with Diffusion Features and Recursive Normalized Cut
by: Couairon, Paul, et al.
Published: (2024) -
Towards Motion Forecasting with Real-World Perception Inputs: Are End-to-End Approaches Competitive?
by: Xu, Yihong, et al.
Published: (2023)