Harnessing Input-Adaptive Inference for Efficient VLN
Fuente:
arXiv
Saved in:
| Main Authors: | Kang, Dongwoo, Perincherry, Akhil, Coalson, Zachary, Gabriel, Aiden, Lee, Stefan, Hong, Sanghyun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Do Visual Imaginations Improve Vision-and-Language Navigation Agents?
by: Perincherry, Akhil, et al.
Published: (2025)
by: Perincherry, Akhil, et al.
Published: (2025)
You Never Know: Quantization Induces Inconsistent Biases in Vision-Language Foundation Models
by: Slyman, Eric, et al.
Published: (2024)
by: Slyman, Eric, et al.
Published: (2024)
Fail-Closed Alignment for Large Language Models
by: Coalson, Zachary, et al.
Published: (2026)
by: Coalson, Zachary, et al.
Published: (2026)
Advanced Knowledge Transfer: Refined Feature Distillation for Zero-Shot Quantization in Edge Computing
by: Hong, Inpyo, et al.
Published: (2024)
by: Hong, Inpyo, et al.
Published: (2024)
Holistic Unlearning Benchmark: A Multi-Faceted Evaluation for Text-to-Image Diffusion Model Unlearning
by: Moon, Saemi, et al.
Published: (2024)
by: Moon, Saemi, et al.
Published: (2024)
Weakly-supervised Camera Localization by Ground-to-satellite Image Registration
by: Shi, Yujiao, et al.
Published: (2024)
by: Shi, Yujiao, et al.
Published: (2024)
Feature Unlearning for Pre-trained GANs and VAEs
by: Moon, Saemi, et al.
Published: (2023)
by: Moon, Saemi, et al.
Published: (2023)
Certified Robustness to Clean-Label Poisoning Using Diffusion Denoising
by: Hong, Sanghyun, et al.
Published: (2024)
by: Hong, Sanghyun, et al.
Published: (2024)
Plane Geometry Problem Solving with Multi-modal Reasoning: A Survey
by: Cho, Seunghyuk, et al.
Published: (2025)
by: Cho, Seunghyuk, et al.
Published: (2025)
S-JEA: Stacked Joint Embedding Architectures for Self-Supervised Visual Representation Learning
by: Manová, Alžběta, et al.
Published: (2023)
by: Manová, Alžběta, et al.
Published: (2023)
Concept Unlearning via Cross-Attention Activation Projection for Diffusion Models
by: Moon, Saemi, et al.
Published: (2026)
by: Moon, Saemi, et al.
Published: (2026)
NavFormer: IGRF Forecasting in Moving Coordinate Frames
by: Hwang, Yoontae, et al.
Published: (2026)
by: Hwang, Yoontae, et al.
Published: (2026)
Single-pass Adaptive Image Tokenization for Minimum Program Search
by: Duggal, Shivam, et al.
Published: (2025)
by: Duggal, Shivam, et al.
Published: (2025)
ADMN: A Layer-Wise Adaptive Multimodal Network for Dynamic Input Noise and Compute Resources
by: Wu, Jason, et al.
Published: (2025)
by: Wu, Jason, et al.
Published: (2025)
HiAP: A Multi-Granular Stochastic Auto-Pruning Framework for Vision Transformers
by: Li, Andy, et al.
Published: (2026)
by: Li, Andy, et al.
Published: (2026)
Harnessing EHRs for Diffusion-based Anomaly Detection on Chest X-rays
by: Kim, Harim, et al.
Published: (2025)
by: Kim, Harim, et al.
Published: (2025)
Input-Adaptive Generative Dynamics in Diffusion Models
by: Xing, Yucheng, et al.
Published: (2024)
by: Xing, Yucheng, et al.
Published: (2024)
Where and How to Perturb: On the Design of Perturbation Guidance in Diffusion and Flow Models
by: Ahn, Donghoon, et al.
Published: (2025)
by: Ahn, Donghoon, et al.
Published: (2025)
EgoCHARM: Resource-Efficient Hierarchical Activity Recognition using an Egocentric IMU Sensor
by: Padmanabha, Akhil, et al.
Published: (2025)
by: Padmanabha, Akhil, et al.
Published: (2025)
BECoTTA: Input-dependent Online Blending of Experts for Continual Test-time Adaptation
by: Lee, Daeun, et al.
Published: (2024)
by: Lee, Daeun, et al.
Published: (2024)
AdaDeDup: Adaptive Hybrid Data Pruning for Efficient Large-Scale Object Detection Training
by: Kang, Feiyang, et al.
Published: (2025)
by: Kang, Feiyang, et al.
Published: (2025)
SynthVision -- Harnessing Minimal Input for Maximal Output in Computer Vision Models using Synthetic Image data
by: Kularathne, Yudara, et al.
Published: (2024)
by: Kularathne, Yudara, et al.
Published: (2024)
Foresight: Adaptive Layer Reuse for Accelerated and High-Quality Text-to-Video Generation
by: Adnan, Muhammad, et al.
Published: (2025)
by: Adnan, Muhammad, et al.
Published: (2025)
Efficient Bayesian Inference from Noisy Pairwise Comparisons
by: Aczel, Till, et al.
Published: (2025)
by: Aczel, Till, et al.
Published: (2025)
RAViT: Resolution-Adaptive Vision Transformer
by: Guidez, Martial, et al.
Published: (2026)
by: Guidez, Martial, et al.
Published: (2026)
MultiFloodSynth: Multi-Annotated Flood Synthetic Dataset Generation
by: Kang, YoonJe, et al.
Published: (2025)
by: Kang, YoonJe, et al.
Published: (2025)
DMesh++: An Efficient Differentiable Mesh for Complex Shapes
by: Son, Sanghyun, et al.
Published: (2024)
by: Son, Sanghyun, et al.
Published: (2024)
Group Relative Augmentation for Data Efficient Action Detection
by: Patel, Deep Anil, et al.
Published: (2025)
by: Patel, Deep Anil, et al.
Published: (2025)
QuantU-Net: Efficient Wearable Medical Imaging Using Bitwidth as a Trainable Parameter
by: Boerkamp, Christiaan, et al.
Published: (2025)
by: Boerkamp, Christiaan, et al.
Published: (2025)
ASAP: Attention-Shift-Aware Pruning for Efficient LVLM Inference
by: Pathak, Surendra, et al.
Published: (2026)
by: Pathak, Surendra, et al.
Published: (2026)
Rate-In: Information-Driven Adaptive Dropout Rates for Improved Inference-Time Uncertainty Estimation
by: Zeevi, Tal, et al.
Published: (2024)
by: Zeevi, Tal, et al.
Published: (2024)
Navigating Conflicting Views: Harnessing Trust for Learning
by: Lu, Jueqing, et al.
Published: (2024)
by: Lu, Jueqing, et al.
Published: (2024)
Harnessing Superclasses for Learning from Hierarchical Databases
by: Urbani, Nicolas, et al.
Published: (2024)
by: Urbani, Nicolas, et al.
Published: (2024)
E-MD3C: Taming Masked Diffusion Transformers for Efficient Zero-Shot Object Customization
by: Pham, Trung X., et al.
Published: (2025)
by: Pham, Trung X., et al.
Published: (2025)
Attention-aware Inference Optimizations for Large Vision-Language Models with Memory-efficient Decoding
by: Ilhan, Fatih, et al.
Published: (2026)
by: Ilhan, Fatih, et al.
Published: (2026)
Bi-ICE: An Inner Interpretable Framework for Image Classification via Bi-directional Interactions between Concept and Input Embeddings
by: Hong, Jinyung, et al.
Published: (2024)
by: Hong, Jinyung, et al.
Published: (2024)
Rethinking Data Input for Point Cloud Upsampling
by: Zhang, Tongxu
Published: (2024)
by: Zhang, Tongxu
Published: (2024)
QIANets: Quantum-Integrated Adaptive Networks for Reduced Latency and Improved Inference Times in CNN Models
by: Balapanov, Zhumazhan, et al.
Published: (2024)
by: Balapanov, Zhumazhan, et al.
Published: (2024)
Spatial-and-Frequency-aware Restoration method for Images based on Diffusion Models
by: Lee, Kyungsung, et al.
Published: (2024)
by: Lee, Kyungsung, et al.
Published: (2024)
Harnessing Data Asymmetry: Manifold Learning in the Finsler World
by: Dagès, Thomas, et al.
Published: (2026)
by: Dagès, Thomas, et al.
Published: (2026)
Similar Items
-
Do Visual Imaginations Improve Vision-and-Language Navigation Agents?
by: Perincherry, Akhil, et al.
Published: (2025) -
You Never Know: Quantization Induces Inconsistent Biases in Vision-Language Foundation Models
by: Slyman, Eric, et al.
Published: (2024) -
Fail-Closed Alignment for Large Language Models
by: Coalson, Zachary, et al.
Published: (2026) -
Advanced Knowledge Transfer: Refined Feature Distillation for Zero-Shot Quantization in Edge Computing
by: Hong, Inpyo, et al.
Published: (2024) -
Holistic Unlearning Benchmark: A Multi-Faceted Evaluation for Text-to-Image Diffusion Model Unlearning
by: Moon, Saemi, et al.
Published: (2024)