Test-Time Canonicalization by Foundation Models for Robust Perception
Fuente:
arXiv
Salvato in:
| Autori principali: | Singhal, Utkarsh, Feng, Ryan, Yu, Stella X., Prakash, Atul |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Co-domain Symmetry for Complex-Valued Deep Learning
di: Singhal, Utkarsh, et al.
Pubblicazione: (2021)
di: Singhal, Utkarsh, et al.
Pubblicazione: (2021)
Learning to Transform for Generalizable Instance-wise Invariance
di: Singhal, Utkarsh, et al.
Pubblicazione: (2023)
di: Singhal, Utkarsh, et al.
Pubblicazione: (2023)
Defending Object Detectors against Patch Attacks with Out-of-Distribution Smoothing
di: Feng, Ryan, et al.
Pubblicazione: (2022)
di: Feng, Ryan, et al.
Pubblicazione: (2022)
3DPCNet: Pose Canonicalization for Robust Viewpoint-Invariant 3D Kinematic Analysis from Monocular RGB cameras
di: Ekanayake, Tharindu, et al.
Pubblicazione: (2025)
di: Ekanayake, Tharindu, et al.
Pubblicazione: (2025)
Local Scale Equivariance with Latent Deep Equilibrium Canonicalizer
di: Rahman, Md Ashiqur, et al.
Pubblicazione: (2025)
di: Rahman, Md Ashiqur, et al.
Pubblicazione: (2025)
RNAS-CL: Robust Neural Architecture Search by Cross-Layer Knowledge Distillation
di: Nath, Utkarsh, et al.
Pubblicazione: (2023)
di: Nath, Utkarsh, et al.
Pubblicazione: (2023)
Pose-Aware Self-Supervised Learning with Viewpoint Trajectory Regularization
di: Wang, Jiayun, et al.
Pubblicazione: (2024)
di: Wang, Jiayun, et al.
Pubblicazione: (2024)
Foundation Model-oriented Robustness: Robust Image Model Evaluation with Pretrained Models
di: Zhang, Peiyan, et al.
Pubblicazione: (2023)
di: Zhang, Peiyan, et al.
Pubblicazione: (2023)
Perception of Visual Content: Differences Between Humans and Foundation Models
di: Pratama, Nardiena A., et al.
Pubblicazione: (2024)
di: Pratama, Nardiena A., et al.
Pubblicazione: (2024)
Ramen: Robust Test-Time Adaptation of Vision-Language Models with Active Sample Selection
di: Bao, Wenxuan, et al.
Pubblicazione: (2026)
di: Bao, Wenxuan, et al.
Pubblicazione: (2026)
CodeMerge: Codebook-Guided Model Merging for Robust Test-Time Adaptation in Autonomous Driving
di: Yang, Huitong, et al.
Pubblicazione: (2025)
di: Yang, Huitong, et al.
Pubblicazione: (2025)
Test-Time Multimodal Backdoor Detection by Contrastive Prompting
di: Niu, Yuwei, et al.
Pubblicazione: (2024)
di: Niu, Yuwei, et al.
Pubblicazione: (2024)
Adapting in the Dark: Efficient and Stable Test-Time Adaptation for Black-Box Models
di: Zhang, Yunbei, et al.
Pubblicazione: (2026)
di: Zhang, Yunbei, et al.
Pubblicazione: (2026)
Car Drag Coefficient Prediction from 3D Point Clouds Using a Slice-Based Surrogate Model
di: Singh, Utkarsh, et al.
Pubblicazione: (2026)
di: Singh, Utkarsh, et al.
Pubblicazione: (2026)
Robust Adaptation of Foundation Models with Black-Box Visual Prompting
di: Oh, Changdae, et al.
Pubblicazione: (2024)
di: Oh, Changdae, et al.
Pubblicazione: (2024)
Lie Algebra Canonicalization: Equivariant Neural Operators under arbitrary Lie Groups
di: Shumaylov, Zakhar, et al.
Pubblicazione: (2024)
di: Shumaylov, Zakhar, et al.
Pubblicazione: (2024)
Towards Universal Fake Image Detectors that Generalize Across Generative Models
di: Ojha, Utkarsh, et al.
Pubblicazione: (2023)
di: Ojha, Utkarsh, et al.
Pubblicazione: (2023)
BiCLIP: Domain Canonicalization via Structured Geometric Transformation
di: Mantini, Pranav, et al.
Pubblicazione: (2026)
di: Mantini, Pranav, et al.
Pubblicazione: (2026)
Extracting Usable Predictions from Quantized Networks through Uncertainty Quantification for OOD Detection
di: Singhal, Rishi, et al.
Pubblicazione: (2024)
di: Singhal, Rishi, et al.
Pubblicazione: (2024)
SEVD: Synthetic Event-based Vision Dataset for Ego and Fixed Traffic Perception
di: Aliminati, Manideep Reddy, et al.
Pubblicazione: (2024)
di: Aliminati, Manideep Reddy, et al.
Pubblicazione: (2024)
Leveraging Hierarchical Feature Sharing for Efficient Dataset Condensation
di: Zheng, Haizhong, et al.
Pubblicazione: (2023)
di: Zheng, Haizhong, et al.
Pubblicazione: (2023)
A General Framework for Inference-time Scaling and Steering of Diffusion Models
di: Singhal, Raghav, et al.
Pubblicazione: (2025)
di: Singhal, Raghav, et al.
Pubblicazione: (2025)
Bridging Modalities via Progressive Re-alignment for Multimodal Test-Time Adaptation
di: Li, Jiacheng, et al.
Pubblicazione: (2025)
di: Li, Jiacheng, et al.
Pubblicazione: (2025)
Spurious Feature Eraser: Stabilizing Test-Time Adaptation for Vision-Language Foundation Model
di: Ma, Huan, et al.
Pubblicazione: (2024)
di: Ma, Huan, et al.
Pubblicazione: (2024)
Test-Time Training on Video Streams
di: Wang, Renhao, et al.
Pubblicazione: (2023)
di: Wang, Renhao, et al.
Pubblicazione: (2023)
Long-Tail Learning with Foundation Model: Heavy Fine-Tuning Hurts
di: Shi, Jiang-Xin, et al.
Pubblicazione: (2023)
di: Shi, Jiang-Xin, et al.
Pubblicazione: (2023)
Mitigating the Bias in the Model for Continual Test-Time Adaptation
di: Chung, Inseop, et al.
Pubblicazione: (2024)
di: Chung, Inseop, et al.
Pubblicazione: (2024)
FairDeFace: Evaluating the Fairness and Adversarial Robustness of Face Obfuscation Methods
di: Khorzooghi, Seyyed Mohammad Sadegh Moosavi, et al.
Pubblicazione: (2025)
di: Khorzooghi, Seyyed Mohammad Sadegh Moosavi, et al.
Pubblicazione: (2025)
An Investigation of Visual Foundation Models Robustness
di: Gupta, Sandeep, et al.
Pubblicazione: (2025)
di: Gupta, Sandeep, et al.
Pubblicazione: (2025)
Advancing Vision-based Human Action Recognition: Exploring Vision-Language CLIP Model for Generalisation in Domain-Independent Tasks
di: Shandilya, Utkarsh, et al.
Pubblicazione: (2025)
di: Shandilya, Utkarsh, et al.
Pubblicazione: (2025)
Calibrated and Robust Foundation Models for Vision-Language and Medical Image Tasks Under Distribution Shift
di: Khan, Behraj, et al.
Pubblicazione: (2025)
di: Khan, Behraj, et al.
Pubblicazione: (2025)
Rethinking Weight Decay for Robust Fine-Tuning of Foundation Models
di: Tian, Junjiao, et al.
Pubblicazione: (2024)
di: Tian, Junjiao, et al.
Pubblicazione: (2024)
Distilling Out-of-Distribution Robustness from Vision-Language Foundation Models
di: Zhou, Andy, et al.
Pubblicazione: (2023)
di: Zhou, Andy, et al.
Pubblicazione: (2023)
Test-Time Iterative Error Correction for Efficient Diffusion Models
di: Zhong, Yunshan, et al.
Pubblicazione: (2025)
di: Zhong, Yunshan, et al.
Pubblicazione: (2025)
Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion Models
di: Eyring, Luca, et al.
Pubblicazione: (2025)
di: Eyring, Luca, et al.
Pubblicazione: (2025)
When Test-Time Adaptation Meets Self-Supervised Models
di: Han, Jisu, et al.
Pubblicazione: (2025)
di: Han, Jisu, et al.
Pubblicazione: (2025)
Efficient Test-Time Scaling for Small Vision-Language Models
di: Kaya, Mehmet Onurcan, et al.
Pubblicazione: (2025)
di: Kaya, Mehmet Onurcan, et al.
Pubblicazione: (2025)
Adaptive Retention & Correction: Test-Time Training for Continual Learning
di: Chen, Haoran, et al.
Pubblicazione: (2024)
di: Chen, Haoran, et al.
Pubblicazione: (2024)
Toward Inherently Robust VLMs Against Visual Perception Attacks
di: MohajerAnsari, Pedram, et al.
Pubblicazione: (2025)
di: MohajerAnsari, Pedram, et al.
Pubblicazione: (2025)
Introducing Visual Perception Token into Multimodal Large Language Model
di: Yu, Runpeng, et al.
Pubblicazione: (2025)
di: Yu, Runpeng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Co-domain Symmetry for Complex-Valued Deep Learning
di: Singhal, Utkarsh, et al.
Pubblicazione: (2021) -
Learning to Transform for Generalizable Instance-wise Invariance
di: Singhal, Utkarsh, et al.
Pubblicazione: (2023) -
Defending Object Detectors against Patch Attacks with Out-of-Distribution Smoothing
di: Feng, Ryan, et al.
Pubblicazione: (2022) -
3DPCNet: Pose Canonicalization for Robust Viewpoint-Invariant 3D Kinematic Analysis from Monocular RGB cameras
di: Ekanayake, Tharindu, et al.
Pubblicazione: (2025) -
Local Scale Equivariance with Latent Deep Equilibrium Canonicalizer
di: Rahman, Md Ashiqur, et al.
Pubblicazione: (2025)