Test-Time Canonicalization by Foundation Models for Robust Perception
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Singhal, Utkarsh, Feng, Ryan, Yu, Stella X., Prakash, Atul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Co-domain Symmetry for Complex-Valued Deep Learning
von: Singhal, Utkarsh, et al.
Veröffentlicht: (2021)
von: Singhal, Utkarsh, et al.
Veröffentlicht: (2021)
Learning to Transform for Generalizable Instance-wise Invariance
von: Singhal, Utkarsh, et al.
Veröffentlicht: (2023)
von: Singhal, Utkarsh, et al.
Veröffentlicht: (2023)
Defending Object Detectors against Patch Attacks with Out-of-Distribution Smoothing
von: Feng, Ryan, et al.
Veröffentlicht: (2022)
von: Feng, Ryan, et al.
Veröffentlicht: (2022)
3DPCNet: Pose Canonicalization for Robust Viewpoint-Invariant 3D Kinematic Analysis from Monocular RGB cameras
von: Ekanayake, Tharindu, et al.
Veröffentlicht: (2025)
von: Ekanayake, Tharindu, et al.
Veröffentlicht: (2025)
Local Scale Equivariance with Latent Deep Equilibrium Canonicalizer
von: Rahman, Md Ashiqur, et al.
Veröffentlicht: (2025)
von: Rahman, Md Ashiqur, et al.
Veröffentlicht: (2025)
RNAS-CL: Robust Neural Architecture Search by Cross-Layer Knowledge Distillation
von: Nath, Utkarsh, et al.
Veröffentlicht: (2023)
von: Nath, Utkarsh, et al.
Veröffentlicht: (2023)
Pose-Aware Self-Supervised Learning with Viewpoint Trajectory Regularization
von: Wang, Jiayun, et al.
Veröffentlicht: (2024)
von: Wang, Jiayun, et al.
Veröffentlicht: (2024)
Foundation Model-oriented Robustness: Robust Image Model Evaluation with Pretrained Models
von: Zhang, Peiyan, et al.
Veröffentlicht: (2023)
von: Zhang, Peiyan, et al.
Veröffentlicht: (2023)
Perception of Visual Content: Differences Between Humans and Foundation Models
von: Pratama, Nardiena A., et al.
Veröffentlicht: (2024)
von: Pratama, Nardiena A., et al.
Veröffentlicht: (2024)
Ramen: Robust Test-Time Adaptation of Vision-Language Models with Active Sample Selection
von: Bao, Wenxuan, et al.
Veröffentlicht: (2026)
von: Bao, Wenxuan, et al.
Veröffentlicht: (2026)
CodeMerge: Codebook-Guided Model Merging for Robust Test-Time Adaptation in Autonomous Driving
von: Yang, Huitong, et al.
Veröffentlicht: (2025)
von: Yang, Huitong, et al.
Veröffentlicht: (2025)
Test-Time Multimodal Backdoor Detection by Contrastive Prompting
von: Niu, Yuwei, et al.
Veröffentlicht: (2024)
von: Niu, Yuwei, et al.
Veröffentlicht: (2024)
Adapting in the Dark: Efficient and Stable Test-Time Adaptation for Black-Box Models
von: Zhang, Yunbei, et al.
Veröffentlicht: (2026)
von: Zhang, Yunbei, et al.
Veröffentlicht: (2026)
Car Drag Coefficient Prediction from 3D Point Clouds Using a Slice-Based Surrogate Model
von: Singh, Utkarsh, et al.
Veröffentlicht: (2026)
von: Singh, Utkarsh, et al.
Veröffentlicht: (2026)
Robust Adaptation of Foundation Models with Black-Box Visual Prompting
von: Oh, Changdae, et al.
Veröffentlicht: (2024)
von: Oh, Changdae, et al.
Veröffentlicht: (2024)
Lie Algebra Canonicalization: Equivariant Neural Operators under arbitrary Lie Groups
von: Shumaylov, Zakhar, et al.
Veröffentlicht: (2024)
von: Shumaylov, Zakhar, et al.
Veröffentlicht: (2024)
Towards Universal Fake Image Detectors that Generalize Across Generative Models
von: Ojha, Utkarsh, et al.
Veröffentlicht: (2023)
von: Ojha, Utkarsh, et al.
Veröffentlicht: (2023)
BiCLIP: Domain Canonicalization via Structured Geometric Transformation
von: Mantini, Pranav, et al.
Veröffentlicht: (2026)
von: Mantini, Pranav, et al.
Veröffentlicht: (2026)
Extracting Usable Predictions from Quantized Networks through Uncertainty Quantification for OOD Detection
von: Singhal, Rishi, et al.
Veröffentlicht: (2024)
von: Singhal, Rishi, et al.
Veröffentlicht: (2024)
SEVD: Synthetic Event-based Vision Dataset for Ego and Fixed Traffic Perception
von: Aliminati, Manideep Reddy, et al.
Veröffentlicht: (2024)
von: Aliminati, Manideep Reddy, et al.
Veröffentlicht: (2024)
Leveraging Hierarchical Feature Sharing for Efficient Dataset Condensation
von: Zheng, Haizhong, et al.
Veröffentlicht: (2023)
von: Zheng, Haizhong, et al.
Veröffentlicht: (2023)
A General Framework for Inference-time Scaling and Steering of Diffusion Models
von: Singhal, Raghav, et al.
Veröffentlicht: (2025)
von: Singhal, Raghav, et al.
Veröffentlicht: (2025)
Bridging Modalities via Progressive Re-alignment for Multimodal Test-Time Adaptation
von: Li, Jiacheng, et al.
Veröffentlicht: (2025)
von: Li, Jiacheng, et al.
Veröffentlicht: (2025)
Spurious Feature Eraser: Stabilizing Test-Time Adaptation for Vision-Language Foundation Model
von: Ma, Huan, et al.
Veröffentlicht: (2024)
von: Ma, Huan, et al.
Veröffentlicht: (2024)
Test-Time Training on Video Streams
von: Wang, Renhao, et al.
Veröffentlicht: (2023)
von: Wang, Renhao, et al.
Veröffentlicht: (2023)
Long-Tail Learning with Foundation Model: Heavy Fine-Tuning Hurts
von: Shi, Jiang-Xin, et al.
Veröffentlicht: (2023)
von: Shi, Jiang-Xin, et al.
Veröffentlicht: (2023)
Mitigating the Bias in the Model for Continual Test-Time Adaptation
von: Chung, Inseop, et al.
Veröffentlicht: (2024)
von: Chung, Inseop, et al.
Veröffentlicht: (2024)
FairDeFace: Evaluating the Fairness and Adversarial Robustness of Face Obfuscation Methods
von: Khorzooghi, Seyyed Mohammad Sadegh Moosavi, et al.
Veröffentlicht: (2025)
von: Khorzooghi, Seyyed Mohammad Sadegh Moosavi, et al.
Veröffentlicht: (2025)
An Investigation of Visual Foundation Models Robustness
von: Gupta, Sandeep, et al.
Veröffentlicht: (2025)
von: Gupta, Sandeep, et al.
Veröffentlicht: (2025)
Advancing Vision-based Human Action Recognition: Exploring Vision-Language CLIP Model for Generalisation in Domain-Independent Tasks
von: Shandilya, Utkarsh, et al.
Veröffentlicht: (2025)
von: Shandilya, Utkarsh, et al.
Veröffentlicht: (2025)
Calibrated and Robust Foundation Models for Vision-Language and Medical Image Tasks Under Distribution Shift
von: Khan, Behraj, et al.
Veröffentlicht: (2025)
von: Khan, Behraj, et al.
Veröffentlicht: (2025)
Rethinking Weight Decay for Robust Fine-Tuning of Foundation Models
von: Tian, Junjiao, et al.
Veröffentlicht: (2024)
von: Tian, Junjiao, et al.
Veröffentlicht: (2024)
Distilling Out-of-Distribution Robustness from Vision-Language Foundation Models
von: Zhou, Andy, et al.
Veröffentlicht: (2023)
von: Zhou, Andy, et al.
Veröffentlicht: (2023)
Test-Time Iterative Error Correction for Efficient Diffusion Models
von: Zhong, Yunshan, et al.
Veröffentlicht: (2025)
von: Zhong, Yunshan, et al.
Veröffentlicht: (2025)
Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion Models
von: Eyring, Luca, et al.
Veröffentlicht: (2025)
von: Eyring, Luca, et al.
Veröffentlicht: (2025)
When Test-Time Adaptation Meets Self-Supervised Models
von: Han, Jisu, et al.
Veröffentlicht: (2025)
von: Han, Jisu, et al.
Veröffentlicht: (2025)
Efficient Test-Time Scaling for Small Vision-Language Models
von: Kaya, Mehmet Onurcan, et al.
Veröffentlicht: (2025)
von: Kaya, Mehmet Onurcan, et al.
Veröffentlicht: (2025)
Adaptive Retention & Correction: Test-Time Training for Continual Learning
von: Chen, Haoran, et al.
Veröffentlicht: (2024)
von: Chen, Haoran, et al.
Veröffentlicht: (2024)
Toward Inherently Robust VLMs Against Visual Perception Attacks
von: MohajerAnsari, Pedram, et al.
Veröffentlicht: (2025)
von: MohajerAnsari, Pedram, et al.
Veröffentlicht: (2025)
Introducing Visual Perception Token into Multimodal Large Language Model
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Co-domain Symmetry for Complex-Valued Deep Learning
von: Singhal, Utkarsh, et al.
Veröffentlicht: (2021) -
Learning to Transform for Generalizable Instance-wise Invariance
von: Singhal, Utkarsh, et al.
Veröffentlicht: (2023) -
Defending Object Detectors against Patch Attacks with Out-of-Distribution Smoothing
von: Feng, Ryan, et al.
Veröffentlicht: (2022) -
3DPCNet: Pose Canonicalization for Robust Viewpoint-Invariant 3D Kinematic Analysis from Monocular RGB cameras
von: Ekanayake, Tharindu, et al.
Veröffentlicht: (2025) -
Local Scale Equivariance with Latent Deep Equilibrium Canonicalizer
von: Rahman, Md Ashiqur, et al.
Veröffentlicht: (2025)