Learning to Transform for Generalizable Instance-wise Invariance
Fuente:
arXiv
Saved in:
| Main Authors: | Singhal, Utkarsh, Esteves, Carlos, Makadia, Ameesh, Yu, Stella X. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Spectrally-Guided Diffusion Noise Schedules
by: Esteves, Carlos, et al.
Published: (2026)
by: Esteves, Carlos, et al.
Published: (2026)
Spectral Image Tokenizer
by: Esteves, Carlos, et al.
Published: (2024)
by: Esteves, Carlos, et al.
Published: (2024)
Single Mesh Diffusion Models with Field Latents for Texture Generation
by: Mitchel, Thomas W., et al.
Published: (2023)
by: Mitchel, Thomas W., et al.
Published: (2023)
Co-domain Symmetry for Complex-Valued Deep Learning
by: Singhal, Utkarsh, et al.
Published: (2021)
by: Singhal, Utkarsh, et al.
Published: (2021)
Factorized Video Autoencoders for Efficient Generative Modelling
by: Suhail, Mohammed, et al.
Published: (2024)
by: Suhail, Mohammed, et al.
Published: (2024)
Test-Time Canonicalization by Foundation Models for Robust Perception
by: Singhal, Utkarsh, et al.
Published: (2025)
by: Singhal, Utkarsh, et al.
Published: (2025)
Wide-Baseline Relative Camera Pose Estimation with Directional Learning
by: Chen, Kefan, et al.
Published: (2021)
by: Chen, Kefan, et al.
Published: (2021)
Instance-wise Supervision-level Optimization in Active Learning
by: Matsuo, Shinnosuke, et al.
Published: (2025)
by: Matsuo, Shinnosuke, et al.
Published: (2025)
Decomposing Private Image Generation via Coarse-to-Fine Wavelet Modeling
by: Bayrooti, Jasmine, et al.
Published: (2026)
by: Bayrooti, Jasmine, et al.
Published: (2026)
3DCodeBench: Benchmarking Agentic Procedural 3D Modeling Via Code
by: Gao, Yipeng, et al.
Published: (2026)
by: Gao, Yipeng, et al.
Published: (2026)
Pose-Aware Self-Supervised Learning with Viewpoint Trajectory Regularization
by: Wang, Jiayun, et al.
Published: (2024)
by: Wang, Jiayun, et al.
Published: (2024)
gen2seg: Generative Models Enable Generalizable Instance Segmentation
by: Khangaonkar, Om, et al.
Published: (2025)
by: Khangaonkar, Om, et al.
Published: (2025)
No Training Wheels: Steering Vectors for Bias Correction at Inference Time
by: Gupta, Aviral, et al.
Published: (2025)
by: Gupta, Aviral, et al.
Published: (2025)
From Pixels to Perception: Interpretable Predictions via Instance-wise Grouped Feature Selection
by: Vandenhirtz, Moritz, et al.
Published: (2025)
by: Vandenhirtz, Moritz, et al.
Published: (2025)
Tailored Transformation Invariance for Industrial Anomaly Detection
by: Schönfeld, Mariette, et al.
Published: (2025)
by: Schönfeld, Mariette, et al.
Published: (2025)
Meta Invariance Defense Towards Generalizable Robustness to Unknown Adversarial Attacks
by: Zhang, Lei, et al.
Published: (2024)
by: Zhang, Lei, et al.
Published: (2024)
Unveil Inversion and Invariance in Flow Transformer for Versatile Image Editing
by: Xu, Pengcheng, et al.
Published: (2024)
by: Xu, Pengcheng, et al.
Published: (2024)
PoissonNet: A Local-Global Approach for Learning on Surfaces
by: Maesumi, Arman, et al.
Published: (2025)
by: Maesumi, Arman, et al.
Published: (2025)
A Novel Shape Guided Transformer Network for Instance Segmentation in Remote Sensing Images
by: Yu, Dawen, et al.
Published: (2024)
by: Yu, Dawen, et al.
Published: (2024)
Instance-Aware Group Quantization for Vision Transformers
by: Moon, Jaehyeon, et al.
Published: (2024)
by: Moon, Jaehyeon, et al.
Published: (2024)
Learning Conditional Invariances through Non-Commutativity
by: Chaudhuri, Abhra, et al.
Published: (2024)
by: Chaudhuri, Abhra, et al.
Published: (2024)
Mitigating Instance Entanglement in Instance-Dependent Partial Label Learning
by: Zhao, Rui, et al.
Published: (2026)
by: Zhao, Rui, et al.
Published: (2026)
Eff-GRot: Efficient and Generalizable Rotation Estimation with Transformers
by: Mathioulakis, Fanis, et al.
Published: (2025)
by: Mathioulakis, Fanis, et al.
Published: (2025)
Class-wise Balancing Data Replay for Federated Class-Incremental Learning
by: Qi, Zhuang, et al.
Published: (2025)
by: Qi, Zhuang, et al.
Published: (2025)
ELSA: Exploiting Layer-wise N:M Sparsity for Vision Transformer Acceleration
by: Huang, Ning-Chi, et al.
Published: (2024)
by: Huang, Ning-Chi, et al.
Published: (2024)
Set2Seq Transformer: Temporal and Position-Aware Set Representations for Sequential Multiple-Instance Learning
by: Efthymiou, Athanasios, et al.
Published: (2024)
by: Efthymiou, Athanasios, et al.
Published: (2024)
VONet: Unsupervised Video Object Learning With Parallel U-Net Attention and Object-wise Sequential VAE
by: Yu, Haonan, et al.
Published: (2024)
by: Yu, Haonan, et al.
Published: (2024)
Accelerating Augmentation Invariance Pretraining
by: Lin, Jinhong, et al.
Published: (2024)
by: Lin, Jinhong, et al.
Published: (2024)
Improving Knowledge Distillation in Transfer Learning with Layer-wise Learning Rates
by: Kokane, Shirley, et al.
Published: (2024)
by: Kokane, Shirley, et al.
Published: (2024)
Beyond Instance Consistency: Investigating View Diversity in Self-supervised Learning
by: Qin, Huaiyuan, et al.
Published: (2025)
by: Qin, Huaiyuan, et al.
Published: (2025)
Federated Learning with Instance-Dependent Noisy Label
by: Wang, Lei, et al.
Published: (2023)
by: Wang, Lei, et al.
Published: (2023)
Fast and Efficient Transformer-based Method for Bird's Eye View Instance Prediction
by: Antunes-García, Miguel, et al.
Published: (2024)
by: Antunes-García, Miguel, et al.
Published: (2024)
GNN-ViTCap: GNN-Enhanced Multiple Instance Learning with Vision Transformers for Whole Slide Image Classification and Captioning
by: Raju, S M Taslim Uddin, et al.
Published: (2025)
by: Raju, S M Taslim Uddin, et al.
Published: (2025)
Extracting Usable Predictions from Quantized Networks through Uncertainty Quantification for OOD Detection
by: Singhal, Rishi, et al.
Published: (2024)
by: Singhal, Rishi, et al.
Published: (2024)
MILD: Modeling the Instance Learning Dynamics for Learning with Noisy Labels
by: Hu, Chuanyang, et al.
Published: (2023)
by: Hu, Chuanyang, et al.
Published: (2023)
Advances in Multiple Instance Learning for Whole Slide Image Analysis: Techniques, Challenges, and Future Directions
by: Wang, Jun, et al.
Published: (2024)
by: Wang, Jun, et al.
Published: (2024)
Impact of Layer Norm on Memorization and Generalization in Transformers
by: Singhal, Rishi, et al.
Published: (2025)
by: Singhal, Rishi, et al.
Published: (2025)
Accelerating Diffusion Transformers with Token-wise Feature Caching
by: Zou, Chang, et al.
Published: (2024)
by: Zou, Chang, et al.
Published: (2024)
Background Invariance Testing According to Semantic Proximity
by: Liao, Zukang, et al.
Published: (2022)
by: Liao, Zukang, et al.
Published: (2022)
Pixel-wise RL on Diffusion Models: Reinforcement Learning from Rich Feedback
by: Kordzanganeh, Mo, et al.
Published: (2024)
by: Kordzanganeh, Mo, et al.
Published: (2024)
Similar Items
-
Spectrally-Guided Diffusion Noise Schedules
by: Esteves, Carlos, et al.
Published: (2026) -
Spectral Image Tokenizer
by: Esteves, Carlos, et al.
Published: (2024) -
Single Mesh Diffusion Models with Field Latents for Texture Generation
by: Mitchel, Thomas W., et al.
Published: (2023) -
Co-domain Symmetry for Complex-Valued Deep Learning
by: Singhal, Utkarsh, et al.
Published: (2021) -
Factorized Video Autoencoders for Efficient Generative Modelling
by: Suhail, Mohammed, et al.
Published: (2024)