Möbius Transform for Mitigating Perspective Distortions in Representation Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Chhipa, Prakash Chandra, Chippa, Meenakshi Subhash, De, Kanjar, Saini, Rajkumar, Liwicki, Marcus, Shah, Mubarak |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LCM: Log Conformal Maps for Robust Representation Learning to Mitigate Perspective Distortion
di: Chippa, Meenakshi Subhash, et al.
Pubblicazione: (2024)
di: Chippa, Meenakshi Subhash, et al.
Pubblicazione: (2024)
Open-Vocabulary Object Detectors: Robustness Challenges under Distribution Shifts
di: Chhipa, Prakash Chandra, et al.
Pubblicazione: (2024)
di: Chhipa, Prakash Chandra, et al.
Pubblicazione: (2024)
A Systematic Performance Analysis of Deep Perceptual Loss Networks: Breaking Transfer Learning Conventions
di: Pihlgren, Gustav Grund, et al.
Pubblicazione: (2023)
di: Pihlgren, Gustav Grund, et al.
Pubblicazione: (2023)
VidTAG: Temporally Aligned Video to GPS Geolocalization with Denoising Sequence Prediction at a Global Scale
di: Kulkarni, Parth Parag, et al.
Pubblicazione: (2026)
di: Kulkarni, Parth Parag, et al.
Pubblicazione: (2026)
Giving each task what it needs -- leveraging structured sparsity for tailored multi-task learning
di: Upadhyay, Richa, et al.
Pubblicazione: (2024)
di: Upadhyay, Richa, et al.
Pubblicazione: (2024)
Meta-Sparsity: Learning Optimal Sparse Structures in Multi-task Networks through Meta-learning
di: Upadhyay, Richa, et al.
Pubblicazione: (2025)
di: Upadhyay, Richa, et al.
Pubblicazione: (2025)
Cross-Language Learning within Arabic Script for Low-Resource HTR
di: Al-azzawi, Sana, et al.
Pubblicazione: (2026)
di: Al-azzawi, Sana, et al.
Pubblicazione: (2026)
Mitigating Perspective Distortion-induced Shape Ambiguity in Image Crops
di: Prakash, Aditya, et al.
Pubblicazione: (2023)
di: Prakash, Aditya, et al.
Pubblicazione: (2023)
Distilling Vision Transformers for Distortion-Robust Representation Learning
di: Alexis, Konstantinos, et al.
Pubblicazione: (2026)
di: Alexis, Konstantinos, et al.
Pubblicazione: (2026)
Enhancing Privacy-Utility Trade-offs to Mitigate Memorization in Diffusion Models
di: Chen, Chen, et al.
Pubblicazione: (2025)
di: Chen, Chen, et al.
Pubblicazione: (2025)
CER-HV: A Human-in-the-Loop Framework for Cleaning Datasets Applied to Arabic-Script HTR
di: Al-azzawi, Sana, et al.
Pubblicazione: (2026)
di: Al-azzawi, Sana, et al.
Pubblicazione: (2026)
GAReT: Cross-view Video Geolocalization with Adapters and Auto-Regressive Transformers
di: Pillai, Manu S, et al.
Pubblicazione: (2024)
di: Pillai, Manu S, et al.
Pubblicazione: (2024)
Rethinking HTG Evaluation: Bridging Generation and Recognition
di: Nikolaidou, Konstantina, et al.
Pubblicazione: (2024)
di: Nikolaidou, Konstantina, et al.
Pubblicazione: (2024)
DiffusionPen: Towards Controlling the Style of Handwritten Text Generation
di: Nikolaidou, Konstantina, et al.
Pubblicazione: (2024)
di: Nikolaidou, Konstantina, et al.
Pubblicazione: (2024)
SemAttNet: Towards Attention-based Semantic Aware Guided Depth Completion
di: Nazir, Danish, et al.
Pubblicazione: (2022)
di: Nazir, Danish, et al.
Pubblicazione: (2022)
PTQ4DiT: Post-training Quantization for Diffusion Transformers
di: Wu, Junyi, et al.
Pubblicazione: (2024)
di: Wu, Junyi, et al.
Pubblicazione: (2024)
Cross-View Open-Vocabulary Object Detection in Aerial Imagery
di: Kini, Jyoti, et al.
Pubblicazione: (2025)
di: Kini, Jyoti, et al.
Pubblicazione: (2025)
Learnability-Guided Diffusion for Dataset Distillation
di: Chan-Santiago, Jeffrey A., et al.
Pubblicazione: (2026)
di: Chan-Santiago, Jeffrey A., et al.
Pubblicazione: (2026)
PackCache: A Training-Free Acceleration Method for Unified Autoregressive Video Generation via Compact KV-Cache
di: Li, Kunyang, et al.
Pubblicazione: (2026)
di: Li, Kunyang, et al.
Pubblicazione: (2026)
TimeLogic: A Temporal Logic Benchmark for Video QA
di: Swetha, Sirnam, et al.
Pubblicazione: (2025)
di: Swetha, Sirnam, et al.
Pubblicazione: (2025)
Understanding Cross-Language Transfer Improvements in Low-Resource HTR: The Role of Sequence Modeling
di: Al-azzawi, Sana, et al.
Pubblicazione: (2026)
di: Al-azzawi, Sana, et al.
Pubblicazione: (2026)
Radially Distorted Homographies, Revisited
di: Wadenbäck, Mårten, et al.
Pubblicazione: (2025)
di: Wadenbäck, Mårten, et al.
Pubblicazione: (2025)
Privacy Beyond Pixels: Latent Anonymization for Privacy-Preserving Video Understanding
di: Fioresi, Joseph, et al.
Pubblicazione: (2025)
di: Fioresi, Joseph, et al.
Pubblicazione: (2025)
Training-Free Multi-Concept Image Editing
di: Foteinopoulou, Niki, et al.
Pubblicazione: (2026)
di: Foteinopoulou, Niki, et al.
Pubblicazione: (2026)
LoRAtorio: An intrinsic approach to LoRA Skill Composition
di: Foteinopoulou, Niki, et al.
Pubblicazione: (2025)
di: Foteinopoulou, Niki, et al.
Pubblicazione: (2025)
ALBAR: Adversarial Learning approach to mitigate Biases in Action Recognition
di: Fioresi, Joseph, et al.
Pubblicazione: (2025)
di: Fioresi, Joseph, et al.
Pubblicazione: (2025)
Dual Orthogonal Guidance for Robust Diffusion-based Handwritten Text Generation
di: Nikolaidou, Konstantina, et al.
Pubblicazione: (2025)
di: Nikolaidou, Konstantina, et al.
Pubblicazione: (2025)
Shape2.5D: A Dataset of Texture-less Surfaces for Depth and Normals Estimation
di: Khan, Muhammad Saif Ullah, et al.
Pubblicazione: (2024)
di: Khan, Muhammad Saif Ullah, et al.
Pubblicazione: (2024)
Safe-LLaVA: A Privacy-Preserving Vision-Language Dataset and Benchmark for Biometric Safety
di: Kim, Younggun, et al.
Pubblicazione: (2025)
di: Kim, Younggun, et al.
Pubblicazione: (2025)
StretchySnake: Flexible SSM Training Unlocks Action Recognition Across Spatio-Temporal Scales
di: Siddiqui, Nyle, et al.
Pubblicazione: (2025)
di: Siddiqui, Nyle, et al.
Pubblicazione: (2025)
SegVG: Transferring Object Bounding Box to Segmentation for Visual Grounding
di: Kang, Weitai, et al.
Pubblicazione: (2024)
di: Kang, Weitai, et al.
Pubblicazione: (2024)
Exploring Local Memorization in Diffusion Models via Bright Ending Attention
di: Chen, Chen, et al.
Pubblicazione: (2024)
di: Chen, Chen, et al.
Pubblicazione: (2024)
CityGuessr: City-Level Video Geo-Localization on a Global Scale
di: Kulkarni, Parth Parag, et al.
Pubblicazione: (2024)
di: Kulkarni, Parth Parag, et al.
Pubblicazione: (2024)
Modified TSception for Analyzing Driver Drowsiness and Mental Workload from EEG
di: Siddhad, Gourav, et al.
Pubblicazione: (2025)
di: Siddhad, Gourav, et al.
Pubblicazione: (2025)
Attend Locally, Remember Linearly: Linear Attention as Cross-Frame Memory for Autoregressive Video Diffusion
di: Li, Kunyang, et al.
Pubblicazione: (2026)
di: Li, Kunyang, et al.
Pubblicazione: (2026)
Weakly-Supervised Spatiotemporal Anomaly Detection
di: Gianchandani, Urvi, et al.
Pubblicazione: (2026)
di: Gianchandani, Urvi, et al.
Pubblicazione: (2026)
Leveraging Pre-Trained Visual Models for AI-Generated Video Detection
di: Veeramachaneni, Keerthi, et al.
Pubblicazione: (2025)
di: Veeramachaneni, Keerthi, et al.
Pubblicazione: (2025)
From Play to Replay: Composed Video Retrieval for Temporally Fine-Grained Videos
di: Gupta, Animesh, et al.
Pubblicazione: (2025)
di: Gupta, Animesh, et al.
Pubblicazione: (2025)
LUCES-MV: A Multi-View Dataset for Near-Field Point Light Source Photometric Stereo
di: Logothetis, Fotios, et al.
Pubblicazione: (2024)
di: Logothetis, Fotios, et al.
Pubblicazione: (2024)
DiaLoc: An Iterative Approach to Embodied Dialog Localization
di: Zhang, Chao, et al.
Pubblicazione: (2024)
di: Zhang, Chao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
LCM: Log Conformal Maps for Robust Representation Learning to Mitigate Perspective Distortion
di: Chippa, Meenakshi Subhash, et al.
Pubblicazione: (2024) -
Open-Vocabulary Object Detectors: Robustness Challenges under Distribution Shifts
di: Chhipa, Prakash Chandra, et al.
Pubblicazione: (2024) -
A Systematic Performance Analysis of Deep Perceptual Loss Networks: Breaking Transfer Learning Conventions
di: Pihlgren, Gustav Grund, et al.
Pubblicazione: (2023) -
VidTAG: Temporally Aligned Video to GPS Geolocalization with Denoising Sequence Prediction at a Global Scale
di: Kulkarni, Parth Parag, et al.
Pubblicazione: (2026) -
Giving each task what it needs -- leveraging structured sparsity for tailored multi-task learning
di: Upadhyay, Richa, et al.
Pubblicazione: (2024)