Mitigating Perspective Distortion-induced Shape Ambiguity in Image Crops
Fuente:
arXiv
Saved in:
| Main Authors: | Prakash, Aditya, Gupta, Arjun, Gupta, Saurabh |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bimanual 3D Hand Motion and Articulation Forecasting in Everyday Images
by: Prakash, Aditya, et al.
Published: (2025)
by: Prakash, Aditya, et al.
Published: (2025)
3D Hand Pose Estimation in Everyday Egocentric Images
by: Prakash, Aditya, et al.
Published: (2023)
by: Prakash, Aditya, et al.
Published: (2023)
Precise Mobile Manipulation of Small Everyday Objects
by: Gupta, Arjun, et al.
Published: (2025)
by: Gupta, Arjun, et al.
Published: (2025)
3D Reconstruction of Objects in Hands without Real World 3D Supervision
by: Prakash, Aditya, et al.
Published: (2023)
by: Prakash, Aditya, et al.
Published: (2023)
Opening Articulated Structures in the Real World
by: Gupta, Arjun, et al.
Published: (2024)
by: Gupta, Arjun, et al.
Published: (2024)
How Do I Do That? Synthesizing 3D Hand Motion and Contacts for Everyday Interactions
by: Prakash, Aditya, et al.
Published: (2025)
by: Prakash, Aditya, et al.
Published: (2025)
PhysGen: Rigid-Body Physics-Grounded Image-to-Video Generation
by: Liu, Shaowei, et al.
Published: (2024)
by: Liu, Shaowei, et al.
Published: (2024)
Push Past Green: Learning to Look Behind Plant Foliage by Moving It
by: Zhang, Xiaoyu, et al.
Published: (2023)
by: Zhang, Xiaoyu, et al.
Published: (2023)
When Better Eyes Lead to Blindness: A Diagnostic Study of the Information Bottleneck in CNN-LSTM Image Captioning Models
by: Gupta, Hitesh Kumar
Published: (2025)
by: Gupta, Hitesh Kumar
Published: (2025)
Diffusion Meets DAgger: Supercharging Eye-in-hand Imitation Learning
by: Zhang, Xiaoyu, et al.
Published: (2024)
by: Zhang, Xiaoyu, et al.
Published: (2024)
Specify and Edit: Overcoming Ambiguity in Text-Based Image Editing
by: Iakovleva, Ekaterina, et al.
Published: (2024)
by: Iakovleva, Ekaterina, et al.
Published: (2024)
Mitigating Bad Ground Truth in Supervised Machine Learning based Crop Classification: A Multi-Level Framework with Sentinel-2 Images
by: A, Sanayya, et al.
Published: (2025)
by: A, Sanayya, et al.
Published: (2025)
Visual Sync: Multi-Camera Synchronization via Cross-View Object Motion
by: Liu, Shaowei, et al.
Published: (2025)
by: Liu, Shaowei, et al.
Published: (2025)
Prototype Guided Backdoor Defense
by: Amula, Venkat Adithya, et al.
Published: (2025)
by: Amula, Venkat Adithya, et al.
Published: (2025)
IDAL: Improved Domain Adaptive Learning for Natural Images Dataset
by: Gupta, Ravi Kant, et al.
Published: (2025)
by: Gupta, Ravi Kant, et al.
Published: (2025)
DistortBench: Benchmarking Vision Language Models on Image Distortion Identification
by: Goyal, Divyanshu, et al.
Published: (2026)
by: Goyal, Divyanshu, et al.
Published: (2026)
An Investigation of Visual Foundation Models Robustness
by: Gupta, Sandeep, et al.
Published: (2025)
by: Gupta, Sandeep, et al.
Published: (2025)
Scalable Whole Slide Image Representation Using K-Mean Clustering and Fisher Vector Aggregation
by: Gupta, Ravi Kant, et al.
Published: (2025)
by: Gupta, Ravi Kant, et al.
Published: (2025)
Towards Ambiguity-Free Spatial Foundation Model: Rethinking and Decoupling Depth Ambiguity
by: Xu, Xiaohao, et al.
Published: (2025)
by: Xu, Xiaohao, et al.
Published: (2025)
An Uncertainty-Aware Loss Function Incorporating Fuzzy Logic: Application to MRI Brain Image Segmentation
by: Verma, Hanuman, et al.
Published: (2026)
by: Verma, Hanuman, et al.
Published: (2026)
Panoptic Pairwise Distortion Graph
by: Janjua, Muhammad Kamran, et al.
Published: (2026)
by: Janjua, Muhammad Kamran, et al.
Published: (2026)
BiPrompt: Bilateral Prompt Optimization for Visual and Textual Debiasing in Vision-Language Models
by: Gupta, Sunny, et al.
Published: (2026)
by: Gupta, Sunny, et al.
Published: (2026)
When Generative Augmentation Hurts: A Benchmark Study of GAN and Diffusion Models for Bias Correction in AI Classification Systems
by: Gupta, Shesh Narayan, et al.
Published: (2026)
by: Gupta, Shesh Narayan, et al.
Published: (2026)
DeepRepViz: Identifying Confounders in Deep Learning Model Predictions
by: Rane, Roshan Prakash, et al.
Published: (2023)
by: Rane, Roshan Prakash, et al.
Published: (2023)
Proactive Gradient Conflict Mitigation in Multi-Task Learning: A Sparse Training Perspective
by: Zhang, Zhi, et al.
Published: (2024)
by: Zhang, Zhi, et al.
Published: (2024)
Face Detection: Present State and Research Directions
by: Prabhat, Purnendu, et al.
Published: (2024)
by: Prabhat, Purnendu, et al.
Published: (2024)
LightPneumoNet: Lightweight Pneumonia Classifier
by: Chauhan, Neilansh, et al.
Published: (2025)
by: Chauhan, Neilansh, et al.
Published: (2025)
Benchmarking Suite for Synthetic Aperture Radar Imagery Anomaly Detection (SARIAD) Algorithms
by: Chauvin, Lucian, et al.
Published: (2025)
by: Chauvin, Lucian, et al.
Published: (2025)
Improved Crop and Weed Detection with Diverse Data Ensemble Learning
by: Asad, Muhammad Hamza, et al.
Published: (2023)
by: Asad, Muhammad Hamza, et al.
Published: (2023)
A Simple and Effective Reinforcement Learning Method for Text-to-Image Diffusion Fine-tuning
by: Gupta, Shashank, et al.
Published: (2025)
by: Gupta, Shashank, et al.
Published: (2025)
Training-free Detection of AI-generated images via Cropping Robustness
by: Choi, Sungik, et al.
Published: (2025)
by: Choi, Sungik, et al.
Published: (2025)
Learned Visual Navigation for Under-Canopy Agricultural Robots
by: Sivakumar, Arun Narenthiran, et al.
Published: (2021)
by: Sivakumar, Arun Narenthiran, et al.
Published: (2021)
Mitigating Catastrophic Forgetting and Mode Collapse in Text-to-Image Diffusion via Latent Replay
by: Otani, Aoi
Published: (2025)
by: Otani, Aoi
Published: (2025)
Diffusion Soup: Model Merging for Text-to-Image Diffusion Models
by: Biggs, Benjamin, et al.
Published: (2024)
by: Biggs, Benjamin, et al.
Published: (2024)
TrackNetV4: Enhancing Fast Sports Object Tracking with Motion Attention Maps
by: Raj, Arjun, et al.
Published: (2024)
by: Raj, Arjun, et al.
Published: (2024)
ShapeWords: Guiding Text-to-Image Synthesis with 3D Shape-Aware Prompts
by: Petrov, Dmitry, et al.
Published: (2024)
by: Petrov, Dmitry, et al.
Published: (2024)
Fair Foundation Models for Medical Image Analysis: Challenges and Perspectives
by: Queiroz, Dilermando, et al.
Published: (2025)
by: Queiroz, Dilermando, et al.
Published: (2025)
Picturing Ambiguity: A Visual Twist on the Winograd Schema Challenge
by: Park, Brendan, et al.
Published: (2024)
by: Park, Brendan, et al.
Published: (2024)
Meta-Adaptive Prompt Distillation for Few-Shot Visual Question Answering
by: Gupta, Akash, et al.
Published: (2025)
by: Gupta, Akash, et al.
Published: (2025)
Robust 3D Shape Reconstruction in Zero-Shot from a Single Image in the Wild
by: Cho, Junhyeong, et al.
Published: (2024)
by: Cho, Junhyeong, et al.
Published: (2024)
Similar Items
-
Bimanual 3D Hand Motion and Articulation Forecasting in Everyday Images
by: Prakash, Aditya, et al.
Published: (2025) -
3D Hand Pose Estimation in Everyday Egocentric Images
by: Prakash, Aditya, et al.
Published: (2023) -
Precise Mobile Manipulation of Small Everyday Objects
by: Gupta, Arjun, et al.
Published: (2025) -
3D Reconstruction of Objects in Hands without Real World 3D Supervision
by: Prakash, Aditya, et al.
Published: (2023) -
Opening Articulated Structures in the Real World
by: Gupta, Arjun, et al.
Published: (2024)