3D Reconstruction from Sketches
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Talwar, Abhimanyu, Laasri, Julien |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Instance Segmentation for Point Sets
von: Talwar, Abhimanyu, et al.
Veröffentlicht: (2025)
von: Talwar, Abhimanyu, et al.
Veröffentlicht: (2025)
Visual-Text Cross Alignment: Refining the Similarity Score in Vision-Language Models
von: Li, Jinhao, et al.
Veröffentlicht: (2024)
von: Li, Jinhao, et al.
Veröffentlicht: (2024)
Harmony: A Joint Self-Supervised and Weakly-Supervised Framework for Learning General Purpose Visual Representations
von: Baharoon, Mohammed, et al.
Veröffentlicht: (2024)
von: Baharoon, Mohammed, et al.
Veröffentlicht: (2024)
Zero-Shot Multi-Criteria Visual Quality Inspection for Semi-Controlled Industrial Environments via Real-Time 3D Digital Twin Simulation
von: Araya-Martinez, Jose Moises, et al.
Veröffentlicht: (2025)
von: Araya-Martinez, Jose Moises, et al.
Veröffentlicht: (2025)
RefineFormer3D: Efficient 3D Medical Image Segmentation via Adaptive Multi-Scale Transformer with Cross Attention Fusion
von: Tyagi, Kavyansh, et al.
Veröffentlicht: (2026)
von: Tyagi, Kavyansh, et al.
Veröffentlicht: (2026)
Image Reconstruction as a Tool for Feature Analysis
von: Allakhverdov, Eduard, et al.
Veröffentlicht: (2025)
von: Allakhverdov, Eduard, et al.
Veröffentlicht: (2025)
Physics-informed Variational Autoencoders for Improved Robustness to Environmental Factors of Variation
von: Thoreau, Romain, et al.
Veröffentlicht: (2022)
von: Thoreau, Romain, et al.
Veröffentlicht: (2022)
The Power of Next-Frame Prediction for Learning Physical Laws
von: Winterbottom, Thomas, et al.
Veröffentlicht: (2024)
von: Winterbottom, Thomas, et al.
Veröffentlicht: (2024)
The MSR-Video to Text Dataset with Clean Annotations
von: Chen, Haoran, et al.
Veröffentlicht: (2021)
von: Chen, Haoran, et al.
Veröffentlicht: (2021)
Sequence Matters: Harnessing Video Models in 3D Super-Resolution
von: Ko, Hyun-kyu, et al.
Veröffentlicht: (2024)
von: Ko, Hyun-kyu, et al.
Veröffentlicht: (2024)
ClustViT: Clustering-based Token Merging for Semantic Segmentation
von: Montello, Fabio, et al.
Veröffentlicht: (2025)
von: Montello, Fabio, et al.
Veröffentlicht: (2025)
Synthetic Industrial Object Detection: GenAI vs. Feature-Based Methods
von: Araya-Martinez, Jose Moises, et al.
Veröffentlicht: (2025)
von: Araya-Martinez, Jose Moises, et al.
Veröffentlicht: (2025)
One-to-Normal: Anomaly Personalization for Few-shot Anomaly Detection
von: Li, Yiyue, et al.
Veröffentlicht: (2025)
von: Li, Yiyue, et al.
Veröffentlicht: (2025)
Multi-scale Temporal Prediction via Incremental Generation and Multi-agent Collaboration
von: Zeng, Zhitao, et al.
Veröffentlicht: (2025)
von: Zeng, Zhitao, et al.
Veröffentlicht: (2025)
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
von: Zeng, Zhitao, et al.
Veröffentlicht: (2025)
von: Zeng, Zhitao, et al.
Veröffentlicht: (2025)
Extrapolating and Decoupling Image-to-Video Generation Models: Motion Modeling is Easier Than You Think
von: Tian, Jie, et al.
Veröffentlicht: (2025)
von: Tian, Jie, et al.
Veröffentlicht: (2025)
Deep Learning Approaches for Human Action Recognition in Video Data
von: Xie, Yufei
Veröffentlicht: (2024)
von: Xie, Yufei
Veröffentlicht: (2024)
LatentForensics: Towards frugal deepfake detection in the StyleGAN latent space
von: Delmas, Matthieu, et al.
Veröffentlicht: (2023)
von: Delmas, Matthieu, et al.
Veröffentlicht: (2023)
PlaneSAM: Multimodal Plane Instance Segmentation Using the Segment Anything Model
von: Deng, Zhongchen, et al.
Veröffentlicht: (2024)
von: Deng, Zhongchen, et al.
Veröffentlicht: (2024)
UrbanAlign: Post-hoc Semantic Calibration for VLM-Human Preference Alignment
von: Zhang, Yecheng, et al.
Veröffentlicht: (2026)
von: Zhang, Yecheng, et al.
Veröffentlicht: (2026)
SynthRender and IRIS: Open-Source Framework and Dataset for Bidirectional Sim-Real Transfer in Industrial Object Perception
von: Araya-Martinez, Jose Moises, et al.
Veröffentlicht: (2026)
von: Araya-Martinez, Jose Moises, et al.
Veröffentlicht: (2026)
DiffYOLO: Object Detection for Anti-Noise via YOLO and Diffusion Models
von: Liu, Yichen, et al.
Veröffentlicht: (2024)
von: Liu, Yichen, et al.
Veröffentlicht: (2024)
Video-CoE: Reinforcing Video Event Prediction via Chain of Events
von: Su, Qile, et al.
Veröffentlicht: (2026)
von: Su, Qile, et al.
Veröffentlicht: (2026)
Light Transport-aware Diffusion Posterior Sampling for Single-View Reconstruction of 3D Volumes
von: Leonard, Ludwic, et al.
Veröffentlicht: (2025)
von: Leonard, Ludwic, et al.
Veröffentlicht: (2025)
Do Generative Metrics Predict YOLO Performance? An Evaluation Across Models, Augmentation Ratios, and Dataset Complexity
von: Marian, Vasile, et al.
Veröffentlicht: (2026)
von: Marian, Vasile, et al.
Veröffentlicht: (2026)
AMANet: Advancing SAR Ship Detection with Adaptive Multi-Hierarchical Attention Network
von: Ma, Xiaolin, et al.
Veröffentlicht: (2024)
von: Ma, Xiaolin, et al.
Veröffentlicht: (2024)
Learning Discriminative Spatio-temporal Representations for Semi-supervised Action Recognition
von: Wang, Yu, et al.
Veröffentlicht: (2024)
von: Wang, Yu, et al.
Veröffentlicht: (2024)
A Spitting Image: Modular Superpixel Tokenization in Vision Transformers
von: Aasan, Marius, et al.
Veröffentlicht: (2024)
von: Aasan, Marius, et al.
Veröffentlicht: (2024)
ShapBPT: Image Feature Attributions Using Data-Aware Binary Partition Trees
von: Rashid, Muhammad, et al.
Veröffentlicht: (2026)
von: Rashid, Muhammad, et al.
Veröffentlicht: (2026)
When Less is Enough: Adaptive Token Reduction for Efficient Image Representation
von: Allakhverdov, Eduard, et al.
Veröffentlicht: (2025)
von: Allakhverdov, Eduard, et al.
Veröffentlicht: (2025)
Unsupervised Anomaly Detection Using Diffusion Trend Analysis for Display Inspection
von: Kim, Eunwoo, et al.
Veröffentlicht: (2024)
von: Kim, Eunwoo, et al.
Veröffentlicht: (2024)
On the Inherent Robustness of One-Stage Object Detection against Out-of-Distribution Data
von: Martinez-Seras, Aitor, et al.
Veröffentlicht: (2024)
von: Martinez-Seras, Aitor, et al.
Veröffentlicht: (2024)
SUGARCREPE++ Dataset: Vision-Language Model Sensitivity to Semantic and Lexical Alterations
von: Dumpala, Sri Harsha, et al.
Veröffentlicht: (2024)
von: Dumpala, Sri Harsha, et al.
Veröffentlicht: (2024)
Non-Robust Features are Not Always Useful in One-Class Classification
von: Lau, Matthew, et al.
Veröffentlicht: (2024)
von: Lau, Matthew, et al.
Veröffentlicht: (2024)
Fast 3D point clouds retrieval for Large-scale 3D Place Recognition
von: Zede, Chahine-Nicolas, et al.
Veröffentlicht: (2025)
von: Zede, Chahine-Nicolas, et al.
Veröffentlicht: (2025)
VA-$π$: Variational Policy Alignment for Pixel-Aware Autoregressive Generation
von: Liao, Xinyao, et al.
Veröffentlicht: (2025)
von: Liao, Xinyao, et al.
Veröffentlicht: (2025)
Supervised Learning Has a Necessary Geometric Blind Spot: Theory, Consequences, and Minimal Repair
von: Rajput, Vishal
Veröffentlicht: (2026)
von: Rajput, Vishal
Veröffentlicht: (2026)
Modular Deep Learning Framework for Assistive Perception: Gaze, Affect, and Speaker Identification
von: Anchan, Akshit Pramod, et al.
Veröffentlicht: (2025)
von: Anchan, Akshit Pramod, et al.
Veröffentlicht: (2025)
Splat and Distill: Augmenting Teachers with Feed-Forward 3D Reconstruction For 3D-Aware Distillation
von: Shavin, David, et al.
Veröffentlicht: (2026)
von: Shavin, David, et al.
Veröffentlicht: (2026)
bit2bit: 1-bit quanta video reconstruction via self-supervised photon prediction
von: Liu, Yehe, et al.
Veröffentlicht: (2024)
von: Liu, Yehe, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Instance Segmentation for Point Sets
von: Talwar, Abhimanyu, et al.
Veröffentlicht: (2025) -
Visual-Text Cross Alignment: Refining the Similarity Score in Vision-Language Models
von: Li, Jinhao, et al.
Veröffentlicht: (2024) -
Harmony: A Joint Self-Supervised and Weakly-Supervised Framework for Learning General Purpose Visual Representations
von: Baharoon, Mohammed, et al.
Veröffentlicht: (2024) -
Zero-Shot Multi-Criteria Visual Quality Inspection for Semi-Controlled Industrial Environments via Real-Time 3D Digital Twin Simulation
von: Araya-Martinez, Jose Moises, et al.
Veröffentlicht: (2025) -
RefineFormer3D: Efficient 3D Medical Image Segmentation via Adaptive Multi-Scale Transformer with Cross Attention Fusion
von: Tyagi, Kavyansh, et al.
Veröffentlicht: (2026)