Gespeichert in:
| Hauptverfasser: | Bharadwaj, Skanda, Collins, Robert, Liu, Yanxi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2412.20666 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FootFormer: Estimating Stability from Visual Input
von: Kraiger, Keaton, et al.
Veröffentlicht: (2025)
von: Kraiger, Keaton, et al.
Veröffentlicht: (2025)
Convex Relaxation for Robust Vanishing Point Estimation in Manhattan World
von: Liao, Bangyan, et al.
Veröffentlicht: (2025)
von: Liao, Bangyan, et al.
Veröffentlicht: (2025)
Vanishing-Point-Guided Video Semantic Segmentation of Driving Scenes
von: Guo, Diandian, et al.
Veröffentlicht: (2024)
von: Guo, Diandian, et al.
Veröffentlicht: (2024)
TAPNext: Tracking Any Point (TAP) as Next Token Prediction
von: Zholus, Artem, et al.
Veröffentlicht: (2025)
von: Zholus, Artem, et al.
Veröffentlicht: (2025)
VPOcc: Exploiting Vanishing Point for 3D Semantic Occupancy Prediction
von: Kim, Junsu, et al.
Veröffentlicht: (2024)
von: Kim, Junsu, et al.
Veröffentlicht: (2024)
ControlVP: Interactive Geometric Refinement of AI-Generated Images with Consistent Vanishing Points
von: Okumura, Ryota, et al.
Veröffentlicht: (2025)
von: Okumura, Ryota, et al.
Veröffentlicht: (2025)
Transfer Learning-Based CNN Models for Plant Species Identification Using Leaf Venation Patterns
von: Bharadwaj, Bandita, et al.
Veröffentlicht: (2025)
von: Bharadwaj, Bandita, et al.
Veröffentlicht: (2025)
Mapping the Vanishing and Transformation of Urban Villages in China
von: Zhang, Wenyu, et al.
Veröffentlicht: (2025)
von: Zhang, Wenyu, et al.
Veröffentlicht: (2025)
RadarSplat: Radar Gaussian Splatting for High-Fidelity Data Synthesis and 3D Reconstruction of Autonomous Driving Scenes
von: Kung, Pou-Chun, et al.
Veröffentlicht: (2025)
von: Kung, Pou-Chun, et al.
Veröffentlicht: (2025)
Understanding Robustness of Visual State Space Models for Image Classification
von: Du, Chengbin, et al.
Veröffentlicht: (2024)
von: Du, Chengbin, et al.
Veröffentlicht: (2024)
Fus-MAE: A cross-attention-based data fusion approach for Masked Autoencoders in remote sensing
von: Chan-To-Hing, Hugo, et al.
Veröffentlicht: (2024)
von: Chan-To-Hing, Hugo, et al.
Veröffentlicht: (2024)
SpheriGait: Enriching Spatial Representation via Spherical Projection for LiDAR-based Gait Recognition
von: Wang, Yanxi, et al.
Veröffentlicht: (2024)
von: Wang, Yanxi, et al.
Veröffentlicht: (2024)
Memory Consolidation Enables Long-Context Video Understanding
von: Balažević, Ivana, et al.
Veröffentlicht: (2024)
von: Balažević, Ivana, et al.
Veröffentlicht: (2024)
BootsTAP: Bootstrapped Training for Tracking-Any-Point
von: Doersch, Carl, et al.
Veröffentlicht: (2024)
von: Doersch, Carl, et al.
Veröffentlicht: (2024)
TAPVid-3D: A Benchmark for Tracking Any Point in 3D
von: Koppula, Skanda, et al.
Veröffentlicht: (2024)
von: Koppula, Skanda, et al.
Veröffentlicht: (2024)
Integrating SAM Supervision for 3D Weakly Supervised Point Cloud Segmentation
von: You, Lechun, et al.
Veröffentlicht: (2025)
von: You, Lechun, et al.
Veröffentlicht: (2025)
CAT: Exploiting Inter-Class Dynamics for Domain Adaptive Object Detection
von: Kennerley, Mikhail, et al.
Veröffentlicht: (2024)
von: Kennerley, Mikhail, et al.
Veröffentlicht: (2024)
HorGait: A Hybrid Model for Accurate Gait Recognition in LiDAR Point Cloud Planar Projections
von: Hao, Jiaxing, et al.
Veröffentlicht: (2024)
von: Hao, Jiaxing, et al.
Veröffentlicht: (2024)
Block-level Text Spotting with LLMs
von: Bannur, Ganesh, et al.
Veröffentlicht: (2024)
von: Bannur, Ganesh, et al.
Veröffentlicht: (2024)
RRCANet: Recurrent Reusable-Convolution Attention Network for Infrared Small Target Detection
von: Liu, Yongxian, et al.
Veröffentlicht: (2025)
von: Liu, Yongxian, et al.
Veröffentlicht: (2025)
Accelerating SfM-based Pose Estimation with Dominating Set
von: Joseph, Joji, et al.
Veröffentlicht: (2025)
von: Joseph, Joji, et al.
Veröffentlicht: (2025)
DepthVanish: Optimizing Adversarial Interval Structures for Stereo-Depth-Invisible Patches
von: Xing, Yun, et al.
Veröffentlicht: (2025)
von: Xing, Yun, et al.
Veröffentlicht: (2025)
Gradually Vanishing Gap in Prototypical Network for Unsupervised Domain Adaptation
von: Wang, Shanshan, et al.
Veröffentlicht: (2024)
von: Wang, Shanshan, et al.
Veröffentlicht: (2024)
Take A Shortcut Back: Mitigating the Gradient Vanishing for Training Spiking Neural Networks
von: Guo, Yufei, et al.
Veröffentlicht: (2024)
von: Guo, Yufei, et al.
Veröffentlicht: (2024)
A Timely Survey on Vision Transformer for Deepfake Detection
von: Wang, Zhikan, et al.
Veröffentlicht: (2024)
von: Wang, Zhikan, et al.
Veröffentlicht: (2024)
Adaptive Multi Scale Document Binarisation Using Vision Mamba
von: Azfar, Mohd., et al.
Veröffentlicht: (2024)
von: Azfar, Mohd., et al.
Veröffentlicht: (2024)
Language-Guided Face Animation by Recurrent StyleGAN-based Generator
von: Hang, Tiankai, et al.
Veröffentlicht: (2022)
von: Hang, Tiankai, et al.
Veröffentlicht: (2022)
Vanish into Thin Air: Cross-prompt Universal Adversarial Attacks for SAM2
von: Zhou, Ziqi, et al.
Veröffentlicht: (2025)
von: Zhou, Ziqi, et al.
Veröffentlicht: (2025)
REPS: Reconstruction-based Point Cloud Sampling
von: Zhang, Guoqing, et al.
Veröffentlicht: (2024)
von: Zhang, Guoqing, et al.
Veröffentlicht: (2024)
Sparse Convolutional Recurrent Learning for Efficient Event-based Neuromorphic Object Detection
von: Wang, Shenqi, et al.
Veröffentlicht: (2025)
von: Wang, Shenqi, et al.
Veröffentlicht: (2025)
Image Generation from Image Captioning -- Invertible Approach
von: Menon, Nandakishore S, et al.
Veröffentlicht: (2024)
von: Menon, Nandakishore S, et al.
Veröffentlicht: (2024)
APC2Mesh: Bridging the gap from occluded building façades to full 3D models
von: Akwensi, Perpetual Hope, et al.
Veröffentlicht: (2024)
von: Akwensi, Perpetual Hope, et al.
Veröffentlicht: (2024)
Fake It To Make It: Virtual Multiviews to Enhance Monocular Indoor Semantic Scene Completion
von: Selvakumar, Anith, et al.
Veröffentlicht: (2025)
von: Selvakumar, Anith, et al.
Veröffentlicht: (2025)
A Recurrent YOLOv8-based framework for Event-Based Object Detection
von: Silva, Diego A., et al.
Veröffentlicht: (2024)
von: Silva, Diego A., et al.
Veröffentlicht: (2024)
Unsupervised Video Highlight Detection by Learning from Audio and Visual Recurrence
von: Islam, Zahidul, et al.
Veröffentlicht: (2024)
von: Islam, Zahidul, et al.
Veröffentlicht: (2024)
TERDNet: Transformer Encoder-Recurrent Decoder Network for Scene Change Detection
von: Yoon, Jiae, et al.
Veröffentlicht: (2026)
von: Yoon, Jiae, et al.
Veröffentlicht: (2026)
Gradient-Driven 3D Segmentation and Affordance Transfer in Gaussian Splatting Using 2D Masks
von: Joseph, Joji, et al.
Veröffentlicht: (2024)
von: Joseph, Joji, et al.
Veröffentlicht: (2024)
A Real-Time Defense Against Object Vanishing Adversarial Patch Attacks for Object Detection in Autonomous Vehicles
von: Mu, Jaden
Veröffentlicht: (2024)
von: Mu, Jaden
Veröffentlicht: (2024)
A Simple Recipe for Contrastively Pre-training Video-First Encoders Beyond 16 Frames
von: Papalampidi, Pinelopi, et al.
Veröffentlicht: (2023)
von: Papalampidi, Pinelopi, et al.
Veröffentlicht: (2023)
Where Do Vision-Language Models Fail? World Scale Analysis for Image Geolocalization
von: Bharadwaj, Siddhant, et al.
Veröffentlicht: (2026)
von: Bharadwaj, Siddhant, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
FootFormer: Estimating Stability from Visual Input
von: Kraiger, Keaton, et al.
Veröffentlicht: (2025) -
Convex Relaxation for Robust Vanishing Point Estimation in Manhattan World
von: Liao, Bangyan, et al.
Veröffentlicht: (2025) -
Vanishing-Point-Guided Video Semantic Segmentation of Driving Scenes
von: Guo, Diandian, et al.
Veröffentlicht: (2024) -
TAPNext: Tracking Any Point (TAP) as Next Token Prediction
von: Zholus, Artem, et al.
Veröffentlicht: (2025) -
VPOcc: Exploiting Vanishing Point for 3D Semantic Occupancy Prediction
von: Kim, Junsu, et al.
Veröffentlicht: (2024)