Reading a Ruler in the Wild
Fuente:
arXiv
Salvato in:
| Autori principali: | Pan, Yimu, Mehta, Manas, Sincerbeaux, Gwen, Goldstein, Jeffery A., Gernand, Alison D., Wang, James Z. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
S2S2: Semantic Stacking for Robust Semantic Segmentation in Medical Imaging
di: Pan, Yimu, et al.
Pubblicazione: (2024)
di: Pan, Yimu, et al.
Pubblicazione: (2024)
VLCD: Vision-Language Contrastive Distillation for Accurate and Efficient Automatic Placenta Analysis
di: Mehta, Manas, et al.
Pubblicazione: (2025)
di: Mehta, Manas, et al.
Pubblicazione: (2025)
Reading Recognition in the Wild
di: Yang, Charig, et al.
Pubblicazione: (2025)
di: Yang, Charig, et al.
Pubblicazione: (2025)
studentSplat: Your Student Model Learns Single-view 3D Gaussian Splatting
di: Pan, Yimu, et al.
Pubblicazione: (2026)
di: Pan, Yimu, et al.
Pubblicazione: (2026)
GDKVM: Echocardiography Video Segmentation via Spatiotemporal Key-Value Memory with Gated Delta Rule
di: Wang, Rui, et al.
Pubblicazione: (2025)
di: Wang, Rui, et al.
Pubblicazione: (2025)
OV-SCAN: Semantically Consistent Alignment for Novel Object Discovery in Open-Vocabulary 3D Object Detection
di: Chow, Adrian, et al.
Pubblicazione: (2025)
di: Chow, Adrian, et al.
Pubblicazione: (2025)
Machine learning identification of maternal inflammatory response and histologic choroamnionitis from placental membrane whole slide images
di: Sharma, Abhishek, et al.
Pubblicazione: (2024)
di: Sharma, Abhishek, et al.
Pubblicazione: (2024)
LEO-MINI: An Efficient Multimodal Large Language Model using Conditional Token Reduction and Mixture of Multi-Modal Experts
di: Wang, Yimu, et al.
Pubblicazione: (2025)
di: Wang, Yimu, et al.
Pubblicazione: (2025)
SOAR: Self-Occluded Avatar Recovery from a Single Video In the Wild
di: Pan, Zhuoyang, et al.
Pubblicazione: (2024)
di: Pan, Zhuoyang, et al.
Pubblicazione: (2024)
Visual Text Generation in the Wild
di: Zhu, Yuanzhi, et al.
Pubblicazione: (2024)
di: Zhu, Yuanzhi, et al.
Pubblicazione: (2024)
Sparse-View 3D Gaussian Splatting in the Wild
di: Park, Wongi, et al.
Pubblicazione: (2026)
di: Park, Wongi, et al.
Pubblicazione: (2026)
WildGaussians: 3D Gaussian Splatting in the Wild
di: Kulhanek, Jonas, et al.
Pubblicazione: (2024)
di: Kulhanek, Jonas, et al.
Pubblicazione: (2024)
Zero-Shot Monocular Scene Flow Estimation in the Wild
di: Liang, Yiqing, et al.
Pubblicazione: (2025)
di: Liang, Yiqing, et al.
Pubblicazione: (2025)
WildDoc: How Far Are We from Achieving Comprehensive and Robust Document Understanding in the Wild?
di: Wang, An-Lan, et al.
Pubblicazione: (2025)
di: Wang, An-Lan, et al.
Pubblicazione: (2025)
HAWAII: Hierarchical Visual Knowledge Transfer for Efficient Vision-Language Models
di: Wang, Yimu, et al.
Pubblicazione: (2025)
di: Wang, Yimu, et al.
Pubblicazione: (2025)
WildVidFit: Video Virtual Try-On in the Wild via Image-Based Controlled Diffusion Models
di: He, Zijian, et al.
Pubblicazione: (2024)
di: He, Zijian, et al.
Pubblicazione: (2024)
WildPose: A Unified Framework for Robust Pose Estimation in the Wild
di: Zheng, Jianhao, et al.
Pubblicazione: (2026)
di: Zheng, Jianhao, et al.
Pubblicazione: (2026)
EmbodMocap: In-the-Wild 4D Human-Scene Reconstruction for Embodied Agents
di: Wang, Wenjia, et al.
Pubblicazione: (2026)
di: Wang, Wenjia, et al.
Pubblicazione: (2026)
Scene Grounding In the Wild
di: Cohen, Tamir, et al.
Pubblicazione: (2026)
di: Cohen, Tamir, et al.
Pubblicazione: (2026)
Forecasting Motion in the Wild
di: Thakkar, Neerja, et al.
Pubblicazione: (2026)
di: Thakkar, Neerja, et al.
Pubblicazione: (2026)
Towards Natural Image Matting in the Wild via Real-Scenario Prior
di: Xia, Ruihao, et al.
Pubblicazione: (2024)
di: Xia, Ruihao, et al.
Pubblicazione: (2024)
DressWild: Feed-Forward Pose-Agnostic Garment Sewing Pattern Generation from In-the-Wild Images
di: Tao, Zeng, et al.
Pubblicazione: (2026)
di: Tao, Zeng, et al.
Pubblicazione: (2026)
WildCAT3D: Appearance-Aware Multi-View Diffusion in the Wild
di: Alper, Morris, et al.
Pubblicazione: (2025)
di: Alper, Morris, et al.
Pubblicazione: (2025)
WildDet3D: Scaling Promptable 3D Detection in the Wild
di: Huang, Weikai, et al.
Pubblicazione: (2026)
di: Huang, Weikai, et al.
Pubblicazione: (2026)
LDFaceNet: Latent Diffusion-based Network for High-Fidelity Deepfake Generation
di: Mehta, Dwij, et al.
Pubblicazione: (2024)
di: Mehta, Dwij, et al.
Pubblicazione: (2024)
Edicho: Consistent Image Editing in the Wild
di: Bai, Qingyan, et al.
Pubblicazione: (2024)
di: Bai, Qingyan, et al.
Pubblicazione: (2024)
Mitigating the Modality Gap: Few-Shot Out-of-Distribution Detection with Multi-modal Prototypes and Image Bias Estimation
di: Wang, Yimu, et al.
Pubblicazione: (2025)
di: Wang, Yimu, et al.
Pubblicazione: (2025)
Survey of Video Diffusion Models: Foundations, Implementations, and Applications
di: Wang, Yimu, et al.
Pubblicazione: (2025)
di: Wang, Yimu, et al.
Pubblicazione: (2025)
Online Adaptation for Implicit Object Tracking and Shape Reconstruction in the Wild
di: Ye, Jianglong, et al.
Pubblicazione: (2021)
di: Ye, Jianglong, et al.
Pubblicazione: (2021)
WildOS: Open-Vocabulary Object Search in the Wild
di: Shah, Hardik, et al.
Pubblicazione: (2026)
di: Shah, Hardik, et al.
Pubblicazione: (2026)
Statewide Visual Geolocalization in the Wild
di: Fervers, Florian, et al.
Pubblicazione: (2024)
di: Fervers, Florian, et al.
Pubblicazione: (2024)
Extreme Rotation Estimation in the Wild
di: Bezalel, Hana, et al.
Pubblicazione: (2024)
di: Bezalel, Hana, et al.
Pubblicazione: (2024)
Font Impression Estimation in the Wild
di: Kitajima, Kazuki, et al.
Pubblicazione: (2024)
di: Kitajima, Kazuki, et al.
Pubblicazione: (2024)
LiWi: Layering in the Wild
di: He, Yu, et al.
Pubblicazione: (2026)
di: He, Yu, et al.
Pubblicazione: (2026)
TeleEgo: Benchmarking Egocentric AI Assistants in the Wild
di: Yan, Jiaqi, et al.
Pubblicazione: (2025)
di: Yan, Jiaqi, et al.
Pubblicazione: (2025)
Gaussian in the Wild: 3D Gaussian Splatting for Unconstrained Image Collections
di: Zhang, Dongbin, et al.
Pubblicazione: (2024)
di: Zhang, Dongbin, et al.
Pubblicazione: (2024)
WildGHand: Learning Anti-Perturbation Gaussian Hand Avatars from Monocular In-the-Wild Videos
di: Li, Hanhui, et al.
Pubblicazione: (2026)
di: Li, Hanhui, et al.
Pubblicazione: (2026)
SAM2LoRA: Composite Loss-Guided, Parameter-Efficient Finetuning of SAM2 for Retinal Fundus Segmentation
di: Mandal, Sayan, et al.
Pubblicazione: (2025)
di: Mandal, Sayan, et al.
Pubblicazione: (2025)
UNIFORM: Unifying Knowledge from Large-scale and Diverse Pre-trained Models
di: Wang, Yimu, et al.
Pubblicazione: (2025)
di: Wang, Yimu, et al.
Pubblicazione: (2025)
PAWS: Perception of Articulation in the Wild at Scale from Egocentric Videos
di: Wang, Yihao, et al.
Pubblicazione: (2026)
di: Wang, Yihao, et al.
Pubblicazione: (2026)
Documenti analoghi
-
S2S2: Semantic Stacking for Robust Semantic Segmentation in Medical Imaging
di: Pan, Yimu, et al.
Pubblicazione: (2024) -
VLCD: Vision-Language Contrastive Distillation for Accurate and Efficient Automatic Placenta Analysis
di: Mehta, Manas, et al.
Pubblicazione: (2025) -
Reading Recognition in the Wild
di: Yang, Charig, et al.
Pubblicazione: (2025) -
studentSplat: Your Student Model Learns Single-view 3D Gaussian Splatting
di: Pan, Yimu, et al.
Pubblicazione: (2026) -
GDKVM: Echocardiography Video Segmentation via Spatiotemporal Key-Value Memory with Gated Delta Rule
di: Wang, Rui, et al.
Pubblicazione: (2025)