Saved in:
| Main Author: | Greene, Michelle R. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2606.02481 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NYC-Event-VPR: A Large-Scale High-Resolution Event-Based Visual Place Recognition Dataset in Dense Urban Environments
by: Pan, Taiyi, et al.
Published: (2024)
by: Pan, Taiyi, et al.
Published: (2024)
RAW-Domain Degradation Models for Realistic Smartphone Super-Resolution
by: Mosleh, Ali, et al.
Published: (2026)
by: Mosleh, Ali, et al.
Published: (2026)
RAW-Diffusion: RGB-Guided Diffusion Models for High-Fidelity RAW Image Generation
by: Reinders, Christoph, et al.
Published: (2024)
by: Reinders, Christoph, et al.
Published: (2024)
The Limits of Learning from Pictures and Text: Vision-Language Models and Embodied Scene Understanding
by: Rosenberg, Gillian, et al.
Published: (2026)
by: Rosenberg, Gillian, et al.
Published: (2026)
HUE Dataset: High-Resolution Event and Frame Sequences for Low-Light Vision
by: Ercan, Burak, et al.
Published: (2024)
by: Ercan, Burak, et al.
Published: (2024)
ReXInTheWild: A Unified Benchmark for Medical Photograph Understanding
by: Banerjee, Oishi, et al.
Published: (2026)
by: Banerjee, Oishi, et al.
Published: (2026)
WildCross: A Cross-Modal Large Scale Benchmark for Place Recognition and Metric Depth Estimation in Natural Environments
by: Knights, Joshua, et al.
Published: (2026)
by: Knights, Joshua, et al.
Published: (2026)
LENVIZ: A High-Resolution Low-Exposure Night Vision Benchmark Dataset
by: Aithal, Manjushree, et al.
Published: (2025)
by: Aithal, Manjushree, et al.
Published: (2025)
RAW-Adapter: Adapting Pre-trained Visual Model to Camera RAW Images and A Benchmark
by: Cui, Ziteng, et al.
Published: (2025)
by: Cui, Ziteng, et al.
Published: (2025)
RAW-Adapter: Adapting Pre-trained Visual Model to Camera RAW Images
by: Cui, Ziteng, et al.
Published: (2024)
by: Cui, Ziteng, et al.
Published: (2024)
NTIRE 2025 Challenge on RAW Image Restoration and Super-Resolution
by: Conde, Marcos V., et al.
Published: (2025)
by: Conde, Marcos V., et al.
Published: (2025)
VINS-120K: Ultra High-Resolution Image Editing with A Large-Scale Dataset
by: Chen, Zhizhou, et al.
Published: (2026)
by: Chen, Zhizhou, et al.
Published: (2026)
WildFake: A Large-scale Challenging Dataset for AI-Generated Images Detection
by: Hong, Yan, et al.
Published: (2024)
by: Hong, Yan, et al.
Published: (2024)
The Photographer Eye: Teaching Multimodal Large Language Models to Understand Image Aesthetics like Photographers
by: Qi, Daiqing, et al.
Published: (2025)
by: Qi, Daiqing, et al.
Published: (2025)
RAW-Flow: Advancing RGB-to-RAW Image Reconstruction with Deterministic Latent Flow Matching
by: Liu, Zhen, et al.
Published: (2026)
by: Liu, Zhen, et al.
Published: (2026)
Edit-aware RAW Reconstruction
by: Punnappurath, Abhijith, et al.
Published: (2025)
by: Punnappurath, Abhijith, et al.
Published: (2025)
Deep RAW Image Super-Resolution. A NTIRE 2024 Challenge Survey
by: Conde, Marcos V., et al.
Published: (2024)
by: Conde, Marcos V., et al.
Published: (2024)
ATRNet-STAR: A Large Dataset and Benchmark Towards Remote Sensing Object Recognition in the Wild
by: Liu, Yongxiang, et al.
Published: (2025)
by: Liu, Yongxiang, et al.
Published: (2025)
RiverScope: High-Resolution River Masking Dataset
by: Daroya, Rangel, et al.
Published: (2025)
by: Daroya, Rangel, et al.
Published: (2025)
Brighteye: Glaucoma Screening with Color Fundus Photographs based on Vision Transformer
by: Lin, Hui, et al.
Published: (2024)
by: Lin, Hui, et al.
Published: (2024)
Large-Scale Dataset and Benchmark for Skin Tone Classification in the Wild
by: Matias, Vitor Pereira, et al.
Published: (2026)
by: Matias, Vitor Pereira, et al.
Published: (2026)
HERO: Rethinking Visual Token Early Dropping in High-Resolution Large Vision-Language Models
by: Li, Xu, et al.
Published: (2025)
by: Li, Xu, et al.
Published: (2025)
Vision-Enhanced Large Language Models for High-Resolution Image Synthesis and Multimodal Data Interpretation
by: KV, Karthikeya
Published: (2025)
by: KV, Karthikeya
Published: (2025)
SmartWilds: Multimodal Wildlife Monitoring Dataset
by: Kline, Jenna, et al.
Published: (2025)
by: Kline, Jenna, et al.
Published: (2025)
PanAf20K: A Large Video Dataset for Wild Ape Detection and Behaviour Recognition
by: Brookes, Otto, et al.
Published: (2024)
by: Brookes, Otto, et al.
Published: (2024)
Removing Reflections from RAW Photos
by: Kee, Eric, et al.
Published: (2024)
by: Kee, Eric, et al.
Published: (2024)
BaboonLand Dataset: Tracking Primates in the Wild and Automating Behaviour Recognition from Drone Videos
by: Duporge, Isla, et al.
Published: (2024)
by: Duporge, Isla, et al.
Published: (2024)
Deblurring in the Wild: A Real-World Image Deblurring Dataset from Smartphone High-Speed Videos
by: Mahmud, Syed Mumtahin, et al.
Published: (2025)
by: Mahmud, Syed Mumtahin, et al.
Published: (2025)
ChildPlay-Hand: A Dataset of Hand Manipulations in the Wild
by: Farkhondeh, Arya, et al.
Published: (2024)
by: Farkhondeh, Arya, et al.
Published: (2024)
WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences
by: Lu, Yujie, et al.
Published: (2024)
by: Lu, Yujie, et al.
Published: (2024)
HiRes-LLaVA: Restoring Fragmentation Input in High-Resolution Large Vision-Language Models
by: Huang, Runhui, et al.
Published: (2024)
by: Huang, Runhui, et al.
Published: (2024)
Global Compression Commander: Plug-and-Play Inference Acceleration for High-Resolution Large Vision-Language Models
by: Liu, Xuyang, et al.
Published: (2025)
by: Liu, Xuyang, et al.
Published: (2025)
LaVPR: Benchmarking Language and Vision for Place Recognition
by: Idan, Ofer, et al.
Published: (2026)
by: Idan, Ofer, et al.
Published: (2026)
Unveiling Hidden Details: A RAW Data-Enhanced Paradigm for Real-World Super-Resolution
by: Peng, Long, et al.
Published: (2024)
by: Peng, Long, et al.
Published: (2024)
Carousel: A High-Resolution Dataset for Multi-Target Automatic Image Cropping
by: Loya, Rafe, et al.
Published: (2025)
by: Loya, Rafe, et al.
Published: (2025)
Towards RAW Object Detection in Diverse Conditions
by: Li, Zhong-Yu, et al.
Published: (2024)
by: Li, Zhong-Yu, et al.
Published: (2024)
FlexAttention for Efficient High-Resolution Vision-Language Models
by: Li, Junyan, et al.
Published: (2024)
by: Li, Junyan, et al.
Published: (2024)
TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On
by: Shao, Dingbao, et al.
Published: (2026)
by: Shao, Dingbao, et al.
Published: (2026)
WildWorld: A Large-Scale Dataset for Dynamic World Modeling with Actions and Explicit State toward Generative ARPG
by: Li, Zhen, et al.
Published: (2026)
by: Li, Zhen, et al.
Published: (2026)
Caltech Aerial RGB-Thermal Dataset in the Wild
by: Lee, Connor, et al.
Published: (2024)
by: Lee, Connor, et al.
Published: (2024)
Similar Items
-
NYC-Event-VPR: A Large-Scale High-Resolution Event-Based Visual Place Recognition Dataset in Dense Urban Environments
by: Pan, Taiyi, et al.
Published: (2024) -
RAW-Domain Degradation Models for Realistic Smartphone Super-Resolution
by: Mosleh, Ali, et al.
Published: (2026) -
RAW-Diffusion: RGB-Guided Diffusion Models for High-Fidelity RAW Image Generation
by: Reinders, Christoph, et al.
Published: (2024) -
The Limits of Learning from Pictures and Text: Vision-Language Models and Embodied Scene Understanding
by: Rosenberg, Gillian, et al.
Published: (2026) -
HUE Dataset: High-Resolution Event and Frame Sequences for Low-Light Vision
by: Ercan, Burak, et al.
Published: (2024)