From Ideal to Real: Unified and Data-Efficient Dense Prediction for Real-World Scenarios
Fuente:
arXiv
Saved in:
| Main Authors: | Xia, Changliang, Jia, Chengyou, Dang, Zhuohang, Luo, Minnan, Li, Zhihui, Chang, Xiaojun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
$\mathrm{D}^\mathrm{3}$-Predictor: Noise-Free Deterministic Diffusion for Dense Prediction
by: Xia, Changliang, et al.
Published: (2025)
by: Xia, Changliang, et al.
Published: (2025)
ChatGen: Automatic Text-to-Image Generation From FreeStyle Chatting
by: Jia, Chengyou, et al.
Published: (2024)
by: Jia, Chengyou, et al.
Published: (2024)
PSDiff: Diffusion Model for Person Search with Iterative and Collaborative Refinement
by: Jia, Chengyou, et al.
Published: (2023)
by: Jia, Chengyou, et al.
Published: (2023)
PaCo-RL: Advancing Reinforcement Learning for Consistent Image Generation with Pairwise Reward Modeling
by: Ping, Bowen, et al.
Published: (2025)
by: Ping, Bowen, et al.
Published: (2025)
Multi-Modal Dataset Distillation in the Wild
by: Dang, Zhuohang, et al.
Published: (2025)
by: Dang, Zhuohang, et al.
Published: (2025)
Disentangled Representation Learning with Transmitted Information Bottleneck
by: Dang, Zhuohang, et al.
Published: (2023)
by: Dang, Zhuohang, et al.
Published: (2023)
SSMG: Spatial-Semantic Map Guided Diffusion Model for Free-form Layout-to-Image Generation
by: Jia, Chengyou, et al.
Published: (2023)
by: Jia, Chengyou, et al.
Published: (2023)
Why Settle for One? Text-to-ImageSet Generation and Evaluation
by: Jia, Chengyou, et al.
Published: (2025)
by: Jia, Chengyou, et al.
Published: (2025)
Disentangled Noisy Correspondence Learning
by: Dang, Zhuohang, et al.
Published: (2024)
by: Dang, Zhuohang, et al.
Published: (2024)
Flow-Factory: A Unified Framework for Reinforcement Learning in Flow-Matching Models
by: Ping, Bowen, et al.
Published: (2026)
by: Ping, Bowen, et al.
Published: (2026)
Beyond Dense Futures: World Models as Structured Planners for Robotic Manipulation
by: Jin, Minghao, et al.
Published: (2026)
by: Jin, Minghao, et al.
Published: (2026)
FDDet: Achieving Data-Efficient Food Defect Detection Under Real-World Scenarios
by: Xu, Ruihao, et al.
Published: (2026)
by: Xu, Ruihao, et al.
Published: (2026)
Remote Photoplethysmography in Real-World and Extreme Lighting Scenarios
by: Shao, Hang, et al.
Published: (2025)
by: Shao, Hang, et al.
Published: (2025)
DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World
by: Li, Xiangtai, et al.
Published: (2025)
by: Li, Xiangtai, et al.
Published: (2025)
Efficient Backdoor Attacks for Deep Neural Networks in Real-world Scenarios
by: Li, Ziqiang, et al.
Published: (2023)
by: Li, Ziqiang, et al.
Published: (2023)
CoFFT: Chain of Foresight-Focus Thought for Visual Language Models
by: Zhang, Xinyu, et al.
Published: (2025)
by: Zhang, Xinyu, et al.
Published: (2025)
Zero-Shot Head Swapping in Real-World Scenarios
by: Kang, Taewoong, et al.
Published: (2025)
by: Kang, Taewoong, et al.
Published: (2025)
Bridging the Gap Between Ideal and Real-world Evaluation: Benchmarking AI-Generated Image Detection in Challenging Scenarios
by: Li, Chunxiao, et al.
Published: (2025)
by: Li, Chunxiao, et al.
Published: (2025)
Exploring Efficient Asymmetric Blind-Spots for Self-Supervised Denoising in Real-World Scenarios
by: Chen, Shiyan, et al.
Published: (2023)
by: Chen, Shiyan, et al.
Published: (2023)
MFFI: Multi-Dimensional Face Forgery Image Dataset for Real-World Scenarios
by: Miao, Changtao, et al.
Published: (2025)
by: Miao, Changtao, et al.
Published: (2025)
MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?
by: Zhang, Yi-Fan, et al.
Published: (2024)
by: Zhang, Yi-Fan, et al.
Published: (2024)
From Ideal to Real: Stable Video Object Removal under Imperfect Conditions
by: Hu, Jiagao, et al.
Published: (2026)
by: Hu, Jiagao, et al.
Published: (2026)
DDL: A Large-Scale Datasets for Deepfake Detection and Localization in Diversified Real-World Scenarios
by: Miao, Changtao, et al.
Published: (2025)
by: Miao, Changtao, et al.
Published: (2025)
Unified Dense Prediction of Video Diffusion
by: Yang, Lehan, et al.
Published: (2025)
by: Yang, Lehan, et al.
Published: (2025)
Learning Object-Centric Representations Based on Slots in Real World Scenarios
by: Akan, Adil Kaan
Published: (2025)
by: Akan, Adil Kaan
Published: (2025)
Your Super Resolution Model is not Enough for Tackling Real-World Scenarios
by: Yoon, Dongsik, et al.
Published: (2025)
by: Yoon, Dongsik, et al.
Published: (2025)
CAT: A Conditional Adaptation Tailor for Efficient and Effective Instance-Specific Pansharpening on Real-World Data
by: Xin, Tianyu, et al.
Published: (2025)
by: Xin, Tianyu, et al.
Published: (2025)
Chameleon: A Data-Efficient Generalist for Dense Visual Prediction in the Wild
by: Kim, Donggyun, et al.
Published: (2024)
by: Kim, Donggyun, et al.
Published: (2024)
RGBT-Ground Benchmark: Visual Grounding Beyond RGB in Complex Real-World Scenarios
by: Zhao, Tianyi, et al.
Published: (2025)
by: Zhao, Tianyi, et al.
Published: (2025)
From Controlled Scenarios to Real-World: Cross-Domain Degradation Pattern Matching for All-in-One Image Restoration
by: Fan, Junyu, et al.
Published: (2025)
by: Fan, Junyu, et al.
Published: (2025)
A Multilevel Strategy to Improve People Tracking in a Real-World Scenario
by: de Oliveira, Cristiano B., et al.
Published: (2024)
by: de Oliveira, Cristiano B., et al.
Published: (2024)
DVD: A Comprehensive Dataset for Advancing Violence Detection in Real-World Scenarios
by: Kollias, Dimitrios, et al.
Published: (2025)
by: Kollias, Dimitrios, et al.
Published: (2025)
Toward Generalizable Deblurring: Leveraging Massive Blur Priors with Linear Attention for Real-World Scenarios
by: Gao, Yuanting, et al.
Published: (2026)
by: Gao, Yuanting, et al.
Published: (2026)
Towards Natural Image Matting in the Wild via Real-Scenario Prior
by: Xia, Ruihao, et al.
Published: (2024)
by: Xia, Ruihao, et al.
Published: (2024)
BADAS: Context Aware Collision Prediction Using Real-World Dashcam Data
by: Goldshmidt, Roni, et al.
Published: (2025)
by: Goldshmidt, Roni, et al.
Published: (2025)
MDPBench: A Benchmark for Multilingual Document Parsing in Real-World Scenarios
by: Li, Zhang, et al.
Published: (2026)
by: Li, Zhang, et al.
Published: (2026)
CBVS: A Large-Scale Chinese Image-Text Benchmark for Real-World Short Video Search Scenarios
by: Qiao, Xiangshuo, et al.
Published: (2024)
by: Qiao, Xiangshuo, et al.
Published: (2024)
Real-time Multi-view Omnidirectional Depth Estimation for Real Scenarios based on Teacher-Student Learning with Unlabeled Data
by: Li, Ming, et al.
Published: (2024)
by: Li, Ming, et al.
Published: (2024)
PhoStream: Benchmarking Real-World Streaming for Omnimodal Assistants in Mobile Scenarios
by: Lu, Xudong, et al.
Published: (2026)
by: Lu, Xudong, et al.
Published: (2026)
From Virtual Games to Real-World Play
by: Sun, Wenqiang, et al.
Published: (2025)
by: Sun, Wenqiang, et al.
Published: (2025)
Similar Items
-
$\mathrm{D}^\mathrm{3}$-Predictor: Noise-Free Deterministic Diffusion for Dense Prediction
by: Xia, Changliang, et al.
Published: (2025) -
ChatGen: Automatic Text-to-Image Generation From FreeStyle Chatting
by: Jia, Chengyou, et al.
Published: (2024) -
PSDiff: Diffusion Model for Person Search with Iterative and Collaborative Refinement
by: Jia, Chengyou, et al.
Published: (2023) -
PaCo-RL: Advancing Reinforcement Learning for Consistent Image Generation with Pairwise Reward Modeling
by: Ping, Bowen, et al.
Published: (2025) -
Multi-Modal Dataset Distillation in the Wild
by: Dang, Zhuohang, et al.
Published: (2025)