Saved in:
| Main Authors: | Lu, Wanglong, Wang, Jikai, Wang, Tao, Zhang, Kaihao, Jiang, Xianta, Zhao, Hanli |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2412.21042 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing
by: Lu, Wanglong, et al.
Published: (2024)
by: Lu, Wanglong, et al.
Published: (2024)
TextDoctor: Unified Document Image Inpainting via Patch Pyramid Diffusion Models
by: Lu, Wanglong, et al.
Published: (2025)
by: Lu, Wanglong, et al.
Published: (2025)
Relightable and Dynamic Gaussian Avatar Reconstruction from Monocular Video
by: Choi, Seonghwa, et al.
Published: (2025)
by: Choi, Seonghwa, et al.
Published: (2025)
Do Inpainting Yourself: Generative Facial Inpainting Guided by Exemplars
by: Lu, Wanglong, et al.
Published: (2022)
by: Lu, Wanglong, et al.
Published: (2022)
Learning Joint Denoising, Demosaicing, and Compression from the Raw Natural Image Noise Dataset
by: Brummer, Benoit, et al.
Published: (2025)
by: Brummer, Benoit, et al.
Published: (2025)
MdaIF: Robust One-Stop Multi-Degradation-Aware Image Fusion with Language-Driven Semantics
by: Li, Jing, et al.
Published: (2025)
by: Li, Jing, et al.
Published: (2025)
Event-ECC: Asynchronous Tracking of Events with Continuous Optimization
by: Zafeiri, Maria, et al.
Published: (2024)
by: Zafeiri, Maria, et al.
Published: (2024)
BG-YOLO: A Bidirectional-Guided Method for Underwater Object Detection
by: Zhang, Jian, et al.
Published: (2024)
by: Zhang, Jian, et al.
Published: (2024)
ForensicFormer: Hierarchical Multi-Scale Reasoning for Cross-Domain Image Forgery Detection
by: Samson, Hema Hariharan
Published: (2026)
by: Samson, Hema Hariharan
Published: (2026)
Improving Visual Object Tracking through Visual Prompting
by: Chen, Shih-Fang, et al.
Published: (2024)
by: Chen, Shih-Fang, et al.
Published: (2024)
The Impact of Image Resolution on Face Detection: A Comparative Analysis of MTCNN, YOLOv XI and YOLOv XII models
by: Ömercikoğlu, Ahmet Can, et al.
Published: (2025)
by: Ömercikoğlu, Ahmet Can, et al.
Published: (2025)
FLD+: Data-efficient Evaluation Metric for Generative Models
by: Jeevan, Pranav, et al.
Published: (2024)
by: Jeevan, Pranav, et al.
Published: (2024)
WaveMixSR-V2: Enhancing Super-resolution with Higher Efficiency
by: Jeevan, Pranav, et al.
Published: (2024)
by: Jeevan, Pranav, et al.
Published: (2024)
Normalizing Flow-Based Metric for Image Generation
by: Jeevan, Pranav, et al.
Published: (2024)
by: Jeevan, Pranav, et al.
Published: (2024)
Learning to Expand Images for Efficient Visual Autoregressive Modeling
by: Yang, Ruiqing, et al.
Published: (2025)
by: Yang, Ruiqing, et al.
Published: (2025)
Motion-Based Sign Language Video Summarization using Curvature and Torsion
by: Sartinas, Evangelos G., et al.
Published: (2023)
by: Sartinas, Evangelos G., et al.
Published: (2023)
Haze-Aware Attention Network for Single-Image Dehazing
by: Tong, Lihan, et al.
Published: (2024)
by: Tong, Lihan, et al.
Published: (2024)
Eleven Primitives and Three Gates: The Universal Structure of Computational Imaging
by: Yang, Chengshuai, et al.
Published: (2026)
by: Yang, Chengshuai, et al.
Published: (2026)
Neural Image Compression Using Masked Sparse Visual Representation
by: Jiang, Wei, et al.
Published: (2023)
by: Jiang, Wei, et al.
Published: (2023)
Few-Class Arena: A Benchmark for Efficient Selection of Vision Models and Dataset Difficulty Measurement
by: Cao, Bryan Bo, et al.
Published: (2024)
by: Cao, Bryan Bo, et al.
Published: (2024)
Detailed Evaluation of Modern Machine Learning Approaches for Optic Plastics Sorting
by: Maheshkar, Vaishali, et al.
Published: (2025)
by: Maheshkar, Vaishali, et al.
Published: (2025)
Evaluation Metric for Quality Control and Generative Models in Histopathology Images
by: Jeevan, Pranav, et al.
Published: (2024)
by: Jeevan, Pranav, et al.
Published: (2024)
BFORE: Butterfly-Firefly Optimized Retinex Enhancement for Low-Light Image Quality Improvement
by: Cherif, Ahmed
Published: (2026)
by: Cherif, Ahmed
Published: (2026)
Model Agnostic Defense against Adversarial Patch Attacks on Object Detection in Unmanned Aerial Vehicles
by: Pathak, Saurabh, et al.
Published: (2024)
by: Pathak, Saurabh, et al.
Published: (2024)
Revealing an Unattractivity Bias in Mental Reconstruction of Occluded Faces using Generative Image Models
by: Riedmann, Frederik, et al.
Published: (2024)
by: Riedmann, Frederik, et al.
Published: (2024)
CADE 2.5 - ZeResFDG: Frequency-Decoupled, Rescaled and Zero-Projected Guidance for SD/SDXL Latent Diffusion Models
by: Rychkovskiy, Denis
Published: (2025)
by: Rychkovskiy, Denis
Published: (2025)
Neural Fields for 3D Tracking of Anatomy and Surgical Instruments in Monocular Laparoscopic Video Clips
by: Gerats, Beerend G. A., et al.
Published: (2024)
by: Gerats, Beerend G. A., et al.
Published: (2024)
StatsMerging: Statistics-Guided Model Merging via Task-Specific Teacher Distillation
by: Merugu, Ranjith, et al.
Published: (2025)
by: Merugu, Ranjith, et al.
Published: (2025)
Enhancing rice leaf images: An overview of image denoising techniques
by: Chutia, Rupjyoti, et al.
Published: (2025)
by: Chutia, Rupjyoti, et al.
Published: (2025)
i-DEQ: A stable inertial deep equilibrium model for image restoration
by: Clerc, Antonin, et al.
Published: (2026)
by: Clerc, Antonin, et al.
Published: (2026)
Under-Canopy Terrain Reconstruction in Dense Forests Using RGB Imaging and Neural 3D Reconstruction
by: Sheffer, Refael, et al.
Published: (2026)
by: Sheffer, Refael, et al.
Published: (2026)
Visual-Instructed Degradation Diffusion for All-in-One Image Restoration
by: Luo, Wenyang, et al.
Published: (2025)
by: Luo, Wenyang, et al.
Published: (2025)
Mapping Tomato Cropping Systems in California Using AlphaEarth Geospatial Embeddings and Deep Learning Analysis
by: Narimani, Mohammadreza, et al.
Published: (2026)
by: Narimani, Mohammadreza, et al.
Published: (2026)
Lightweight Complementary-Cue Fusion for Robust Video Face Forgery Detection
by: Baek, Sunghwan, et al.
Published: (2026)
by: Baek, Sunghwan, et al.
Published: (2026)
A unified Benchmark for Multi-Frame Image Restoration under Severe Refractive Warping
by: Shugaev, Maxim V., et al.
Published: (2026)
by: Shugaev, Maxim V., et al.
Published: (2026)
HieraEdgeNet: A Multi-Scale Edge-Enhanced Framework for Automated Pollen Recognition
by: Long, Yuchong, et al.
Published: (2025)
by: Long, Yuchong, et al.
Published: (2025)
Beyond RGB: Leveraging Vision Transformers for Thermal Weapon Segmentation
by: Kambhatla, Akhila, et al.
Published: (2025)
by: Kambhatla, Akhila, et al.
Published: (2025)
Digital analysis of early color photographs taken using regular color screen processes
by: Hubička, Jan, et al.
Published: (2023)
by: Hubička, Jan, et al.
Published: (2023)
N-DriverMotion: Driver motion learning and prediction using an event-based camera and directly trained spiking neural networks on Loihi 2
by: Chung, Hyo Jong, et al.
Published: (2024)
by: Chung, Hyo Jong, et al.
Published: (2024)
Out-of-Sight Embodied Agents: Multimodal Tracking, Sensor Fusion, and Trajectory Forecasting
by: Zhang, Haichao, et al.
Published: (2025)
by: Zhang, Haichao, et al.
Published: (2025)
Similar Items
-
FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing
by: Lu, Wanglong, et al.
Published: (2024) -
TextDoctor: Unified Document Image Inpainting via Patch Pyramid Diffusion Models
by: Lu, Wanglong, et al.
Published: (2025) -
Relightable and Dynamic Gaussian Avatar Reconstruction from Monocular Video
by: Choi, Seonghwa, et al.
Published: (2025) -
Do Inpainting Yourself: Generative Facial Inpainting Guided by Exemplars
by: Lu, Wanglong, et al.
Published: (2022) -
Learning Joint Denoising, Demosaicing, and Compression from the Raw Natural Image Noise Dataset
by: Brummer, Benoit, et al.
Published: (2025)