A 2-Stage Model for Vehicle Class and Orientation Detection with Photo-Realistic Image Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Youngmin, Kang, Donghwa, Baek, Hyeongboo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CR-QAT: Curriculum Relational Quantization-Aware Training for Open-Vocabulary Object Detection
by: Park, Jinyeong, et al.
Published: (2026)
by: Park, Jinyeong, et al.
Published: (2026)
Fluently Lying: Adversarial Robustness Can Be Substrate-Dependent
by: Kang, Daye, et al.
Published: (2026)
by: Kang, Daye, et al.
Published: (2026)
CF-DETR: Coarse-to-Fine Transformer for Real-Time Object Detection
by: Shin, Woojin, et al.
Published: (2025)
by: Shin, Woojin, et al.
Published: (2025)
AT-SNN: Adaptive Tokens for Vision Transformer on Spiking Neural Network
by: Kang, Donghwa, et al.
Published: (2024)
by: Kang, Donghwa, et al.
Published: (2024)
Timestep-Compressed Attack on Spiking Neural Networks through Timestep-Level Backpropagation
by: Kang, Donghwa, et al.
Published: (2025)
by: Kang, Donghwa, et al.
Published: (2025)
BankTweak: Adversarial Attack against Multi-Object Trackers by Manipulating Feature Banks
by: Shin, Woojin, et al.
Published: (2024)
by: Shin, Woojin, et al.
Published: (2024)
Real Time Scheduling Framework for Multi Object Detection via Spiking Neural Networks
by: Kang, Donghwa, et al.
Published: (2025)
by: Kang, Donghwa, et al.
Published: (2025)
STAS: Spatio-Temporal Adaptive Computation Time for Spiking Transformers
by: Kang, Donghwa, et al.
Published: (2025)
by: Kang, Donghwa, et al.
Published: (2025)
Event-based Facial Keypoint Alignment via Cross-Modal Fusion Attention and Self-Supervised Multi-Event Representation Learning
by: Kang, Donghwa, et al.
Published: (2025)
by: Kang, Donghwa, et al.
Published: (2025)
FYI: Flip Your Images for Dataset Distillation
by: Son, Byunggwan, et al.
Published: (2024)
by: Son, Byunggwan, et al.
Published: (2024)
Photo-Realistic Image Restoration in the Wild with Controlled Vision-Language Models
by: Luo, Ziwei, et al.
Published: (2024)
by: Luo, Ziwei, et al.
Published: (2024)
Reward-Instruct: A Reward-Centric Approach to Fast Photo-Realistic Image Generation
by: Luo, Yihong, et al.
Published: (2025)
by: Luo, Yihong, et al.
Published: (2025)
SDDGR: Stable Diffusion-based Deep Generative Replay for Class Incremental Object Detection
by: Kim, Junsu, et al.
Published: (2024)
by: Kim, Junsu, et al.
Published: (2024)
Scaling Up to Excellence: Practicing Model Scaling for Photo-Realistic Image Restoration In the Wild
by: Yu, Fanghua, et al.
Published: (2024)
by: Yu, Fanghua, et al.
Published: (2024)
VLM-PL: Advanced Pseudo Labeling Approach for Class Incremental Object Detection via Vision-Language Model
by: Kim, Junsu, et al.
Published: (2024)
by: Kim, Junsu, et al.
Published: (2024)
MonoWAD: Weather-Adaptive Diffusion Model for Robust Monocular 3D Object Detection
by: Oh, Youngmin, et al.
Published: (2024)
by: Oh, Youngmin, et al.
Published: (2024)
Dust to Tower: Coarse-to-Fine Photo-Realistic Scene Reconstruction from Sparse Uncalibrated Images
by: Cai, Xudong, et al.
Published: (2024)
by: Cai, Xudong, et al.
Published: (2024)
Finding NeMo: Negative-mined Mosaic Augmentation for Referring Image Segmentation
by: Ha, Seongsu, et al.
Published: (2024)
by: Ha, Seongsu, et al.
Published: (2024)
UCMNet: Uncertainty-Aware Context Memory Network for Under-Display Camera Image Restoration
by: Kim, Daehyun, et al.
Published: (2026)
by: Kim, Daehyun, et al.
Published: (2026)
Can Synthetic Images Conquer Forgetting? Beyond Unexplored Doubts in Few-Shot Class-Incremental Learning
by: Kim, Junsu, et al.
Published: (2025)
by: Kim, Junsu, et al.
Published: (2025)
Scalp Diagnostic System With Label-Free Segmentation and Training-Free Image Translation
by: Kim, Youngmin, et al.
Published: (2024)
by: Kim, Youngmin, et al.
Published: (2024)
LucidFlux: Caption-Free Photo-Realistic Image Restoration via a Large-Scale Diffusion Transformer
by: Fei, Song, et al.
Published: (2025)
by: Fei, Song, et al.
Published: (2025)
Diff-Palm: Realistic Palmprint Generation with Polynomial Creases and Intra-Class Variation Controllable Diffusion Models
by: Jin, Jianlong, et al.
Published: (2025)
by: Jin, Jianlong, et al.
Published: (2025)
Subnet-Aware Dynamic Supernet Training for Neural Architecture Search
by: Jeon, Jeimin, et al.
Published: (2025)
by: Jeon, Jeimin, et al.
Published: (2025)
Emulating Self-attention with Convolution for Efficient Image Super-Resolution
by: Lee, Dongheon, et al.
Published: (2025)
by: Lee, Dongheon, et al.
Published: (2025)
Implicit Grid Convolution for Multi-Scale Image Super-Resolution
by: Lee, Dongheon, et al.
Published: (2024)
by: Lee, Dongheon, et al.
Published: (2024)
Semantic-Aware Reconstruction Error for Detecting AI-Generated Images
by: Kang, Ju Yeon, et al.
Published: (2025)
by: Kang, Ju Yeon, et al.
Published: (2025)
H2G: Hierarchy-Aware Hyperbolic Grouping for 3D Scenes
by: Ko, ByungHa, et al.
Published: (2026)
by: Ko, ByungHa, et al.
Published: (2026)
DefectFill: Realistic Defect Generation with Inpainting Diffusion Model for Visual Inspection
by: Song, Jaewoo, et al.
Published: (2025)
by: Song, Jaewoo, et al.
Published: (2025)
InstantFamily: Masked Attention for Zero-shot Multi-ID Image Generation
by: Kim, Chanran, et al.
Published: (2024)
by: Kim, Chanran, et al.
Published: (2024)
4Real-Video: Learning Generalizable Photo-Realistic 4D Video Diffusion
by: Wang, Chaoyang, et al.
Published: (2024)
by: Wang, Chaoyang, et al.
Published: (2024)
Refracting Reality: Generating Images with Realistic Transparent Objects
by: Yin, Yue, et al.
Published: (2025)
by: Yin, Yue, et al.
Published: (2025)
Pixel-aligned RGB-NIR Stereo Imaging and Dataset for Robot Vision
by: Kim, Jinnyeong, et al.
Published: (2024)
by: Kim, Jinnyeong, et al.
Published: (2024)
UNCOVER: Unknown Class Object Detection for Autonomous Vehicles in Real-time
by: Schmarje, Lars, et al.
Published: (2024)
by: Schmarje, Lars, et al.
Published: (2024)
Orientation Matters: Making 3D Generative Models Orientation-Aligned
by: Lu, Yichong, et al.
Published: (2025)
by: Lu, Yichong, et al.
Published: (2025)
ViKey: Enhancing Temporal Understanding in Videos via Visual Prompting
by: Lee, Yeonkyung, et al.
Published: (2026)
by: Lee, Yeonkyung, et al.
Published: (2026)
GOOD: Towards Domain Generalized Orientated Object Detection
by: Bi, Qi, et al.
Published: (2024)
by: Bi, Qi, et al.
Published: (2024)
ActionSwitch: Class-agnostic Detection of Simultaneous Actions in Streaming Videos
by: Kang, Hyolim, et al.
Published: (2024)
by: Kang, Hyolim, et al.
Published: (2024)
HUGSIM: A Real-Time, Photo-Realistic and Closed-Loop Simulator for Autonomous Driving
by: Zhou, Hongyu, et al.
Published: (2024)
by: Zhou, Hongyu, et al.
Published: (2024)
MIRAGE: Model-agnostic Industrial Realistic Anomaly Generation and Evaluation for Visual Anomaly Detection
by: Hu, Jinwei, et al.
Published: (2026)
by: Hu, Jinwei, et al.
Published: (2026)
Similar Items
-
CR-QAT: Curriculum Relational Quantization-Aware Training for Open-Vocabulary Object Detection
by: Park, Jinyeong, et al.
Published: (2026) -
Fluently Lying: Adversarial Robustness Can Be Substrate-Dependent
by: Kang, Daye, et al.
Published: (2026) -
CF-DETR: Coarse-to-Fine Transformer for Real-Time Object Detection
by: Shin, Woojin, et al.
Published: (2025) -
AT-SNN: Adaptive Tokens for Vision Transformer on Spiking Neural Network
by: Kang, Donghwa, et al.
Published: (2024) -
Timestep-Compressed Attack on Spiking Neural Networks through Timestep-Level Backpropagation
by: Kang, Donghwa, et al.
Published: (2025)