ControlVP: Interactive Geometric Refinement of AI-Generated Images with Consistent Vanishing Points
Fuente:
arXiv
Saved in:
| Main Authors: | Okumura, Ryota, Shiohara, Kaede, Yamasaki, Toshihiko |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unified Vector Floorplan Generation via Markup Representation
by: Shiohara, Kaede, et al.
Published: (2026)
by: Shiohara, Kaede, et al.
Published: (2026)
Face2Diffusion for Fast and Editable Face Personalization
by: Shiohara, Kaede, et al.
Published: (2024)
by: Shiohara, Kaede, et al.
Published: (2024)
ExposeAnyone: Personalized Audio-to-Expression Diffusion Models Are Robust Zero-Shot Face Forgery Detectors
by: Shiohara, Kaede, et al.
Published: (2026)
by: Shiohara, Kaede, et al.
Published: (2026)
Robust Deepfake Detection for Electronic Know Your Customer Systems Using Registered Images
by: Amada, Takuma, et al.
Published: (2025)
by: Amada, Takuma, et al.
Published: (2025)
PetFace: A Large-Scale Dataset and Benchmark for Animal Identification
by: Shinoda, Risa, et al.
Published: (2024)
by: Shinoda, Risa, et al.
Published: (2024)
OpenAnimalTracks: A Dataset for Animal Track Recognition
by: Shinoda, Risa, et al.
Published: (2024)
by: Shinoda, Risa, et al.
Published: (2024)
Iterative Self-Improvement of Vision Language Models for Image Scoring and Self-Explanation
by: Tanji, Naoto, et al.
Published: (2025)
by: Tanji, Naoto, et al.
Published: (2025)
BioVITA: Biological Dataset, Model, and Benchmark for Visual-Textual-Acoustic Alignment
by: Shinoda, Risa, et al.
Published: (2026)
by: Shinoda, Risa, et al.
Published: (2026)
A Multihead Continual Learning Framework for Fine-Grained Fashion Image Retrieval with Contrastive Learning and Exponential Moving Average Distillation
by: Xiao, Ling, et al.
Published: (2026)
by: Xiao, Ling, et al.
Published: (2026)
CE-FAM: Concept-Based Explanation via Fusion of Activation Maps
by: Kuroki, Michihiro, et al.
Published: (2025)
by: Kuroki, Michihiro, et al.
Published: (2025)
Bias Beyond Demographics: Probing Decision Boundaries in Black-Box LVLMs via Counterfactual VQA
by: Zhao, Zaiying, et al.
Published: (2025)
by: Zhao, Zaiying, et al.
Published: (2025)
BSED: Baseline Shapley-Based Explainable Detector
by: Kuroki, Michihiro, et al.
Published: (2023)
by: Kuroki, Michihiro, et al.
Published: (2023)
Reward Incremental Learning in Text-to-Image Generation
by: Wang, Maorong, et al.
Published: (2024)
by: Wang, Maorong, et al.
Published: (2024)
Difficulty Controlled Diffusion Model for Synthesizing Effective Training Data
by: Wang, Zerun, et al.
Published: (2024)
by: Wang, Zerun, et al.
Published: (2024)
Language-guided Detection and Mitigation of Unknown Dataset Bias
by: Zhao, Zaiying, et al.
Published: (2024)
by: Zhao, Zaiying, et al.
Published: (2024)
Mirai: Autoregressive Visual Generation Needs Foresight
by: Yu, Yonghao, et al.
Published: (2026)
by: Yu, Yonghao, et al.
Published: (2026)
Weatherproofing Retrieval for Localization with Generative AI and Geometric Consistency
by: Kalantidis, Yannis, et al.
Published: (2024)
by: Kalantidis, Yannis, et al.
Published: (2024)
Recurrence-based Vanishing Point Detection
by: Bharadwaj, Skanda, et al.
Published: (2024)
by: Bharadwaj, Skanda, et al.
Published: (2024)
Joint Fusion and Encoding: Advancing Multimodal Retrieval from the Ground Up
by: Huang, Lang, et al.
Published: (2025)
by: Huang, Lang, et al.
Published: (2025)
From Obstacles to Resources: Semi-supervised Learning Faces Synthetic Data Contamination
by: Wang, Zerun, et al.
Published: (2024)
by: Wang, Zerun, et al.
Published: (2024)
Language-Guided Self-Supervised Video Summarization Using Text Semantic Matching Considering the Diversity of the Video
by: Sugihara, Tomoya, et al.
Published: (2024)
by: Sugihara, Tomoya, et al.
Published: (2024)
Adversarial Training from Mean Field Perspective
by: Kumano, Soichiro, et al.
Published: (2025)
by: Kumano, Soichiro, et al.
Published: (2025)
Adversarially Pretrained Transformers May Be Universally Robust In-Context Learners
by: Kumano, Soichiro, et al.
Published: (2025)
by: Kumano, Soichiro, et al.
Published: (2025)
Auto-Comp: An Automated Pipeline for Scalable Compositional Probing of Contrastive Vision-Language Models
by: Sbrolli, Cristian, et al.
Published: (2026)
by: Sbrolli, Cristian, et al.
Published: (2026)
Theoretical Understanding of Learning from Adversarial Perturbations
by: Kumano, Soichiro, et al.
Published: (2024)
by: Kumano, Soichiro, et al.
Published: (2024)
Wide Two-Layer Networks can Learn from Adversarial Perturbations
by: Kumano, Soichiro, et al.
Published: (2024)
by: Kumano, Soichiro, et al.
Published: (2024)
Geometric Consistency Refinement for Single Image Novel View Synthesis via Test-Time Adaptation of Diffusion Models
by: Bengtson, Josef, et al.
Published: (2025)
by: Bengtson, Josef, et al.
Published: (2025)
SJD-VP: Speculative Jacobi Decoding with Verification Prediction for Autoregressive Image Generation
by: Shan, Bingqi, et al.
Published: (2026)
by: Shan, Bingqi, et al.
Published: (2026)
Attribute-Guided Multi-Level Attention Network for Fine-Grained Fashion Retrieval
by: Xiao, Ling, et al.
Published: (2022)
by: Xiao, Ling, et al.
Published: (2022)
Online Open-set Semi-supervised Object Detection with Dual Competing Head
by: Wang, Zerun, et al.
Published: (2023)
by: Wang, Zerun, et al.
Published: (2023)
Spectral Probing of Feature Upsamplers in 2D-to-3D Scene Reconstruction
by: Xiao, Ling, et al.
Published: (2026)
by: Xiao, Ling, et al.
Published: (2026)
TextTIGER: Text-based Intelligent Generation with Entity Prompt Refinement for Text-to-Image Generation
by: Ozaki, Shintaro, et al.
Published: (2025)
by: Ozaki, Shintaro, et al.
Published: (2025)
Convex Relaxation for Robust Vanishing Point Estimation in Manhattan World
by: Liao, Bangyan, et al.
Published: (2025)
by: Liao, Bangyan, et al.
Published: (2025)
Vanishing-Point-Guided Video Semantic Segmentation of Driving Scenes
by: Guo, Diandian, et al.
Published: (2024)
by: Guo, Diandian, et al.
Published: (2024)
Sora Generates Videos with Stunning Geometrical Consistency
by: Li, Xuanyi, et al.
Published: (2024)
by: Li, Xuanyi, et al.
Published: (2024)
Continual Distillation of Teachers from Different Domains
by: Michel, Nicolas, et al.
Published: (2026)
by: Michel, Nicolas, et al.
Published: (2026)
Dealing with Synthetic Data Contamination in Online Continual Learning
by: Wang, Maorong, et al.
Published: (2024)
by: Wang, Maorong, et al.
Published: (2024)
TDRI: Two-Phase Dialogue Refinement and Co-Adaptation for Interactive Image Generation
by: Feng, Yuheng, et al.
Published: (2025)
by: Feng, Yuheng, et al.
Published: (2025)
Refining Segmentation On-the-Fly: An Interactive Framework for Point Cloud Semantic Segmentation
by: Zhang, Peng, et al.
Published: (2024)
by: Zhang, Peng, et al.
Published: (2024)
Grab-3D: Detecting AI-Generated Videos from 3D Geometric Temporal Consistency
by: Chen, Wenhan, et al.
Published: (2025)
by: Chen, Wenhan, et al.
Published: (2025)
Similar Items
-
Unified Vector Floorplan Generation via Markup Representation
by: Shiohara, Kaede, et al.
Published: (2026) -
Face2Diffusion for Fast and Editable Face Personalization
by: Shiohara, Kaede, et al.
Published: (2024) -
ExposeAnyone: Personalized Audio-to-Expression Diffusion Models Are Robust Zero-Shot Face Forgery Detectors
by: Shiohara, Kaede, et al.
Published: (2026) -
Robust Deepfake Detection for Electronic Know Your Customer Systems Using Registered Images
by: Amada, Takuma, et al.
Published: (2025) -
PetFace: A Large-Scale Dataset and Benchmark for Animal Identification
by: Shinoda, Risa, et al.
Published: (2024)