Region-Wise Correspondence Prediction between Manga Line Art Images
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Yingxuan, Mao, Jiafeng, Qiu, Qianru, Matsui, Yusuke |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MangaDiT: Reference-Guided Line Art Colorization with Hierarchical Attention in Diffusion Transformers
von: Qiu, Qianru, et al.
Veröffentlicht: (2025)
von: Qiu, Qianru, et al.
Veröffentlicht: (2025)
Noisy Label Refinement with Semantically Reliable Synthetic Images
von: Li, Yingxuan, et al.
Veröffentlicht: (2025)
von: Li, Yingxuan, et al.
Veröffentlicht: (2025)
Manga109Dialog: A Large-scale Dialogue Dataset for Comics Speaker Detection
von: Li, Yingxuan, et al.
Veröffentlicht: (2023)
von: Li, Yingxuan, et al.
Veröffentlicht: (2023)
Zero-Shot Character Identification and Speaker Prediction in Comics via Iterative Multimodal Fusion
von: Li, Yingxuan, et al.
Veröffentlicht: (2024)
von: Li, Yingxuan, et al.
Veröffentlicht: (2024)
Exploring Palette based Color Guidance in Diffusion Models
von: Qiu, Qianru, et al.
Veröffentlicht: (2025)
von: Qiu, Qianru, et al.
Veröffentlicht: (2025)
MangaNinja: Line Art Colorization with Precise Reference Following
von: Liu, Zhiheng, et al.
Veröffentlicht: (2025)
von: Liu, Zhiheng, et al.
Veröffentlicht: (2025)
High-Frequency Anti-DreamBooth: Robust Defense against Personalized Image Synthesis
von: Onikubo, Takuto, et al.
Veröffentlicht: (2024)
von: Onikubo, Takuto, et al.
Veröffentlicht: (2024)
LORE: Latent Optimization for Precise Semantic Control in Rectified Flow-based Image Editing
von: Ouyang, Liangyang, et al.
Veröffentlicht: (2025)
von: Ouyang, Liangyang, et al.
Veröffentlicht: (2025)
Guided Image Synthesis via Initial Image Editing in Diffusion Model
von: Mao, Jiafeng, et al.
Veröffentlicht: (2023)
von: Mao, Jiafeng, et al.
Veröffentlicht: (2023)
ZoDi: Zero-Shot Domain Adaptation with Diffusion-Based Image Transfer
von: Azuma, Hiroki, et al.
Veröffentlicht: (2024)
von: Azuma, Hiroki, et al.
Veröffentlicht: (2024)
Inference-time Trajectory Optimization for Manga Image Editing
von: Furuta, Ryosuke
Veröffentlicht: (2026)
von: Furuta, Ryosuke
Veröffentlicht: (2026)
MangaFlow: An End-to-End Agentic Framework for Controllable Story to Manga Generation
von: Wang, Muyao, et al.
Veröffentlicht: (2026)
von: Wang, Muyao, et al.
Veröffentlicht: (2026)
RouteExtract: A Modular Pipeline for Extracting Routes from Paper Maps
von: Kremser, Bjoern, et al.
Veröffentlicht: (2025)
von: Kremser, Bjoern, et al.
Veröffentlicht: (2025)
SVGEditBench: A Benchmark Dataset for Quantitative Assessment of LLM's SVG Editing Capabilities
von: Nishina, Kunato, et al.
Veröffentlicht: (2024)
von: Nishina, Kunato, et al.
Veröffentlicht: (2024)
Adversarial Doodles: Interpretable and Human-drawable Attacks Provide Describable Insights
von: Nara, Ryoya, et al.
Veröffentlicht: (2023)
von: Nara, Ryoya, et al.
Veröffentlicht: (2023)
MangaUB: A Manga Understanding Benchmark for Large Multimodal Models
von: Ikuta, Hikaru, et al.
Veröffentlicht: (2024)
von: Ikuta, Hikaru, et al.
Veröffentlicht: (2024)
MangaVQA and MangaLMM: A Benchmark and Specialized Model for Multimodal Manga Understanding
von: Baek, Jeonghun, et al.
Veröffentlicht: (2025)
von: Baek, Jeonghun, et al.
Veröffentlicht: (2025)
Manga109-v2026: Revisiting Manga109 Annotations for Modern Manga Understanding
von: Baek, Jeonghun, et al.
Veröffentlicht: (2026)
von: Baek, Jeonghun, et al.
Veröffentlicht: (2026)
ShapeMoiré: Channel-Wise Shape-Guided Network for Image Demoiréing
von: Cao, Jinming, et al.
Veröffentlicht: (2024)
von: Cao, Jinming, et al.
Veröffentlicht: (2024)
Manga Generation via Layout-controllable Diffusion
von: Chen, Siyu, et al.
Veröffentlicht: (2024)
von: Chen, Siyu, et al.
Veröffentlicht: (2024)
WiseEdit: Benchmarking Cognition- and Creativity-Informed Image Editing
von: Pan, Kaihang, et al.
Veröffentlicht: (2025)
von: Pan, Kaihang, et al.
Veröffentlicht: (2025)
LotusFilter: Fast Diverse Nearest Neighbor Search via a Learned Cutoff Table
von: Matsui, Yusuke
Veröffentlicht: (2025)
von: Matsui, Yusuke
Veröffentlicht: (2025)
Weakly-Supervised Semantic Segmentation with Image-Level Labels: from Traditional Models to Foundation Models
von: Chen, Zhaozheng, et al.
Veröffentlicht: (2023)
von: Chen, Zhaozheng, et al.
Veröffentlicht: (2023)
The Manga Whisperer: Automatically Generating Transcriptions for Comics
von: Sachdeva, Ragav, et al.
Veröffentlicht: (2024)
von: Sachdeva, Ragav, et al.
Veröffentlicht: (2024)
Reducing Class-Wise Performance Disparity via Margin Regularization
von: Zhu, Beier, et al.
Veröffentlicht: (2026)
von: Zhu, Beier, et al.
Veröffentlicht: (2026)
Number it: Temporal Grounding Videos like Flipping Manga
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
The Lottery Ticket Hypothesis in Denoising: Towards Semantic-Driven Initialization
von: Mao, Jiafeng, et al.
Veröffentlicht: (2023)
von: Mao, Jiafeng, et al.
Veröffentlicht: (2023)
Probabilistic 3D Correspondence Prediction from Sparse Unsegmented Images
von: Iyer, Krithika, et al.
Veröffentlicht: (2024)
von: Iyer, Krithika, et al.
Veröffentlicht: (2024)
Reward Incremental Learning in Text-to-Image Generation
von: Wang, Maorong, et al.
Veröffentlicht: (2024)
von: Wang, Maorong, et al.
Veröffentlicht: (2024)
Cobra: Efficient Line Art COlorization with BRoAder References
von: Zhuang, Junhao, et al.
Veröffentlicht: (2025)
von: Zhuang, Junhao, et al.
Veröffentlicht: (2025)
Training-Free Sketch-Guided Diffusion with Latent Optimization
von: Ding, Sandra Zhang, et al.
Veröffentlicht: (2024)
von: Ding, Sandra Zhang, et al.
Veröffentlicht: (2024)
DiffSensei: Bridging Multi-Modal LLMs and Diffusion Models for Customized Manga Generation
von: Wu, Jianzong, et al.
Veröffentlicht: (2024)
von: Wu, Jianzong, et al.
Veröffentlicht: (2024)
Tails Tell Tales: Chapter-Wide Manga Transcriptions with Character Names
von: Sachdeva, Ragav, et al.
Veröffentlicht: (2024)
von: Sachdeva, Ragav, et al.
Veröffentlicht: (2024)
In-Context Translation: Towards Unifying Image Recognition, Processing, and Generation
von: Xue, Han, et al.
Veröffentlicht: (2024)
von: Xue, Han, et al.
Veröffentlicht: (2024)
OWT: A Foundational Organ-Wise Tokenization Framework for Medical Imaging
von: Song, Sifan, et al.
Veröffentlicht: (2025)
von: Song, Sifan, et al.
Veröffentlicht: (2025)
Insertion Network for Image Sequence Correspondence
von: Su, Dingjie, et al.
Veröffentlicht: (2026)
von: Su, Dingjie, et al.
Veröffentlicht: (2026)
Re:Verse -- Can Your VLM Read a Manga?
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2025)
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2025)
Difficulty Controlled Diffusion Model for Synthesizing Effective Training Data
von: Wang, Zerun, et al.
Veröffentlicht: (2024)
von: Wang, Zerun, et al.
Veröffentlicht: (2024)
From Obstacles to Resources: Semi-supervised Learning Faces Synthetic Data Contamination
von: Wang, Zerun, et al.
Veröffentlicht: (2024)
von: Wang, Zerun, et al.
Veröffentlicht: (2024)
How Panel Layouts Define Manga: Insights from Visual Ablation Experiments
von: Feng, Siyuan, et al.
Veröffentlicht: (2024)
von: Feng, Siyuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MangaDiT: Reference-Guided Line Art Colorization with Hierarchical Attention in Diffusion Transformers
von: Qiu, Qianru, et al.
Veröffentlicht: (2025) -
Noisy Label Refinement with Semantically Reliable Synthetic Images
von: Li, Yingxuan, et al.
Veröffentlicht: (2025) -
Manga109Dialog: A Large-scale Dialogue Dataset for Comics Speaker Detection
von: Li, Yingxuan, et al.
Veröffentlicht: (2023) -
Zero-Shot Character Identification and Speaker Prediction in Comics via Iterative Multimodal Fusion
von: Li, Yingxuan, et al.
Veröffentlicht: (2024) -
Exploring Palette based Color Guidance in Diffusion Models
von: Qiu, Qianru, et al.
Veröffentlicht: (2025)