Global-Local Aware Scene Text Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Fuxiang, Su, Tonghua, Di, Donglin, Chen, Yin, Wu, Xiangqian, Wang, Zhongjie, Fan, Lei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DiTalker: A Unified DiT-based Framework for High-Quality and Speaking Styles Controllable Portrait Animation
von: Feng, He, et al.
Veröffentlicht: (2025)
von: Feng, He, et al.
Veröffentlicht: (2025)
FOCA: Frequency-Oriented Cross-Domain Forgery Detection, Localization and Explanation via Multi-Modal Large Language Model
von: Liu, Zhou, et al.
Veröffentlicht: (2026)
von: Liu, Zhou, et al.
Veröffentlicht: (2026)
EFDiT: Efficient Fine-grained Image Generation Using Diffusion Transformer Models
von: Wang, Kun, et al.
Veröffentlicht: (2025)
von: Wang, Kun, et al.
Veröffentlicht: (2025)
Adams Bashforth Moulton Solver for Inversion and Editing in Rectified Flow
von: Ma, Yongjia, et al.
Veröffentlicht: (2025)
von: Ma, Yongjia, et al.
Veröffentlicht: (2025)
CogPortrait: Fine-Grained Eye-Region Control in Portrait Animation via Hierarchical Agent Planning
von: Feng, He, et al.
Veröffentlicht: (2026)
von: Feng, He, et al.
Veröffentlicht: (2026)
GUSLO: General and Unified Structured Light Optimization
von: Wan, Tinglei, et al.
Veröffentlicht: (2025)
von: Wan, Tinglei, et al.
Veröffentlicht: (2025)
Spatially Gene Expression Prediction using Dual-Scale Contrastive Learning
von: Qu, Mingcheng, et al.
Veröffentlicht: (2025)
von: Qu, Mingcheng, et al.
Veröffentlicht: (2025)
Chain of World: World Model Thinking in Latent Motion
von: Yang, Fuxiang, et al.
Veröffentlicht: (2026)
von: Yang, Fuxiang, et al.
Veröffentlicht: (2026)
Multimodal Cancer Survival Analysis via Hypergraph Learning with Cross-Modality Rebalance
von: Qu, Mingcheng, et al.
Veröffentlicht: (2025)
von: Qu, Mingcheng, et al.
Veröffentlicht: (2025)
Memory-Augmented Incomplete Multimodal Survival Prediction via Cross-Slide and Gene-Attentive Hypergraph Learning
von: Qu, Mingcheng, et al.
Veröffentlicht: (2025)
von: Qu, Mingcheng, et al.
Veröffentlicht: (2025)
One-Shot Pose-Driving Face Animation Platform
von: Feng, He, et al.
Veröffentlicht: (2024)
von: Feng, He, et al.
Veröffentlicht: (2024)
Real Face Video Animation Platform
von: Chen, Xiaokai, et al.
Veröffentlicht: (2024)
von: Chen, Xiaokai, et al.
Veröffentlicht: (2024)
DH-FaceVid-1K: A Large-Scale High-Quality Dataset for Face Video Generation
von: Di, Donglin, et al.
Veröffentlicht: (2024)
von: Di, Donglin, et al.
Veröffentlicht: (2024)
Balancing Stability and Plasticity in Pretrained Detector: A Dual-Path Framework for Incremental Object Detection
von: Li, Songze, et al.
Veröffentlicht: (2025)
von: Li, Songze, et al.
Veröffentlicht: (2025)
DUKAE: DUal-level Knowledge Accumulation and Ensemble for Pre-Trained Model-Based Continual Learning
von: Li, Songze, et al.
Veröffentlicht: (2025)
von: Li, Songze, et al.
Veröffentlicht: (2025)
Multimodal Continual Instruction Tuning with Dynamic Gradient Guidance
von: Li, Songze, et al.
Veröffentlicht: (2025)
von: Li, Songze, et al.
Veröffentlicht: (2025)
Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion
von: Ma, Yongjia, et al.
Veröffentlicht: (2025)
von: Ma, Yongjia, et al.
Veröffentlicht: (2025)
GAOT: Generating Articulated Objects Through Text-Guided Diffusion Models
von: Sun, Hao, et al.
Veröffentlicht: (2025)
von: Sun, Hao, et al.
Veröffentlicht: (2025)
Salvaging the Overlooked: Leveraging Class-Aware Contrastive Learning for Multi-Class Anomaly Detection
von: Fan, Lei, et al.
Veröffentlicht: (2024)
von: Fan, Lei, et al.
Veröffentlicht: (2024)
Recognition-Synergistic Scene Text Editing
von: Fang, Zhengyao, et al.
Veröffentlicht: (2025)
von: Fang, Zhengyao, et al.
Veröffentlicht: (2025)
TextSculptor: Training and Benchmarking Scene Text Editing
von: Lin, Yiheng, et al.
Veröffentlicht: (2026)
von: Lin, Yiheng, et al.
Veröffentlicht: (2026)
TrAME: Trajectory-Anchored Multi-View Editing for Text-Guided 3D Gaussian Splatting Manipulation
von: Luo, Chaofan, et al.
Veröffentlicht: (2024)
von: Luo, Chaofan, et al.
Veröffentlicht: (2024)
Watch Your Steps: Local Image and Scene Editing by Text Instructions
von: Mirzaei, Ashkan, et al.
Veröffentlicht: (2023)
von: Mirzaei, Ashkan, et al.
Veröffentlicht: (2023)
LatentEditor: Text Driven Local Editing of 3D Scenes
von: Khalid, Umar, et al.
Veröffentlicht: (2023)
von: Khalid, Umar, et al.
Veröffentlicht: (2023)
Mono4DEditor: Text-Driven 4D Scene Editing from Monocular Video via Point-Level Localization of Language-Embedded Gaussians
von: Shi, Jin-Chuan, et al.
Veröffentlicht: (2025)
von: Shi, Jin-Chuan, et al.
Veröffentlicht: (2025)
TextCtrl: Diffusion-based Scene Text Editing with Prior Guidance Control
von: Zeng, Weichao, et al.
Veröffentlicht: (2024)
von: Zeng, Weichao, et al.
Veröffentlicht: (2024)
MANTA: A Large-Scale Multi-View and Visual-Text Anomaly Detection Dataset for Tiny Objects
von: Fan, Lei, et al.
Veröffentlicht: (2024)
von: Fan, Lei, et al.
Veröffentlicht: (2024)
FLUX-Text: A Simple and Advanced Diffusion Transformer Baseline for Scene Text Editing
von: Lan, Rui, et al.
Veröffentlicht: (2025)
von: Lan, Rui, et al.
Veröffentlicht: (2025)
Towards Training-Free Scene Text Editing
von: Li, Yubo, et al.
Veröffentlicht: (2026)
von: Li, Yubo, et al.
Veröffentlicht: (2026)
Gaze into the Details: Locality-Sensitive Enhancement for OCTA Retinal Vessel Segmentation
von: Huang, Tuopusen, et al.
Veröffentlicht: (2026)
von: Huang, Tuopusen, et al.
Veröffentlicht: (2026)
GlyphMastero: A Glyph Encoder for High-Fidelity Scene Text Editing
von: Wang, Tong, et al.
Veröffentlicht: (2025)
von: Wang, Tong, et al.
Veröffentlicht: (2025)
Localized Gaussian Splatting Editing with Contextual Awareness
von: Xiao, Hanyuan, et al.
Veröffentlicht: (2024)
von: Xiao, Hanyuan, et al.
Veröffentlicht: (2024)
Hypergraph Tversky-Aware Domain Incremental Learning for Brain Tumor Segmentation with Missing Modalities
von: Wang, Junze, et al.
Veröffentlicht: (2025)
von: Wang, Junze, et al.
Veröffentlicht: (2025)
Physics-Aware 3D Gaussian Editing for Driving Scene Generation
von: Zhou, Feng, et al.
Veröffentlicht: (2026)
von: Zhou, Feng, et al.
Veröffentlicht: (2026)
Edit Fidelity Field: Semantics-Aware Region Isolation for Training-Free Scene Text Editing
von: Li, Guandong, et al.
Veröffentlicht: (2026)
von: Li, Guandong, et al.
Veröffentlicht: (2026)
SyncNoise: Geometrically Consistent Noise Prediction for Text-based 3D Scene Editing
von: Li, Ruihuang, et al.
Veröffentlicht: (2024)
von: Li, Ruihuang, et al.
Veröffentlicht: (2024)
TextMastero: Mastering High-Quality Scene Text Editing in Diverse Languages and Styles
von: Wang, Tong, et al.
Veröffentlicht: (2024)
von: Wang, Tong, et al.
Veröffentlicht: (2024)
Free-Editor: Zero-shot Text-driven 3D Scene Editing
von: Karim, Nazmul, et al.
Veröffentlicht: (2023)
von: Karim, Nazmul, et al.
Veröffentlicht: (2023)
ToCoAD: Two-Stage Contrastive Learning for Industrial Anomaly Detection
von: Liang, Yun, et al.
Veröffentlicht: (2024)
von: Liang, Yun, et al.
Veröffentlicht: (2024)
Self-Prompting Diffusion Transformer for Open-Vocabulary Scene Text Editing via In-Context Learning
von: Li, Hongxi, et al.
Veröffentlicht: (2026)
von: Li, Hongxi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
DiTalker: A Unified DiT-based Framework for High-Quality and Speaking Styles Controllable Portrait Animation
von: Feng, He, et al.
Veröffentlicht: (2025) -
FOCA: Frequency-Oriented Cross-Domain Forgery Detection, Localization and Explanation via Multi-Modal Large Language Model
von: Liu, Zhou, et al.
Veröffentlicht: (2026) -
EFDiT: Efficient Fine-grained Image Generation Using Diffusion Transformer Models
von: Wang, Kun, et al.
Veröffentlicht: (2025) -
Adams Bashforth Moulton Solver for Inversion and Editing in Rectified Flow
von: Ma, Yongjia, et al.
Veröffentlicht: (2025) -
CogPortrait: Fine-Grained Eye-Region Control in Portrait Animation via Hierarchical Agent Planning
von: Feng, He, et al.
Veröffentlicht: (2026)