Infinite-Story: A Training-Free Consistent Text-to-Image Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Park, Jihun, Lee, Kyoungmin, Gim, Jongmin, Jo, Hyeonseo, Oh, Minseok, Choi, Wonhyeok, Hwang, Kyumin, Kim, Jaeyeul, Choi, Minwoo, Im, Sunghoon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Training-Free Style-Personalization via SVD-Based Feature Decomposition
by: Lee, Kyoungmin, et al.
Published: (2025)
by: Lee, Kyoungmin, et al.
Published: (2025)
A Training-Free Style-aligned Image Generation with Scale-wise Autoregressive Model
by: Park, Jihun, et al.
Published: (2025)
by: Park, Jihun, et al.
Published: (2025)
Self-supervised Monocular Depth Estimation Robust to Reflective Surface Leveraged by Triplet Mining
by: Choi, Wonhyeok, et al.
Published: (2025)
by: Choi, Wonhyeok, et al.
Published: (2025)
A Review of Online Diffusion Policy RL Algorithms for Scalable Robotic Control
by: Choi, Wonhyeok, et al.
Published: (2026)
by: Choi, Wonhyeok, et al.
Published: (2026)
Style-Editor: Text-driven object-centric style editing
by: Park, Jihun, et al.
Published: (2024)
by: Park, Jihun, et al.
Published: (2024)
Intrinsic Image Decomposition for Robust Self-supervised Monocular Depth Estimation on Reflective Surfaces
by: Choi, Wonhyeok, et al.
Published: (2025)
by: Choi, Wonhyeok, et al.
Published: (2025)
Scale-invariant and View-relational Representation Learning for Full Surround Monocular Depth
by: Hwang, Kyumin, et al.
Published: (2025)
by: Hwang, Kyumin, et al.
Published: (2025)
Depth-discriminative Metric Learning for Monocular 3D Object Detection
by: Choi, Wonhyeok, et al.
Published: (2024)
by: Choi, Wonhyeok, et al.
Published: (2024)
Bridging Geometric and Semantic Foundation Models for Generalized Monocular Depth Estimation
by: Ma, Sanggyun, et al.
Published: (2025)
by: Ma, Sanggyun, et al.
Published: (2025)
Rethinking LiDAR Domain Generalization: Single Source as Multiple Density Domains
by: Kim, Jaeyeul, et al.
Published: (2023)
by: Kim, Jaeyeul, et al.
Published: (2023)
Flow4D: Leveraging 4D Voxel Network for LiDAR Scene Flow Estimation
by: Kim, Jaeyeul, et al.
Published: (2024)
by: Kim, Jaeyeul, et al.
Published: (2024)
Multi-task Learning for Real-time Autonomous Driving Leveraging Task-adaptive Attention Generator
by: Choi, Wonhyeok, et al.
Published: (2024)
by: Choi, Wonhyeok, et al.
Published: (2024)
CAVIS: Context-Aware Video Instance Segmentation
by: Lee, Seunghun, et al.
Published: (2024)
by: Lee, Seunghun, et al.
Published: (2024)
Latest Object Memory Management for Temporally Consistent Video Instance Segmentation
by: Lee, Seunghun, et al.
Published: (2025)
by: Lee, Seunghun, et al.
Published: (2025)
CookingSense: A Culinary Knowledgebase with Multidisciplinary Assertions
by: Choi, Donghee, et al.
Published: (2024)
by: Choi, Donghee, et al.
Published: (2024)
Loss-Optimized Reconfigurable Nonlocal Metasurface-aided Cavity Antenna
by: Cho, Minwoo, et al.
Published: (2026)
by: Cho, Minwoo, et al.
Published: (2026)
Prompt-Based Safety Guidance Is Ineffective for Unlearned Text-to-Image Diffusion Models
by: Shin, Jiwoo, et al.
Published: (2025)
by: Shin, Jiwoo, et al.
Published: (2025)
Conditional Unbalanced Optimal Transport Maps: An Outlier-Robust Framework for Conditional Generative Modeling
by: Yoon, Jiwoo, et al.
Published: (2026)
by: Yoon, Jiwoo, et al.
Published: (2026)
FIFO-Diffusion: Generating Infinite Videos from Text without Training
by: Kim, Jihwan, et al.
Published: (2024)
by: Kim, Jihwan, et al.
Published: (2024)
Temporal Grounding as a Learning Signal for Referring Video Object Segmentation
by: Lee, Seunghun, et al.
Published: (2025)
by: Lee, Seunghun, et al.
Published: (2025)
StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization
by: Gaur, Gopalji, et al.
Published: (2025)
by: Gaur, Gopalji, et al.
Published: (2025)
Bond-Strength-Based Understanding of Oxygen Vacancy Migration Barriers in Rutile Oxides
by: Kim, Inseo, et al.
Published: (2026)
by: Kim, Inseo, et al.
Published: (2026)
Partially Equivariant Reinforcement Learning in Symmetry-Breaking Environments
by: Chang, Junwoo, et al.
Published: (2025)
by: Chang, Junwoo, et al.
Published: (2025)
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
by: Park, Sangha, et al.
Published: (2025)
by: Park, Sangha, et al.
Published: (2025)
Spatial Transport Optimization by Repositioning Attention Map for Training-Free Text-to-Image Synthesis
by: Han, Woojung, et al.
Published: (2025)
by: Han, Woojung, et al.
Published: (2025)
ParCo-SDF: Learning Prior-Free Partial-to-Complete Signed Distance Fields of Deformable Objects
by: Hwang, Deokmin, et al.
Published: (2026)
by: Hwang, Deokmin, et al.
Published: (2026)
Training-Free Consistent Text-to-Image Generation
by: Tewel, Yoad, et al.
Published: (2024)
by: Tewel, Yoad, et al.
Published: (2024)
Functional Fibers in Soft Robotics: Advances in Material, Structural, and Systemic Tactics
by: Joonhee Won, et al.
Published: (2026)
by: Joonhee Won, et al.
Published: (2026)
Development of a Validation and Inspection Tool for Armband-based Lifelog Data (VITAL) to Facilitate the Clinical Use of Wearable Data: A Prototype and Usability Evaluation
by: Eunyoung, Im, et al.
Published: (2025)
by: Eunyoung, Im, et al.
Published: (2025)
UniHENN: Designing Faster and More Versatile Homomorphic Encryption-based CNNs without im2col
by: Choi, Hyunmin, et al.
Published: (2024)
by: Choi, Hyunmin, et al.
Published: (2024)
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images
by: Lee, Jaeseong, et al.
Published: (2025)
by: Lee, Jaeseong, et al.
Published: (2025)
Long-Tailed Recognition on Binary Networks by Calibrating A Pre-trained Model
by: Kim, Jihun, et al.
Published: (2024)
by: Kim, Jihun, et al.
Published: (2024)
TextGuider: Training-Free Guidance for Text Rendering via Attention Alignment
by: Baek, Kanghyun, et al.
Published: (2025)
by: Baek, Kanghyun, et al.
Published: (2025)
Piece of Table: A Divide-and-Conquer Approach for Selecting Subtables in Table Question Answering
by: Lee, Wonjin, et al.
Published: (2024)
by: Lee, Wonjin, et al.
Published: (2024)
OmniText: A Training-Free Generalist for Controllable Text-Image Manipulation
by: Gunawan, Agus, et al.
Published: (2025)
by: Gunawan, Agus, et al.
Published: (2025)
LabTOP: A Unified Model for Lab Test Outcome Prediction on Electronic Health Records
by: Im, Sujeong, et al.
Published: (2025)
by: Im, Sujeong, et al.
Published: (2025)
Effective HA‐Loading Strategy in Internal Dome‐Assisted Suction Cups for Enhanced Transdermal Delivery Under Vertical Adhesion and Vibration
by: Minjin Kim, et al.
Published: (2025)
by: Minjin Kim, et al.
Published: (2025)
Effective HA‐Loading Strategy in Internal Dome‐Assisted Suction Cups for Enhanced Transdermal Delivery Under Vertical Adhesion and Vibration (Adv. Funct. Mater. 20/2026)
by: Minjin Kim, et al.
Published: (2026)
by: Minjin Kim, et al.
Published: (2026)
Probing small-scale power spectrum with gravitational-wave diffractive lensing
by: Kim, Sungjung, et al.
Published: (2025)
by: Kim, Sungjung, et al.
Published: (2025)
WAVE: Warp-Based View Guidance for Consistent Novel View Synthesis Using a Single Image
by: Park, Jiwoo, et al.
Published: (2025)
by: Park, Jiwoo, et al.
Published: (2025)
Similar Items
-
A Training-Free Style-Personalization via SVD-Based Feature Decomposition
by: Lee, Kyoungmin, et al.
Published: (2025) -
A Training-Free Style-aligned Image Generation with Scale-wise Autoregressive Model
by: Park, Jihun, et al.
Published: (2025) -
Self-supervised Monocular Depth Estimation Robust to Reflective Surface Leveraged by Triplet Mining
by: Choi, Wonhyeok, et al.
Published: (2025) -
A Review of Online Diffusion Policy RL Algorithms for Scalable Robotic Control
by: Choi, Wonhyeok, et al.
Published: (2026) -
Style-Editor: Text-driven object-centric style editing
by: Park, Jihun, et al.
Published: (2024)