Untwisting RoPE: Frequency Control for Shared Attention in DiTs
Fuente:
arXiv
Saved in:
| Main Authors: | Mikaeili, Aryan, Patashnik, Or, Tagliasacchi, Andrea, Cohen-Or, Daniel, Mahdavi-Amiri, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
In-Context Sync-LoRA for Portrait Video Editing
by: Polaczek, Sagi, et al.
Published: (2025)
by: Polaczek, Sagi, et al.
Published: (2025)
Sound Sparks Motion: Audio and Text Tuning for Video Editing
by: Razlighi, AmirHossein Naghi, et al.
Published: (2026)
by: Razlighi, AmirHossein Naghi, et al.
Published: (2026)
In-2-4D: Inbetweening from Two Single-View Images to 4D Generation
by: Nag, Sauradip, et al.
Published: (2025)
by: Nag, Sauradip, et al.
Published: (2025)
Griffin: Generative Reference and Layout Guided Image Composition
by: Mikaeili, Aryan, et al.
Published: (2025)
by: Mikaeili, Aryan, et al.
Published: (2025)
Nested Attention: Semantic-aware Attention Values for Concept Personalization
by: Patashnik, Or, et al.
Published: (2025)
by: Patashnik, Or, et al.
Published: (2025)
HiMat: DiT-based Ultra-High Resolution SVBRDF Generation
by: Wang, Zixiong, et al.
Published: (2025)
by: Wang, Zixiong, et al.
Published: (2025)
TurboEdit: Text-Based Image Editing Using Few-Step Diffusion Models
by: Deutch, Gilad, et al.
Published: (2024)
by: Deutch, Gilad, et al.
Published: (2024)
Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation
by: Dahary, Omer, et al.
Published: (2024)
by: Dahary, Omer, et al.
Published: (2024)
Advances in Neural 3D Mesh Texturing: A Survey
by: Perla, Sai Raj Kishore, et al.
Published: (2026)
by: Perla, Sai Raj Kishore, et al.
Published: (2026)
Consolidating Attention Features for Multi-view Image Editing
by: Patashnik, Or, et al.
Published: (2024)
by: Patashnik, Or, et al.
Published: (2024)
LooseRoPE: Content-aware Attention Manipulation for Semantic Harmonization
by: Sella, Etai, et al.
Published: (2026)
by: Sella, Etai, et al.
Published: (2026)
SweepNet: Unsupervised Learning Shape Abstraction via Neural Sweepers
by: Zhao, Mingrui, et al.
Published: (2024)
by: Zhao, Mingrui, et al.
Published: (2024)
Continuous Control of Editing Models via Adaptive-Origin Guidance
by: Wolf, Alon, et al.
Published: (2026)
by: Wolf, Alon, et al.
Published: (2026)
Tight Inversion: Image-Conditioned Inversion for Real Image Editing
by: Kadosh, Edo, et al.
Published: (2025)
by: Kadosh, Edo, et al.
Published: (2025)
Volumetric Rendering with Baked Quadrature Fields
by: Sharma, Gopal, et al.
Published: (2023)
by: Sharma, Gopal, et al.
Published: (2023)
LCM-Lookahead for Encoder-based Text-to-Image Personalization
by: Gal, Rinon, et al.
Published: (2024)
by: Gal, Rinon, et al.
Published: (2024)
ReRoPE: Repurposing RoPE for Relative Camera Control
by: Li, Chunyang, et al.
Published: (2026)
by: Li, Chunyang, et al.
Published: (2026)
Style Aligned Image Generation via Shared Attention
by: Hertz, Amir, et al.
Published: (2023)
by: Hertz, Amir, et al.
Published: (2023)
SAEdit: Token-level control for continuous image editing via Sparse AutoEncoder
by: Kamenetsky, Ronen, et al.
Published: (2025)
by: Kamenetsky, Ronen, et al.
Published: (2025)
Be Decisive: Noise-Induced Layouts for Multi-Subject Generation
by: Dahary, Omer, et al.
Published: (2025)
by: Dahary, Omer, et al.
Published: (2025)
VLM-Guided Adaptive Negative Prompting for Creative Generation
by: Golan, Shelly, et al.
Published: (2025)
by: Golan, Shelly, et al.
Published: (2025)
Cora: Correspondence-aware image editing using few step diffusion
by: Alimohammadi, Amirhossein, et al.
Published: (2025)
by: Alimohammadi, Amirhossein, et al.
Published: (2025)
Power Foam: Unifying Real-Time Differentiable Ray Tracing and Rasterization
by: Govindarajan, Shrisudhan, et al.
Published: (2026)
by: Govindarajan, Shrisudhan, et al.
Published: (2026)
Image Generation from Contextually-Contradictory Prompts
by: Huberman, Saar, et al.
Published: (2025)
by: Huberman, Saar, et al.
Published: (2025)
JUST-DUB-IT: Video Dubbing via Joint Audio-Visual Diffusion
by: Chen, Anthony, et al.
Published: (2026)
by: Chen, Anthony, et al.
Published: (2026)
Stable Flow: Vital Layers for Training-Free Image Editing
by: Avrahami, Omri, et al.
Published: (2024)
by: Avrahami, Omri, et al.
Published: (2024)
ReNoise: Real Image Inversion Through Iterative Noising
by: Garibi, Daniel, et al.
Published: (2024)
by: Garibi, Daniel, et al.
Published: (2024)
Dynamic Concepts Personalization from Single Videos
by: Abdal, Rameen, et al.
Published: (2025)
by: Abdal, Rameen, et al.
Published: (2025)
Zero-Shot Dynamic Concept Personalization with Grid-Based LoRA
by: Abdal, Rameen, et al.
Published: (2025)
by: Abdal, Rameen, et al.
Published: (2025)
StochasticSplats: Stochastic Rasterization for Sorting-Free 3D Gaussian Splatting
by: Kheradmand, Shakiba, et al.
Published: (2025)
by: Kheradmand, Shakiba, et al.
Published: (2025)
PointNeRF++: A multi-scale, point-based Neural Radiance Field
by: Sun, Weiwei, et al.
Published: (2023)
by: Sun, Weiwei, et al.
Published: (2023)
Object-level Visual Prompts for Compositional Image Generation
by: Parmar, Gaurav, et al.
Published: (2025)
by: Parmar, Gaurav, et al.
Published: (2025)
MultiAct: Text-to-Motion Generation from Composite Text via Tailored Attention Guidance
by: Sala, Nathan, et al.
Published: (2026)
by: Sala, Nathan, et al.
Published: (2026)
Colorful-Noise: Training-Free Low-Frequency Noise Manipulation for Color-Based Conditional Image Generation
by: Cohen, Nadav Z., et al.
Published: (2026)
by: Cohen, Nadav Z., et al.
Published: (2026)
Functionalization via Structure Completion and Motion Rectification
by: Zhao, Mingrui, et al.
Published: (2026)
by: Zhao, Mingrui, et al.
Published: (2026)
DiLightNet: Fine-grained Lighting Control for Diffusion-based Image Generation
by: Zeng, Chong, et al.
Published: (2024)
by: Zeng, Chong, et al.
Published: (2024)
TriTex: Learning Texture from a Single Mesh via Triplane Semantic Features
by: Cohen-Bar, Dana, et al.
Published: (2025)
by: Cohen-Bar, Dana, et al.
Published: (2025)
IP-Composer: Semantic Composition of Visual Concepts
by: Dorfman, Sara, et al.
Published: (2025)
by: Dorfman, Sara, et al.
Published: (2025)
RoMo: A Robust Solver for Full-body Unlabeled Optical Motion Capture
by: Pan, Xiaoyu, et al.
Published: (2024)
by: Pan, Xiaoyu, et al.
Published: (2024)
RoCoTex: A Robust Method for Consistent Texture Synthesis with Diffusion Models
by: Kim, Jangyeong, et al.
Published: (2024)
by: Kim, Jangyeong, et al.
Published: (2024)
Similar Items
-
In-Context Sync-LoRA for Portrait Video Editing
by: Polaczek, Sagi, et al.
Published: (2025) -
Sound Sparks Motion: Audio and Text Tuning for Video Editing
by: Razlighi, AmirHossein Naghi, et al.
Published: (2026) -
In-2-4D: Inbetweening from Two Single-View Images to 4D Generation
by: Nag, Sauradip, et al.
Published: (2025) -
Griffin: Generative Reference and Layout Guided Image Composition
by: Mikaeili, Aryan, et al.
Published: (2025) -
Nested Attention: Semantic-aware Attention Values for Concept Personalization
by: Patashnik, Or, et al.
Published: (2025)