LooseRoPE: Content-aware Attention Manipulation for Semantic Harmonization
Fuente:
arXiv
Saved in:
| Main Authors: | Sella, Etai, Baron, Yoav, Averbuch-Elor, Hadar, Cohen-Or, Daniel, Patashnik, Or |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prox-E: Fine-Grained 3D Shape Editing via Primitive-Based Abstractions
by: Sella, Etai, et al.
Published: (2026)
by: Sella, Etai, et al.
Published: (2026)
Spice-E : Structural Priors in 3D Diffusion using Cross-Entity Attention
by: Sella, Etai, et al.
Published: (2023)
by: Sella, Etai, et al.
Published: (2023)
Blended Point Cloud Diffusion for Localized Text-guided Shape Editing
by: Sella, Etai, et al.
Published: (2025)
by: Sella, Etai, et al.
Published: (2025)
ReNoise: Real Image Inversion Through Iterative Noising
by: Garibi, Daniel, et al.
Published: (2024)
by: Garibi, Daniel, et al.
Published: (2024)
Untwisting RoPE: Frequency Control for Shared Attention in DiTs
by: Mikaeili, Aryan, et al.
Published: (2026)
by: Mikaeili, Aryan, et al.
Published: (2026)
InstanceGen: Image Generation with Instance-level Instructions
by: Sella, Etai, et al.
Published: (2025)
by: Sella, Etai, et al.
Published: (2025)
Let it Snow! Animating 3D Gaussian Scenes with Dynamic Weather Effects via Physics-Guided Score Distillation
by: Fiebelman, Gal, et al.
Published: (2025)
by: Fiebelman, Gal, et al.
Published: (2025)
4‐LEGS: 4D Language Embedded Gaussian Splatting
by: Gal Fiebelman, et al.
Published: (2025)
by: Gal Fiebelman, et al.
Published: (2025)
4-LEGS: 4D Language Embedded Gaussian Splatting
by: Fiebelman, Gal, et al.
Published: (2024)
by: Fiebelman, Gal, et al.
Published: (2024)
Lang3D-XL: Language Embedded 3D Gaussians for Large-scale Scenes
by: Krakovsky, Shai, et al.
Published: (2025)
by: Krakovsky, Shai, et al.
Published: (2025)
Nested Attention: Semantic-aware Attention Values for Concept Personalization
by: Patashnik, Or, et al.
Published: (2025)
by: Patashnik, Or, et al.
Published: (2025)
HaLo-NeRF: Learning Geometry-Guided Semantics for Exploring Unconstrained Photo Collections
by: Dudai, Chen, et al.
Published: (2024)
by: Dudai, Chen, et al.
Published: (2024)
MeshOn: Intersection-Free Mesh-to-Mesh Composition
by: Kim, Hyunwoo, et al.
Published: (2026)
by: Kim, Hyunwoo, et al.
Published: (2026)
Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation
by: Dahary, Omer, et al.
Published: (2024)
by: Dahary, Omer, et al.
Published: (2024)
Supercharging Floorplan Localization with Semantic Rays
by: Grader, Yuval, et al.
Published: (2025)
by: Grader, Yuval, et al.
Published: (2025)
Emergent Visual-Semantic Hierarchies in Image-Text Representations
by: Alper, Morris, et al.
Published: (2024)
by: Alper, Morris, et al.
Published: (2024)
A Joint Study of Phrase Grounding and Task Performance in Vision and Language Models
by: Kojima, Noriyuki, et al.
Published: (2023)
by: Kojima, Noriyuki, et al.
Published: (2023)
Consolidating Attention Features for Multi-view Image Editing
by: Patashnik, Or, et al.
Published: (2024)
by: Patashnik, Or, et al.
Published: (2024)
TurboEdit: Text-Based Image Editing Using Few-Step Diffusion Models
by: Deutch, Gilad, et al.
Published: (2024)
by: Deutch, Gilad, et al.
Published: (2024)
In-Context Sync-LoRA for Portrait Video Editing
by: Polaczek, Sagi, et al.
Published: (2025)
by: Polaczek, Sagi, et al.
Published: (2025)
Tight Inversion: Image-Conditioned Inversion for Real Image Editing
by: Kadosh, Edo, et al.
Published: (2025)
by: Kadosh, Edo, et al.
Published: (2025)
Be Decisive: Noise-Induced Layouts for Multi-Subject Generation
by: Dahary, Omer, et al.
Published: (2025)
by: Dahary, Omer, et al.
Published: (2025)
SAEdit: Token-level control for continuous image editing via Sparse AutoEncoder
by: Kamenetsky, Ronen, et al.
Published: (2025)
by: Kamenetsky, Ronen, et al.
Published: (2025)
Raster2Seq: Polygon Sequence Generation for Floorplan Reconstruction
by: Phung, Hao, et al.
Published: (2026)
by: Phung, Hao, et al.
Published: (2026)
Kiki or Bouba? Sound Symbolism in Vision-and-Language Models
by: Alper, Morris, et al.
Published: (2023)
by: Alper, Morris, et al.
Published: (2023)
Image Generation from Contextually-Contradictory Prompts
by: Huberman, Saar, et al.
Published: (2025)
by: Huberman, Saar, et al.
Published: (2025)
LCM-Lookahead for Encoder-based Text-to-Image Personalization
by: Gal, Rinon, et al.
Published: (2024)
by: Gal, Rinon, et al.
Published: (2024)
A Neural-Network-Based Approach for Loose-Fitting Clothing
by: Jin, Yongxu, et al.
Published: (2024)
by: Jin, Yongxu, et al.
Published: (2024)
Stable Flow: Vital Layers for Training-Free Image Editing
by: Avrahami, Omri, et al.
Published: (2024)
by: Avrahami, Omri, et al.
Published: (2024)
IP-Composer: Semantic Composition of Visual Concepts
by: Dorfman, Sara, et al.
Published: (2025)
by: Dorfman, Sara, et al.
Published: (2025)
Continuous Control of Editing Models via Adaptive-Origin Guidance
by: Wolf, Alon, et al.
Published: (2026)
by: Wolf, Alon, et al.
Published: (2026)
VLM-Guided Adaptive Negative Prompting for Creative Generation
by: Golan, Shelly, et al.
Published: (2025)
by: Golan, Shelly, et al.
Published: (2025)
Real‐Time Per‐Garment Virtual Try‐On with Temporal Consistency for Loose‐Fitting Garments
by: Zaiqiang Wu, et al.
Published: (2025)
by: Zaiqiang Wu, et al.
Published: (2025)
Dynamic Concepts Personalization from Single Videos
by: Abdal, Rameen, et al.
Published: (2025)
by: Abdal, Rameen, et al.
Published: (2025)
Zero-Shot Dynamic Concept Personalization with Grid-Based LoRA
by: Abdal, Rameen, et al.
Published: (2025)
by: Abdal, Rameen, et al.
Published: (2025)
TriTex: Learning Texture from a Single Mesh via Triplane Semantic Features
by: Cohen-Bar, Dana, et al.
Published: (2025)
by: Cohen-Bar, Dana, et al.
Published: (2025)
JUST-DUB-IT: Video Dubbing via Joint Audio-Visual Diffusion
by: Chen, Anthony, et al.
Published: (2026)
by: Chen, Anthony, et al.
Published: (2026)
Object-level Visual Prompts for Compositional Image Generation
by: Parmar, Gaurav, et al.
Published: (2025)
by: Parmar, Gaurav, et al.
Published: (2025)
Composing People Together: Iterative Pose-Image Generation for Multi-Person Interaction Scenes
by: Peng, Wenxuan, et al.
Published: (2026)
by: Peng, Wenxuan, et al.
Published: (2026)
Dynamic Scene Understanding from Vision-Language Representations
by: Pruss, Shahaf, et al.
Published: (2025)
by: Pruss, Shahaf, et al.
Published: (2025)
Similar Items
-
Prox-E: Fine-Grained 3D Shape Editing via Primitive-Based Abstractions
by: Sella, Etai, et al.
Published: (2026) -
Spice-E : Structural Priors in 3D Diffusion using Cross-Entity Attention
by: Sella, Etai, et al.
Published: (2023) -
Blended Point Cloud Diffusion for Localized Text-guided Shape Editing
by: Sella, Etai, et al.
Published: (2025) -
ReNoise: Real Image Inversion Through Iterative Noising
by: Garibi, Daniel, et al.
Published: (2024) -
Untwisting RoPE: Frequency Control for Shared Attention in DiTs
by: Mikaeili, Aryan, et al.
Published: (2026)