Mapping the Mind of an Instruction-based Image Editing using SMILE
Fuente:
arXiv
Saved in:
| Main Authors: | Dehghani, Zeinab, Aslansefat, Koorosh, Khan, Adil, Rivera, Adín Ramírez, George, Franky, Khalid, Muhammad |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Explaining Large Language Models with gSMILE
by: Dehghani, Zeinab, et al.
Published: (2025)
by: Dehghani, Zeinab, et al.
Published: (2025)
CoEditor++: Instruction-based Visual Editing via Cognitive Reasoning
by: Ni, Minheng, et al.
Published: (2026)
by: Ni, Minheng, et al.
Published: (2026)
Contrastive-SDXL: Annotation-Preserving Night-Time Augmentation for Pedestrian Detection
by: George, Franky, et al.
Published: (2026)
by: George, Franky, et al.
Published: (2026)
VideoMap: Supporting Video Editing Exploration, Brainstorming, and Prototyping in the Latent Space
by: Lin, David Chuan-En, et al.
Published: (2022)
by: Lin, David Chuan-En, et al.
Published: (2022)
FreeDrag: Feature Dragging for Reliable Point-based Image Editing
by: Ling, Pengyang, et al.
Published: (2023)
by: Ling, Pengyang, et al.
Published: (2023)
Putting Our Minds Together: Iterative Exploration for Collaborative Mind Mapping
by: Yang, Ying, et al.
Published: (2024)
by: Yang, Ying, et al.
Published: (2024)
PromptArtisan: Multi-instruction Image Editing in Single Pass with Complete Attention Control
by: Swami, Kunal, et al.
Published: (2025)
by: Swami, Kunal, et al.
Published: (2025)
LEDITS++: Limitless Image Editing using Text-to-Image Models
by: Brack, Manuel, et al.
Published: (2023)
by: Brack, Manuel, et al.
Published: (2023)
InstructEdit: Instruction-based Knowledge Editing for Large Language Models
by: Zhang, Ningyu, et al.
Published: (2024)
by: Zhang, Ningyu, et al.
Published: (2024)
Instruction-Guided Editing Controls for Images and Multimedia: A Survey in LLM era
by: Nguyen, Thanh Tam, et al.
Published: (2024)
by: Nguyen, Thanh Tam, et al.
Published: (2024)
ClassMind: Scaling Classroom Observation and Instructional Feedback with Multimodal AI
by: Qu, Ao, et al.
Published: (2025)
by: Qu, Ao, et al.
Published: (2025)
Psychlysis: Towards the Creation of a Questionnaire-based Machine Learning Tool to Analyze States of Mind
by: Jani, Hemakshi, et al.
Published: (2025)
by: Jani, Hemakshi, et al.
Published: (2025)
"A Great Start, But...": Evaluating LLM-Generated Mind Maps for Information Mapping in Video-Based Design
by: He, Tianhao, et al.
Published: (2025)
by: He, Tianhao, et al.
Published: (2025)
Mitigating Individual Skin Tone Bias in Skin Lesion Classification through Distribution-Aware Reweighting
by: Paxton, Kuniko, et al.
Published: (2025)
by: Paxton, Kuniko, et al.
Published: (2025)
HIVE: Harnessing Human Feedback for Instructional Visual Editing
by: Zhang, Shu, et al.
Published: (2023)
by: Zhang, Shu, et al.
Published: (2023)
Prompt Engineering for Large Language Model-assisted Inductive Thematic Analysis
by: Khalid, Muhammad Talal, et al.
Published: (2025)
by: Khalid, Muhammad Talal, et al.
Published: (2025)
See Through Their Minds: Learning Transferable Neural Representation from Cross-Subject fMRI
by: Liu, Yulong, et al.
Published: (2024)
by: Liu, Yulong, et al.
Published: (2024)
Connecting Dreams with Visual Brainstorming Instruction
by: Sun, Yasheng, et al.
Published: (2024)
by: Sun, Yasheng, et al.
Published: (2024)
MindSpeech: Continuous Imagined Speech Decoding using High-Density fNIRS and Prompt Tuning for Advanced Human-AI Interaction
by: Zhang, Suyi, et al.
Published: (2024)
by: Zhang, Suyi, et al.
Published: (2024)
Vitron: A Unified Pixel-level Vision LLM for Understanding, Generating, Segmenting, Editing
by: Fei, Hao, et al.
Published: (2024)
by: Fei, Hao, et al.
Published: (2024)
Mind the Dark: A Gamified Exploration of Deceptive Design Awareness for Children in the Digital Age
by: Khan, Noverah, et al.
Published: (2025)
by: Khan, Noverah, et al.
Published: (2025)
Dreamcrafter: Immersive Editing of 3D Radiance Fields Through Flexible, Generative Inputs and Outputs
by: Vachha, Cyrus, et al.
Published: (2025)
by: Vachha, Cyrus, et al.
Published: (2025)
Prompting with Sign Parameters for Low-resource Sign Language Instruction Generation
by: Tariquzzaman, Md, et al.
Published: (2025)
by: Tariquzzaman, Md, et al.
Published: (2025)
Point and Instruct: Enabling Precise Image Editing by Unifying Direct Manipulation and Text Instructions
by: Helbling, Alec, et al.
Published: (2024)
by: Helbling, Alec, et al.
Published: (2024)
Regressor-Guided Generative Image Editing Balances User Emotions to Reduce Time Spent Online
by: Gebhardt, Christoph, et al.
Published: (2025)
by: Gebhardt, Christoph, et al.
Published: (2025)
Editing Physiological Signals in Videos Using Latent Representations
by: Zhou, Tianwen, et al.
Published: (2025)
by: Zhou, Tianwen, et al.
Published: (2025)
CADReasoner: Iterative Program Editing for CAD Reverse Engineering
by: Kabisov, Soslan, et al.
Published: (2026)
by: Kabisov, Soslan, et al.
Published: (2026)
mEBAL2 Database and Benchmark: Image-based Multispectral Eyeblink Detection
by: Daza, Roberto, et al.
Published: (2023)
by: Daza, Roberto, et al.
Published: (2023)
MindCine: Multimodal EEG-to-Video Reconstruction with Large-Scale Pretrained Models
by: Zhou, Tian-Yi, et al.
Published: (2026)
by: Zhou, Tian-Yi, et al.
Published: (2026)
Has the Virtualization of the Face Changed Facial Perception? A Study of the Impact of Photo Editing and Augmented Reality on Facial Perception
by: Conwill, Louisa, et al.
Published: (2023)
by: Conwill, Louisa, et al.
Published: (2023)
A Survey on Conceptual model of Enterprise ontology
by: Rajabi, Zeinab, et al.
Published: (2025)
by: Rajabi, Zeinab, et al.
Published: (2025)
A Monocular SLAM-based Multi-User Positioning System with Image Occlusion in Augmented Reality
by: Lien, Wei-Hsiang, et al.
Published: (2024)
by: Lien, Wei-Hsiang, et al.
Published: (2024)
Is Medieval Distant Viewing Possible? : Extending and Enriching Annotation of Legacy Image Collections using Visual Analytics
by: Meinecke, Christofer, et al.
Published: (2022)
by: Meinecke, Christofer, et al.
Published: (2022)
Supervised Contrastive Learning for Ordinal Engagement Measurement
by: Safa, Sadaf, et al.
Published: (2025)
by: Safa, Sadaf, et al.
Published: (2025)
Deep Learning in Mild Cognitive Impairment Diagnosis using Eye Movements and Image Content in Visual Memory Tasks
by: Rocha, Tomás Silva Santos, et al.
Published: (2025)
by: Rocha, Tomás Silva Santos, et al.
Published: (2025)
Simplifying Integration of Custom Controllers in Exergames
by: Khan, Hassan Ali, et al.
Published: (2024)
by: Khan, Hassan Ali, et al.
Published: (2024)
MindGPT: Advancing Human-AI Interaction with Non-Invasive fNIRS-Based Imagined Speech Decoding
by: Zhang, Suyi, et al.
Published: (2024)
by: Zhang, Suyi, et al.
Published: (2024)
Sparse Activation Editing for Reliable Instruction Following in Narratives
by: Zhao, Runcong, et al.
Published: (2025)
by: Zhao, Runcong, et al.
Published: (2025)
Exploring Text-based Realistic Building Facades Editing Applicaiton
by: Wang, Jing, et al.
Published: (2024)
by: Wang, Jing, et al.
Published: (2024)
MindCross: Fast New Subject Adaptation with Limited Data for Cross-subject Video Reconstruction from Brain Signals
by: Liu, Xuan-Hao, et al.
Published: (2025)
by: Liu, Xuan-Hao, et al.
Published: (2025)
Similar Items
-
Explaining Large Language Models with gSMILE
by: Dehghani, Zeinab, et al.
Published: (2025) -
CoEditor++: Instruction-based Visual Editing via Cognitive Reasoning
by: Ni, Minheng, et al.
Published: (2026) -
Contrastive-SDXL: Annotation-Preserving Night-Time Augmentation for Pedestrian Detection
by: George, Franky, et al.
Published: (2026) -
VideoMap: Supporting Video Editing Exploration, Brainstorming, and Prototyping in the Latent Space
by: Lin, David Chuan-En, et al.
Published: (2022) -
FreeDrag: Feature Dragging for Reliable Point-based Image Editing
by: Ling, Pengyang, et al.
Published: (2023)