ReFACT: Updating Text-to-Image Models by Editing the Text Encoder
Fuente:
arXiv
Saved in:
| Main Authors: | Arad, Dana, Orgad, Hadas, Belinkov, Yonatan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Diffusion Lens: Interpreting Text Encoders in Text-to-Image Pipelines
by: Toker, Michael, et al.
Published: (2024)
by: Toker, Michael, et al.
Published: (2024)
Position-aware Automatic Circuit Discovery
by: Haklay, Tal, et al.
Published: (2025)
by: Haklay, Tal, et al.
Published: (2025)
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
by: Orgad, Hadas, et al.
Published: (2024)
by: Orgad, Hadas, et al.
Published: (2024)
Same Task, Different Circuits: Disentangling Modality-Specific Mechanisms in VLMs
by: Nikankin, Yaniv, et al.
Published: (2025)
by: Nikankin, Yaniv, et al.
Published: (2025)
STRICT: Stress Test of Rendering Images Containing Text
by: Zhang, Tianyu, et al.
Published: (2025)
by: Zhang, Tianyu, et al.
Published: (2025)
jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
by: Koukounas, Andreas, et al.
Published: (2024)
by: Koukounas, Andreas, et al.
Published: (2024)
Enhancing OCR for Sino-Vietnamese Language Processing via Fine-tuned PaddleOCRv5
by: Nguyen, Minh Hoang, et al.
Published: (2025)
by: Nguyen, Minh Hoang, et al.
Published: (2025)
Multiplication in Multimodal LLMs: Computation with Text, Image, and Audio Inputs
by: Balter, Samuel G., et al.
Published: (2026)
by: Balter, Samuel G., et al.
Published: (2026)
Jina CLIP: Your CLIP Model Is Also Your Text Retriever
by: Koukounas, Andreas, et al.
Published: (2024)
by: Koukounas, Andreas, et al.
Published: (2024)
Profiling German Text Simplification with Interpretable Model-Fingerprints
by: Klöser, Lars, et al.
Published: (2026)
by: Klöser, Lars, et al.
Published: (2026)
A Survey of Text Watermarking in the Era of Large Language Models
by: Liu, Aiwei, et al.
Published: (2023)
by: Liu, Aiwei, et al.
Published: (2023)
Accurate Retraining-free Pruning for Pretrained Encoder-based Language Models
by: Park, Seungcheol, et al.
Published: (2023)
by: Park, Seungcheol, et al.
Published: (2023)
Contrasting Linguistic Patterns in Human and LLM-Generated News Text
by: Muñoz-Ortiz, Alberto, et al.
Published: (2023)
by: Muñoz-Ortiz, Alberto, et al.
Published: (2023)
jina-vlm: Small Multilingual Vision Language Model
by: Koukounas, Andreas, et al.
Published: (2025)
by: Koukounas, Andreas, et al.
Published: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
by: Ashuach, Tomer, et al.
Published: (2025)
by: Ashuach, Tomer, et al.
Published: (2025)
Beyond RNNs: Benchmarking Attention-Based Image Captioning Models
by: Yanambakkam, Hemanth Teja, et al.
Published: (2025)
by: Yanambakkam, Hemanth Teja, et al.
Published: (2025)
Prior-based Noisy Text Data Filtering: Fast and Strong Alternative For Perplexity
by: Seo, Yeongbin, et al.
Published: (2025)
by: Seo, Yeongbin, et al.
Published: (2025)
The MSR-Video to Text Dataset with Clean Annotations
by: Chen, Haoran, et al.
Published: (2021)
by: Chen, Haoran, et al.
Published: (2021)
Exploiting Pre-trained Encoder-Decoder Transformers for Sequence-to-Sequence Constituent Parsing
by: Fernández-González, Daniel, et al.
Published: (2026)
by: Fernández-González, Daniel, et al.
Published: (2026)
Controlled Automatic Task-Specific Synthetic Data Generation for Hallucination Detection
by: Xie, Yong, et al.
Published: (2024)
by: Xie, Yong, et al.
Published: (2024)
MemeIntel: Explainable Detection of Propagandistic and Hateful Memes
by: Kmainasi, Mohamed Bayan, et al.
Published: (2025)
by: Kmainasi, Mohamed Bayan, et al.
Published: (2025)
ArMeme: Propagandistic Content in Arabic Memes
by: Alam, Firoj, et al.
Published: (2024)
by: Alam, Firoj, et al.
Published: (2024)
BitMar: Low-Bit Multimodal Fusion with Episodic Memory for Edge Devices
by: Aman, Euhid, et al.
Published: (2025)
by: Aman, Euhid, et al.
Published: (2025)
RONA: Pragmatically Diverse Image Captioning with Coherence Relations
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
Bonsai: Interpretable Tree-Adaptive Grounded Reasoning
by: Sanders, Kate, et al.
Published: (2025)
by: Sanders, Kate, et al.
Published: (2025)
Continuous Latent Diffusion Language Model
by: Guo, Hongcan, et al.
Published: (2026)
by: Guo, Hongcan, et al.
Published: (2026)
SUGARCREPE++ Dataset: Vision-Language Model Sensitivity to Semantic and Lexical Alterations
by: Dumpala, Sri Harsha, et al.
Published: (2024)
by: Dumpala, Sri Harsha, et al.
Published: (2024)
Explainable Image Captioning using CNN- CNN architecture and Hierarchical Attention
by: Mohan, Rishi Kesav, et al.
Published: (2024)
by: Mohan, Rishi Kesav, et al.
Published: (2024)
A Comparative Analysis of Noise Reduction Methods in Sentiment Analysis on Noisy Bangla Texts
by: Elahi, Kazi Toufique, et al.
Published: (2024)
by: Elahi, Kazi Toufique, et al.
Published: (2024)
Mechanisms of Prompt-Induced Hallucination in Vision-Language Models
by: Rudman, William, et al.
Published: (2026)
by: Rudman, William, et al.
Published: (2026)
IRONIC: Coherence-Aware Reasoning Chains for Multi-Modal Sarcasm Detection
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
Arithmetic Without Algorithms: Language Models Solve Math With a Bag of Heuristics
by: Nikankin, Yaniv, et al.
Published: (2024)
by: Nikankin, Yaniv, et al.
Published: (2024)
Technical Report on the Pangram AI-Generated Text Classifier
by: Emi, Bradley, et al.
Published: (2024)
by: Emi, Bradley, et al.
Published: (2024)
Understanding the Effects of RLHF on the Quality and Detectability of LLM-Generated Texts
by: Xu, Beining, et al.
Published: (2025)
by: Xu, Beining, et al.
Published: (2025)
NOTAI.AI: Explainable Detection of Machine-Generated Text via Curvature and Feature Attribution
by: Breneur, Oleksandr Marchenko, et al.
Published: (2026)
by: Breneur, Oleksandr Marchenko, et al.
Published: (2026)
Co-NAML-LSTUR: A Combined Model with Attentive Multi-View Learning and Long- and Short-term User Representations for News Recommendation
by: Nguyen, Minh Hoang, et al.
Published: (2025)
by: Nguyen, Minh Hoang, et al.
Published: (2025)
Raw Text is All you Need: Knowledge-intensive Multi-turn Instruction Tuning for Large Language Model
by: Hou, Xia, et al.
Published: (2024)
by: Hou, Xia, et al.
Published: (2024)
Can Large Language Models (or Humans) Disentangle Text?
by: de Pieuchon, Nicolas Audinet, et al.
Published: (2024)
by: de Pieuchon, Nicolas Audinet, et al.
Published: (2024)
Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark
by: Cheng, Ziming, et al.
Published: (2025)
by: Cheng, Ziming, et al.
Published: (2025)
Embedding Compression via Spherical Coordinates
by: Xiao, Han
Published: (2026)
by: Xiao, Han
Published: (2026)
Similar Items
-
Diffusion Lens: Interpreting Text Encoders in Text-to-Image Pipelines
by: Toker, Michael, et al.
Published: (2024) -
Position-aware Automatic Circuit Discovery
by: Haklay, Tal, et al.
Published: (2025) -
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
by: Orgad, Hadas, et al.
Published: (2024) -
Same Task, Different Circuits: Disentangling Modality-Specific Mechanisms in VLMs
by: Nikankin, Yaniv, et al.
Published: (2025) -
STRICT: Stress Test of Rendering Images Containing Text
by: Zhang, Tianyu, et al.
Published: (2025)