BRAT: Bonus oRthogonAl Token for Architecture Agnostic Textual Inversion
Fuente:
arXiv
Guardado en:
| Autor principal: | Baker, James |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multi-Class Textual-Inversion Secretly Yields a Semantic-Agnostic Classifier
por: Wang, Kai, et al.
Publicado: (2024)
por: Wang, Kai, et al.
Publicado: (2024)
Zero-Shot Personalization of Objects via Textual Inversion
por: Roy, Aniket, et al.
Publicado: (2026)
por: Roy, Aniket, et al.
Publicado: (2026)
Model-Agnostic Human Preference Inversion in Diffusion Models
por: Kim, Jeeyung, et al.
Publicado: (2024)
por: Kim, Jeeyung, et al.
Publicado: (2024)
Textual Inversion and Self-supervised Refinement for Radiology Report Generation
por: Luo, Yuanjiang, et al.
Publicado: (2024)
por: Luo, Yuanjiang, et al.
Publicado: (2024)
CapeX: Category-Agnostic Pose Estimation from Textual Point Explanation
por: Rusanovsky, Matan, et al.
Publicado: (2024)
por: Rusanovsky, Matan, et al.
Publicado: (2024)
Textual Inversion for Efficient Adaptation of Open-Vocabulary Object Detectors Without Forgetting
por: Ruis, Frank, et al.
Publicado: (2025)
por: Ruis, Frank, et al.
Publicado: (2025)
Directional Textual Inversion for Personalized Text-to-Image Generation
por: Kim, Kunhee, et al.
Publicado: (2025)
por: Kim, Kunhee, et al.
Publicado: (2025)
Prompt-driven Transferable Adversarial Attack on Person Re-Identification with Attribute-aware Textual Inversion
por: Bian, Yuan, et al.
Publicado: (2025)
por: Bian, Yuan, et al.
Publicado: (2025)
Fine-grained Textual Inversion Network for Zero-Shot Composed Image Retrieval
por: Lin, Haoqiang, et al.
Publicado: (2025)
por: Lin, Haoqiang, et al.
Publicado: (2025)
SEA: Supervised Embedding Alignment for Token-Level Visual-Textual Integration in MLLMs
por: Yin, Yuanyang, et al.
Publicado: (2024)
por: Yin, Yuanyang, et al.
Publicado: (2024)
ID-EA: Identity-driven Text Enhancement and Adaptation with Textual Inversion for Personalized Text-to-Image Generation
por: Jin, Hyun-Jun, et al.
Publicado: (2025)
por: Jin, Hyun-Jun, et al.
Publicado: (2025)
Mull-Tokens: Modality-Agnostic Latent Thinking
por: Ray, Arijit, et al.
Publicado: (2025)
por: Ray, Arijit, et al.
Publicado: (2025)
FADE: A Task-Agnostic Upsampling Operator for Encoder-Decoder Architectures
por: Lu, Hao, et al.
Publicado: (2024)
por: Lu, Hao, et al.
Publicado: (2024)
iSEARLE: Improving Textual Inversion for Zero-Shot Composed Image Retrieval
por: Agnolucci, Lorenzo, et al.
Publicado: (2024)
por: Agnolucci, Lorenzo, et al.
Publicado: (2024)
Difference Inversion: Interpolate and Isolate the Difference with Token Consistency for Image Analogy Generation
por: Kim, Hyunsoo, et al.
Publicado: (2025)
por: Kim, Hyunsoo, et al.
Publicado: (2025)
MiniGPT4-Video: Advancing Multimodal LLMs for Video Understanding with Interleaved Visual-Textual Tokens
por: Ataallah, Kirolos, et al.
Publicado: (2024)
por: Ataallah, Kirolos, et al.
Publicado: (2024)
Medical diffusion on a budget: Textual Inversion for medical image generation
por: de Wilde, Bram, et al.
Publicado: (2023)
por: de Wilde, Bram, et al.
Publicado: (2023)
Style Ambiguity Loss Using CLIP
por: Baker, James
Publicado: (2024)
por: Baker, James
Publicado: (2024)
MONKEY: Masking ON KEY-Value Activation Adapter for Personalization
por: Baker, James
Publicado: (2025)
por: Baker, James
Publicado: (2025)
MTFusion: Reconstructing Any 3D Object from Single Image Using Multi-word Textual Inversion
por: Liu, Yu, et al.
Publicado: (2024)
por: Liu, Yu, et al.
Publicado: (2024)
Viewpoint Textual Inversion: Discovering Scene Representations and 3D View Control in 2D Diffusion Models
por: Burgess, James, et al.
Publicado: (2023)
por: Burgess, James, et al.
Publicado: (2023)
Explaining Chest X-ray Pathology Models using Textual Concepts
por: Sadashivaiah, Vijay, et al.
Publicado: (2024)
por: Sadashivaiah, Vijay, et al.
Publicado: (2024)
A Cognitive Process-Inspired Architecture for Subject-Agnostic Brain Visual Decoding
por: Lu, Jingyu, et al.
Publicado: (2025)
por: Lu, Jingyu, et al.
Publicado: (2025)
Using Multimodal Foundation Models and Clustering for Improved Style Ambiguity Loss
por: Baker, James
Publicado: (2024)
por: Baker, James
Publicado: (2024)
Visual Prompt-Agnostic Evolution
por: Wang, Junze, et al.
Publicado: (2026)
por: Wang, Junze, et al.
Publicado: (2026)
Joint Architecture-Token-Bitwidth Multi-Axis Optimization of Vision Transformers for Semiconductor IC Packaging
por: Nguyen, Phat, et al.
Publicado: (2026)
por: Nguyen, Phat, et al.
Publicado: (2026)
Advancing Textual Prompt Learning with Anchored Attributes
por: Li, Zheng, et al.
Publicado: (2024)
por: Li, Zheng, et al.
Publicado: (2024)
Visual Textualization for Image Prompted Object Detection
por: Wu, Yongjian, et al.
Publicado: (2025)
por: Wu, Yongjian, et al.
Publicado: (2025)
Dual-Schedule Inversion: Training- and Tuning-Free Inversion for Real Image Editing
por: Huang, Jiancheng, et al.
Publicado: (2024)
por: Huang, Jiancheng, et al.
Publicado: (2024)
SimInversion: A Simple Framework for Inversion-Based Text-to-Image Editing
por: Qian, Qi, et al.
Publicado: (2024)
por: Qian, Qi, et al.
Publicado: (2024)
Oscillation Inversion: Understand the structure of Large Flow Model through the Lens of Inversion Method
por: Zheng, Yan, et al.
Publicado: (2024)
por: Zheng, Yan, et al.
Publicado: (2024)
Negative-prompt Inversion: Fast Image Inversion for Editing with Text-guided Diffusion Models
por: Miyake, Daiki, et al.
Publicado: (2023)
por: Miyake, Daiki, et al.
Publicado: (2023)
Architecture-Agnostic Modality-Isolated Gated Fusion for Robust Multi-Modal Prostate MRI Segmentation
por: Shu, Yongbo, et al.
Publicado: (2026)
por: Shu, Yongbo, et al.
Publicado: (2026)
Aligning Actions and Walking to LLM-Generated Textual Descriptions
por: Chivereanu, Radu, et al.
Publicado: (2024)
por: Chivereanu, Radu, et al.
Publicado: (2024)
ReGround: Improving Textual and Spatial Grounding at No Cost
por: Lee, Phillip Y., et al.
Publicado: (2024)
por: Lee, Phillip Y., et al.
Publicado: (2024)
OrienText: Surface Oriented Textual Image Generation
por: Paliwal, Shubham Singh, et al.
Publicado: (2025)
por: Paliwal, Shubham Singh, et al.
Publicado: (2025)
Precise Parameter Localization for Textual Generation in Diffusion Models
por: Staniszewski, Łukasz, et al.
Publicado: (2025)
por: Staniszewski, Łukasz, et al.
Publicado: (2025)
EDITS: Enhancing Dataset Distillation with Implicit Textual Semantics
por: Xia, Qianxin, et al.
Publicado: (2025)
por: Xia, Qianxin, et al.
Publicado: (2025)
Visual and Textual Prompts in VLLMs for Enhancing Emotion Recognition
por: Wang, Zhifeng, et al.
Publicado: (2025)
por: Wang, Zhifeng, et al.
Publicado: (2025)
Medal S: Spatio-Textual Prompt Model for Medical Segmentation
por: Shi, Pengcheng, et al.
Publicado: (2025)
por: Shi, Pengcheng, et al.
Publicado: (2025)
Ejemplares similares
-
Multi-Class Textual-Inversion Secretly Yields a Semantic-Agnostic Classifier
por: Wang, Kai, et al.
Publicado: (2024) -
Zero-Shot Personalization of Objects via Textual Inversion
por: Roy, Aniket, et al.
Publicado: (2026) -
Model-Agnostic Human Preference Inversion in Diffusion Models
por: Kim, Jeeyung, et al.
Publicado: (2024) -
Textual Inversion and Self-supervised Refinement for Radiology Report Generation
por: Luo, Yuanjiang, et al.
Publicado: (2024) -
CapeX: Category-Agnostic Pose Estimation from Textual Point Explanation
por: Rusanovsky, Matan, et al.
Publicado: (2024)