MPN: Leveraging Multilingual Patch Neuron for Cross-lingual Model Editing
Fuente:
arXiv
Guardado en:
| Autores principales: | Si, Nianwen, Zhang, Hao, Zhang, Weiqiang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Leveraging LLMs for On-the-Fly Instruction Guided Image Editing
por: Santos, Rodrigo, et al.
Publicado: (2024)
por: Santos, Rodrigo, et al.
Publicado: (2024)
CLAIM: Mitigating Multilingual Object Hallucination in Large Vision-Language Models with Cross-Lingual Attention Intervention
por: Ye, Zekai, et al.
Publicado: (2025)
por: Ye, Zekai, et al.
Publicado: (2025)
MULTITEXTEDIT: Benchmarking Cross-Lingual Degradation in Text-in-Image Editing
por: Cheng, Liwei, et al.
Publicado: (2026)
por: Cheng, Liwei, et al.
Publicado: (2026)
Cross-lingual Editing in Multilingual Language Models
por: Beniwal, Himanshu, et al.
Publicado: (2024)
por: Beniwal, Himanshu, et al.
Publicado: (2024)
ETCHR: Editing To Clarify and Harness Reasoning
por: Zhang, Beichen, et al.
Publicado: (2026)
por: Zhang, Beichen, et al.
Publicado: (2026)
Error-Driven Scene Editing for 3D Grounding in Large Language Models
por: Zhang, Yue, et al.
Publicado: (2025)
por: Zhang, Yue, et al.
Publicado: (2025)
Mitigating Multilingual Hallucination in Large Vision-Language Models
por: Qu, Xiaoye, et al.
Publicado: (2024)
por: Qu, Xiaoye, et al.
Publicado: (2024)
VisionGraph: Leveraging Large Multimodal Models for Graph Theory Problems in Visual Context
por: Li, Yunxin, et al.
Publicado: (2024)
por: Li, Yunxin, et al.
Publicado: (2024)
MoKus: Leveraging Cross-Modal Knowledge Transfer for Knowledge-Aware Concept Customization
por: Zhu, Chenyang, et al.
Publicado: (2026)
por: Zhu, Chenyang, et al.
Publicado: (2026)
Cross-modal Information Flow in Multimodal Large Language Models
por: Zhang, Zhi, et al.
Publicado: (2024)
por: Zhang, Zhi, et al.
Publicado: (2024)
The Mirage of Performance Gains: Why Contrastive Decoding Fails to Mitigate Object Hallucinations in MLLMs?
por: Yin, Hao, et al.
Publicado: (2025)
por: Yin, Hao, et al.
Publicado: (2025)
Interpretable and Reliable Detection of AI-Generated Images via Grounded Reasoning in MLLMs
por: Ji, Yikun, et al.
Publicado: (2025)
por: Ji, Yikun, et al.
Publicado: (2025)
VEGGIE: Instructional Editing and Reasoning Video Concepts with Grounded Generation
por: Yu, Shoubin, et al.
Publicado: (2025)
por: Yu, Shoubin, et al.
Publicado: (2025)
Bayesian Optimization for Controlled Image Editing via LLMs
por: Cai, Chengkun, et al.
Publicado: (2025)
por: Cai, Chengkun, et al.
Publicado: (2025)
POLYCHARTQA: Benchmarking Large Vision-Language Models with Multilingual Chart Question Answering
por: Xu, Yichen, et al.
Publicado: (2025)
por: Xu, Yichen, et al.
Publicado: (2025)
Bifrost-1: Bridging Multimodal LLMs and Diffusion Models with Patch-level CLIP Latents
por: Lin, Han, et al.
Publicado: (2025)
por: Lin, Han, et al.
Publicado: (2025)
MagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing
por: Zhang, Kai, et al.
Publicado: (2023)
por: Zhang, Kai, et al.
Publicado: (2023)
VLKEB: A Large Vision-Language Model Knowledge Editing Benchmark
por: Huang, Han, et al.
Publicado: (2024)
por: Huang, Han, et al.
Publicado: (2024)
XL-HeadTags: Leveraging Multimodal Retrieval Augmentation for the Multilingual Generation of News Headlines and Tags
por: Shohan, Faisal Tareque, et al.
Publicado: (2024)
por: Shohan, Faisal Tareque, et al.
Publicado: (2024)
EditReward: A Human-Aligned Reward Model for Instruction-Guided Image Editing
por: Wu, Keming, et al.
Publicado: (2025)
por: Wu, Keming, et al.
Publicado: (2025)
Image-Text Relation Prediction for Multilingual Tweets
por: Rikters, Matīss, et al.
Publicado: (2025)
por: Rikters, Matīss, et al.
Publicado: (2025)
Leveraging Open Knowledge for Advancing Task Expertise in Large Language Models
por: Yang, Yuncheng, et al.
Publicado: (2024)
por: Yang, Yuncheng, et al.
Publicado: (2024)
GRR-CoCa: Leveraging LLM Mechanisms in Multimodal Model Architectures
por: Patock, Jake R., et al.
Publicado: (2025)
por: Patock, Jake R., et al.
Publicado: (2025)
DART: Leveraging Multi-Agent Disagreement for Tool Recruitment in Multimodal Reasoning
por: Sivakumaran, Nithin, et al.
Publicado: (2025)
por: Sivakumaran, Nithin, et al.
Publicado: (2025)
MAMI: Multi-Attentional Mutual-Information for Long Sequence Neuron Captioning
por: Fauzulhaq, Alfirsa Damasyifa, et al.
Publicado: (2024)
por: Fauzulhaq, Alfirsa Damasyifa, et al.
Publicado: (2024)
VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effects
por: Gao, Xiangbo, et al.
Publicado: (2026)
por: Gao, Xiangbo, et al.
Publicado: (2026)
Law of the Weakest Link: Cross Capabilities of Large Language Models
por: Zhong, Ming, et al.
Publicado: (2024)
por: Zhong, Ming, et al.
Publicado: (2024)
MENTOR: Efficient Multimodal-Conditioned Tuning for Autoregressive Vision Generation Models
por: Zhao, Haozhe, et al.
Publicado: (2025)
por: Zhao, Haozhe, et al.
Publicado: (2025)
Uni3D-LLM: Unifying Point Cloud Perception, Generation and Editing with Large Language Models
por: Liu, Dingning, et al.
Publicado: (2024)
por: Liu, Dingning, et al.
Publicado: (2024)
CoPE-VideoLM: Leveraging Codec Primitives For Efficient Video Language Modeling
por: Sarkar, Sayan Deb, et al.
Publicado: (2026)
por: Sarkar, Sayan Deb, et al.
Publicado: (2026)
NeoBabel: A Multilingual Open Tower for Visual Generation
por: Derakhshani, Mohammad Mahdi, et al.
Publicado: (2025)
por: Derakhshani, Mohammad Mahdi, et al.
Publicado: (2025)
Chitrakshara: A Large Multilingual Multimodal Dataset for Indian languages
por: Khan, Shaharukh, et al.
Publicado: (2026)
por: Khan, Shaharukh, et al.
Publicado: (2026)
Concept Lancet: Image Editing with Compositional Representation Transplant
por: Luo, Jinqi, et al.
Publicado: (2025)
por: Luo, Jinqi, et al.
Publicado: (2025)
ZINA: Multimodal Fine-grained Hallucination Detection and Editing
por: Wada, Yuiga, et al.
Publicado: (2025)
por: Wada, Yuiga, et al.
Publicado: (2025)
Multimodal LLM Enhanced Cross-lingual Cross-modal Retrieval
por: Wang, Yabing, et al.
Publicado: (2024)
por: Wang, Yabing, et al.
Publicado: (2024)
Beyond Translation: Cross-Cultural Meme Transcreation with Vision-Language Models
por: Zhao, Yuming, et al.
Publicado: (2026)
por: Zhao, Yuming, et al.
Publicado: (2026)
Annotations on a Budget: Leveraging Geo-Data Similarity to Balance Model Performance and Annotation Cost
por: Ignat, Oana, et al.
Publicado: (2024)
por: Ignat, Oana, et al.
Publicado: (2024)
MotionEdit: Benchmarking and Learning Motion-Centric Image Editing
por: Wan, Yixin, et al.
Publicado: (2025)
por: Wan, Yixin, et al.
Publicado: (2025)
Cross-Modal Adapter for Vision-Language Retrieval
por: Jiang, Haojun, et al.
Publicado: (2022)
por: Jiang, Haojun, et al.
Publicado: (2022)
HNC: Leveraging Hard Negative Captions towards Models with Fine-Grained Visual-Linguistic Comprehension Capabilities
por: Dönmez, Esra, et al.
Publicado: (2026)
por: Dönmez, Esra, et al.
Publicado: (2026)
Ejemplares similares
-
Leveraging LLMs for On-the-Fly Instruction Guided Image Editing
por: Santos, Rodrigo, et al.
Publicado: (2024) -
CLAIM: Mitigating Multilingual Object Hallucination in Large Vision-Language Models with Cross-Lingual Attention Intervention
por: Ye, Zekai, et al.
Publicado: (2025) -
MULTITEXTEDIT: Benchmarking Cross-Lingual Degradation in Text-in-Image Editing
por: Cheng, Liwei, et al.
Publicado: (2026) -
Cross-lingual Editing in Multilingual Language Models
por: Beniwal, Himanshu, et al.
Publicado: (2024) -
ETCHR: Editing To Clarify and Harness Reasoning
por: Zhang, Beichen, et al.
Publicado: (2026)