MPN: Leveraging Multilingual Patch Neuron for Cross-lingual Model Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Si, Nianwen, Zhang, Hao, Zhang, Weiqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Leveraging LLMs for On-the-Fly Instruction Guided Image Editing
von: Santos, Rodrigo, et al.
Veröffentlicht: (2024)
von: Santos, Rodrigo, et al.
Veröffentlicht: (2024)
CLAIM: Mitigating Multilingual Object Hallucination in Large Vision-Language Models with Cross-Lingual Attention Intervention
von: Ye, Zekai, et al.
Veröffentlicht: (2025)
von: Ye, Zekai, et al.
Veröffentlicht: (2025)
MULTITEXTEDIT: Benchmarking Cross-Lingual Degradation in Text-in-Image Editing
von: Cheng, Liwei, et al.
Veröffentlicht: (2026)
von: Cheng, Liwei, et al.
Veröffentlicht: (2026)
Cross-lingual Editing in Multilingual Language Models
von: Beniwal, Himanshu, et al.
Veröffentlicht: (2024)
von: Beniwal, Himanshu, et al.
Veröffentlicht: (2024)
ETCHR: Editing To Clarify and Harness Reasoning
von: Zhang, Beichen, et al.
Veröffentlicht: (2026)
von: Zhang, Beichen, et al.
Veröffentlicht: (2026)
Error-Driven Scene Editing for 3D Grounding in Large Language Models
von: Zhang, Yue, et al.
Veröffentlicht: (2025)
von: Zhang, Yue, et al.
Veröffentlicht: (2025)
Mitigating Multilingual Hallucination in Large Vision-Language Models
von: Qu, Xiaoye, et al.
Veröffentlicht: (2024)
von: Qu, Xiaoye, et al.
Veröffentlicht: (2024)
VisionGraph: Leveraging Large Multimodal Models for Graph Theory Problems in Visual Context
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
MoKus: Leveraging Cross-Modal Knowledge Transfer for Knowledge-Aware Concept Customization
von: Zhu, Chenyang, et al.
Veröffentlicht: (2026)
von: Zhu, Chenyang, et al.
Veröffentlicht: (2026)
Cross-modal Information Flow in Multimodal Large Language Models
von: Zhang, Zhi, et al.
Veröffentlicht: (2024)
von: Zhang, Zhi, et al.
Veröffentlicht: (2024)
The Mirage of Performance Gains: Why Contrastive Decoding Fails to Mitigate Object Hallucinations in MLLMs?
von: Yin, Hao, et al.
Veröffentlicht: (2025)
von: Yin, Hao, et al.
Veröffentlicht: (2025)
Interpretable and Reliable Detection of AI-Generated Images via Grounded Reasoning in MLLMs
von: Ji, Yikun, et al.
Veröffentlicht: (2025)
von: Ji, Yikun, et al.
Veröffentlicht: (2025)
VEGGIE: Instructional Editing and Reasoning Video Concepts with Grounded Generation
von: Yu, Shoubin, et al.
Veröffentlicht: (2025)
von: Yu, Shoubin, et al.
Veröffentlicht: (2025)
Bayesian Optimization for Controlled Image Editing via LLMs
von: Cai, Chengkun, et al.
Veröffentlicht: (2025)
von: Cai, Chengkun, et al.
Veröffentlicht: (2025)
POLYCHARTQA: Benchmarking Large Vision-Language Models with Multilingual Chart Question Answering
von: Xu, Yichen, et al.
Veröffentlicht: (2025)
von: Xu, Yichen, et al.
Veröffentlicht: (2025)
Bifrost-1: Bridging Multimodal LLMs and Diffusion Models with Patch-level CLIP Latents
von: Lin, Han, et al.
Veröffentlicht: (2025)
von: Lin, Han, et al.
Veröffentlicht: (2025)
MagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing
von: Zhang, Kai, et al.
Veröffentlicht: (2023)
von: Zhang, Kai, et al.
Veröffentlicht: (2023)
VLKEB: A Large Vision-Language Model Knowledge Editing Benchmark
von: Huang, Han, et al.
Veröffentlicht: (2024)
von: Huang, Han, et al.
Veröffentlicht: (2024)
XL-HeadTags: Leveraging Multimodal Retrieval Augmentation for the Multilingual Generation of News Headlines and Tags
von: Shohan, Faisal Tareque, et al.
Veröffentlicht: (2024)
von: Shohan, Faisal Tareque, et al.
Veröffentlicht: (2024)
EditReward: A Human-Aligned Reward Model for Instruction-Guided Image Editing
von: Wu, Keming, et al.
Veröffentlicht: (2025)
von: Wu, Keming, et al.
Veröffentlicht: (2025)
Image-Text Relation Prediction for Multilingual Tweets
von: Rikters, Matīss, et al.
Veröffentlicht: (2025)
von: Rikters, Matīss, et al.
Veröffentlicht: (2025)
Leveraging Open Knowledge for Advancing Task Expertise in Large Language Models
von: Yang, Yuncheng, et al.
Veröffentlicht: (2024)
von: Yang, Yuncheng, et al.
Veröffentlicht: (2024)
GRR-CoCa: Leveraging LLM Mechanisms in Multimodal Model Architectures
von: Patock, Jake R., et al.
Veröffentlicht: (2025)
von: Patock, Jake R., et al.
Veröffentlicht: (2025)
DART: Leveraging Multi-Agent Disagreement for Tool Recruitment in Multimodal Reasoning
von: Sivakumaran, Nithin, et al.
Veröffentlicht: (2025)
von: Sivakumaran, Nithin, et al.
Veröffentlicht: (2025)
MAMI: Multi-Attentional Mutual-Information for Long Sequence Neuron Captioning
von: Fauzulhaq, Alfirsa Damasyifa, et al.
Veröffentlicht: (2024)
von: Fauzulhaq, Alfirsa Damasyifa, et al.
Veröffentlicht: (2024)
VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effects
von: Gao, Xiangbo, et al.
Veröffentlicht: (2026)
von: Gao, Xiangbo, et al.
Veröffentlicht: (2026)
Law of the Weakest Link: Cross Capabilities of Large Language Models
von: Zhong, Ming, et al.
Veröffentlicht: (2024)
von: Zhong, Ming, et al.
Veröffentlicht: (2024)
MENTOR: Efficient Multimodal-Conditioned Tuning for Autoregressive Vision Generation Models
von: Zhao, Haozhe, et al.
Veröffentlicht: (2025)
von: Zhao, Haozhe, et al.
Veröffentlicht: (2025)
Uni3D-LLM: Unifying Point Cloud Perception, Generation and Editing with Large Language Models
von: Liu, Dingning, et al.
Veröffentlicht: (2024)
von: Liu, Dingning, et al.
Veröffentlicht: (2024)
CoPE-VideoLM: Leveraging Codec Primitives For Efficient Video Language Modeling
von: Sarkar, Sayan Deb, et al.
Veröffentlicht: (2026)
von: Sarkar, Sayan Deb, et al.
Veröffentlicht: (2026)
NeoBabel: A Multilingual Open Tower for Visual Generation
von: Derakhshani, Mohammad Mahdi, et al.
Veröffentlicht: (2025)
von: Derakhshani, Mohammad Mahdi, et al.
Veröffentlicht: (2025)
Chitrakshara: A Large Multilingual Multimodal Dataset for Indian languages
von: Khan, Shaharukh, et al.
Veröffentlicht: (2026)
von: Khan, Shaharukh, et al.
Veröffentlicht: (2026)
Concept Lancet: Image Editing with Compositional Representation Transplant
von: Luo, Jinqi, et al.
Veröffentlicht: (2025)
von: Luo, Jinqi, et al.
Veröffentlicht: (2025)
ZINA: Multimodal Fine-grained Hallucination Detection and Editing
von: Wada, Yuiga, et al.
Veröffentlicht: (2025)
von: Wada, Yuiga, et al.
Veröffentlicht: (2025)
Multimodal LLM Enhanced Cross-lingual Cross-modal Retrieval
von: Wang, Yabing, et al.
Veröffentlicht: (2024)
von: Wang, Yabing, et al.
Veröffentlicht: (2024)
Beyond Translation: Cross-Cultural Meme Transcreation with Vision-Language Models
von: Zhao, Yuming, et al.
Veröffentlicht: (2026)
von: Zhao, Yuming, et al.
Veröffentlicht: (2026)
Annotations on a Budget: Leveraging Geo-Data Similarity to Balance Model Performance and Annotation Cost
von: Ignat, Oana, et al.
Veröffentlicht: (2024)
von: Ignat, Oana, et al.
Veröffentlicht: (2024)
MotionEdit: Benchmarking and Learning Motion-Centric Image Editing
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
Cross-Modal Adapter for Vision-Language Retrieval
von: Jiang, Haojun, et al.
Veröffentlicht: (2022)
von: Jiang, Haojun, et al.
Veröffentlicht: (2022)
HNC: Leveraging Hard Negative Captions towards Models with Fine-Grained Visual-Linguistic Comprehension Capabilities
von: Dönmez, Esra, et al.
Veröffentlicht: (2026)
von: Dönmez, Esra, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Leveraging LLMs for On-the-Fly Instruction Guided Image Editing
von: Santos, Rodrigo, et al.
Veröffentlicht: (2024) -
CLAIM: Mitigating Multilingual Object Hallucination in Large Vision-Language Models with Cross-Lingual Attention Intervention
von: Ye, Zekai, et al.
Veröffentlicht: (2025) -
MULTITEXTEDIT: Benchmarking Cross-Lingual Degradation in Text-in-Image Editing
von: Cheng, Liwei, et al.
Veröffentlicht: (2026) -
Cross-lingual Editing in Multilingual Language Models
von: Beniwal, Himanshu, et al.
Veröffentlicht: (2024) -
ETCHR: Editing To Clarify and Harness Reasoning
von: Zhang, Beichen, et al.
Veröffentlicht: (2026)