Routing to the Right Expertise: A Trustworthy Judge for Instruction-based Image Editing
Fuente:
arXiv
Guardado en:
| Autores principales: | Sun, Chenxi, Zhang, Hongzhi, Wang, Qi, Zhang, Fuzheng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Capybara-OMNI: An Efficient Paradigm for Building Omni-Modal Language Models
por: Ji, Xingguang, et al.
Publicado: (2025)
por: Ji, Xingguang, et al.
Publicado: (2025)
Data Metabolism: An Efficient Data Design Schema For Vision Language Model
por: Zhang, Jingyuan, et al.
Publicado: (2025)
por: Zhang, Jingyuan, et al.
Publicado: (2025)
Decoding at the Speed of Thought: Harnessing Parallel Decoding of Lexical Units for LLMs
por: Sun, Chenxi, et al.
Publicado: (2024)
por: Sun, Chenxi, et al.
Publicado: (2024)
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
por: Fu, Jia, et al.
Publicado: (2025)
por: Fu, Jia, et al.
Publicado: (2025)
Klear-AgentForge: Forging Agentic Intelligence through Posttraining Scaling
por: Wang, Qi, et al.
Publicado: (2025)
por: Wang, Qi, et al.
Publicado: (2025)
Instruction-based Image Editing with Planning, Reasoning, and Generation
por: Ji, Liya, et al.
Publicado: (2026)
por: Ji, Liya, et al.
Publicado: (2026)
Leanabell-Prover-V2: Verifier-integrated Reasoning for Formal Theorem Proving via Reinforcement Learning
por: Ji, Xingguang, et al.
Publicado: (2025)
por: Ji, Xingguang, et al.
Publicado: (2025)
Chain-of-Specificity: An Iteratively Refining Method for Eliciting Knowledge from Large Language Models
por: Wei, Kaiwen, et al.
Publicado: (2024)
por: Wei, Kaiwen, et al.
Publicado: (2024)
Inductive-Deductive Strategy Reuse for Multi-Turn Instructional Dialogues
por: Ou, Jiao, et al.
Publicado: (2024)
por: Ou, Jiao, et al.
Publicado: (2024)
I2EBench: A Comprehensive Benchmark for Instruction-based Image Editing
por: Ma, Yiwei, et al.
Publicado: (2024)
por: Ma, Yiwei, et al.
Publicado: (2024)
Who Judges the Judge? LLM Jury-on-Demand: Building Trustworthy LLM Evaluation Systems
por: Li, Xiaochuan, et al.
Publicado: (2025)
por: Li, Xiaochuan, et al.
Publicado: (2025)
CLIPDrag: Combining Text-based and Drag-based Instructions for Image Editing
por: Jiang, Ziqi, et al.
Publicado: (2024)
por: Jiang, Ziqi, et al.
Publicado: (2024)
ProJudge: A Multi-Modal Multi-Discipline Benchmark and Instruction-Tuning Dataset for MLLM-based Process Judges
por: Ai, Jiaxin, et al.
Publicado: (2025)
por: Ai, Jiaxin, et al.
Publicado: (2025)
TAGE: Trustworthy Attribute Group Editing for Stable Few-shot Image Generation
por: Zhang, Ruicheng, et al.
Publicado: (2024)
por: Zhang, Ruicheng, et al.
Publicado: (2024)
HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing
por: Hui, Mude, et al.
Publicado: (2024)
por: Hui, Mude, et al.
Publicado: (2024)
MagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing
por: Zhang, Kai, et al.
Publicado: (2023)
por: Zhang, Kai, et al.
Publicado: (2023)
The Myth of Expert Specialization in MoEs: Why Routing Reflects Geometry, Not Necessarily Domain Expertise
por: Wang, Xi, et al.
Publicado: (2026)
por: Wang, Xi, et al.
Publicado: (2026)
DLEBench: Evaluating Small-scale Object Editing Ability for Instruction-based Image Editing Model
por: Hong, Shibo, et al.
Publicado: (2026)
por: Hong, Shibo, et al.
Publicado: (2026)
Kontinuous Kontext: Continuous Strength Control for Instruction-based Image Editing
por: Parihar, Rishubh, et al.
Publicado: (2025)
por: Parihar, Rishubh, et al.
Publicado: (2025)
Instruction-Guided Editing Controls for Images and Multimedia: A Survey in LLM era
por: Nguyen, Thanh Tam, et al.
Publicado: (2024)
por: Nguyen, Thanh Tam, et al.
Publicado: (2024)
Are We on the Right Way to Assessing LLM-as-a-Judge?
por: Feng, Yuanning, et al.
Publicado: (2025)
por: Feng, Yuanning, et al.
Publicado: (2025)
AgentInit: Initializing LLM-based Multi-Agent Systems via Diversity and Expertise Orchestration for Effective and Efficient Collaboration
por: Tian, Chunhao, et al.
Publicado: (2025)
por: Tian, Chunhao, et al.
Publicado: (2025)
Reversible Lifelong Model Editing via Semantic Routing-Based LoRA
por: Luo, Haihua, et al.
Publicado: (2026)
por: Luo, Haihua, et al.
Publicado: (2026)
Reasoning Is Not Free: Robust Adaptive Cost-Efficient Routing for LLM-as-a-Judge
por: Zhang, Wenbo, et al.
Publicado: (2026)
por: Zhang, Wenbo, et al.
Publicado: (2026)
RubricEval: A Rubric-Level Meta-Evaluation Benchmark for LLM Judges in Instruction Following
por: Pan, Tianjun, et al.
Publicado: (2026)
por: Pan, Tianjun, et al.
Publicado: (2026)
Mapping the Mind of an Instruction-based Image Editing using SMILE
por: Dehghani, Zeinab, et al.
Publicado: (2024)
por: Dehghani, Zeinab, et al.
Publicado: (2024)
MCIE: Multimodal LLM-Driven Complex Instruction Image Editing with Spatial Guidance
por: Bai, Xuehai, et al.
Publicado: (2026)
por: Bai, Xuehai, et al.
Publicado: (2026)
FewFedPIT: Towards Privacy-preserving and Few-shot Federated Instruction Tuning
por: Zhang, Zhuo, et al.
Publicado: (2024)
por: Zhang, Zhuo, et al.
Publicado: (2024)
Leanabell-Prover: Posttraining Scaling in Formal Reasoning
por: Zhang, Jingyuan, et al.
Publicado: (2025)
por: Zhang, Jingyuan, et al.
Publicado: (2025)
FreeEdit: Mask-free Reference-based Image Editing with Multi-modal Instruction
por: He, Runze, et al.
Publicado: (2024)
por: He, Runze, et al.
Publicado: (2024)
OpenGPT-4o-Image: A Comprehensive Dataset for Advanced Image Generation and Editing
por: Chen, Zhihong, et al.
Publicado: (2025)
por: Chen, Zhihong, et al.
Publicado: (2025)
Speaking the Right Language: The Impact of Expertise Alignment in User-AI Interactions
por: Palta, Shramay, et al.
Publicado: (2025)
por: Palta, Shramay, et al.
Publicado: (2025)
EditWorld: Simulating World Dynamics for Instruction-Following Image Editing
por: Yang, Ling, et al.
Publicado: (2024)
por: Yang, Ling, et al.
Publicado: (2024)
Balancing Preservation and Modification: A Region and Semantic Aware Metric for Instruction-Based Image Editing
por: Li, Zhuoying, et al.
Publicado: (2025)
por: Li, Zhuoying, et al.
Publicado: (2025)
How Trustworthy Are LLM-as-Judge Ratings for Interpretive Responses? Implications for Qualitative Research Workflows
por: Han, Songhee, et al.
Publicado: (2026)
por: Han, Songhee, et al.
Publicado: (2026)
What Makes a Good Reasoning Chain? Uncovering Structural Patterns in Long Chain-of-Thought Reasoning
por: Jiang, Gangwei, et al.
Publicado: (2025)
por: Jiang, Gangwei, et al.
Publicado: (2025)
From Physician Expertise to Clinical Agents: Preserving, Standardizing, and Scaling Physicians' Medical Expertise with Lightweight LLM
por: Luo, Chanyong, et al.
Publicado: (2026)
por: Luo, Chanyong, et al.
Publicado: (2026)
Judging the Judges: Human Validation of Multi-LLM Evaluation for High-Quality K--12 Science Instructional Materials
por: He, Peng, et al.
Publicado: (2026)
por: He, Peng, et al.
Publicado: (2026)
SPPD: Self-training with Process Preference Learning Using Dynamic Value Margin
por: Yi, Hao, et al.
Publicado: (2025)
por: Yi, Hao, et al.
Publicado: (2025)
ChartM$^3$: Benchmarking Chart Editing with Multimodal Instructions
por: Yang, Donglu, et al.
Publicado: (2025)
por: Yang, Donglu, et al.
Publicado: (2025)
Ejemplares similares
-
Capybara-OMNI: An Efficient Paradigm for Building Omni-Modal Language Models
por: Ji, Xingguang, et al.
Publicado: (2025) -
Data Metabolism: An Efficient Data Design Schema For Vision Language Model
por: Zhang, Jingyuan, et al.
Publicado: (2025) -
Decoding at the Speed of Thought: Harnessing Parallel Decoding of Lexical Units for LLMs
por: Sun, Chenxi, et al.
Publicado: (2024) -
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
por: Fu, Jia, et al.
Publicado: (2025) -
Klear-AgentForge: Forging Agentic Intelligence through Posttraining Scaling
por: Wang, Qi, et al.
Publicado: (2025)