VDE Bench: Evaluating The Capability of Image Editing Models to Modify Visual Documents
Fuente:
arXiv
Salvato in:
| Autori principali: | Yi, Hongzhu, Yang, Yujia, Wang, Yuanxiang, Li, Tong, Guan, Zhenyu, Zong, Tianyu, Chen, Jiahuan, Bao, Chenxi, Yang, Tiankun, Jin, Haopeng, Yuan, Yixuan, Wang, Xinming, Yu, Tao, Gao, Ruilin, Tao, Ruiwen, Liang, Haijin, Ma, Jin, Luo, Jinwen, Yeshani, Zuo, Xinyu, Xu, Jungang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Omni IIE Bench: Benchmarking the Practical Capabilities of Image Editing Models
di: Yang, Yujia, et al.
Pubblicazione: (2026)
di: Yang, Yujia, et al.
Pubblicazione: (2026)
RPO:Reinforcement Fine-Tuning with Partial Reasoning Optimization
di: Yi, Hongzhu, et al.
Pubblicazione: (2026)
di: Yi, Hongzhu, et al.
Pubblicazione: (2026)
Beyond Closed-Pool Video Retrieval: A Benchmark and Agent Framework for Real-World Video Search and Moment Localization
di: Yu, Tao, et al.
Pubblicazione: (2026)
di: Yu, Tao, et al.
Pubblicazione: (2026)
JTCSE: Joint Tensor-Modulus Constraints and Cross-Attention for Unsupervised Contrastive Learning of Sentence Embeddings
di: Zong, Tianyu, et al.
Pubblicazione: (2025)
di: Zong, Tianyu, et al.
Pubblicazione: (2025)
HY-Himmel Technical Report: Hierarchical Interleaved Multi-stream Motion Encoding for Long Video Understanding
di: Jin, Haopeng, et al.
Pubblicazione: (2026)
di: Jin, Haopeng, et al.
Pubblicazione: (2026)
Dynamic Deep Graph Learning for Incomplete Multi-View Clustering with Masked Graph Reconstruction Loss
di: Zhang, Zhenghao, et al.
Pubblicazione: (2025)
di: Zhang, Zhenghao, et al.
Pubblicazione: (2025)
TNCSE: Tensor's Norm Constraints for Unsupervised Contrastive Learning of Sentence Embeddings
di: Zong, Tianyu, et al.
Pubblicazione: (2025)
di: Zong, Tianyu, et al.
Pubblicazione: (2025)
A novel method for satellite prediction of VDE‐TER service traffic information based on multilayers bidirectional LSTM
di: Ruiwen Wu, et al.
Pubblicazione: (2024)
di: Ruiwen Wu, et al.
Pubblicazione: (2024)
FAIRGAMER: Evaluating Social Biases in LLM-Based Video Game NPCs
di: Shi, Bingkang, et al.
Pubblicazione: (2025)
di: Shi, Bingkang, et al.
Pubblicazione: (2025)
PaperX: A Unified Framework for Multimodal Academic Presentation Generation with Scholar DAG
di: Yu, Tao, et al.
Pubblicazione: (2026)
di: Yu, Tao, et al.
Pubblicazione: (2026)
MoCha-LD: Temporal-Chunked Latent Diffusion with Motion-Aware Consistency Regularization for Long Video Generation
di: Jin, Haopeng
Pubblicazione: (2026)
di: Jin, Haopeng
Pubblicazione: (2026)
CARL-MoE: Communication-Aware Adaptive Routing with Load-Balanced Expert Parallelism for Efficient Mixture-of-Experts Training
di: Jin, Haopeng
Pubblicazione: (2026)
di: Jin, Haopeng
Pubblicazione: (2026)
FreqFormer: Hierarchical Frequency-Domain Attention with Adaptive Spectral Routing for Long-Sequence Video Diffusion Transformers
di: Jin, Haopeng
Pubblicazione: (2026)
di: Jin, Haopeng
Pubblicazione: (2026)
C$^{3}$Bench: A Comprehensive Classical Chinese Understanding Benchmark for Large Language Models
di: Cao, Jiahuan, et al.
Pubblicazione: (2024)
di: Cao, Jiahuan, et al.
Pubblicazione: (2024)
ShotFinder: Imagination-Driven Open-Domain Video Shot Retrieval via Web Search
di: Yu, Tao, et al.
Pubblicazione: (2026)
di: Yu, Tao, et al.
Pubblicazione: (2026)
MPCI-Bench: A Benchmark for Multimodal Pairwise Contextual Integrity Evaluation of Language Model Agents
di: Wang, Shouju, et al.
Pubblicazione: (2026)
di: Wang, Shouju, et al.
Pubblicazione: (2026)
Towards Principled Dataset Distillation: A Spectral Distribution Perspective
di: Wu, Ruixi, et al.
Pubblicazione: (2026)
di: Wu, Ruixi, et al.
Pubblicazione: (2026)
Beyond the All-in-One Agent: Benchmarking Role-Specialized Multi-Agent Collaboration in Enterprise Workflows
di: Yu, Tao, et al.
Pubblicazione: (2026)
di: Yu, Tao, et al.
Pubblicazione: (2026)
Omni-DeepSearch: A Benchmark for Audio-Driven Omni-Modal Deep Search
di: Yu, Tao, et al.
Pubblicazione: (2026)
di: Yu, Tao, et al.
Pubblicazione: (2026)
WeatherBench: A Real-World Benchmark Dataset for All-in-One Adverse Weather Image Restoration
di: Guan, Qiyuan, et al.
Pubblicazione: (2025)
di: Guan, Qiyuan, et al.
Pubblicazione: (2025)
VDE: Training-Free Accelerating Rectified Flow Model via Velocity Decomposition and Estimation
di: Tan, Junwen, et al.
Pubblicazione: (2026)
di: Tan, Junwen, et al.
Pubblicazione: (2026)
V2Rho-FNO: Fourier Neural Operator for Electronic Density Prediction
di: Jin, Yingdi, et al.
Pubblicazione: (2026)
di: Jin, Yingdi, et al.
Pubblicazione: (2026)
AICA-Bench: Holistically Examining the Capabilities of VLMs in Affective Image Content Analysis
di: She, Dong, et al.
Pubblicazione: (2026)
di: She, Dong, et al.
Pubblicazione: (2026)
InfiBench: Evaluating the Question-Answering Capabilities of Code Large Language Models
di: Li, Linyi, et al.
Pubblicazione: (2024)
di: Li, Linyi, et al.
Pubblicazione: (2024)
BrowserAgent: Building Web Agents with Human-Inspired Web Browsing Actions
di: Yu, Tao, et al.
Pubblicazione: (2025)
di: Yu, Tao, et al.
Pubblicazione: (2025)
BattleAgentBench: A Benchmark for Evaluating Cooperation and Competition Capabilities of Language Models in Multi-Agent Systems
di: Wang, Wei, et al.
Pubblicazione: (2024)
di: Wang, Wei, et al.
Pubblicazione: (2024)
Multimodal Video Emotion Recognition with Reliable Reasoning Priors
di: Wang, Zhepeng, et al.
Pubblicazione: (2025)
di: Wang, Zhepeng, et al.
Pubblicazione: (2025)
Stop Before You Fail: Operational Capability Boundaries for Mitigating Unproductive Reasoning in Large Reasoning Models
di: Zhang, Qingjie, et al.
Pubblicazione: (2025)
di: Zhang, Qingjie, et al.
Pubblicazione: (2025)
Few-Shot Image Classification and Segmentation as Visual Question Answering Using Vision-Language Models
di: Meng, Tian, et al.
Pubblicazione: (2024)
di: Meng, Tian, et al.
Pubblicazione: (2024)
The exact group-sparse recovery for block diagonal matrices with subexponential entries
di: Dai, Guozheng, et al.
Pubblicazione: (2025)
di: Dai, Guozheng, et al.
Pubblicazione: (2025)
MatFormBench: A Benchmarking Evaluation Framework for Target-Driven Materials Formulation
di: Wu, Linhan, et al.
Pubblicazione: (2026)
di: Wu, Linhan, et al.
Pubblicazione: (2026)
Amphiregulin Promotes Proliferation and Migration of the Damaged Endothelial Cells in Kawasaki Disease Cell Models
di: Jiawen Xu, et al.
Pubblicazione: (2025)
di: Jiawen Xu, et al.
Pubblicazione: (2025)
Cognitive Aesthetic Evaluation Model - Implementation Code
di: Chenxi, Jin
Pubblicazione: (2026)
di: Chenxi, Jin
Pubblicazione: (2026)
RainGaugeNet: CSI-Based Sub-6 GHz Rainfall Attenuation Measurement and Classification for ISAC Applications
di: Li, Yan, et al.
Pubblicazione: (2025)
di: Li, Yan, et al.
Pubblicazione: (2025)
How Does Eco‐Migrant Relocation Affect the Eco‐Friendly Behaviors of Rural Residents? Evidence From Hainan Province, China
di: Weiqin Li, et al.
Pubblicazione: (2026)
di: Weiqin Li, et al.
Pubblicazione: (2026)
OCRGenBench: A Comprehensive Benchmark for Evaluating OCR Generative Capabilities
di: Zhang, Peirong, et al.
Pubblicazione: (2025)
di: Zhang, Peirong, et al.
Pubblicazione: (2025)
FastMTP: Accelerating LLM Inference with Enhanced Multi-Token Prediction
di: Cai, Yuxuan, et al.
Pubblicazione: (2025)
di: Cai, Yuxuan, et al.
Pubblicazione: (2025)
Four Eyes Are Better Than Two: Harnessing the Collaborative Potential of Large Models via Differentiated Thinking and Complementary Ensembles
di: Xie, Jun, et al.
Pubblicazione: (2025)
di: Xie, Jun, et al.
Pubblicazione: (2025)
Team of One: Cracking Complex Video QA with Model Synergy
di: Xie, Jun, et al.
Pubblicazione: (2025)
di: Xie, Jun, et al.
Pubblicazione: (2025)
Speech digital biomarker combined with fluid biomarkers can predict Alzheimer’s type cognitive impairment through machine learning
di: Gang Wang, et al.
Pubblicazione: (2025)
di: Gang Wang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Omni IIE Bench: Benchmarking the Practical Capabilities of Image Editing Models
di: Yang, Yujia, et al.
Pubblicazione: (2026) -
RPO:Reinforcement Fine-Tuning with Partial Reasoning Optimization
di: Yi, Hongzhu, et al.
Pubblicazione: (2026) -
Beyond Closed-Pool Video Retrieval: A Benchmark and Agent Framework for Real-World Video Search and Moment Localization
di: Yu, Tao, et al.
Pubblicazione: (2026) -
JTCSE: Joint Tensor-Modulus Constraints and Cross-Attention for Unsupervised Contrastive Learning of Sentence Embeddings
di: Zong, Tianyu, et al.
Pubblicazione: (2025) -
HY-Himmel Technical Report: Hierarchical Interleaved Multi-stream Motion Encoding for Long Video Understanding
di: Jin, Haopeng, et al.
Pubblicazione: (2026)