WorldGenBench: A World-Knowledge-Integrated Benchmark for Reasoning-Driven Text-to-Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Daoan, Jiang, Che, Xu, Ruoshi, Chen, Biaoxiang, Jin, Zijian, Lu, Yutian, Zhang, Jianguo, Yong, Liang, Luo, Jiebo, Luo, Shengda |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Sphinx: Benchmarking and Modeling for LLM-Driven Pull Request Review
von: Zhang, Daoan, et al.
Veröffentlicht: (2026)
von: Zhang, Daoan, et al.
Veröffentlicht: (2026)
CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs
von: Zhang, Daoan, et al.
Veröffentlicht: (2024)
von: Zhang, Daoan, et al.
Veröffentlicht: (2024)
Learning Brain Tumor Representation in 3D High-Resolution MR Images via Interpretable State Space Models
von: Hu, Qingqiao, et al.
Veröffentlicht: (2024)
von: Hu, Qingqiao, et al.
Veröffentlicht: (2024)
FINEMATCH: Aspect-based Fine-grained Image and Text Mismatch Detection and Correction
von: Hua, Hang, et al.
Veröffentlicht: (2024)
von: Hua, Hang, et al.
Veröffentlicht: (2024)
T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation
von: Chen, Yubin, et al.
Veröffentlicht: (2025)
von: Chen, Yubin, et al.
Veröffentlicht: (2025)
Gradient-Guided Modality Decoupling for Missing-Modality Robustness
von: Wang, Hao, et al.
Veröffentlicht: (2024)
von: Wang, Hao, et al.
Veröffentlicht: (2024)
Reinforcing 3D Understanding in Point-VLMs via Geometric Reward Credit Assignment
von: Chen, Jingkun, et al.
Veröffentlicht: (2026)
von: Chen, Jingkun, et al.
Veröffentlicht: (2026)
VisualActBench: Can VLMs See and Act like a Human?
von: Zhang, Daoan, et al.
Veröffentlicht: (2025)
von: Zhang, Daoan, et al.
Veröffentlicht: (2025)
World-To-Image: Grounding Text-to-Image Generation with Agent-Driven World Knowledge
von: Son, Moo Hyun, et al.
Veröffentlicht: (2025)
von: Son, Moo Hyun, et al.
Veröffentlicht: (2025)
A Versatile Multimodal Agent for Multimedia Content Generation
von: Zhang, Daoan, et al.
Veröffentlicht: (2026)
von: Zhang, Daoan, et al.
Veröffentlicht: (2026)
LAST: LeArning to Think in Space and Time for Generalist Vision-Language Models
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
R2I-Bench: Benchmarking Reasoning-Driven Text-to-Image Generation
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
Integrating Text and Image Pre-training for Multi-modal Algorithmic Reasoning
von: Zhang, Zijian, et al.
Veröffentlicht: (2024)
von: Zhang, Zijian, et al.
Veröffentlicht: (2024)
RSA-Bench: Benchmarking Audio Large Models in Real-World Acoustic Scenarios
von: Zhang, Yibo, et al.
Veröffentlicht: (2026)
von: Zhang, Yibo, et al.
Veröffentlicht: (2026)
Code to Think, Think to Code: A Survey on Code-Enhanced Reasoning and Reasoning-Driven Code Intelligence in LLMs
von: Yang, Dayu, et al.
Veröffentlicht: (2025)
von: Yang, Dayu, et al.
Veröffentlicht: (2025)
MIRA: Multimodal Iterative Reasoning Agent for Image Editing
von: Zeng, Ziyun, et al.
Veröffentlicht: (2025)
von: Zeng, Ziyun, et al.
Veröffentlicht: (2025)
CBVS: A Large-Scale Chinese Image-Text Benchmark for Real-World Short Video Search Scenarios
von: Qiao, Xiangshuo, et al.
Veröffentlicht: (2024)
von: Qiao, Xiangshuo, et al.
Veröffentlicht: (2024)
HWE-Bench: Benchmarking LLM Agents on Real-World Hardware Bug Repair Tasks
von: Cui, Fan, et al.
Veröffentlicht: (2026)
von: Cui, Fan, et al.
Veröffentlicht: (2026)
Can World Simulators Reason? Gen-ViRe: A Generative Visual Reasoning Benchmark
von: Liu, Xinxin, et al.
Veröffentlicht: (2025)
von: Liu, Xinxin, et al.
Veröffentlicht: (2025)
SOK-Bench: A Situated Video Reasoning Benchmark with Aligned Open-World Knowledge
von: Wang, Andong, et al.
Veröffentlicht: (2024)
von: Wang, Andong, et al.
Veröffentlicht: (2024)
ChronoMagic-Bench: A Benchmark for Metamorphic Evaluation of Text-to-Time-lapse Video Generation
von: Yuan, Shenghai, et al.
Veröffentlicht: (2024)
von: Yuan, Shenghai, et al.
Veröffentlicht: (2024)
UniGenBench++: A Unified Semantic Evaluation Benchmark for Text-to-Image Generation
von: Wang, Yibin, et al.
Veröffentlicht: (2025)
von: Wang, Yibin, et al.
Veröffentlicht: (2025)
WorldEdit: Towards Open-World Image Editing with a Knowledge-Informed Benchmark
von: Lin, Wang, et al.
Veröffentlicht: (2026)
von: Lin, Wang, et al.
Veröffentlicht: (2026)
WorldCoder-Bench: Benchmarking Physically Grounded 3D World Synthesis
von: Lu, Shuo, et al.
Veröffentlicht: (2026)
von: Lu, Shuo, et al.
Veröffentlicht: (2026)
RecipeGen: A Benchmark for Real-World Recipe Image Generation
von: Zhang, Ruoxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Ruoxuan, et al.
Veröffentlicht: (2025)
PhyBench: A Physical Commonsense Benchmark for Evaluating Text-to-Image Models
von: Meng, Fanqing, et al.
Veröffentlicht: (2024)
von: Meng, Fanqing, et al.
Veröffentlicht: (2024)
SafeGen-Bench: Benchmarking Safety in Image-Conditioned Text-to-Video Generation
von: Ma, Yingzi, et al.
Veröffentlicht: (2026)
von: Ma, Yingzi, et al.
Veröffentlicht: (2026)
Text2World: Benchmarking Large Language Models for Symbolic World Model Generation
von: Hu, Mengkang, et al.
Veröffentlicht: (2025)
von: Hu, Mengkang, et al.
Veröffentlicht: (2025)
Beyond Words and Pixels: A Benchmark for Implicit World Knowledge Reasoning in Generative Models
von: Han, Tianyang, et al.
Veröffentlicht: (2025)
von: Han, Tianyang, et al.
Veröffentlicht: (2025)
UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images
von: Qu, Yiting, et al.
Veröffentlicht: (2024)
von: Qu, Yiting, et al.
Veröffentlicht: (2024)
ShoppingBench: A Real-World Intent-Grounded Shopping Benchmark for LLM-based Agents
von: Wang, Jiangyuan, et al.
Veröffentlicht: (2025)
von: Wang, Jiangyuan, et al.
Veröffentlicht: (2025)
GenColorBench: A Color Evaluation Benchmark for Text-to-Image Generation Models
von: Butt, Muhammad Atif, et al.
Veröffentlicht: (2025)
von: Butt, Muhammad Atif, et al.
Veröffentlicht: (2025)
Downstream-Pretext Domain Knowledge Traceback for Active Learning
von: Zhang, Beichen, et al.
Veröffentlicht: (2024)
von: Zhang, Beichen, et al.
Veröffentlicht: (2024)
Hengqin-RA-v1: Advanced Large Language Model for Diagnosis and Treatment of Rheumatoid Arthritis with Dataset based Traditional Chinese Medicine
von: Liu, Yishen, et al.
Veröffentlicht: (2025)
von: Liu, Yishen, et al.
Veröffentlicht: (2025)
CriticBench: Benchmarking LLMs for Critique-Correct Reasoning
von: Lin, Zicheng, et al.
Veröffentlicht: (2024)
von: Lin, Zicheng, et al.
Veröffentlicht: (2024)
BizFinBench: A Business-Driven Real-World Financial Benchmark for Evaluating LLMs
von: Lu, Guilong, et al.
Veröffentlicht: (2025)
von: Lu, Guilong, et al.
Veröffentlicht: (2025)
GaussianStyle: Gaussian Head Avatar via StyleGAN
von: Liu, Pinxin, et al.
Veröffentlicht: (2024)
von: Liu, Pinxin, et al.
Veröffentlicht: (2024)
SpreadsheetBench: Towards Challenging Real World Spreadsheet Manipulation
von: Ma, Zeyao, et al.
Veröffentlicht: (2024)
von: Ma, Zeyao, et al.
Veröffentlicht: (2024)
iWorld-Bench: A Benchmark for Interactive World Models with a Unified Action Generation Framework
von: Fang, Jianjie, et al.
Veröffentlicht: (2026)
von: Fang, Jianjie, et al.
Veröffentlicht: (2026)
WorldGen: From Text to Traversable and Interactive 3D Worlds
von: Wang, Dilin, et al.
Veröffentlicht: (2025)
von: Wang, Dilin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Sphinx: Benchmarking and Modeling for LLM-Driven Pull Request Review
von: Zhang, Daoan, et al.
Veröffentlicht: (2026) -
CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs
von: Zhang, Daoan, et al.
Veröffentlicht: (2024) -
Learning Brain Tumor Representation in 3D High-Resolution MR Images via Interpretable State Space Models
von: Hu, Qingqiao, et al.
Veröffentlicht: (2024) -
FINEMATCH: Aspect-based Fine-grained Image and Text Mismatch Detection and Correction
von: Hua, Hang, et al.
Veröffentlicht: (2024) -
T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation
von: Chen, Yubin, et al.
Veröffentlicht: (2025)