FLUX-Reason-6M & PRISM-Bench: A Million-Scale Text-to-Image Reasoning Dataset and Comprehensive Benchmark
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fang, Rongyao, Yu, Aldrich, Duan, Chengqi, Huang, Linjiang, Bai, Shuai, Cai, Yuxuan, Wang, Kun, Liu, Si, Liu, Xihui, Li, Hongsheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
T2I-ReasonBench: Benchmarking Reasoning-Informed Text-to-Image Generation
von: Sun, Kaiyue, et al.
Veröffentlicht: (2025)
von: Sun, Kaiyue, et al.
Veröffentlicht: (2025)
GoT-R1: Unleashing Reasoning Capability of MLLM for Visual Generation with Reinforcement Learning
von: Duan, Chengqi, et al.
Veröffentlicht: (2025)
von: Duan, Chengqi, et al.
Veröffentlicht: (2025)
GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing
von: Fang, Rongyao, et al.
Veröffentlicht: (2025)
von: Fang, Rongyao, et al.
Veröffentlicht: (2025)
T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation
von: Huang, Kaiyi, et al.
Veröffentlicht: (2023)
von: Huang, Kaiyi, et al.
Veröffentlicht: (2023)
MathCanvas: Intrinsic Visual Chain-of-Thought for Multimodal Mathematical Reasoning
von: Shi, Weikang, et al.
Veröffentlicht: (2025)
von: Shi, Weikang, et al.
Veröffentlicht: (2025)
FouriScale: A Frequency Perspective on Training-Free High-Resolution Image Synthesis
von: Huang, Linjiang, et al.
Veröffentlicht: (2024)
von: Huang, Linjiang, et al.
Veröffentlicht: (2024)
CodePlot-CoT: Mathematical Visual Reasoning by Thinking with Code-Driven Images
von: Duan, Chengqi, et al.
Veröffentlicht: (2025)
von: Duan, Chengqi, et al.
Veröffentlicht: (2025)
PUMA: Empowering Unified MLLM with Multi-granular Visual Generation
von: Fang, Rongyao, et al.
Veröffentlicht: (2024)
von: Fang, Rongyao, et al.
Veröffentlicht: (2024)
BioProBench: Comprehensive Dataset and Benchmark in Biological Protocol Understanding and Reasoning
von: Liu, Yuyang, et al.
Veröffentlicht: (2025)
von: Liu, Yuyang, et al.
Veröffentlicht: (2025)
T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation
von: Sun, Kaiyue, et al.
Veröffentlicht: (2024)
von: Sun, Kaiyue, et al.
Veröffentlicht: (2024)
SOLVE: Synergy of Language-Vision and End-to-End Networks for Autonomous Driving
von: Chen, Xuesong, et al.
Veröffentlicht: (2025)
von: Chen, Xuesong, et al.
Veröffentlicht: (2025)
T2S-Bench & Structure-of-Thought: Benchmarking and Prompting Comprehensive Text-to-Structure Reasoning
von: Wang, Qinsi, et al.
Veröffentlicht: (2026)
von: Wang, Qinsi, et al.
Veröffentlicht: (2026)
MTR-Bench: A Comprehensive Benchmark for Multi-Turn Reasoning Evaluation
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2025)
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2025)
Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning
von: Shao, Hao, et al.
Veröffentlicht: (2024)
von: Shao, Hao, et al.
Veröffentlicht: (2024)
FinanceReasoning: Benchmarking Financial Numerical Reasoning More Credible, Comprehensive and Challenging
von: Tang, Zichen, et al.
Veröffentlicht: (2025)
von: Tang, Zichen, et al.
Veröffentlicht: (2025)
GDI-Bench: A Benchmark for General Document Intelligence with Vision and Reasoning Decoupling
von: Li, Siqi, et al.
Veröffentlicht: (2025)
von: Li, Siqi, et al.
Veröffentlicht: (2025)
FLUX-Text: A Simple and Advanced Diffusion Transformer Baseline for Scene Text Editing
von: Lan, Rui, et al.
Veröffentlicht: (2025)
von: Lan, Rui, et al.
Veröffentlicht: (2025)
EditThinker: Unlocking Iterative Reasoning for Any Image Editor
von: Li, Hongyu, et al.
Veröffentlicht: (2025)
von: Li, Hongyu, et al.
Veröffentlicht: (2025)
R2I-Bench: Benchmarking Reasoning-Driven Text-to-Image Generation
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities
von: Zhu, Chenming, et al.
Veröffentlicht: (2024)
von: Zhu, Chenming, et al.
Veröffentlicht: (2024)
TIR-Bench: A Comprehensive Benchmark for Agentic Thinking-with-Images Reasoning
von: Li, Ming, et al.
Veröffentlicht: (2025)
von: Li, Ming, et al.
Veröffentlicht: (2025)
GaussianPainter: Painting Point Cloud into 3D Gaussians with Normal Guidance
von: Zhou, Jingqiu, et al.
Veröffentlicht: (2024)
von: Zhou, Jingqiu, et al.
Veröffentlicht: (2024)
MER-Bench: A Comprehensive Benchmark for Multimodal Meme Reappraisal
von: Nie, Yiqi, et al.
Veröffentlicht: (2026)
von: Nie, Yiqi, et al.
Veröffentlicht: (2026)
HighlightBench: Benchmarking Markup-Driven Table Reasoning in Scientific Documents
von: Wang, Lexin, et al.
Veröffentlicht: (2026)
von: Wang, Lexin, et al.
Veröffentlicht: (2026)
MMMG: A Massive, Multidisciplinary, Multi-Tier Generation Benchmark for Text-to-Image Reasoning
von: Luo, Yuxuan, et al.
Veröffentlicht: (2025)
von: Luo, Yuxuan, et al.
Veröffentlicht: (2025)
ScholarBench: A Bilingual Benchmark for Abstraction, Comprehension, and Reasoning Evaluation in Academic Contexts
von: Noh, Dongwon, et al.
Veröffentlicht: (2025)
von: Noh, Dongwon, et al.
Veröffentlicht: (2025)
GIR-Bench: Versatile Benchmark for Generating Images with Reasoning
von: Li, Hongxiang, et al.
Veröffentlicht: (2025)
von: Li, Hongxiang, et al.
Veröffentlicht: (2025)
GreekBarBench: A Challenging Benchmark for Free-Text Legal Reasoning and Citations
von: Chlapanis, Odysseas S., et al.
Veröffentlicht: (2025)
von: Chlapanis, Odysseas S., et al.
Veröffentlicht: (2025)
TAD-Bench: A Comprehensive Benchmark for Embedding-Based Text Anomaly Detection
von: Cao, Yang, et al.
Veröffentlicht: (2025)
von: Cao, Yang, et al.
Veröffentlicht: (2025)
MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs
von: Yuan, Jiakang, et al.
Veröffentlicht: (2025)
von: Yuan, Jiakang, et al.
Veröffentlicht: (2025)
PRISM: Probing Reasoning, Instruction, and Source Memory in LLM Hallucinations
von: Wu, Yuhe, et al.
Veröffentlicht: (2026)
von: Wu, Yuhe, et al.
Veröffentlicht: (2026)
PunchBench: Benchmarking MLLMs in Multimodal Punchline Comprehension
von: Ouyang, Kun, et al.
Veröffentlicht: (2024)
von: Ouyang, Kun, et al.
Veröffentlicht: (2024)
TextReasoningBench: Does Reasoning Really Improve Text Classification in Large Language Models?
von: Guo, Xinyu, et al.
Veröffentlicht: (2026)
von: Guo, Xinyu, et al.
Veröffentlicht: (2026)
PredBench: Benchmarking Spatio-Temporal Prediction across Diverse Disciplines
von: Wang, ZiDong, et al.
Veröffentlicht: (2024)
von: Wang, ZiDong, et al.
Veröffentlicht: (2024)
DermaBench: A Clinician-Annotated Benchmark Dataset for Dermatology Visual Question Answering and Reasoning
von: Yilmaz, Abdurrahim, et al.
Veröffentlicht: (2026)
von: Yilmaz, Abdurrahim, et al.
Veröffentlicht: (2026)
MixRea: Benchmarking Explicit-Implicit Reasoning in Large Language Models
von: Cai, Yuanqing, et al.
Veröffentlicht: (2026)
von: Cai, Yuanqing, et al.
Veröffentlicht: (2026)
CriticBench: Benchmarking LLMs for Critique-Correct Reasoning
von: Lin, Zicheng, et al.
Veröffentlicht: (2024)
von: Lin, Zicheng, et al.
Veröffentlicht: (2024)
RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension
von: Gao, Tianyi, et al.
Veröffentlicht: (2025)
von: Gao, Tianyi, et al.
Veröffentlicht: (2025)
ATR-Bench: A Federated Learning Benchmark for Adaptation, Trust, and Reasoning
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
V-ReasonBench: Toward Unified Reasoning Benchmark Suite for Video Generation Models
von: Luo, Yang, et al.
Veröffentlicht: (2025)
von: Luo, Yang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
T2I-ReasonBench: Benchmarking Reasoning-Informed Text-to-Image Generation
von: Sun, Kaiyue, et al.
Veröffentlicht: (2025) -
GoT-R1: Unleashing Reasoning Capability of MLLM for Visual Generation with Reinforcement Learning
von: Duan, Chengqi, et al.
Veröffentlicht: (2025) -
GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing
von: Fang, Rongyao, et al.
Veröffentlicht: (2025) -
T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation
von: Huang, Kaiyi, et al.
Veröffentlicht: (2023) -
MathCanvas: Intrinsic Visual Chain-of-Thought for Multimodal Mathematical Reasoning
von: Shi, Weikang, et al.
Veröffentlicht: (2025)