OIDA-QA: A Multimodal Benchmark for Analyzing the Opioid Industry Documents Archive
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shen, Xuan, Wingenroth, Brian, Wang, Zichao, Kuen, Jason, Zhu, Wanrong, Zhang, Ruiyi, Wang, Yiwei, Ma, Lichun, Liu, Anqi, Liu, Hongfu, Sun, Tong, Hawkins, Kevin S., Tasker, Kate, Alexander, G. Caleb, Gu, Jiuxiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Toffee: Efficient Million-Scale Dataset Construction for Subject-Driven Text-to-Image Generation
von: Zhou, Yufan, et al.
Veröffentlicht: (2024)
von: Zhou, Yufan, et al.
Veröffentlicht: (2024)
Customization Assistant for Text-to-image Generation
von: Zhou, Yufan, et al.
Veröffentlicht: (2023)
von: Zhou, Yufan, et al.
Veröffentlicht: (2023)
OIDA International Journal of Sustainable Development
Veröffentlicht: (2017)
Veröffentlicht: (2017)
SOHES: Self-supervised Open-world Hierarchical Entity Segmentation
von: Cao, Shengcao, et al.
Veröffentlicht: (2024)
von: Cao, Shengcao, et al.
Veröffentlicht: (2024)
SNCE: Geometry-Aware Supervision for Scalable Discrete Image Generation
von: Li, Shufan, et al.
Veröffentlicht: (2026)
von: Li, Shufan, et al.
Veröffentlicht: (2026)
Towards Visual Text Grounding of Multimodal Large Language Model
von: Li, Ming, et al.
Veröffentlicht: (2025)
von: Li, Ming, et al.
Veröffentlicht: (2025)
Automatic Layout Planning for Visually-Rich Documents with Instruction-Following Models
von: Zhu, Wanrong, et al.
Veröffentlicht: (2024)
von: Zhu, Wanrong, et al.
Veröffentlicht: (2024)
Soldiers' Stories
von: Tasker, Yvonne
Veröffentlicht: (2018)
von: Tasker, Yvonne
Veröffentlicht: (2018)
VipAct: Visual-Perception Enhancement via Specialized VLM Agent Collaboration and Tool-use
von: Zhang, Zhehao, et al.
Veröffentlicht: (2024)
von: Zhang, Zhehao, et al.
Veröffentlicht: (2024)
Sparse-LaViDa: Sparse Multimodal Discrete Diffusion Language Models
von: Li, Shufan, et al.
Veröffentlicht: (2025)
von: Li, Shufan, et al.
Veröffentlicht: (2025)
Lavida-O: Elastic Large Masked Diffusion Models for Unified Multimodal Understanding and Generation
von: Li, Shufan, et al.
Veröffentlicht: (2025)
von: Li, Shufan, et al.
Veröffentlicht: (2025)
Automatic Method Illustration Generation for AI Scientific Papers via Drawing Middleware Creation, Evolution, and Orchestration
von: Li, Zhuoling, et al.
Veröffentlicht: (2026)
von: Li, Zhuoling, et al.
Veröffentlicht: (2026)
MiLDEdit: Reasoning-Based Multi-Layer Design Document Editing
von: Lin, Zihao, et al.
Veröffentlicht: (2026)
von: Lin, Zihao, et al.
Veröffentlicht: (2026)
Improve Temporal Awareness of LLMs for Sequential Recommendation
von: Chu, Zhendong, et al.
Veröffentlicht: (2024)
von: Chu, Zhendong, et al.
Veröffentlicht: (2024)
Methemoglobinemia Secondary to Zinc Foreign Body Ingestion in a Dog
von: Kate Tasker, et al.
Veröffentlicht: (2026)
von: Kate Tasker, et al.
Veröffentlicht: (2026)
Building a virtual milky way
von: E. J. Tasker
Veröffentlicht: (2007)
von: E. J. Tasker
Veröffentlicht: (2007)
SUGAR: Subject-Driven Video Customization in a Zero-Shot Manner
von: Zhou, Yufan, et al.
Veröffentlicht: (2024)
von: Zhou, Yufan, et al.
Veröffentlicht: (2024)
LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
DiffGraph: An Automated Agent-driven Model Merging Framework for In-the-Wild Text-to-Image Generation
von: Li, Zhuoling, et al.
Veröffentlicht: (2026)
von: Li, Zhuoling, et al.
Veröffentlicht: (2026)
XQ-GAN: An Open-source Image Tokenization Framework for Autoregressive Generation
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
ImageFolder: Autoregressive Image Generation with Folded Tokens
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
ARTIST: Improving the Generation of Text-rich Images with Disentangled Diffusion Models and Large Language Models
von: Zhang, Jianyi, et al.
Veröffentlicht: (2024)
von: Zhang, Jianyi, et al.
Veröffentlicht: (2024)
LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding
von: Zhang, Yanzhe, et al.
Veröffentlicht: (2023)
von: Zhang, Yanzhe, et al.
Veröffentlicht: (2023)
TRINS: Towards Multimodal Language Models that Can Read
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
LaViDa-R1: Advancing Reasoning for Unified Multimodal Diffusion Language Models
von: Li, Shufan, et al.
Veröffentlicht: (2026)
von: Li, Shufan, et al.
Veröffentlicht: (2026)
Refer to Any Segmentation Mask Group With Vision-Language Prompts
von: Cao, Shengcao, et al.
Veröffentlicht: (2025)
von: Cao, Shengcao, et al.
Veröffentlicht: (2025)
Advancing Test-Time Adaptation in Wild Acoustic Test Settings
von: Liu, Hongfu, et al.
Veröffentlicht: (2023)
von: Liu, Hongfu, et al.
Veröffentlicht: (2023)
METAL: A Multi-Agent Framework for Chart Generation with Test-Time Scaling
von: Li, Bingxuan, et al.
Veröffentlicht: (2025)
von: Li, Bingxuan, et al.
Veröffentlicht: (2025)
Output Feedback Control for T‐S Fuzzy Markov Jump Systems Subjected to Parameter‐Dependent Dissipative Performance
von: Jian Wang, et al.
Veröffentlicht: (2024)
von: Jian Wang, et al.
Veröffentlicht: (2024)
MMR: Evaluating Reading Ability of Large Multimodal Models
von: Chen, Jian, et al.
Veröffentlicht: (2024)
von: Chen, Jian, et al.
Veröffentlicht: (2024)
Salutary Labeling with Zero Human Annotation
von: Xiao, Wenxiao, et al.
Veröffentlicht: (2024)
von: Xiao, Wenxiao, et al.
Veröffentlicht: (2024)
Strong convergence of a fully discrete scheme for stochastic Burgers equation with fractional-type noise
von: Wang, Yibo, et al.
Veröffentlicht: (2024)
von: Wang, Yibo, et al.
Veröffentlicht: (2024)
Approximation of the invariant measure for stochastic Allen-Cahn equation via an explicit fully discrete scheme
von: Wang, Yibo, et al.
Veröffentlicht: (2024)
von: Wang, Yibo, et al.
Veröffentlicht: (2024)
SV-RAG: LoRA-Contextualizing Adaptation of MLLMs for Long Document Understanding
von: Chen, Jian, et al.
Veröffentlicht: (2024)
von: Chen, Jian, et al.
Veröffentlicht: (2024)
Estimating Exoplanet Mass using Machine Learning on Incomplete Datasets
von: Lalande, Florian, et al.
Veröffentlicht: (2024)
von: Lalande, Florian, et al.
Veröffentlicht: (2024)
Liouville theorem for fully nonlinear elliptic equations with the small oscillation and the periodicity in $x$ and the periodic right hand term
von: Liang, Lichun
Veröffentlicht: (2026)
von: Liang, Lichun
Veröffentlicht: (2026)
The asymptotic behavior for divergence elliptic equations in exterior domains with periodic coefficients
von: Liang, Lichun
Veröffentlicht: (2026)
von: Liang, Lichun
Veröffentlicht: (2026)
RELOOP: Recursive Retrieval with Multi-Hop Reasoner and Planners for Heterogeneous QA
von: Yang, Ruiyi, et al.
Veröffentlicht: (2025)
von: Yang, Ruiyi, et al.
Veröffentlicht: (2025)
Numerical Pruning for Efficient Autoregressive Models
von: Shen, Xuan, et al.
Veröffentlicht: (2024)
von: Shen, Xuan, et al.
Veröffentlicht: (2024)
DraftAttention: Fast Video Diffusion via Low-Resolution Attention Guidance
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Toffee: Efficient Million-Scale Dataset Construction for Subject-Driven Text-to-Image Generation
von: Zhou, Yufan, et al.
Veröffentlicht: (2024) -
Customization Assistant for Text-to-image Generation
von: Zhou, Yufan, et al.
Veröffentlicht: (2023) -
OIDA International Journal of Sustainable Development
Veröffentlicht: (2017) -
SOHES: Self-supervised Open-world Hierarchical Entity Segmentation
von: Cao, Shengcao, et al.
Veröffentlicht: (2024) -
SNCE: Geometry-Aware Supervision for Scalable Discrete Image Generation
von: Li, Shufan, et al.
Veröffentlicht: (2026)