Cardiverse: Harnessing LLMs for Novel Card Game Prototyping
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Danrui, Zhang, Sen, Sohn, Sam S., Hu, Kaidong, Usman, Muhammad, Kapadia, Mubbasir |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Words to Worlds: Transforming One-line Prompt into Immersive Multi-modal Digital Stories with Communicative LLM Agent
von: Sohn, Samuel S., et al.
Veröffentlicht: (2024)
von: Sohn, Samuel S., et al.
Veröffentlicht: (2024)
GeoGuess: Multimodal Reasoning based on Hierarchy of Visual Information in Street View
von: Cheng, Fenghua, et al.
Veröffentlicht: (2025)
von: Cheng, Fenghua, et al.
Veröffentlicht: (2025)
OmniTrace: A Unified Framework for Generation-Time Attribution in Omni-Modal LLMs
von: Yan, Qianqi, et al.
Veröffentlicht: (2026)
von: Yan, Qianqi, et al.
Veröffentlicht: (2026)
Narrative-to-Scene Generation: An LLM-Driven Pipeline for 2D Game Environments
von: Chen, Yi-Chun, et al.
Veröffentlicht: (2025)
von: Chen, Yi-Chun, et al.
Veröffentlicht: (2025)
OmnixR: Evaluating Omni-modality Language Models on Reasoning across Modalities
von: Chen, Lichang, et al.
Veröffentlicht: (2024)
von: Chen, Lichang, et al.
Veröffentlicht: (2024)
FineFake: A Knowledge-Enriched Dataset for Fine-Grained Multi-Domain Fake News Detection
von: Zhou, Ziyi, et al.
Veröffentlicht: (2024)
von: Zhou, Ziyi, et al.
Veröffentlicht: (2024)
Can Large Language Models Help Multimodal Language Analysis? MMLA: A Comprehensive Benchmark
von: Zhang, Hanlei, et al.
Veröffentlicht: (2025)
von: Zhang, Hanlei, et al.
Veröffentlicht: (2025)
Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
Towards Robust Multimodal Sentiment Analysis with Incomplete Data
von: Zhang, Haoyu, et al.
Veröffentlicht: (2024)
von: Zhang, Haoyu, et al.
Veröffentlicht: (2024)
Interpretable Multimodal Misinformation Detection with Logic Reasoning
von: Liu, Hui, et al.
Veröffentlicht: (2023)
von: Liu, Hui, et al.
Veröffentlicht: (2023)
Shapley Value-based Contrastive Alignment for Multimodal Information Extraction
von: Luo, Wen, et al.
Veröffentlicht: (2024)
von: Luo, Wen, et al.
Veröffentlicht: (2024)
Reasoning LLMs are Wandering Solution Explorers
von: Lu, Jiahao, et al.
Veröffentlicht: (2025)
von: Lu, Jiahao, et al.
Veröffentlicht: (2025)
LLM-Guided Semantic Relational Reasoning for Multimodal Intent Recognition
von: Zhou, Qianrui, et al.
Veröffentlicht: (2025)
von: Zhou, Qianrui, et al.
Veröffentlicht: (2025)
Unsupervised Multimodal Clustering for Semantics Discovery in Multimodal Utterances
von: Zhang, Hanlei, et al.
Veröffentlicht: (2024)
von: Zhang, Hanlei, et al.
Veröffentlicht: (2024)
PTA: Enhancing Multimodal Sentiment Analysis through Pipelined Prediction and Translation-based Alignment
von: Song, Shezheng, et al.
Veröffentlicht: (2024)
von: Song, Shezheng, et al.
Veröffentlicht: (2024)
Multi-agent Undercover Gaming: Hallucination Removal via Counterfactual Test for Multimodal Reasoning
von: Liang, Dayong, et al.
Veröffentlicht: (2025)
von: Liang, Dayong, et al.
Veröffentlicht: (2025)
Knowledge-Guided Dynamic Modality Attention Fusion Framework for Multimodal Sentiment Analysis
von: Feng, Xinyu, et al.
Veröffentlicht: (2024)
von: Feng, Xinyu, et al.
Veröffentlicht: (2024)
Towards Better Text-to-Image Generation Alignment via Attention Modulation
von: Wu, Yihang, et al.
Veröffentlicht: (2024)
von: Wu, Yihang, et al.
Veröffentlicht: (2024)
A Survey on Image-text Multimodal Models
von: Guo, Ruifeng, et al.
Veröffentlicht: (2023)
von: Guo, Ruifeng, et al.
Veröffentlicht: (2023)
Tailored Teaching with Balanced Difficulty: Elevating Reasoning in Multimodal Chain-of-Thought via Prompt Curriculum
von: Yang, Xinglong, et al.
Veröffentlicht: (2025)
von: Yang, Xinglong, et al.
Veröffentlicht: (2025)
HeGTa: Leveraging Heterogeneous Graph-enhanced Large Language Models for Few-shot Complex Table Understanding
von: Jin, Rihui, et al.
Veröffentlicht: (2024)
von: Jin, Rihui, et al.
Veröffentlicht: (2024)
SIN-Bench: Tracing Native Evidence Chains in Long-Context Multimodal Scientific Interleaved Literature
von: Ren, Yiming, et al.
Veröffentlicht: (2026)
von: Ren, Yiming, et al.
Veröffentlicht: (2026)
TimeLens: Rethinking Video Temporal Grounding with Multimodal LLMs
von: Zhang, Jun, et al.
Veröffentlicht: (2025)
von: Zhang, Jun, et al.
Veröffentlicht: (2025)
GameTileNet: A Semantic Dataset for Low-Resolution Game Art in Procedural Content Generation
von: Chen, Yi-Chun, et al.
Veröffentlicht: (2025)
von: Chen, Yi-Chun, et al.
Veröffentlicht: (2025)
Traj-MLLM: Can Multimodal Large Language Models Reform Trajectory Data Mining?
von: Liu, Shuo, et al.
Veröffentlicht: (2025)
von: Liu, Shuo, et al.
Veröffentlicht: (2025)
SlideTailor: Personalized Presentation Slide Generation for Scientific Papers
von: Zeng, Wenzheng, et al.
Veröffentlicht: (2025)
von: Zeng, Wenzheng, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Generation for Electrocardiogram-Language Models
von: Song, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Song, Xiaoyu, et al.
Veröffentlicht: (2025)
Prolonged Reasoning Is Not All You Need: Certainty-Based Adaptive Routing for Efficient LLM/MLLM Reasoning
von: Lu, Jinghui, et al.
Veröffentlicht: (2025)
von: Lu, Jinghui, et al.
Veröffentlicht: (2025)
Beyond Spurious Signals: Debiasing Multimodal Large Language Models via Counterfactual Inference and Adaptive Expert Routing
von: Wu, Zichen, et al.
Veröffentlicht: (2025)
von: Wu, Zichen, et al.
Veröffentlicht: (2025)
A Survey of Generative Categories and Techniques in Multimodal Generative Models
von: Han, Longzhen, et al.
Veröffentlicht: (2025)
von: Han, Longzhen, et al.
Veröffentlicht: (2025)
MultimodalHugs: Enabling Sign Language Processing in Hugging Face
von: Sant, Gerard, et al.
Veröffentlicht: (2025)
von: Sant, Gerard, et al.
Veröffentlicht: (2025)
A Multimodal Framework for Explainable Evaluation of Soft Skills in Educational Environments
von: Guerrero-Sosa, Jared D. T., et al.
Veröffentlicht: (2025)
von: Guerrero-Sosa, Jared D. T., et al.
Veröffentlicht: (2025)
A Highly Clean Recipe Dataset with Ingredient States Annotation for State Probing Task
von: Toyooka, Mashiro, et al.
Veröffentlicht: (2025)
von: Toyooka, Mashiro, et al.
Veröffentlicht: (2025)
Memory-Centric Embodied Question Answering
von: Zhai, Mingliang, et al.
Veröffentlicht: (2025)
von: Zhai, Mingliang, et al.
Veröffentlicht: (2025)
CMMU: A Benchmark for Chinese Multi-modal Multi-type Question Understanding and Reasoning
von: He, Zheqi, et al.
Veröffentlicht: (2024)
von: He, Zheqi, et al.
Veröffentlicht: (2024)
"Is This It?": Towards Ecologically Valid Benchmarks for Situated Collaboration
von: Bohus, Dan, et al.
Veröffentlicht: (2024)
von: Bohus, Dan, et al.
Veröffentlicht: (2024)
Enhancing Multimodal Affective Analysis with Learned Live Comment Features
von: Deng, Zhaoyuan, et al.
Veröffentlicht: (2024)
von: Deng, Zhaoyuan, et al.
Veröffentlicht: (2024)
SemEval-2024 Task 3: Multimodal Emotion Cause Analysis in Conversations
von: Wang, Fanfan, et al.
Veröffentlicht: (2024)
von: Wang, Fanfan, et al.
Veröffentlicht: (2024)
History-Guided Iterative Visual Reasoning with Self-Correction
von: Yang, Xinglong, et al.
Veröffentlicht: (2026)
von: Yang, Xinglong, et al.
Veröffentlicht: (2026)
A Bounding Box is Worth One Token: Interleaving Layout and Text in a Large Language Model for Document Understanding
von: Lu, Jinghui, et al.
Veröffentlicht: (2024)
von: Lu, Jinghui, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
From Words to Worlds: Transforming One-line Prompt into Immersive Multi-modal Digital Stories with Communicative LLM Agent
von: Sohn, Samuel S., et al.
Veröffentlicht: (2024) -
GeoGuess: Multimodal Reasoning based on Hierarchy of Visual Information in Street View
von: Cheng, Fenghua, et al.
Veröffentlicht: (2025) -
OmniTrace: A Unified Framework for Generation-Time Attribution in Omni-Modal LLMs
von: Yan, Qianqi, et al.
Veröffentlicht: (2026) -
Narrative-to-Scene Generation: An LLM-Driven Pipeline for 2D Game Environments
von: Chen, Yi-Chun, et al.
Veröffentlicht: (2025) -
OmnixR: Evaluating Omni-modality Language Models on Reasoning across Modalities
von: Chen, Lichang, et al.
Veröffentlicht: (2024)