MAGID: An Automated Pipeline for Generating Synthetic Multi-modal Datasets
Fuente:
arXiv
Saved in:
| Main Authors: | Aboutalebi, Hossein, Song, Hwanjun, Xie, Yusheng, Gupta, Arshit, Sun, Justin, Su, Hang, Shalyminov, Igor, Pappas, Nikolaos, Singh, Siffi, Mansour, Saab |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can Your Model Tell a Negation from an Implicature? Unravelling Challenges With Intent Encoders
by: Zhang, Yuwei, et al.
Published: (2024)
by: Zhang, Yuwei, et al.
Published: (2024)
FineSurE: Fine-grained Summarization Evaluation using LLMs
by: Song, Hwanjun, et al.
Published: (2024)
by: Song, Hwanjun, et al.
Published: (2024)
Semi-Supervised Dialogue Abstractive Summarization via High-Quality Pseudolabel Selection
by: He, Jianfeng, et al.
Published: (2024)
by: He, Jianfeng, et al.
Published: (2024)
Controllable Conversational Theme Detection Track at DSTC 12
by: Shalyminov, Igor, et al.
Published: (2025)
by: Shalyminov, Igor, et al.
Published: (2025)
CERET: Cost-Effective Extrinsic Refinement for Text Generation
by: Cai, Jason, et al.
Published: (2024)
by: Cai, Jason, et al.
Published: (2024)
TofuEval: Evaluating Hallucinations of LLMs on Topic-Focused Dialogue Summarization
by: Tang, Liyan, et al.
Published: (2024)
by: Tang, Liyan, et al.
Published: (2024)
Controllable Contextualized Image Captioning: Directing the Visual Narrative through User-Defined Highlights
by: Mao, Shunqi, et al.
Published: (2024)
by: Mao, Shunqi, et al.
Published: (2024)
FLAP: Flow-Adhering Planning with Constrained Decoding in LLMs
by: Roy, Shamik, et al.
Published: (2024)
by: Roy, Shamik, et al.
Published: (2024)
GLEAN: Active Generalized Category Discovery with Diverse LLM Feedback
by: Zou, Henry Peng, et al.
Published: (2025)
by: Zou, Henry Peng, et al.
Published: (2025)
DeAL: Decoding-time Alignment for Large Language Models
by: Huang, James Y., et al.
Published: (2024)
by: Huang, James Y., et al.
Published: (2024)
Faithful, Unfaithful or Ambiguous? Multi-Agent Debate with Initial Stance for Summary Evaluation
by: Koupaee, Mahnaz, et al.
Published: (2025)
by: Koupaee, Mahnaz, et al.
Published: (2025)
Alignment Tuning for Large Language Models: A Data-Centric Lens on Alignment Data Pipelines
by: Song, Hwanjun
Published: (2026)
by: Song, Hwanjun
Published: (2026)
Eliciting Better Multilingual Structured Reasoning from LLMs through Code
by: Li, Bryan, et al.
Published: (2024)
by: Li, Bryan, et al.
Published: (2024)
InsTALL: Context-aware Instructional Task Assistance with Multi-modal Large Language Models
by: Nguyen, Pha, et al.
Published: (2025)
by: Nguyen, Pha, et al.
Published: (2025)
DDQN-BASED ADAPTIVE LIGHTWEIGHT HONEYPOT FRAMEWORK FOR INTELLIGENT CYBER THREAT DETECTION IN SMALL AND MEDIUM ENTERPRISES
by: Arshit Rawat
Published: (2025)
by: Arshit Rawat
Published: (2025)
UniSumEval: Towards Unified, Fine-Grained, Multi-Dimensional Summarization Evaluation for LLMs
by: Lee, Yuho, et al.
Published: (2024)
by: Lee, Yuho, et al.
Published: (2024)
DeepfakeArt Challenge: A Benchmark Dataset for Generative AI Art Forgery and Data Poisoning Detection
by: Aboutalebi, Hossein, et al.
Published: (2023)
by: Aboutalebi, Hossein, et al.
Published: (2023)
Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence
by: Bang, Seunghwan, et al.
Published: (2026)
by: Bang, Seunghwan, et al.
Published: (2026)
LLM-based User Profile Management for Recommender System
by: Bang, Seunghwan, et al.
Published: (2025)
by: Bang, Seunghwan, et al.
Published: (2025)
CppPerf: An Automated Pipeline and Dataset for Performance-Improving C++ Commits
by: Ho, Tommy, et al.
Published: (2026)
by: Ho, Tommy, et al.
Published: (2026)
Exploiting Data Significance in Remote Estimation of Discrete-State Markov Sources
by: Luo, Jiping, et al.
Published: (2024)
by: Luo, Jiping, et al.
Published: (2024)
From AoI to QVAoI: Query-Based Semantics-Aware Scheduling for Energy-Harvesting IoT Systems
by: Delfani, Erfan, et al.
Published: (2024)
by: Delfani, Erfan, et al.
Published: (2024)
From Timestamps to Versions: Version AoI in Single- and Multi-Hop Networks
by: Delfani, Erfan, et al.
Published: (2025)
by: Delfani, Erfan, et al.
Published: (2025)
Real-Time Reconstruction and Actuation Error Analysis for Markov Sources over MPR Channels
by: Elessawy, Pansee S., et al.
Published: (2026)
by: Elessawy, Pansee S., et al.
Published: (2026)
Goal-oriented Estimation of Multiple Markov Sources in Resource-constrained Systems
by: Luo, Jiping, et al.
Published: (2023)
by: Luo, Jiping, et al.
Published: (2023)
Semantics-Aware Updates from Remote Energy Harvesting Devices to Interconnected LEO Satellites
by: Delfani, Erfan, et al.
Published: (2025)
by: Delfani, Erfan, et al.
Published: (2025)
Semantic-Aware Remote Estimation of Multiple Markov Sources Under Constraints
by: Luo, Jiping, et al.
Published: (2024)
by: Luo, Jiping, et al.
Published: (2024)
Computation-aware Energy-harvesting Federated Learning: Cyclic Scheduling with Selective Participation
by: Jeong, Eunjeong, et al.
Published: (2025)
by: Jeong, Eunjeong, et al.
Published: (2025)
Battery-aware Cyclic Scheduling in Energy-harvesting Federated Learning
by: Jeong, Eunjeong, et al.
Published: (2025)
by: Jeong, Eunjeong, et al.
Published: (2025)
On the Role of Age and Semantics of Information in Remote Estimation of Markov Sources
by: Luo, Jiping, et al.
Published: (2025)
by: Luo, Jiping, et al.
Published: (2025)
Technology‐Enabled Competitiveness and Experiences in Tourism: A Transformative Era
by: Eleni Michopoulou, et al.
Published: (2025)
by: Eleni Michopoulou, et al.
Published: (2025)
Computing the Exact Pareto Front in Average-Cost Multi-Objective Markov Decision Processes
by: Luo, Jiping, et al.
Published: (2026)
by: Luo, Jiping, et al.
Published: (2026)
Optimizing Version AoI in Energy-Harvesting IoT: Model-Based and Learning-Based Approaches
by: Delfani, Erfan, et al.
Published: (2025)
by: Delfani, Erfan, et al.
Published: (2025)
On the Cost of Consecutive Estimation Error: Significance-Aware Non-linear Aging
by: Luo, Jiping, et al.
Published: (2024)
by: Luo, Jiping, et al.
Published: (2024)
Joint Accuracy and Confidentiality in Semantic-Aware Secure Remote Reconstruction
by: Li, Bowen, et al.
Published: (2026)
by: Li, Bowen, et al.
Published: (2026)
Optimizing Information Freshness in IoT Systems with Update Rate Constraints: A Token-Based Approach
by: Delfani, Erfan, et al.
Published: (2024)
by: Delfani, Erfan, et al.
Published: (2024)
Learning to Summarize from LLM-generated Feedback
by: Song, Hwanjun, et al.
Published: (2024)
by: Song, Hwanjun, et al.
Published: (2024)
Aligning Extraction and Generation for Robust Retrieval-Augmented Generation
by: Song, Hwanjun, et al.
Published: (2025)
by: Song, Hwanjun, et al.
Published: (2025)
Quantile-Free Uncertainty Quantification in Graph Neural Networks
by: park, Soyoung, et al.
Published: (2026)
by: park, Soyoung, et al.
Published: (2026)
DeepPersona: A Generative Engine for Scaling Deep Synthetic Personas
by: Wang, Zhen, et al.
Published: (2025)
by: Wang, Zhen, et al.
Published: (2025)
Similar Items
-
Can Your Model Tell a Negation from an Implicature? Unravelling Challenges With Intent Encoders
by: Zhang, Yuwei, et al.
Published: (2024) -
FineSurE: Fine-grained Summarization Evaluation using LLMs
by: Song, Hwanjun, et al.
Published: (2024) -
Semi-Supervised Dialogue Abstractive Summarization via High-Quality Pseudolabel Selection
by: He, Jianfeng, et al.
Published: (2024) -
Controllable Conversational Theme Detection Track at DSTC 12
by: Shalyminov, Igor, et al.
Published: (2025) -
CERET: Cost-Effective Extrinsic Refinement for Text Generation
by: Cai, Jason, et al.
Published: (2024)