MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Joshi, Siddharth, Nushi, Besmira, Balachandran, Vidhisha, Chandrasekaran, Varun, Vineet, Vibhav, Joshi, Neel, Mirzasoleiman, Baharan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BenchAgents: Multi-Agent Systems for Structured Benchmark Creation
von: Butt, Natasha, et al.
Veröffentlicht: (2024)
von: Butt, Natasha, et al.
Veröffentlicht: (2024)
Unearthing Skill-Level Insights for Understanding Trade-Offs of Foundation Models
von: Moayeri, Mazda, et al.
Veröffentlicht: (2024)
von: Moayeri, Mazda, et al.
Veröffentlicht: (2024)
Eureka: Evaluating and Understanding Large Foundation Models
von: Balachandran, Vidhisha, et al.
Veröffentlicht: (2024)
von: Balachandran, Vidhisha, et al.
Veröffentlicht: (2024)
Inference-Time Scaling for Complex Tasks: Where We Stand and What Lies Ahead
von: Balachandran, Vidhisha, et al.
Veröffentlicht: (2025)
von: Balachandran, Vidhisha, et al.
Veröffentlicht: (2025)
Attention Speaks Volumes: Localizing and Mitigating Bias in Language Models
von: Adiga, Rishabh, et al.
Veröffentlicht: (2024)
von: Adiga, Rishabh, et al.
Veröffentlicht: (2024)
Improving Instruction-Following in Language Models through Activation Steering
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
Just Do It!? Computer-Use Agents Exhibit Blind Goal-Directedness
von: Shayegani, Erfan, et al.
Veröffentlicht: (2025)
von: Shayegani, Erfan, et al.
Veröffentlicht: (2025)
What MLLMs Learn about When they Learn about Multimodal Reasoning
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
Diversity of Thought Improves Reasoning Abilities of LLMs
von: Naik, Ranjita, et al.
Veröffentlicht: (2023)
von: Naik, Ranjita, et al.
Veröffentlicht: (2023)
Data-Efficient Contrastive Self-supervised Learning: Most Beneficial Examples for Supervised Learning Contribute the Least
von: Joshi, Siddharth, et al.
Veröffentlicht: (2023)
von: Joshi, Siddharth, et al.
Veröffentlicht: (2023)
Data-Efficient Contrastive Language-Image Pretraining: Prioritizing Data Quality over Quantity
von: Joshi, Siddharth, et al.
Veröffentlicht: (2024)
von: Joshi, Siddharth, et al.
Veröffentlicht: (2024)
Few-shot Adaptation to Distribution Shifts By Mixing Source and Target Embeddings
von: Xue, Yihao, et al.
Veröffentlicht: (2023)
von: Xue, Yihao, et al.
Veröffentlicht: (2023)
Attention Satisfies: A Constraint-Satisfaction Lens on Factual Errors of Language Models
von: Yuksekgonul, Mert, et al.
Veröffentlicht: (2023)
von: Yuksekgonul, Mert, et al.
Veröffentlicht: (2023)
Beyond Semantic Entropy: Boosting LLM Uncertainty Quantification with Pairwise Semantic Similarity
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
Phi-4-reasoning Technical Report
von: Abdin, Marah, et al.
Veröffentlicht: (2025)
von: Abdin, Marah, et al.
Veröffentlicht: (2025)
Investigating the Benefits of Projection Head for Representation Learning
von: Xue, Yihao, et al.
Veröffentlicht: (2024)
von: Xue, Yihao, et al.
Veröffentlicht: (2024)
Dataset Distillation via Knowledge Distillation: Towards Efficient Self-Supervised Pre-Training of Deep Networks
von: Joshi, Siddharth, et al.
Veröffentlicht: (2024)
von: Joshi, Siddharth, et al.
Veröffentlicht: (2024)
Elephants Never Forget: Memorization and Learning of Tabular Data in Large Language Models
von: Bordt, Sebastian, et al.
Veröffentlicht: (2024)
von: Bordt, Sebastian, et al.
Veröffentlicht: (2024)
Detecting Data Contamination in LLMs via In-Context Learning
von: Zawalski, Michał, et al.
Veröffentlicht: (2025)
von: Zawalski, Michał, et al.
Veröffentlicht: (2025)
Verify when Uncertain: Beyond Self-Consistency in Black Box Hallucination Detection
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
SmallToLarge (S2L): Scalable Data Selection for Fine-tuning Large Language Models by Summarizing Training Trajectories of Small Models
von: Yang, Yu, et al.
Veröffentlicht: (2024)
von: Yang, Yu, et al.
Veröffentlicht: (2024)
Mini-batch Coresets for Memory-efficient Language Model Training on Data Mixtures
von: Nguyen, Dang, et al.
Veröffentlicht: (2024)
von: Nguyen, Dang, et al.
Veröffentlicht: (2024)
From Hallucinations to Facts: Enhancing Language Models with Curated Knowledge Graphs
von: Joshi, Ratnesh Kumar, et al.
Veröffentlicht: (2024)
von: Joshi, Ratnesh Kumar, et al.
Veröffentlicht: (2024)
Better Safe than Sorry: Pre-training CLIP against Targeted Data Poisoning and Backdoor Attacks
von: Yang, Wenhan, et al.
Veröffentlicht: (2023)
von: Yang, Wenhan, et al.
Veröffentlicht: (2023)
Tracing the Traces: Latent Temporal Signals for Efficient and Accurate Reasoning
von: Vilas, Martina G., et al.
Veröffentlicht: (2025)
von: Vilas, Martina G., et al.
Veröffentlicht: (2025)
Understanding the Robustness of Multi-modal Contrastive Learning to Distribution Shift
von: Xue, Yihao, et al.
Veröffentlicht: (2023)
von: Xue, Yihao, et al.
Veröffentlicht: (2023)
Do We Need All the Synthetic Data? Targeted Image Augmentation via Diffusion Models
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
Synthetic Text Generation for Training Large Language Models via Gradient Matching
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
TarGEN: Targeted Data Generation with Large Language Models
von: Gupta, Himanshu, et al.
Veröffentlicht: (2023)
von: Gupta, Himanshu, et al.
Veröffentlicht: (2023)
Exposing the Achilles' Heel: Evaluating LLMs Ability to Handle Mistakes in Mathematical Reasoning
von: Singh, Joykirat, et al.
Veröffentlicht: (2024)
von: Singh, Joykirat, et al.
Veröffentlicht: (2024)
Don't Hallucinate, Abstain: Identifying LLM Knowledge Gaps via Multi-LLM Collaboration
von: Feng, Shangbin, et al.
Veröffentlicht: (2024)
von: Feng, Shangbin, et al.
Veröffentlicht: (2024)
Knowledge Card: Filling LLMs' Knowledge Gaps with Plug-in Specialized Language Models
von: Feng, Shangbin, et al.
Veröffentlicht: (2023)
von: Feng, Shangbin, et al.
Veröffentlicht: (2023)
Reasoning Up the Instruction Ladder for Controllable Language Models
von: Zheng, Zishuo, et al.
Veröffentlicht: (2025)
von: Zheng, Zishuo, et al.
Veröffentlicht: (2025)
FACTS&EVIDENCE: An Interactive Tool for Transparent Fine-Grained Factual Verification of Machine-Generated Text
von: Boonsanong, Varich, et al.
Veröffentlicht: (2025)
von: Boonsanong, Varich, et al.
Veröffentlicht: (2025)
Infinity-MM: Scaling Multimodal Performance with Large-Scale and High-Quality Instruction Data
von: Gu, Shuhao, et al.
Veröffentlicht: (2024)
von: Gu, Shuhao, et al.
Veröffentlicht: (2024)
KGQuiz: Evaluating the Generalization of Encoded Knowledge in Large Language Models
von: Bai, Yuyang, et al.
Veröffentlicht: (2023)
von: Bai, Yuyang, et al.
Veröffentlicht: (2023)
Fine-grained Hallucination Detection and Editing for Language Models
von: Mishra, Abhika, et al.
Veröffentlicht: (2024)
von: Mishra, Abhika, et al.
Veröffentlicht: (2024)
Resolving Knowledge Conflicts in Large Language Models
von: Wang, Yike, et al.
Veröffentlicht: (2023)
von: Wang, Yike, et al.
Veröffentlicht: (2023)
Sample More to Think Less: Group Filtered Policy Optimization for Concise Reasoning
von: Shrivastava, Vaishnavi, et al.
Veröffentlicht: (2025)
von: Shrivastava, Vaishnavi, et al.
Veröffentlicht: (2025)
Curating Stopwords in Marathi: A TF-IDF Approach for Improved Text Analysis and Information Retrieval
von: Chavan, Rohan, et al.
Veröffentlicht: (2024)
von: Chavan, Rohan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
BenchAgents: Multi-Agent Systems for Structured Benchmark Creation
von: Butt, Natasha, et al.
Veröffentlicht: (2024) -
Unearthing Skill-Level Insights for Understanding Trade-Offs of Foundation Models
von: Moayeri, Mazda, et al.
Veröffentlicht: (2024) -
Eureka: Evaluating and Understanding Large Foundation Models
von: Balachandran, Vidhisha, et al.
Veröffentlicht: (2024) -
Inference-Time Scaling for Complex Tasks: Where We Stand and What Lies Ahead
von: Balachandran, Vidhisha, et al.
Veröffentlicht: (2025) -
Attention Speaks Volumes: Localizing and Mitigating Bias in Language Models
von: Adiga, Rishabh, et al.
Veröffentlicht: (2024)