Saved in:
| Main Authors: | Qi, Xuan, He, Luxi, Roth, Dan, Fu, Xingyu |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2603.19688 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FamiCom: Further Demystifying Prompts for Language Models with Task-Agnostic Performance Estimation
by: Li, Bangzheng, et al.
Published: (2024)
by: Li, Bangzheng, et al.
Published: (2024)
Commonsense-T2I Challenge: Can Text-to-Image Generation Models Understand Commonsense?
by: Fu, Xingyu, et al.
Published: (2024)
by: Fu, Xingyu, et al.
Published: (2024)
What is in Your Safe Data? Identifying Benign Data that Breaks Safety
by: He, Luxi, et al.
Published: (2024)
by: He, Luxi, et al.
Published: (2024)
Conflicts in Texts: Data, Implications and Challenges
by: Liu, Siyi, et al.
Published: (2025)
by: Liu, Siyi, et al.
Published: (2025)
Demystifying CLIP Data
by: Xu, Hu, et al.
Published: (2023)
by: Xu, Hu, et al.
Published: (2023)
Learning Human-Perceived Fakeness in AI-Generated Videos via Multimodal LLMs
by: Fu, Xingyu, et al.
Published: (2025)
by: Fu, Xingyu, et al.
Published: (2025)
Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models
by: Hu, Yushi, et al.
Published: (2024)
by: Hu, Yushi, et al.
Published: (2024)
Demystifying Domain-adaptive Post-training for Financial LLMs
by: Ke, Zixuan, et al.
Published: (2025)
by: Ke, Zixuan, et al.
Published: (2025)
Dynamic Clue Bottlenecks: Towards Interpretable-by-Design Visual Question Answering
by: Fu, Xingyu, et al.
Published: (2023)
by: Fu, Xingyu, et al.
Published: (2023)
Influential Training Data Retrieval for Explaining Verbalized Confidence of LLMs
by: Xia, Yuxi, et al.
Published: (2026)
by: Xia, Yuxi, et al.
Published: (2026)
Deceptive Semantic Shortcuts on Reasoning Chains: How Far Can Models Go without Hallucination?
by: Li, Bangzheng, et al.
Published: (2023)
by: Li, Bangzheng, et al.
Published: (2023)
On the Calibration of Multilingual Question Answering LLMs
by: Yang, Yahan, et al.
Published: (2023)
by: Yang, Yahan, et al.
Published: (2023)
UEval: A Benchmark for Unified Multimodal Generation
by: Li, Bo, et al.
Published: (2026)
by: Li, Bo, et al.
Published: (2026)
Demystifying When Pruning Works via Representation Hierarchies
by: He, Shwai, et al.
Published: (2026)
by: He, Shwai, et al.
Published: (2026)
EVALALIGN: Supervised Fine-Tuning Multimodal LLMs with Human-Aligned Data for Evaluating Text-to-Image Models
by: Tan, Zhiyu, et al.
Published: (2024)
by: Tan, Zhiyu, et al.
Published: (2024)
CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs
by: Wang, Zirui, et al.
Published: (2024)
by: Wang, Zirui, et al.
Published: (2024)
Demystifying Long Chain-of-Thought Reasoning in LLMs
by: Yeo, Edward, et al.
Published: (2025)
by: Yeo, Edward, et al.
Published: (2025)
LLM-as-a-Prophet: Understanding Predictive Intelligence with Prophet Arena
by: Yang, Qingchuan, et al.
Published: (2025)
by: Yang, Qingchuan, et al.
Published: (2025)
Can LLMs Narrate Tabular Data? An Evaluation Framework for Natural Language Representations of Text-to-SQL System Outputs
by: Singh, Jyotika, et al.
Published: (2025)
by: Singh, Jyotika, et al.
Published: (2025)
BLINK: Multimodal Large Language Models Can See but Not Perceive
by: Fu, Xingyu, et al.
Published: (2024)
by: Fu, Xingyu, et al.
Published: (2024)
Multimodal Misinformation Detection by Learning from Synthetic Data with Multimodal LLMs
by: Zeng, Fengzhu, et al.
Published: (2024)
by: Zeng, Fengzhu, et al.
Published: (2024)
Multimodal Coreference Resolution for Chinese Social Media Dialogues: Dataset and Benchmark Approach
by: Li, Xingyu, et al.
Published: (2025)
by: Li, Xingyu, et al.
Published: (2025)
Demystifying Scientific Problem-Solving in LLMs by Probing Knowledge and Reasoning
by: Li, Alan, et al.
Published: (2025)
by: Li, Alan, et al.
Published: (2025)
Chart-RL: Generalized Chart Comprehension via Reinforcement Learning with Verifiable Rewards
by: Zhang, Xin, et al.
Published: (2026)
by: Zhang, Xin, et al.
Published: (2026)
Arctic-SnowCoder: Demystifying High-Quality Data in Code Pretraining
by: Wei, Yuxiang, et al.
Published: (2024)
by: Wei, Yuxiang, et al.
Published: (2024)
OPDAI at SemEval-2024 Task 6: Small LLMs can Accelerate Hallucination Detection with Weakly Supervised Data
by: Wei, Chengcheng, et al.
Published: (2024)
by: Wei, Chengcheng, et al.
Published: (2024)
GraphGen: Enhancing Supervised Fine-Tuning for LLMs with Knowledge-Driven Synthetic Data Generation
by: Chen, Zihong, et al.
Published: (2025)
by: Chen, Zihong, et al.
Published: (2025)
AIDE: Attribute-Guided MultI-Hop Data Expansion for Data Scarcity in Task-Specific Fine-tuning
by: Li, Jiayu, et al.
Published: (2024)
by: Li, Jiayu, et al.
Published: (2024)
Evaluating LLMs' Mathematical Reasoning in Financial Document Question Answering
by: Srivastava, Pragya, et al.
Published: (2024)
by: Srivastava, Pragya, et al.
Published: (2024)
Mask Tokens as Prophet: Fine-Grained Cache Eviction for Efficient dLLM Inference
by: Huang, Jianuo, et al.
Published: (2025)
by: Huang, Jianuo, et al.
Published: (2025)
SocREval: Large Language Models with the Socratic Method for Reference-Free Reasoning Evaluation
by: He, Hangfeng, et al.
Published: (2023)
by: He, Hangfeng, et al.
Published: (2023)
From Measurement Instruments to Data: Leveraging Theory-Driven Synthetic Training Data for Classifying Social Constructs
by: Birkenmaier, Lukas, et al.
Published: (2024)
by: Birkenmaier, Lukas, et al.
Published: (2024)
LLMs are Single-threaded Reasoners: Demystifying the Working Mechanism of Soft Thinking
by: Wu, Junhong, et al.
Published: (2025)
by: Wu, Junhong, et al.
Published: (2025)
Reasoning is about giving reasons
by: Shah, Krunal, et al.
Published: (2025)
by: Shah, Krunal, et al.
Published: (2025)
Data-efficient LLM Fine-tuning for Code Generation
by: Lv, Weijie, et al.
Published: (2025)
by: Lv, Weijie, et al.
Published: (2025)
ReFocus: Visual Editing as a Chain of Thought for Structured Image Understanding
by: Fu, Xingyu, et al.
Published: (2025)
by: Fu, Xingyu, et al.
Published: (2025)
Enhancing Temporal Understanding in LLMs for Semi-structured Tables
by: Deng, Irwin, et al.
Published: (2024)
by: Deng, Irwin, et al.
Published: (2024)
Ziya2: Data-centric Learning is All LLMs Need
by: Gan, Ruyi, et al.
Published: (2023)
by: Gan, Ruyi, et al.
Published: (2023)
Fine-Tuning LLMs for Report Summarization: Analysis on Supervised and Unsupervised Data
by: Rallapalli, Swati, et al.
Published: (2025)
by: Rallapalli, Swati, et al.
Published: (2025)
Budget-Aware Anytime Reasoning with LLM-Synthesized Preference Data
by: Zhang, Xuanming, et al.
Published: (2026)
by: Zhang, Xuanming, et al.
Published: (2026)
Similar Items
-
FamiCom: Further Demystifying Prompts for Language Models with Task-Agnostic Performance Estimation
by: Li, Bangzheng, et al.
Published: (2024) -
Commonsense-T2I Challenge: Can Text-to-Image Generation Models Understand Commonsense?
by: Fu, Xingyu, et al.
Published: (2024) -
What is in Your Safe Data? Identifying Benign Data that Breaks Safety
by: He, Luxi, et al.
Published: (2024) -
Conflicts in Texts: Data, Implications and Challenges
by: Liu, Siyi, et al.
Published: (2025) -
Demystifying CLIP Data
by: Xu, Hu, et al.
Published: (2023)