Multi-LLM Collaborative Caption Generation in Scientific Documents
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Jaeyoung, Lee, Jongho, Choi, Hong-Jun, Hsu, Ting-Yao, Huang, Chieh-Yang, Kim, Sungchul, Rossi, Ryan, Yu, Tong, Giles, Clyde Lee, Huang, Ting-Hao 'Kenneth', Choi, Sungchul |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SciCapenter: Supporting Caption Composition for Scientific Figures with Machine-Generated Captions and Ratings
by: Hsu, Ting-Yao, et al.
Published: (2024)
by: Hsu, Ting-Yao, et al.
Published: (2024)
Five Years of SciCap: What We Learned and Future Directions for Scientific Figure Captioning
by: Huang, Ting-Hao 'Kenneth', et al.
Published: (2025)
by: Huang, Ting-Hao 'Kenneth', et al.
Published: (2025)
Understanding How Paper Writers Use AI-Generated Captions in Figure Caption Writing
by: Yin, Ho, et al.
Published: (2025)
by: Yin, Ho, et al.
Published: (2025)
Do Large Multimodal Models Solve Caption Generation for Scientific Figures? Lessons Learned from SciCap Challenge 2023
by: Hsu, Ting-Yao E., et al.
Published: (2025)
by: Hsu, Ting-Yao E., et al.
Published: (2025)
Personalized Scientific Figure Caption Generation: An Empirical Study on Author-Specific Writing Style Transfer
by: Kim, Jaeyoung, et al.
Published: (2025)
by: Kim, Jaeyoung, et al.
Published: (2025)
LaMP-Cap: Personalized Figure Caption Generation With Multimodal Figure Profiles
by: Ng, Ho Yin 'Sam', et al.
Published: (2025)
by: Ng, Ho Yin 'Sam', et al.
Published: (2025)
Learning to Reduce: Optimal Representations of Structured Data in Prompting Large Language Models
by: Lee, Younghun, et al.
Published: (2024)
by: Lee, Younghun, et al.
Published: (2024)
Learning to Reduce: Towards Improving Performance of Large Language Models on Structured Data
by: Lee, Younghun, et al.
Published: (2024)
by: Lee, Younghun, et al.
Published: (2024)
DALDA: Data Augmentation Leveraging Diffusion Model and LLM with Adaptive Guidance Scaling
by: Jung, Kyuheon, et al.
Published: (2024)
by: Jung, Kyuheon, et al.
Published: (2024)
Multi-Agent Collaborative Filtering: Orchestrating Users and Items for Agentic Recommendations
by: Xia, Yu, et al.
Published: (2025)
by: Xia, Yu, et al.
Published: (2025)
PILOT-Bench: A Benchmark for Legal Reasoning in the Patent Domain with IRAC-Aligned Classification Tasks
by: Jang, Yehoon, et al.
Published: (2026)
by: Jang, Yehoon, et al.
Published: (2026)
SAND: Boosting LLM Agents with Self-Taught Action Deliberation
by: Xia, Yu, et al.
Published: (2025)
by: Xia, Yu, et al.
Published: (2025)
Charts Are Not Images: On the Challenges of Scientific Chart Editing
by: Li, Shawn, et al.
Published: (2025)
by: Li, Shawn, et al.
Published: (2025)
STELLAR: Scene Text Editor for Low-Resource Languages and Real-World Data
by: Seo, Yongdeuk, et al.
Published: (2025)
by: Seo, Yongdeuk, et al.
Published: (2025)
Forecasting VIX using interpretable Kolmogorov-Arnold networks
by: Cho, So-Yoon, et al.
Published: (2025)
by: Cho, So-Yoon, et al.
Published: (2025)
Improving the performance of optical inverse design of multilayer thin films using CNN-LSTM tandem neural networks
by: Jung, Uijun, et al.
Published: (2025)
by: Jung, Uijun, et al.
Published: (2025)
Diversify-verify-adapt: Efficient and Robust Retrieval-Augmented Ambiguous Question Answering
by: In, Yeonjun, et al.
Published: (2024)
by: In, Yeonjun, et al.
Published: (2024)
Uniform Pessimistic Risk and its Optimal Portfolio
by: Hong, Sungchul, et al.
Published: (2023)
by: Hong, Sungchul, et al.
Published: (2023)
VipAct: Visual-Perception Enhancement via Specialized VLM Agent Collaboration and Tool-use
by: Zhang, Zhehao, et al.
Published: (2024)
by: Zhang, Zhehao, et al.
Published: (2024)
FigCaps-HF: A Figure-to-Caption Generative Framework and Benchmark with Human Feedback
by: Singh, Ashish, et al.
Published: (2023)
by: Singh, Ashish, et al.
Published: (2023)
A Multi-LLM Debiasing Framework
by: Owens, Deonna M., et al.
Published: (2024)
by: Owens, Deonna M., et al.
Published: (2024)
Knowledge-Aware Query Expansion with Large Language Models for Textual and Relational Retrieval
by: Xia, Yu, et al.
Published: (2024)
by: Xia, Yu, et al.
Published: (2024)
QUICK: Quantization-aware Interleaving and Conflict-free Kernel for efficient LLM inference
by: Kim, Taesu, et al.
Published: (2024)
by: Kim, Taesu, et al.
Published: (2024)
Augment before You Try: Knowledge-Enhanced Table Question Answering via Table Expansion
by: Liu, Yujian, et al.
Published: (2024)
by: Liu, Yujian, et al.
Published: (2024)
Hallucination Diversity-Aware Active Learning for Text Summarization
by: Xia, Yu, et al.
Published: (2024)
by: Xia, Yu, et al.
Published: (2024)
Sharpness of convolution bounds for measures
by: Lee, Sanghyuk, et al.
Published: (2026)
by: Lee, Sanghyuk, et al.
Published: (2026)
Self-Debiasing Large Language Models: Zero-Shot Recognition and Reduction of Stereotypes
by: Gallegos, Isabel O., et al.
Published: (2024)
by: Gallegos, Isabel O., et al.
Published: (2024)
Generating Educational Materials with Different Levels of Readability using LLMs
by: Huang, Chieh-Yang, et al.
Published: (2024)
by: Huang, Chieh-Yang, et al.
Published: (2024)
VIVID: Human-AI Collaborative Authoring of Vicarious Dialogues from Lecture Videos
by: Choi, Seulgi, et al.
Published: (2024)
by: Choi, Seulgi, et al.
Published: (2024)
MAGNET: Autonomous Expert Model Generation via Decentralized Autoresearch and BitNet Training
by: Kim, Yongwan, et al.
Published: (2026)
by: Kim, Yongwan, et al.
Published: (2026)
Bias and Fairness in Large Language Models: A Survey
by: Gallegos, Isabel O., et al.
Published: (2023)
by: Gallegos, Isabel O., et al.
Published: (2023)
Interpretable Water Level Forecaster with Spatiotemporal Causal Attention Mechanisms
by: Hong, Sungchul, et al.
Published: (2023)
by: Hong, Sungchul, et al.
Published: (2023)
Masked Language Modeling Becomes Conditional Density Estimation for Tabular Data Synthesis
by: An, Seunghwan, et al.
Published: (2024)
by: An, Seunghwan, et al.
Published: (2024)
Is Safety Standard Same for Everyone? User-Specific Safety Evaluation of Large Language Models
by: In, Yeonjun, et al.
Published: (2025)
by: In, Yeonjun, et al.
Published: (2025)
How Does Conversation Length Impact User's Satisfaction? A Case Study of Length-Controlled Conversations with LLM-Powered Chatbots
by: Huang, Shih-Hong, et al.
Published: (2024)
by: Huang, Shih-Hong, et al.
Published: (2024)
EggHand: A Multimodal Foundation Model for Egocentric Hand Pose Forecasting
by: Choi, Jaeyoung, et al.
Published: (2026)
by: Choi, Jaeyoung, et al.
Published: (2026)
Examining Identity Drift in Conversations of LLM Agents
by: Choi, Junhyuk, et al.
Published: (2024)
by: Choi, Junhyuk, et al.
Published: (2024)
Optimizing Data Delivery: Insights from User Preferences on Visuals, Tables, and Text
by: Luera, Reuben, et al.
Published: (2024)
by: Luera, Reuben, et al.
Published: (2024)
Infusing Environmental Captions for Long-Form Video Language Grounding
by: Lee, Hyogun, et al.
Published: (2024)
by: Lee, Hyogun, et al.
Published: (2024)
Relevance to Utility: Process-Supervised Rewrite for RAG
by: Kim, Jaeyoung, et al.
Published: (2025)
by: Kim, Jaeyoung, et al.
Published: (2025)
Similar Items
-
SciCapenter: Supporting Caption Composition for Scientific Figures with Machine-Generated Captions and Ratings
by: Hsu, Ting-Yao, et al.
Published: (2024) -
Five Years of SciCap: What We Learned and Future Directions for Scientific Figure Captioning
by: Huang, Ting-Hao 'Kenneth', et al.
Published: (2025) -
Understanding How Paper Writers Use AI-Generated Captions in Figure Caption Writing
by: Yin, Ho, et al.
Published: (2025) -
Do Large Multimodal Models Solve Caption Generation for Scientific Figures? Lessons Learned from SciCap Challenge 2023
by: Hsu, Ting-Yao E., et al.
Published: (2025) -
Personalized Scientific Figure Caption Generation: An Empirical Study on Author-Specific Writing Style Transfer
by: Kim, Jaeyoung, et al.
Published: (2025)