Evaluation of LLMs for Process Model Analysis and Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Kumar, Akhil, Zhao, Jianliang Leon, Dobariya, Om |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mind Your Tone: Investigating How Prompt Politeness Affects LLM Accuracy (short paper)
by: Dobariya, Om, et al.
Published: (2025)
by: Dobariya, Om, et al.
Published: (2025)
Evaluating LLMs' Assessment of Mixed-Context Hallucination Through the Lens of Summarization
by: Qi, Siya, et al.
Published: (2025)
by: Qi, Siya, et al.
Published: (2025)
Taxonomy and Analysis of Sensitive User Queries in Generative AI Search
by: Jo, Hwiyeol, et al.
Published: (2024)
by: Jo, Hwiyeol, et al.
Published: (2024)
CausalCite: A Causal Formulation of Paper Citations
by: Kumar, Ishan, et al.
Published: (2023)
by: Kumar, Ishan, et al.
Published: (2023)
Epistemic Diversity and Knowledge Collapse in Large Language Models
by: Wright, Dustin, et al.
Published: (2025)
by: Wright, Dustin, et al.
Published: (2025)
From Data to Behavior: Predicting Unintended Model Behaviors Before Training
by: Wang, Mengru, et al.
Published: (2026)
by: Wang, Mengru, et al.
Published: (2026)
Culinary Crossroads: A RAG Framework for Enhancing Diversity in Cross-Cultural Recipe Adaptation
by: Hu, Tianyi, et al.
Published: (2025)
by: Hu, Tianyi, et al.
Published: (2025)
Retrieval Improvements Do Not Guarantee Better Answers: A Study of RAG for AI Policy QA
by: Mathur, Saahil, et al.
Published: (2026)
by: Mathur, Saahil, et al.
Published: (2026)
Cross-Platform Digital Discourse Analysis of the Israel-Hamas Conflict: Sentiment, Topics, and Event Dynamics
by: Antonakaki, Despoina, et al.
Published: (2025)
by: Antonakaki, Despoina, et al.
Published: (2025)
Ground-Truth Subgraphs for Better Training and Evaluation of Knowledge Graph Augmented LLMs
by: Cattaneo, Alberto, et al.
Published: (2025)
by: Cattaneo, Alberto, et al.
Published: (2025)
From Facts to Conclusions : Integrating Deductive Reasoning in Retrieval-Augmented LLMs
by: Mishra, Shubham, et al.
Published: (2025)
by: Mishra, Shubham, et al.
Published: (2025)
Transit Pulse: Utilizing Social Media as a Source for Customer Feedback and Information Extraction with Large Language Model
by: Wang, Jiahao, et al.
Published: (2024)
by: Wang, Jiahao, et al.
Published: (2024)
Uncertainty and Fairness Awareness in LLM-Based Recommendation Systems
by: Sah, Chandan Kumar, et al.
Published: (2026)
by: Sah, Chandan Kumar, et al.
Published: (2026)
NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models
by: Lee, Chankyu, et al.
Published: (2024)
by: Lee, Chankyu, et al.
Published: (2024)
DiffETM: Diffusion Process Enhanced Embedded Topic Model
by: Shao, Wei, et al.
Published: (2025)
by: Shao, Wei, et al.
Published: (2025)
NyayaAnumana & INLegalLlama: The Largest Indian Legal Judgment Prediction Dataset and Specialized Language Model for Enhanced Decision Analysis
by: Nigam, Shubham Kumar, et al.
Published: (2024)
by: Nigam, Shubham Kumar, et al.
Published: (2024)
Monitoring the evolution of antisemitic discourse on extremist social media using BERT
by: Mustafa, Raza Ul, et al.
Published: (2024)
by: Mustafa, Raza Ul, et al.
Published: (2024)
Consumer-side Fairness in Recommender Systems: A Systematic Survey of Methods and Evaluation
by: Vassøy, Bjørnar, et al.
Published: (2023)
by: Vassøy, Bjørnar, et al.
Published: (2023)
Quantifying Document Impact in RAG-LLMs
by: Gerami, Armin, et al.
Published: (2025)
by: Gerami, Armin, et al.
Published: (2025)
Knowledge Conflicts for LLMs: A Survey
by: Xu, Rongwu, et al.
Published: (2024)
by: Xu, Rongwu, et al.
Published: (2024)
PARROT: A Benchmark for Evaluating LLMs in Cross-System SQL Translation
by: Zhou, Wei, et al.
Published: (2025)
by: Zhou, Wei, et al.
Published: (2025)
Using GPT Models for Qualitative and Quantitative News Analytics in the 2024 US Presidental Election Process
by: Pavlyshenko, Bohdan M.
Published: (2024)
by: Pavlyshenko, Bohdan M.
Published: (2024)
StealthRank: LLM Ranking Manipulation via Stealthy Prompt Optimization
by: Tang, Yiming, et al.
Published: (2025)
by: Tang, Yiming, et al.
Published: (2025)
Evaluating and Enhancing Large Language Models for Novelty Assessment in Scholarly Publications
by: Lin, Ethan, et al.
Published: (2024)
by: Lin, Ethan, et al.
Published: (2024)
Agentic Entropy-Balanced Policy Optimization
by: Dong, Guanting, et al.
Published: (2025)
by: Dong, Guanting, et al.
Published: (2025)
Do LLMs Understand Collaborative Signals? Diagnosis and Repair
by: Pouryousef, Shahrooz, et al.
Published: (2025)
by: Pouryousef, Shahrooz, et al.
Published: (2025)
A Comparative Study of Specialized LLMs as Dense Retrievers
by: Zhang, Hengran, et al.
Published: (2025)
by: Zhang, Hengran, et al.
Published: (2025)
CALRec: Contrastive Alignment of Generative LLMs for Sequential Recommendation
by: Li, Yaoyiran, et al.
Published: (2024)
by: Li, Yaoyiran, et al.
Published: (2024)
Emotional RAG LLMs: Reading Comprehension for the Open Internet
by: Reichman, Benjamin, et al.
Published: (2024)
by: Reichman, Benjamin, et al.
Published: (2024)
New Curriculum, New Chance -- Retrieval Augmented Generation for Lesson Planning in Ugandan Secondary Schools. Prototype Quality Evaluation
by: Kloker, Simon, et al.
Published: (2024)
by: Kloker, Simon, et al.
Published: (2024)
Personalized Benchmarking: Evaluating LLMs by Individual Preferences
by: Garbacea, Cristina, et al.
Published: (2026)
by: Garbacea, Cristina, et al.
Published: (2026)
OpenSanctions Pairs: Large-Scale Entity Matching with LLMs
by: Smith, Chandler, et al.
Published: (2026)
by: Smith, Chandler, et al.
Published: (2026)
A Systematic Evaluation of LLM Strategies for Mental Health Text Analysis: Fine-tuning vs. Prompt Engineering vs. RAG
by: Kermani, Arshia, et al.
Published: (2025)
by: Kermani, Arshia, et al.
Published: (2025)
Rethinking Legal Judgement Prediction in a Realistic Scenario in the Era of Large Language Models
by: Nigam, Shubham Kumar, et al.
Published: (2024)
by: Nigam, Shubham Kumar, et al.
Published: (2024)
An Iterative Utility Judgment Framework Inspired by Philosophical Relevance via LLMs
by: Zhang, Hengran, et al.
Published: (2024)
by: Zhang, Hengran, et al.
Published: (2024)
RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs
by: Yu, Yue, et al.
Published: (2024)
by: Yu, Yue, et al.
Published: (2024)
RAG Foundry: A Framework for Enhancing LLMs for Retrieval Augmented Generation
by: Fleischer, Daniel, et al.
Published: (2024)
by: Fleischer, Daniel, et al.
Published: (2024)
Knowledge-Augmented Large Language Models for Personalized Contextual Query Suggestion
by: Baek, Jinheon, et al.
Published: (2023)
by: Baek, Jinheon, et al.
Published: (2023)
Synthetic Knowledge Ingestion: Towards Knowledge Refinement and Injection for Enhancing Large Language Models
by: Zhang, Jiaxin, et al.
Published: (2024)
by: Zhang, Jiaxin, et al.
Published: (2024)
Is Implicit Knowledge Enough for LLMs? A RAG Approach for Tree-based Structures
by: Gupte, Mihir, et al.
Published: (2025)
by: Gupte, Mihir, et al.
Published: (2025)
Similar Items
-
Mind Your Tone: Investigating How Prompt Politeness Affects LLM Accuracy (short paper)
by: Dobariya, Om, et al.
Published: (2025) -
Evaluating LLMs' Assessment of Mixed-Context Hallucination Through the Lens of Summarization
by: Qi, Siya, et al.
Published: (2025) -
Taxonomy and Analysis of Sensitive User Queries in Generative AI Search
by: Jo, Hwiyeol, et al.
Published: (2024) -
CausalCite: A Causal Formulation of Paper Citations
by: Kumar, Ishan, et al.
Published: (2023) -
Epistemic Diversity and Knowledge Collapse in Large Language Models
by: Wright, Dustin, et al.
Published: (2025)