Trustworthy LLM-Mediated Communication: Evaluating Information Fidelity in LLM as a Communicator (LAAC) Framework in Multiple Application Domains
Fuente:
arXiv
Saved in:
| Main Authors: | Rafi, Mohammed Musthafa, Krishnamurthy, Adarsh, Balu, Aditya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Algorithmic Fragility and Persona Bias in LLM-Generated Autistic Communication
by: Rizvi, Naba, et al.
Published: (2026)
by: Rizvi, Naba, et al.
Published: (2026)
Trustworthy Reasoning: Evaluating and Enhancing Factual Accuracy in LLM Intermediate Thought Processes
by: Jiao, Rui, et al.
Published: (2025)
by: Jiao, Rui, et al.
Published: (2025)
Communication and Verification in LLM Agents towards Collaboration under Information Asymmetry
by: Peng, Run, et al.
Published: (2025)
by: Peng, Run, et al.
Published: (2025)
WorkRB: A Community-Driven Evaluation Framework for AI in the Work Domain
by: De Lange, Matthias, et al.
Published: (2026)
by: De Lange, Matthias, et al.
Published: (2026)
Adapting LLM Agents with Universal Feedback in Communication
by: Wang, Kuan, et al.
Published: (2023)
by: Wang, Kuan, et al.
Published: (2023)
LLM Prompt Evaluation for Educational Applications
by: Holmes, Langdon, et al.
Published: (2026)
by: Holmes, Langdon, et al.
Published: (2026)
Evaluating LLM Reasoning in the Operations Research Domain with ORQA
by: Mostajabdaveh, Mahdi, et al.
Published: (2024)
by: Mostajabdaveh, Mahdi, et al.
Published: (2024)
Communication is All You Need: Persuasion Dataset Construction via Multi-LLM Communication
by: Ma, Weicheng, et al.
Published: (2025)
by: Ma, Weicheng, et al.
Published: (2025)
The Challenges of Evaluating LLM Applications: An Analysis of Automated, Human, and LLM-Based Approaches
by: Abeysinghe, Bhashithe, et al.
Published: (2024)
by: Abeysinghe, Bhashithe, et al.
Published: (2024)
keepitsimple at SemEval-2025 Task 3: LLM-Uncertainty based Approach for Multilingual Hallucination Span Detection
by: Vemula, Saketh Reddy, et al.
Published: (2025)
by: Vemula, Saketh Reddy, et al.
Published: (2025)
EVE: A Domain-Specific LLM Framework for Earth Intelligence
by: Atrio, Àlex R., et al.
Published: (2026)
by: Atrio, Àlex R., et al.
Published: (2026)
Protect: Towards Robust Guardrailing Stack for Trustworthy Enterprise LLM Systems
by: Avinash, Karthik, et al.
Published: (2025)
by: Avinash, Karthik, et al.
Published: (2025)
Integrated Framework for LLM Evaluation with Answer Generation
by: Lee, Sujeong, et al.
Published: (2025)
by: Lee, Sujeong, et al.
Published: (2025)
LLM4Sweat: A Trustworthy Large Language Model for Hyperhidrosis Support
by: Lin, Wenjie, et al.
Published: (2025)
by: Lin, Wenjie, et al.
Published: (2025)
Towards Multi-dimensional Evaluation of LLM Summarization across Domains and Languages
by: Min, Hyangsuk, et al.
Published: (2025)
by: Min, Hyangsuk, et al.
Published: (2025)
Breaking the Ceiling of the LLM Community by Treating Token Generation as a Classification for Ensembling
by: Yu, Yao-Ching, et al.
Published: (2024)
by: Yu, Yao-Ching, et al.
Published: (2024)
Polypersona: Persona-Grounded LLM for Synthetic Survey Responses
by: Dash, Tejaswani, et al.
Published: (2025)
by: Dash, Tejaswani, et al.
Published: (2025)
Simulating LLM-to-LLM Tutoring for Multilingual Math Feedback
by: Tonga, Junior Cedric, et al.
Published: (2025)
by: Tonga, Junior Cedric, et al.
Published: (2025)
Modeling Community Attitude through Reaction Tone: A Human-AI Collaborative Framework for Evaluating LLM Alignment with Linguistic Behaviors in Online Communities
by: Wen, Nuan, et al.
Published: (2026)
by: Wen, Nuan, et al.
Published: (2026)
How Trustworthy Are LLM-as-Judge Ratings for Interpretive Responses? Implications for Qualitative Research Workflows
by: Han, Songhee, et al.
Published: (2026)
by: Han, Songhee, et al.
Published: (2026)
Communication Compression for Tensor Parallel LLM Inference
by: Hansen-Palmus, Jan, et al.
Published: (2024)
by: Hansen-Palmus, Jan, et al.
Published: (2024)
Dynamic Generation of Multi-LLM Agents Communication Topologies with Graph Diffusion Models
by: Jiang, Eric Hanchen, et al.
Published: (2025)
by: Jiang, Eric Hanchen, et al.
Published: (2025)
CLAVE: An Adaptive Framework for Evaluating Values of LLM Generated Responses
by: Yao, Jing, et al.
Published: (2024)
by: Yao, Jing, et al.
Published: (2024)
SemBench: A Universal Semantic Framework for LLM Evaluation
by: Zubillaga, Mikel, et al.
Published: (2026)
by: Zubillaga, Mikel, et al.
Published: (2026)
STELLAR-E: a Synthetic, Tailored, End-to-end LLM Application Rigorous Evaluator
by: Sordo, Alessio, et al.
Published: (2026)
by: Sordo, Alessio, et al.
Published: (2026)
MEQA: A Meta-Evaluation Framework for Question & Answer LLM Benchmarks
by: Veuthey, Jaime Raldua, et al.
Published: (2025)
by: Veuthey, Jaime Raldua, et al.
Published: (2025)
Socio-Culturally Aware Evaluation Framework for LLM-Based Content Moderation
by: Kumar, Shanu, et al.
Published: (2024)
by: Kumar, Shanu, et al.
Published: (2024)
RvLLM: LLM Runtime Verification with Domain Knowledge
by: Zhang, Yedi, et al.
Published: (2025)
by: Zhang, Yedi, et al.
Published: (2025)
LELA: An End-to-end LLM-based Entity Linking Framework with Zero-shot Domain Adaptation
by: Haffoudhi, Samy, et al.
Published: (2026)
by: Haffoudhi, Samy, et al.
Published: (2026)
Epistemic Alignment: A Mediating Framework for User-LLM Knowledge Delivery
by: Clark, Nicholas, et al.
Published: (2025)
by: Clark, Nicholas, et al.
Published: (2025)
MEDIC: Comprehensive Evaluation of Leading Indicators for LLM Safety and Utility in Clinical Applications
by: Kanithi, Praveenkumar, et al.
Published: (2024)
by: Kanithi, Praveenkumar, et al.
Published: (2024)
GuideLLM: Exploring LLM-Guided Conversation with Applications in Autobiography Interviewing
by: Duan, Jinhao, et al.
Published: (2025)
by: Duan, Jinhao, et al.
Published: (2025)
FreeEval: A Modular Framework for Trustworthy and Efficient Evaluation of Large Language Models
by: Yu, Zhuohao, et al.
Published: (2024)
by: Yu, Zhuohao, et al.
Published: (2024)
AI Security Beyond Core Domains: Resume Screening as a Case Study of Adversarial Vulnerabilities in Specialized LLM Applications
by: Mu, Honglin, et al.
Published: (2025)
by: Mu, Honglin, et al.
Published: (2025)
Multiple LLM Agents Debate for Equitable Cultural Alignment
by: Ki, Dayeon, et al.
Published: (2025)
by: Ki, Dayeon, et al.
Published: (2025)
STED and Consistency Scoring: A Framework for Evaluating LLM Structured Output Reliability
by: Wang, Guanghui, et al.
Published: (2025)
by: Wang, Guanghui, et al.
Published: (2025)
Locomo-Plus: Beyond-Factual Cognitive Memory Evaluation Framework for LLM Agents
by: Li, Yifei, et al.
Published: (2026)
by: Li, Yifei, et al.
Published: (2026)
AGENTiGraph: A Multi-Agent Knowledge Graph Framework for Interactive, Domain-Specific LLM Chatbots
by: Zhao, Xinjie, et al.
Published: (2025)
by: Zhao, Xinjie, et al.
Published: (2025)
SLearnLLM: A Self-Learning Framework for Efficient Domain-Specific Adaptation of Large Language Models
by: Liu, Xiang, et al.
Published: (2025)
by: Liu, Xiang, et al.
Published: (2025)
LeMAJ (Legal LLM-as-a-Judge): Bridging Legal Reasoning and LLM Evaluation
by: Enguehard, Joseph, et al.
Published: (2025)
by: Enguehard, Joseph, et al.
Published: (2025)
Similar Items
-
Algorithmic Fragility and Persona Bias in LLM-Generated Autistic Communication
by: Rizvi, Naba, et al.
Published: (2026) -
Trustworthy Reasoning: Evaluating and Enhancing Factual Accuracy in LLM Intermediate Thought Processes
by: Jiao, Rui, et al.
Published: (2025) -
Communication and Verification in LLM Agents towards Collaboration under Information Asymmetry
by: Peng, Run, et al.
Published: (2025) -
WorkRB: A Community-Driven Evaluation Framework for AI in the Work Domain
by: De Lange, Matthias, et al.
Published: (2026) -
Adapting LLM Agents with Universal Feedback in Communication
by: Wang, Kuan, et al.
Published: (2023)