EVINCE: Optimizing Multi-LLM Dialogues Using Conditional Statistics and Information Theory
Fuente:
arXiv
Guardado en:
| Autor principal: | Chang, Edward Y. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Ensuring Ground Truth Accuracy in Healthcare with the EVINCE framework
por: Chang, Edward Y.
Publicado: (2024)
por: Chang, Edward Y.
Publicado: (2024)
SocraSynth: Multi-LLM Reasoning with Conditional Statistics
por: Chang, Edward Y.
Publicado: (2024)
por: Chang, Edward Y.
Publicado: (2024)
SagaLLM: Context Management, Validation, and Transaction Guarantees for Multi-Agent LLM Planning
por: Chang, Edward Y., et al.
Publicado: (2025)
por: Chang, Edward Y., et al.
Publicado: (2025)
ALAS: A Stateful Multi-LLM Agent Framework for Disruption-Aware Planning
por: Chang, Edward Y., et al.
Publicado: (2025)
por: Chang, Edward Y., et al.
Publicado: (2025)
Diagnosing and Mitigating Sycophancy and Skepticism in LLM Causal Judgment
por: Chang, Edward Y.
Publicado: (2026)
por: Chang, Edward Y.
Publicado: (2026)
The Unified Cognitive Consciousness Theory for Language Models: Anchoring Semantics, Thresholds of Activation, and Emergent Reasoning
por: Chang, Edward Y., et al.
Publicado: (2025)
por: Chang, Edward Y., et al.
Publicado: (2025)
Unlocking the Wisdom of Large Language Models: An Introduction to The Path to Artificial General Intelligence
por: Chang, Edward Y.
Publicado: (2024)
por: Chang, Edward Y.
Publicado: (2024)
RAudit: A Blind Auditing Protocol for Large Language Model Reasoning
por: Chang, Edward Y., et al.
Publicado: (2026)
por: Chang, Edward Y., et al.
Publicado: (2026)
Integrating Emotional and Linguistic Models for Ethical Compliance in Large Language Models
por: Chang, Edward Y.
Publicado: (2024)
por: Chang, Edward Y.
Publicado: (2024)
Internal Reasoning vs. External Control: A Thermodynamic Analysis of Sycophancy in Large Language Models
por: Chang, Edward Y.
Publicado: (2025)
por: Chang, Edward Y.
Publicado: (2025)
Understanding LLM Evaluator Behavior: A Structured Multi-Evaluator Framework for Merchant Risk Assessment
por: Wang, Liang, et al.
Publicado: (2026)
por: Wang, Liang, et al.
Publicado: (2026)
Uncovering Biases with Reflective Large Language Models
por: Chang, Edward Y.
Publicado: (2024)
por: Chang, Edward Y.
Publicado: (2024)
CausalT5K: Diagnosing and Informing Refusal for Trustworthy Causal Reasoning of Skepticism, Sycophancy, Detection-Correction, and Rung Collapse
por: Geng, Longling, et al.
Publicado: (2026)
por: Geng, Longling, et al.
Publicado: (2026)
CoE: Collaborative Entropy for Uncertainty Quantification in Agentic Multi-LLM Systems
por: Sun, Kangkang, et al.
Publicado: (2026)
por: Sun, Kangkang, et al.
Publicado: (2026)
LLM-Based SQL Generation: Prompting, Self-Refinement, and Adaptive Weighted Majority Voting
por: Yang, Yu-Jie, et al.
Publicado: (2026)
por: Yang, Yu-Jie, et al.
Publicado: (2026)
A Library of LLM Intrinsics for Retrieval-Augmented Generation
por: Danilevsky, Marina, et al.
Publicado: (2025)
por: Danilevsky, Marina, et al.
Publicado: (2025)
Evaluating LLM Metrics Through Real-World Capabilities
por: Miller, Justin K, et al.
Publicado: (2025)
por: Miller, Justin K, et al.
Publicado: (2025)
Beyond the Black Box: A Statistical Model for LLM Reasoning and Inference
por: Dalal, Siddhartha, et al.
Publicado: (2024)
por: Dalal, Siddhartha, et al.
Publicado: (2024)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
por: Saji, Alan, et al.
Publicado: (2025)
por: Saji, Alan, et al.
Publicado: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
por: Peters, Sydney, et al.
Publicado: (2025)
por: Peters, Sydney, et al.
Publicado: (2025)
Exploring Collatz Dynamics with Human-LLM Collaboration
por: Chang, Edward Y.
Publicado: (2026)
por: Chang, Edward Y.
Publicado: (2026)
LLM-based Automated Theorem Proving Hinges on Scalable Synthetic Data Generation
por: Lai, Junyu, et al.
Publicado: (2025)
por: Lai, Junyu, et al.
Publicado: (2025)
Process Supervision-Guided Policy Optimization for Code Generation
por: Dai, Ning, et al.
Publicado: (2024)
por: Dai, Ning, et al.
Publicado: (2024)
Active Context Compression: Autonomous Memory Management in LLM Agents
por: Verma, Nikhil
Publicado: (2026)
por: Verma, Nikhil
Publicado: (2026)
SOCIA-EVO: Automated Simulator Construction via Dual-Anchored Bi-Level Optimization
por: Hua, Yuncheng, et al.
Publicado: (2026)
por: Hua, Yuncheng, et al.
Publicado: (2026)
End-to-End Optimization of LLM-Driven Multi-Agent Search Systems via Heterogeneous-Group-Based Reinforcement Learning
por: Chen, Guanzhong, et al.
Publicado: (2025)
por: Chen, Guanzhong, et al.
Publicado: (2025)
HR-MultiWOZ: A Task Oriented Dialogue (TOD) Dataset for HR LLM Agent
por: Xu, Weijie, et al.
Publicado: (2024)
por: Xu, Weijie, et al.
Publicado: (2024)
ScrapMem: A Bio-inspired Framework for On-device Personalized Agent Memory via Optical Forgetting
por: Chang, Jiale, et al.
Publicado: (2026)
por: Chang, Jiale, et al.
Publicado: (2026)
Applying Cognitive Design Patterns to General LLM Agents
por: Wray, Robert E., et al.
Publicado: (2025)
por: Wray, Robert E., et al.
Publicado: (2025)
Cultural Benchmarking of LLMs in Standard and Dialectal Arabic Dialogues
por: Kautsar, Muhammad Dehan Al, et al.
Publicado: (2026)
por: Kautsar, Muhammad Dehan Al, et al.
Publicado: (2026)
An Explainable Collaborative Dialogue System using a Theory of Mind
por: Cohen, Philip R., et al.
Publicado: (2023)
por: Cohen, Philip R., et al.
Publicado: (2023)
RADD: Retrieval-Augmented Discrete Diffusion for Multi-Modal Knowledge Graph Completion
por: Niu, Guanglin, et al.
Publicado: (2026)
por: Niu, Guanglin, et al.
Publicado: (2026)
Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs
por: Li, Chang, et al.
Publicado: (2025)
por: Li, Chang, et al.
Publicado: (2025)
OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind
por: Srishty, Sharmin Sultana, et al.
Publicado: (2026)
por: Srishty, Sharmin Sultana, et al.
Publicado: (2026)
SOCIA-Nabla: Textual Gradient Meets Multi-Agent Orchestration for Automated Simulator Generation
por: Hua, Yuncheng, et al.
Publicado: (2025)
por: Hua, Yuncheng, et al.
Publicado: (2025)
SOCIA-$\nabla$: Textual Gradient Meets Multi-Agent Orchestration for Automated Simulator Generation
por: Hua, Yuncheng, et al.
Publicado: (2025)
por: Hua, Yuncheng, et al.
Publicado: (2025)
Semantic Delta: An Interpretable Signal Differentiating Human and LLMs Dialogue
por: Scantamburlo, Riccardo, et al.
Publicado: (2026)
por: Scantamburlo, Riccardo, et al.
Publicado: (2026)
Deciphering Digital Detectives: Understanding LLM Behaviors and Capabilities in Multi-Agent Mystery Games
por: Wu, Dekun, et al.
Publicado: (2023)
por: Wu, Dekun, et al.
Publicado: (2023)
Reasoning-Based AI for Startup Evaluation (R.A.I.S.E.): A Memory-Augmented, Multi-Step Decision Framework
por: Preuveneers, Jack, et al.
Publicado: (2025)
por: Preuveneers, Jack, et al.
Publicado: (2025)
Assistive Large Language Model Agents for Socially-Aware Negotiation Dialogues
por: Hua, Yuncheng, et al.
Publicado: (2024)
por: Hua, Yuncheng, et al.
Publicado: (2024)
Ejemplares similares
-
Ensuring Ground Truth Accuracy in Healthcare with the EVINCE framework
por: Chang, Edward Y.
Publicado: (2024) -
SocraSynth: Multi-LLM Reasoning with Conditional Statistics
por: Chang, Edward Y.
Publicado: (2024) -
SagaLLM: Context Management, Validation, and Transaction Guarantees for Multi-Agent LLM Planning
por: Chang, Edward Y., et al.
Publicado: (2025) -
ALAS: A Stateful Multi-LLM Agent Framework for Disruption-Aware Planning
por: Chang, Edward Y., et al.
Publicado: (2025) -
Diagnosing and Mitigating Sycophancy and Skepticism in LLM Causal Judgment
por: Chang, Edward Y.
Publicado: (2026)