Enhancing Confidence Estimation in Telco LLMs via Twin-Pass CoT-Ensembling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Saenko, Anton, Gajjar, Pranshav, Ganiyu, Abiodun, Shah, Vijay K. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AI5GTest: AI-Driven Specification-Aware Automated Testing and Validation of 5G O-RAN Components
von: Ganiyu, Abiodun, et al.
Veröffentlicht: (2025)
von: Ganiyu, Abiodun, et al.
Veröffentlicht: (2025)
ORANSight-2.0: Foundational LLMs for O-RAN
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2025)
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2025)
ORAN-Bench-13K: An Open Source Benchmark for Assessing LLMs in Open Radio Access Networks
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2024)
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2024)
TeleEmbedBench: A Multi-Corpus Embedding Benchmark for RAG in Telecommunications
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2026)
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2026)
TeleResilienceBench: Quantifying Resilience for LLM Reasoning in Telecommunications
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2026)
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2026)
Preserving Data Privacy for ML-driven Applications in Open Radio Access Networks
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2024)
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2024)
LLM-AUG: Robust Wireless Data Augmentation with In-Context Learning in Large Language Models
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2026)
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2026)
Agents Should Replace Narrow Predictive AI as the Orchestrator in 6G AI-RAN
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2026)
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2026)
Black-Box Evasion Attacks on Data-Driven Open RAN Apps: Tailored Design and Experimental Evaluation
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2025)
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2025)
Tele-LLM-Hub: Building Context-Aware Multi-Agent LLM Systems for Telecom Networks
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2025)
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2025)
Advanced AI Service Provisioning in O-RAN through LLM Engine Integration
von: Natanzi, Seyed Bagher Hashemi, et al.
Veröffentlicht: (2026)
von: Natanzi, Seyed Bagher Hashemi, et al.
Veröffentlicht: (2026)
From Explicit CoT to Implicit CoT: Learning to Internalize CoT Step by Step
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
The Illusion of Reasoning: Exposing Evasive Data Contamination in LLMs via Zero-CoT Truncation
von: Lan, Yifan, et al.
Veröffentlicht: (2026)
von: Lan, Yifan, et al.
Veröffentlicht: (2026)
Cardiac Stability Theory: An Axiomatically Grounded Framework for Continuous Cardiac Health Monitoring via Smartphone Photoplethysmography
von: Oladunni, Timothy, et al.
Veröffentlicht: (2026)
von: Oladunni, Timothy, et al.
Veröffentlicht: (2026)
CER: Confidence Enhanced Reasoning in LLMs
von: Razghandi, Ali, et al.
Veröffentlicht: (2025)
von: Razghandi, Ali, et al.
Veröffentlicht: (2025)
The Effects of Data Augmentation on Confidence Estimation for LLMs
von: Wang, Rui, et al.
Veröffentlicht: (2025)
von: Wang, Rui, et al.
Veröffentlicht: (2025)
How Likely Do LLMs with CoT Mimic Human Reasoning?
von: Bao, Guangsheng, et al.
Veröffentlicht: (2024)
von: Bao, Guangsheng, et al.
Veröffentlicht: (2024)
CAKE: Confidence in Assignments via K-partition Ensembles
von: Semoglou, Aggelos, et al.
Veröffentlicht: (2026)
von: Semoglou, Aggelos, et al.
Veröffentlicht: (2026)
To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning
von: Sprague, Zayne, et al.
Veröffentlicht: (2024)
von: Sprague, Zayne, et al.
Veröffentlicht: (2024)
CoT-UQ: Improving Response-wise Uncertainty Quantification in LLMs with Chain-of-Thought
von: Zhang, Boxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Boxuan, et al.
Veröffentlicht: (2025)
Divide-and-Conquer CoT: RL for Reducing Latency via Parallel Reasoning
von: Mahankali, Arvind, et al.
Veröffentlicht: (2026)
von: Mahankali, Arvind, et al.
Veröffentlicht: (2026)
Context-Aware Graph Attention for Unsupervised Telco Anomaly Detection
von: Malacarne, Sara, et al.
Veröffentlicht: (2026)
von: Malacarne, Sara, et al.
Veröffentlicht: (2026)
Unveiling and Causalizing CoT: A Causal Pespective
von: Fu, Jiarun, et al.
Veröffentlicht: (2025)
von: Fu, Jiarun, et al.
Veröffentlicht: (2025)
Enhance GNNs with Reliable Confidence Estimation via Adversarial Calibration Learning
von: Wang, Yilong, et al.
Veröffentlicht: (2025)
von: Wang, Yilong, et al.
Veröffentlicht: (2025)
Confidence-Credibility Aware Weighted Ensembles of Small LLMs Outperform Large LLMs in Emotion Detection
von: Elgabry, Menna, et al.
Veröffentlicht: (2025)
von: Elgabry, Menna, et al.
Veröffentlicht: (2025)
Attractor-Vascular Coupling Theory: Formal Grounding and Empirical Validation for AAMI-Standard Cuffless Blood Pressure Estimation from Smartphone Photoplethysmography
von: Oladunni, Timothy, et al.
Veröffentlicht: (2026)
von: Oladunni, Timothy, et al.
Veröffentlicht: (2026)
Jamming Smarter, Not Harder: Exploiting O-RAN Y1 RAN Analytics for Efficient Interference
von: Ganiyu, Abiodun, et al.
Veröffentlicht: (2025)
von: Ganiyu, Abiodun, et al.
Veröffentlicht: (2025)
Training on Documents About Monitoring Leads to CoT Obfuscation
von: Haskins, Reilly, et al.
Veröffentlicht: (2026)
von: Haskins, Reilly, et al.
Veröffentlicht: (2026)
Demonstrations, CoT, and Prompting: A Theoretical Analysis of ICL
von: Tong, Xuhan, et al.
Veröffentlicht: (2026)
von: Tong, Xuhan, et al.
Veröffentlicht: (2026)
Self-Verifying Reflection Helps Transformers with CoT Reasoning
von: Yu, Zhongwei, et al.
Veröffentlicht: (2025)
von: Yu, Zhongwei, et al.
Veröffentlicht: (2025)
Meta-CoT: Enhancing Granularity and Generalization in Image Editing
von: Zhang, Shiyi, et al.
Veröffentlicht: (2026)
von: Zhang, Shiyi, et al.
Veröffentlicht: (2026)
Factual Confidence of LLMs: on Reliability and Robustness of Current Estimators
von: Mahaut, Matéo, et al.
Veröffentlicht: (2024)
von: Mahaut, Matéo, et al.
Veröffentlicht: (2024)
CDW-CoT: Clustered Distance-Weighted Chain-of-Thoughts Reasoning
von: Fang, Yuanheng, et al.
Veröffentlicht: (2025)
von: Fang, Yuanheng, et al.
Veröffentlicht: (2025)
Uncertainty Estimation via Hyperspherical Confidence Mapping
von: Choi, Eunseo, et al.
Veröffentlicht: (2026)
von: Choi, Eunseo, et al.
Veröffentlicht: (2026)
Confidence Estimation via Sequential Likelihood Mixing
von: Kirschner, Johannes, et al.
Veröffentlicht: (2025)
von: Kirschner, Johannes, et al.
Veröffentlicht: (2025)
Self-Consistency from Only Two Samples: CoT-PoT Ensembling for Efficient LLM Reasoning
von: Saparkhan, Raman, et al.
Veröffentlicht: (2026)
von: Saparkhan, Raman, et al.
Veröffentlicht: (2026)
LayerCraft: Enhancing Text-to-Image Generation with CoT Reasoning and Layered Object Integration
von: Zhang, Yuyao, et al.
Veröffentlicht: (2025)
von: Zhang, Yuyao, et al.
Veröffentlicht: (2025)
Data Shifts Hurt CoT: A Theoretical Study
von: Yin, Lang, et al.
Veröffentlicht: (2025)
von: Yin, Lang, et al.
Veröffentlicht: (2025)
Visual CoT Makes VLMs Smarter but More Fragile
von: Xu, Chunxue, et al.
Veröffentlicht: (2025)
von: Xu, Chunxue, et al.
Veröffentlicht: (2025)
What Characterizes Effective Reasoning? Revisiting Length, Review, and Structure of CoT
von: Feng, Yunzhen, et al.
Veröffentlicht: (2025)
von: Feng, Yunzhen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AI5GTest: AI-Driven Specification-Aware Automated Testing and Validation of 5G O-RAN Components
von: Ganiyu, Abiodun, et al.
Veröffentlicht: (2025) -
ORANSight-2.0: Foundational LLMs for O-RAN
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2025) -
ORAN-Bench-13K: An Open Source Benchmark for Assessing LLMs in Open Radio Access Networks
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2024) -
TeleEmbedBench: A Multi-Corpus Embedding Benchmark for RAG in Telecommunications
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2026) -
TeleResilienceBench: Quantifying Resilience for LLM Reasoning in Telecommunications
von: Gajjar, Pranshav, et al.
Veröffentlicht: (2026)