CIRCUS: Circuit Consensus under Uncertainty via Stability Ensembles
Fuente:
arXiv
Salvato in:
| Autore principale: | Parekh, Swapnil |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Drop the Act: Probe-Filtered RL for Faithful Chain-of-Thought Reasoning
di: Parekh, Swapnil
Pubblicazione: (2026)
di: Parekh, Swapnil
Pubblicazione: (2026)
Thinking Wrong in Silence: Backdoor Attacks on Continuous Latent Reasoning
di: Parekh, Swapnil
Pubblicazione: (2026)
di: Parekh, Swapnil
Pubblicazione: (2026)
Probabilistic Consensus through Ensemble Validation: A Framework for LLM Reliability
di: Naik, Ninad
Pubblicazione: (2024)
di: Naik, Ninad
Pubblicazione: (2024)
Uncertainty-Aware Fusion: An Ensemble Framework for Mitigating Hallucinations in Large Language Models
di: Dey, Prasenjit, et al.
Pubblicazione: (2025)
di: Dey, Prasenjit, et al.
Pubblicazione: (2025)
When Chain-of-Thought Fails, the Solution Hides in the Hidden States
di: Mehrafarin, Houman, et al.
Pubblicazione: (2026)
di: Mehrafarin, Houman, et al.
Pubblicazione: (2026)
DiCoRe: Enhancing Zero-shot Event Detection via Divergent-Convergent LLM Reasoning
di: Parekh, Tanmay, et al.
Pubblicazione: (2025)
di: Parekh, Tanmay, et al.
Pubblicazione: (2025)
Uncertainty Quantification for Language Models: A Suite of Black-Box, White-Box, LLM Judge, and Ensemble Scorers
di: Bouchard, Dylan, et al.
Pubblicazione: (2025)
di: Bouchard, Dylan, et al.
Pubblicazione: (2025)
Measuring Uncertainty in Transformer Circuits with Effective Information Consistency
di: Krasnovsky, Anatoly A.
Pubblicazione: (2025)
di: Krasnovsky, Anatoly A.
Pubblicazione: (2025)
Exploring Response Uncertainty in MLLMs: An Empirical Evaluation under Misleading Scenarios
di: Dang, Yunkai, et al.
Pubblicazione: (2024)
di: Dang, Yunkai, et al.
Pubblicazione: (2024)
Dynamic Strategy Planning for Efficient Question Answering with Large Language Models
di: Parekh, Tanmay, et al.
Pubblicazione: (2024)
di: Parekh, Tanmay, et al.
Pubblicazione: (2024)
ReConcile: Round-Table Conference Improves Reasoning via Consensus among Diverse LLMs
di: Chen, Justin Chih-Yao, et al.
Pubblicazione: (2023)
di: Chen, Justin Chih-Yao, et al.
Pubblicazione: (2023)
Towards Understanding Fine-Tuning Mechanisms of LLMs via Circuit Analysis
di: Wang, Xu, et al.
Pubblicazione: (2025)
di: Wang, Xu, et al.
Pubblicazione: (2025)
The Role of Ambiguity in Error Prediction via Uncertainty Quantification
di: Staliūnaitė, Ieva Raminta, et al.
Pubblicazione: (2026)
di: Staliūnaitė, Ieva Raminta, et al.
Pubblicazione: (2026)
CoVerRL: Breaking the Consensus Trap in Label-Free Reasoning via Generator-Verifier Co-Evolution
di: Pan, Teng, et al.
Pubblicazione: (2026)
di: Pan, Teng, et al.
Pubblicazione: (2026)
Improving Uncertainty Quantification in Large Language Models via Semantic Embeddings
di: Grewal, Yashvir S., et al.
Pubblicazione: (2024)
di: Grewal, Yashvir S., et al.
Pubblicazione: (2024)
Ensembles of Low-Rank Expert Adapters
di: Li, Yinghao, et al.
Pubblicazione: (2025)
di: Li, Yinghao, et al.
Pubblicazione: (2025)
Ensemble Distillation for Unsupervised Constituency Parsing
di: Shayegh, Behzad, et al.
Pubblicazione: (2023)
di: Shayegh, Behzad, et al.
Pubblicazione: (2023)
Transformer Circuit Faithfulness Metrics are not Robust
di: Miller, Joseph, et al.
Pubblicazione: (2024)
di: Miller, Joseph, et al.
Pubblicazione: (2024)
Trustworthy Summarization via Uncertainty Quantification and Risk Awareness in Large Language Models
di: Pan, Shuaidong, et al.
Pubblicazione: (2025)
di: Pan, Shuaidong, et al.
Pubblicazione: (2025)
SelectIT: Selective Instruction Tuning for LLMs via Uncertainty-Aware Self-Reflection
di: Liu, Liangxin, et al.
Pubblicazione: (2024)
di: Liu, Liangxin, et al.
Pubblicazione: (2024)
Enhancing Trust in Large Language Models via Uncertainty-Calibrated Fine-Tuning
di: Krishnan, Ranganath, et al.
Pubblicazione: (2024)
di: Krishnan, Ranganath, et al.
Pubblicazione: (2024)
Ensembling Language Models with Sequential Monte Carlo
di: Chan, Robin Shing Moon, et al.
Pubblicazione: (2026)
di: Chan, Robin Shing Moon, et al.
Pubblicazione: (2026)
Enhancing Annotated Bibliography Generation with LLM Ensembles
di: Bermejo, Sergio
Pubblicazione: (2024)
di: Bermejo, Sergio
Pubblicazione: (2024)
Fact-Checking the Output of Large Language Models via Token-Level Uncertainty Quantification
di: Fadeeva, Ekaterina, et al.
Pubblicazione: (2024)
di: Fadeeva, Ekaterina, et al.
Pubblicazione: (2024)
ESI: Epistemic Uncertainty Quantification via Semantic-preserving Intervention for Large Language Models
di: Li, Mingda, et al.
Pubblicazione: (2025)
di: Li, Mingda, et al.
Pubblicazione: (2025)
Sparse Attention Decomposition Applied to Circuit Tracing
di: Franco, Gabriel, et al.
Pubblicazione: (2024)
di: Franco, Gabriel, et al.
Pubblicazione: (2024)
Circuit Insights: Towards Interpretability Beyond Activations
di: Golimblevskaia, Elena, et al.
Pubblicazione: (2025)
di: Golimblevskaia, Elena, et al.
Pubblicazione: (2025)
Stabilizing LLM Supervised Fine-Tuning via Explicit Distributional Control
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
AVSD: Adaptive-View Self-Distillation by Balancing Consensus and Teacher-Specific Privileged Signals
di: Nguyen, Duy, et al.
Pubblicazione: (2026)
di: Nguyen, Duy, et al.
Pubblicazione: (2026)
BayesAgent: Bayesian Agentic Reasoning Under Uncertainty via Verbalized Probabilistic Graphical Modeling
di: Huang, Hengguan, et al.
Pubblicazione: (2024)
di: Huang, Hengguan, et al.
Pubblicazione: (2024)
Uncertainty of Thoughts: Uncertainty-Aware Planning Enhances Information Seeking in Large Language Models
di: Hu, Zhiyuan, et al.
Pubblicazione: (2024)
di: Hu, Zhiyuan, et al.
Pubblicazione: (2024)
The Computational Limits of State-Space Models and Mamba via the Lens of Circuit Complexity
di: Chen, Yifang, et al.
Pubblicazione: (2024)
di: Chen, Yifang, et al.
Pubblicazione: (2024)
DFPE: A Diverse Fingerprint Ensemble for Enhancing LLM Performance
di: Cohen, Seffi, et al.
Pubblicazione: (2025)
di: Cohen, Seffi, et al.
Pubblicazione: (2025)
Domain Gating Ensemble Networks for AI-Generated Text Detection
di: Tripathi, Arihant, et al.
Pubblicazione: (2025)
di: Tripathi, Arihant, et al.
Pubblicazione: (2025)
Learning to Trust the Crowd: A Multi-Model Consensus Reasoning Engine for Large Language Models
di: Kallem, Pranav
Pubblicazione: (2026)
di: Kallem, Pranav
Pubblicazione: (2026)
TourPlanner: A Competitive Consensus Framework with Constraint-Gated Reinforcement Learning for Travel Planning
di: Wang, Yinuo, et al.
Pubblicazione: (2026)
di: Wang, Yinuo, et al.
Pubblicazione: (2026)
GTPO: Stabilizing Group Relative Policy Optimization via Gradient and Entropy Control
di: Simoni, Marco, et al.
Pubblicazione: (2025)
di: Simoni, Marco, et al.
Pubblicazione: (2025)
On Uncertainty In Natural Language Processing
di: Ulmer, Dennis
Pubblicazione: (2024)
di: Ulmer, Dennis
Pubblicazione: (2024)
Beyond Components: Singular Vector-Based Interpretability of Transformer Circuits
di: Ahmad, Areeb, et al.
Pubblicazione: (2025)
di: Ahmad, Areeb, et al.
Pubblicazione: (2025)
Efficient Automated Circuit Discovery in Transformers using Contextual Decomposition
di: Hsu, Aliyah R., et al.
Pubblicazione: (2024)
di: Hsu, Aliyah R., et al.
Pubblicazione: (2024)
Documenti analoghi
-
Drop the Act: Probe-Filtered RL for Faithful Chain-of-Thought Reasoning
di: Parekh, Swapnil
Pubblicazione: (2026) -
Thinking Wrong in Silence: Backdoor Attacks on Continuous Latent Reasoning
di: Parekh, Swapnil
Pubblicazione: (2026) -
Probabilistic Consensus through Ensemble Validation: A Framework for LLM Reliability
di: Naik, Ninad
Pubblicazione: (2024) -
Uncertainty-Aware Fusion: An Ensemble Framework for Mitigating Hallucinations in Large Language Models
di: Dey, Prasenjit, et al.
Pubblicazione: (2025) -
When Chain-of-Thought Fails, the Solution Hides in the Hidden States
di: Mehrafarin, Houman, et al.
Pubblicazione: (2026)