C-voting: Confidence-Based Test-Time Voting without Explicit Energy Functions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kubo, Kenji, Kamiya, Shunsuke, Koyama, Masanori, Hayashi, Kohei, Iwasawa, Yusuke, Matsuo, Yutaka |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Thinking While Listening: Fast-Slow Recurrence for Long-Horizon Sequential Modeling
von: Takashiro, Shota, et al.
Veröffentlicht: (2026)
von: Takashiro, Shota, et al.
Veröffentlicht: (2026)
Language Models Do Hard Arithmetic Tasks Easily and Hardly Do Easy Arithmetic Tasks
von: Gambardella, Andrew, et al.
Veröffentlicht: (2024)
von: Gambardella, Andrew, et al.
Veröffentlicht: (2024)
Safe Transformer: An Explicit Safety Bit For Interpretable And Controllable Alignment
von: Feng, Jingyuan, et al.
Veröffentlicht: (2026)
von: Feng, Jingyuan, et al.
Veröffentlicht: (2026)
Leave No Observation Behind: Real-time Correction for VLA Action Chunks
von: Sendai, Kohei, et al.
Veröffentlicht: (2025)
von: Sendai, Kohei, et al.
Veröffentlicht: (2025)
Large Language Models as Theory of Mind Aware Generative Agents with Counterfactual Reflection
von: Yang, Bo, et al.
Veröffentlicht: (2025)
von: Yang, Bo, et al.
Veröffentlicht: (2025)
Towards Empirical Interpretation of Internal Circuits and Properties in Grokked Transformers on Modular Polynomials
von: Furuta, Hiroki, et al.
Veröffentlicht: (2024)
von: Furuta, Hiroki, et al.
Veröffentlicht: (2024)
Self-Harmony: Learning to Harmonize Self-Supervision and Self-Play in Test-Time Reinforcement Learning
von: Wang, Ru, et al.
Veröffentlicht: (2025)
von: Wang, Ru, et al.
Veröffentlicht: (2025)
Inconsistent Tokenizations Cause Language Models to be Perplexed by Japanese Grammar
von: Gambardella, Andrew, et al.
Veröffentlicht: (2025)
von: Gambardella, Andrew, et al.
Veröffentlicht: (2025)
Rethinking Evaluation of Sparse Autoencoders through the Representation of Polysemous Words
von: Minegishi, Gouki, et al.
Veröffentlicht: (2025)
von: Minegishi, Gouki, et al.
Veröffentlicht: (2025)
Topology of Reasoning: Understanding Large Reasoning Models through Reasoning Graph Properties
von: Minegishi, Gouki, et al.
Veröffentlicht: (2025)
von: Minegishi, Gouki, et al.
Veröffentlicht: (2025)
ClinDet-Bench: Beyond Abstention, Evaluating Judgment Determinability of LLMs in Clinical Decision-Making
von: Watanabe, Yusuke, et al.
Veröffentlicht: (2026)
von: Watanabe, Yusuke, et al.
Veröffentlicht: (2026)
Understanding Emergent Misalignment via Feature Superposition Geometry
von: Minegishi, Gouki, et al.
Veröffentlicht: (2026)
von: Minegishi, Gouki, et al.
Veröffentlicht: (2026)
Beyond Induction Heads: In-Context Meta Learning Induces Multi-Phase Circuit Emergence
von: Minegishi, Gouki, et al.
Veröffentlicht: (2025)
von: Minegishi, Gouki, et al.
Veröffentlicht: (2025)
Zipping the Thought: When and How Compressed Reasoning Data Works in LLM Post-Training
von: Matsutani, Kohsei, et al.
Veröffentlicht: (2026)
von: Matsutani, Kohsei, et al.
Veröffentlicht: (2026)
RL Squeezes, SFT Expands: A Comparative Study of Reasoning LLMs
von: Matsutani, Kohsei, et al.
Veröffentlicht: (2025)
von: Matsutani, Kohsei, et al.
Veröffentlicht: (2025)
Semantic Token Clustering for Efficient Uncertainty Quantification in Large Language Models
von: Cao, Qi, et al.
Veröffentlicht: (2026)
von: Cao, Qi, et al.
Veröffentlicht: (2026)
Residual Koopman Spectral Profiling for Predicting and Preventing Transformer Training Instability
von: Kim, Bum Jun, et al.
Veröffentlicht: (2026)
von: Kim, Bum Jun, et al.
Veröffentlicht: (2026)
Suspicion-Agent: Playing Imperfect Information Games with Theory of Mind Aware GPT-4
von: Guo, Jiaxian, et al.
Veröffentlicht: (2023)
von: Guo, Jiaxian, et al.
Veröffentlicht: (2023)
Which Programming Language and What Features at Pre-training Stage Affect Downstream Logical Inference Performance?
von: Uchiyama, Fumiya, et al.
Veröffentlicht: (2024)
von: Uchiyama, Fumiya, et al.
Veröffentlicht: (2024)
SAIL: Test-Time Scaling for In-Context Imitation Learning with VLM
von: Sato, Makoto, et al.
Veröffentlicht: (2026)
von: Sato, Makoto, et al.
Veröffentlicht: (2026)
AQA-TTRL: Self-Adaptation in Audio Question Answering with Test-Time Reinforcement Learning
von: Zhang, Haoyu, et al.
Veröffentlicht: (2025)
von: Zhang, Haoyu, et al.
Veröffentlicht: (2025)
ReAgent: Reversible Multi-Agent Reasoning for Knowledge-Enhanced Multi-Hop QA
von: Zhao, Xinjie, et al.
Veröffentlicht: (2025)
von: Zhao, Xinjie, et al.
Veröffentlicht: (2025)
GenDOM: Generalizable One-shot Deformable Object Manipulation with Parameter-Aware Policy
von: Kuroki, So, et al.
Veröffentlicht: (2023)
von: Kuroki, So, et al.
Veröffentlicht: (2023)
From Chains to Graphs: Self-Structured Reasoning for General-Domain LLMs
von: Chen, Yingjian, et al.
Veröffentlicht: (2026)
von: Chen, Yingjian, et al.
Veröffentlicht: (2026)
DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation
von: Shing, Makoto, et al.
Veröffentlicht: (2025)
von: Shing, Makoto, et al.
Veröffentlicht: (2025)
GenORM: Generalizable One-shot Rope Manipulation with Parameter-Aware Policy
von: Kuroki, So, et al.
Veröffentlicht: (2023)
von: Kuroki, So, et al.
Veröffentlicht: (2023)
JMedEthicBench: A Multi-Turn Conversational Benchmark for Evaluating Medical Safety in Japanese Large Language Models
von: Liu, Junyu, et al.
Veröffentlicht: (2026)
von: Liu, Junyu, et al.
Veröffentlicht: (2026)
Neural Fourier Transform: A General Approach to Equivariant Representation Learning
von: Koyama, Masanori, et al.
Veröffentlicht: (2023)
von: Koyama, Masanori, et al.
Veröffentlicht: (2023)
Bridging Lottery Ticket and Grokking: Understanding Grokking from Inner Structure of Networks
von: Minegishi, Gouki, et al.
Veröffentlicht: (2023)
von: Minegishi, Gouki, et al.
Veröffentlicht: (2023)
The Embodied World Model Based on LLM with Visual Information and Prediction-Oriented Prompts
von: Haijima, Wakana, et al.
Veröffentlicht: (2024)
von: Haijima, Wakana, et al.
Veröffentlicht: (2024)
A Comprehensive Survey on Physical Risk Control in the Era of Foundation Model-enabled Robotics
von: Kojima, Takeshi, et al.
Veröffentlicht: (2025)
von: Kojima, Takeshi, et al.
Veröffentlicht: (2025)
Pairwise Optimal Transports for Training All-to-All Flow-Based Condition Transfer Model
von: Ikeda, Kotaro, et al.
Veröffentlicht: (2025)
von: Ikeda, Kotaro, et al.
Veröffentlicht: (2025)
Enhancing Unimodal Latent Representations in Multimodal VAEs through Iterative Amortized Inference
von: Oshima, Yuta, et al.
Veröffentlicht: (2024)
von: Oshima, Yuta, et al.
Veröffentlicht: (2024)
When to Vote, When to Rewrite: Disagreement-Guided Strategy Routing for Test-Time Scaling
von: Lin, Zhimin, et al.
Veröffentlicht: (2026)
von: Lin, Zhimin, et al.
Veröffentlicht: (2026)
Lost and Found in Translation: Variational Diagnostics for Neural Codebook Channels
von: Hayashi, Yusuke
Veröffentlicht: (2026)
von: Hayashi, Yusuke
Veröffentlicht: (2026)
Test-Time Augmentation for Traveling Salesperson Problem
von: Ishiyama, Ryo, et al.
Veröffentlicht: (2024)
von: Ishiyama, Ryo, et al.
Veröffentlicht: (2024)
Retrieve-Augmented Generation for Speeding up Diffusion Policy without Additional Training
von: Odonchimed, Sodtavilan, et al.
Veröffentlicht: (2025)
von: Odonchimed, Sodtavilan, et al.
Veröffentlicht: (2025)
When Object-Centric World Models Meet Policy Learning: From Pixels to Policies, and Where It Breaks
von: Ferraro, Stefano, et al.
Veröffentlicht: (2025)
von: Ferraro, Stefano, et al.
Veröffentlicht: (2025)
Omanic: Towards Step-wise Evaluation of Multi-hop Reasoning in Large Language Models
von: Gu, Xiaojie, et al.
Veröffentlicht: (2026)
von: Gu, Xiaojie, et al.
Veröffentlicht: (2026)
StatWhy: Formal Verification Tool for Statistical Hypothesis Testing Programs
von: Kawamoto, Yusuke, et al.
Veröffentlicht: (2024)
von: Kawamoto, Yusuke, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Thinking While Listening: Fast-Slow Recurrence for Long-Horizon Sequential Modeling
von: Takashiro, Shota, et al.
Veröffentlicht: (2026) -
Language Models Do Hard Arithmetic Tasks Easily and Hardly Do Easy Arithmetic Tasks
von: Gambardella, Andrew, et al.
Veröffentlicht: (2024) -
Safe Transformer: An Explicit Safety Bit For Interpretable And Controllable Alignment
von: Feng, Jingyuan, et al.
Veröffentlicht: (2026) -
Leave No Observation Behind: Real-time Correction for VLA Action Chunks
von: Sendai, Kohei, et al.
Veröffentlicht: (2025) -
Large Language Models as Theory of Mind Aware Generative Agents with Counterfactual Reflection
von: Yang, Bo, et al.
Veröffentlicht: (2025)