Latent Space Chain-of-Embedding Enables Output-free LLM Self-Evaluation
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Yiming, Zhang, Pei, Yang, Baosong, Wong, Derek F., Wang, Rui |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Embedding Trajectory for Out-of-Distribution Detection in Mathematical Reasoning
di: Wang, Yiming, et al.
Pubblicazione: (2024)
di: Wang, Yiming, et al.
Pubblicazione: (2024)
Output Embedding Centering for Stable LLM Pretraining
di: Stollenwerk, Felix, et al.
Pubblicazione: (2026)
di: Stollenwerk, Felix, et al.
Pubblicazione: (2026)
Sampling-Efficient Test-Time Scaling: Self-Estimating the Best-of-N Sampling in Early Decoding
di: Wang, Yiming, et al.
Pubblicazione: (2025)
di: Wang, Yiming, et al.
Pubblicazione: (2025)
From Reasoning Chains to Verifiable Subproblems: Curriculum Reinforcement Learning Enables Credit Assignment for LLM Reasoning
di: Jiang, Xitai, et al.
Pubblicazione: (2026)
di: Jiang, Xitai, et al.
Pubblicazione: (2026)
On the Overscaling Curse of Parallel Thinking: System Efficacy Contradicts Sample Efficiency
di: Wang, Yiming, et al.
Pubblicazione: (2026)
di: Wang, Yiming, et al.
Pubblicazione: (2026)
SelectIT: Selective Instruction Tuning for LLMs via Uncertainty-Aware Self-Reflection
di: Liu, Liangxin, et al.
Pubblicazione: (2024)
di: Liu, Liangxin, et al.
Pubblicazione: (2024)
Understanding Token Probability Encoding in Output Embeddings
di: Cho, Hakaze, et al.
Pubblicazione: (2024)
di: Cho, Hakaze, et al.
Pubblicazione: (2024)
Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces
di: Yu, Shixing, et al.
Pubblicazione: (2026)
di: Yu, Shixing, et al.
Pubblicazione: (2026)
DLM-Scope: Mechanistic Interpretability of Diffusion Language Models via Sparse Autoencoders
di: Wang, Xu, et al.
Pubblicazione: (2026)
di: Wang, Xu, et al.
Pubblicazione: (2026)
InT: Self-Proposed Interventions Enable Credit Assignment in LLM Reasoning
di: Yang, Matthew Y. R., et al.
Pubblicazione: (2026)
di: Yang, Matthew Y. R., et al.
Pubblicazione: (2026)
An Evaluation on Large Language Model Outputs: Discourse and Memorization
di: de Wynter, Adrian, et al.
Pubblicazione: (2023)
di: de Wynter, Adrian, et al.
Pubblicazione: (2023)
Latent Chain-of-Thought? Decoding the Depth-Recurrent Transformer
di: Lu, Wenquan, et al.
Pubblicazione: (2025)
di: Lu, Wenquan, et al.
Pubblicazione: (2025)
Steer LLM Latents for Hallucination Detection
di: Park, Seongheon, et al.
Pubblicazione: (2025)
di: Park, Seongheon, et al.
Pubblicazione: (2025)
Understanding LLM Embeddings for Regression
di: Tang, Eric, et al.
Pubblicazione: (2024)
di: Tang, Eric, et al.
Pubblicazione: (2024)
Internalizing LLM Reasoning via Discovery and Replay of Latent Actions
di: Shi, Zhenning, et al.
Pubblicazione: (2026)
di: Shi, Zhenning, et al.
Pubblicazione: (2026)
SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning
di: Chen, Jiaqi, et al.
Pubblicazione: (2025)
di: Chen, Jiaqi, et al.
Pubblicazione: (2025)
DOGe: Defensive Output Generation for LLM Protection Against Knowledge Distillation
di: Li, Pingzhi, et al.
Pubblicazione: (2025)
di: Li, Pingzhi, et al.
Pubblicazione: (2025)
GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation
di: Li, Sijia, et al.
Pubblicazione: (2026)
di: Li, Sijia, et al.
Pubblicazione: (2026)
Symbolic Learning Enables Self-Evolving Agents
di: Zhou, Wangchunshu, et al.
Pubblicazione: (2024)
di: Zhou, Wangchunshu, et al.
Pubblicazione: (2024)
Training-free LLM Merging for Multi-task Learning
di: Fu, Zichuan, et al.
Pubblicazione: (2025)
di: Fu, Zichuan, et al.
Pubblicazione: (2025)
Are LLM Evaluators Really Narcissists? Sanity Checking Self-Preference Evaluations
di: Roytburg, Dani, et al.
Pubblicazione: (2026)
di: Roytburg, Dani, et al.
Pubblicazione: (2026)
LatentBreak: Jailbreaking Large Language Models through Latent Space Feedback
di: Mura, Raffaele, et al.
Pubblicazione: (2025)
di: Mura, Raffaele, et al.
Pubblicazione: (2025)
A Formal Comparison Between Chain of Thought and Latent Thought
di: Xu, Kevin, et al.
Pubblicazione: (2025)
di: Xu, Kevin, et al.
Pubblicazione: (2025)
Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks
di: Li, Miaomiao, et al.
Pubblicazione: (2025)
di: Li, Miaomiao, et al.
Pubblicazione: (2025)
Training Proactive and Personalized LLM Agents
di: Sun, Weiwei, et al.
Pubblicazione: (2025)
di: Sun, Weiwei, et al.
Pubblicazione: (2025)
Control the Temperature: Selective Sampling for Diverse and High-Quality LLM Outputs
di: Troshin, Sergey, et al.
Pubblicazione: (2025)
di: Troshin, Sergey, et al.
Pubblicazione: (2025)
Confidence-aware Self-Semantic Distillation on Knowledge Graph Embedding
di: Liu, Yichen, et al.
Pubblicazione: (2022)
di: Liu, Yichen, et al.
Pubblicazione: (2022)
Empowering Diffusion Models on the Embedding Space for Text Generation
di: Gao, Zhujin, et al.
Pubblicazione: (2022)
di: Gao, Zhujin, et al.
Pubblicazione: (2022)
ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces
di: Xu, Xin, et al.
Pubblicazione: (2026)
di: Xu, Xin, et al.
Pubblicazione: (2026)
User-LLM: Efficient LLM Contextualization with User Embeddings
di: Ning, Lin, et al.
Pubblicazione: (2024)
di: Ning, Lin, et al.
Pubblicazione: (2024)
Self-Imagine: Effective Unimodal Reasoning with Multimodal Models using Self-Imagination
di: Akter, Syeda Nahida, et al.
Pubblicazione: (2024)
di: Akter, Syeda Nahida, et al.
Pubblicazione: (2024)
GradAlign: Gradient-Aligned Data Selection for LLM Reinforcement Learning
di: Yang, Ningyuan, et al.
Pubblicazione: (2026)
di: Yang, Ningyuan, et al.
Pubblicazione: (2026)
Measuring Faithfulness Depends on How You Measure: Classifier Sensitivity in LLM Chain-of-Thought Evaluation
di: Young, Richard J.
Pubblicazione: (2026)
di: Young, Richard J.
Pubblicazione: (2026)
Enhancing LLM Evaluations: The Garbling Trick
di: Bradley, William F.
Pubblicazione: (2024)
di: Bradley, William F.
Pubblicazione: (2024)
Think When You Need: Self-Adaptive Chain-of-Thought Learning
di: Yang, Junjie, et al.
Pubblicazione: (2025)
di: Yang, Junjie, et al.
Pubblicazione: (2025)
LLM-NEO: Parameter Efficient Knowledge Distillation for Large Language Models
di: Yang, Runming, et al.
Pubblicazione: (2024)
di: Yang, Runming, et al.
Pubblicazione: (2024)
Think-Augmented Function Calling: Improving LLM Parameter Accuracy Through Embedded Reasoning
di: Wei, Lei, et al.
Pubblicazione: (2026)
di: Wei, Lei, et al.
Pubblicazione: (2026)
LLMs as Zero-shot Graph Learners: Alignment of GNN Representations with LLM Token Embeddings
di: Wang, Duo, et al.
Pubblicazione: (2024)
di: Wang, Duo, et al.
Pubblicazione: (2024)
Scaling Embeddings Outperforms Scaling Experts in Language Models
di: Liu, Hong, et al.
Pubblicazione: (2026)
di: Liu, Hong, et al.
Pubblicazione: (2026)
Deliberation in Latent Space via Differentiable Cache Augmentation
di: Liu, Luyang, et al.
Pubblicazione: (2024)
di: Liu, Luyang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Embedding Trajectory for Out-of-Distribution Detection in Mathematical Reasoning
di: Wang, Yiming, et al.
Pubblicazione: (2024) -
Output Embedding Centering for Stable LLM Pretraining
di: Stollenwerk, Felix, et al.
Pubblicazione: (2026) -
Sampling-Efficient Test-Time Scaling: Self-Estimating the Best-of-N Sampling in Early Decoding
di: Wang, Yiming, et al.
Pubblicazione: (2025) -
From Reasoning Chains to Verifiable Subproblems: Curriculum Reinforcement Learning Enables Credit Assignment for LLM Reasoning
di: Jiang, Xitai, et al.
Pubblicazione: (2026) -
On the Overscaling Curse of Parallel Thinking: System Efficacy Contradicts Sample Efficiency
di: Wang, Yiming, et al.
Pubblicazione: (2026)