Improving Self Consistency in LLMs through Probabilistic Tokenization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sathe, Ashutosh, Aggarwal, Divyanshu, Sitaram, Sunayana |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploring Pretraining via Active Forgetting for Improving Cross Lingual Transfer for Decoder Language Models
von: Aggarwal, Divyanshu, et al.
Veröffentlicht: (2024)
von: Aggarwal, Divyanshu, et al.
Veröffentlicht: (2024)
MAPLE: Multilingual Evaluation of Parameter Efficient Finetuning of Large Language Models
von: Aggarwal, Divyanshu, et al.
Veröffentlicht: (2024)
von: Aggarwal, Divyanshu, et al.
Veröffentlicht: (2024)
Efficient Training of Language Models with Compact and Consistent Next Token Distributions
von: Sathe, Ashutosh, et al.
Veröffentlicht: (2024)
von: Sathe, Ashutosh, et al.
Veröffentlicht: (2024)
A Unified Framework and Dataset for Assessing Societal Bias in Vision-Language Models
von: Sathe, Ashutosh, et al.
Veröffentlicht: (2024)
von: Sathe, Ashutosh, et al.
Veröffentlicht: (2024)
MAFIA: Multi-Adapter Fused Inclusive LanguAge Models
von: Jain, Prachi, et al.
Veröffentlicht: (2024)
von: Jain, Prachi, et al.
Veröffentlicht: (2024)
Exploring Continual Fine-Tuning for Enhancing Language Ability in Large Language Model
von: Aggarwal, Divyanshu, et al.
Veröffentlicht: (2024)
von: Aggarwal, Divyanshu, et al.
Veröffentlicht: (2024)
Bridging the Gap: Dynamic Learning Strategies for Improving Multilingual Performance in LLMs
von: Kumar, Somnath, et al.
Veröffentlicht: (2024)
von: Kumar, Somnath, et al.
Veröffentlicht: (2024)
MEGAVERSE: Benchmarking Large Language Models Across Languages, Modalities, Models and Tasks
von: Ahuja, Sanchit, et al.
Veröffentlicht: (2023)
von: Ahuja, Sanchit, et al.
Veröffentlicht: (2023)
CultureLLM: Incorporating Cultural Differences into Large Language Models
von: Li, Cheng, et al.
Veröffentlicht: (2024)
von: Li, Cheng, et al.
Veröffentlicht: (2024)
Geometry of Decision Making in Language Models
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
Self-Training Meets Consistency: Improving LLMs' Reasoning with Consistency-Driven Rationale Evaluation
von: Lee, Jaehyeok, et al.
Veröffentlicht: (2024)
von: Lee, Jaehyeok, et al.
Veröffentlicht: (2024)
M5 -- A Diverse Benchmark to Assess the Performance of Large Multimodal Models Across Multilingual and Multicultural Vision-Language Tasks
von: Schneider, Florian, et al.
Veröffentlicht: (2024)
von: Schneider, Florian, et al.
Veröffentlicht: (2024)
On Next-Token Prediction in LLMs: How End Goals Determine the Consistency of Decoding Algorithms
von: Trauger, Jacob, et al.
Veröffentlicht: (2025)
von: Trauger, Jacob, et al.
Veröffentlicht: (2025)
Initial Decoding with Minimally Augmented Language Model for Improved Lattice Rescoring in Low Resource ASR
von: Murthy, Savitha, et al.
Veröffentlicht: (2024)
von: Murthy, Savitha, et al.
Veröffentlicht: (2024)
Temporally Consistent Factuality Probing for Large Language Models
von: Bajpai, Ashutosh, et al.
Veröffentlicht: (2024)
von: Bajpai, Ashutosh, et al.
Veröffentlicht: (2024)
Towards Logically Consistent Language Models via Probabilistic Reasoning
von: Calanzone, Diego, et al.
Veröffentlicht: (2024)
von: Calanzone, Diego, et al.
Veröffentlicht: (2024)
Found in the Middle: Permutation Self-Consistency Improves Listwise Ranking in Large Language Models
von: Tang, Raphael, et al.
Veröffentlicht: (2023)
von: Tang, Raphael, et al.
Veröffentlicht: (2023)
Toward a Theory of Tokenization in LLMs
von: Rajaraman, Nived, et al.
Veröffentlicht: (2024)
von: Rajaraman, Nived, et al.
Veröffentlicht: (2024)
Soft Self-Consistency Improves Language Model Agents
von: Wang, Han, et al.
Veröffentlicht: (2024)
von: Wang, Han, et al.
Veröffentlicht: (2024)
Probabilistic Reasoning with LLMs for k-anonymity Estimation
von: Zheng, Jonathan, et al.
Veröffentlicht: (2025)
von: Zheng, Jonathan, et al.
Veröffentlicht: (2025)
Don't Throw Away Your Beams: Improving Consistency-based Uncertainties in LLMs via Beam Search
von: Fadeeva, Ekaterina, et al.
Veröffentlicht: (2025)
von: Fadeeva, Ekaterina, et al.
Veröffentlicht: (2025)
ConsistRM: Improving Generative Reward Models via Consistency-Aware Self-Training
von: Liang, Yu, et al.
Veröffentlicht: (2026)
von: Liang, Yu, et al.
Veröffentlicht: (2026)
Contamination Report for Multilingual Benchmarks
von: Ahuja, Sanchit, et al.
Veröffentlicht: (2024)
von: Ahuja, Sanchit, et al.
Veröffentlicht: (2024)
Improving Diffusion Language Model Decoding through Joint Search in Generation Order and Token Space
von: Shen, Yangyi, et al.
Veröffentlicht: (2026)
von: Shen, Yangyi, et al.
Veröffentlicht: (2026)
Finding the Cracks: Improving LLMs Reasoning with Paraphrastic Probing and Consistency Verification
von: Shi, Weili, et al.
Veröffentlicht: (2026)
von: Shi, Weili, et al.
Veröffentlicht: (2026)
Persuasion Tokens for Editing Factual Knowledge in LLMs
von: Youssef, Paul, et al.
Veröffentlicht: (2026)
von: Youssef, Paul, et al.
Veröffentlicht: (2026)
Silent Tokens, Loud Effects: Padding in LLMs
von: Himelstein, Rom, et al.
Veröffentlicht: (2025)
von: Himelstein, Rom, et al.
Veröffentlicht: (2025)
Fluent but Foreign: Even Regional LLMs Lack Cultural Alignment
von: Agarwal, Dhruv, et al.
Veröffentlicht: (2025)
von: Agarwal, Dhruv, et al.
Veröffentlicht: (2025)
Improving Sequence-to-Sequence Models for Abstractive Text Summarization Using Meta Heuristic Approaches
von: Saxena, Aditya, et al.
Veröffentlicht: (2024)
von: Saxena, Aditya, et al.
Veröffentlicht: (2024)
Tokenization counts: the impact of tokenization on arithmetic in frontier LLMs
von: Singh, Aaditya K., et al.
Veröffentlicht: (2024)
von: Singh, Aaditya K., et al.
Veröffentlicht: (2024)
Probabilistic Token Alignment for Large Language Model Fusion
von: Zeng, Runjia, et al.
Veröffentlicht: (2025)
von: Zeng, Runjia, et al.
Veröffentlicht: (2025)
Self-Consistency via Marginal Sharpening
von: Arzhantsev, Aleksei, et al.
Veröffentlicht: (2026)
von: Arzhantsev, Aleksei, et al.
Veröffentlicht: (2026)
LoRMA: Low-Rank Multiplicative Adaptation for LLMs
von: Bihany, Harsh, et al.
Veröffentlicht: (2025)
von: Bihany, Harsh, et al.
Veröffentlicht: (2025)
Calibration Across Layers: Understanding Calibration Evolution in LLMs
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
Multi-Token Prediction via Self-Distillation
von: Kirchenbauer, John, et al.
Veröffentlicht: (2026)
von: Kirchenbauer, John, et al.
Veröffentlicht: (2026)
IITK at SemEval-2024 Task 2: Exploring the Capabilities of LLMs for Safe Biomedical Natural Language Inference for Clinical Trials
von: Mandal, Shreyasi, et al.
Veröffentlicht: (2024)
von: Mandal, Shreyasi, et al.
Veröffentlicht: (2024)
Trained on Tokens, Calibrated on Concepts: The Emergence of Semantic Calibration in LLMs
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025)
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025)
Curiosity-Driven LLM-as-a-judge for Personalized Creative Judgment
von: Kumar, Vanya Bannihatti, et al.
Veröffentlicht: (2025)
von: Kumar, Vanya Bannihatti, et al.
Veröffentlicht: (2025)
Speech Representation Learning Revisited: The Necessity of Separate Learnable Parameters and Robust Data Augmentation
von: Yadav, Hemant, et al.
Veröffentlicht: (2024)
von: Yadav, Hemant, et al.
Veröffentlicht: (2024)
MS-HuBERT: Mitigating Pre-training and Inference Mismatch in Masked Language Modelling methods for learning Speech Representations
von: Yadav, Hemant, et al.
Veröffentlicht: (2024)
von: Yadav, Hemant, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Exploring Pretraining via Active Forgetting for Improving Cross Lingual Transfer for Decoder Language Models
von: Aggarwal, Divyanshu, et al.
Veröffentlicht: (2024) -
MAPLE: Multilingual Evaluation of Parameter Efficient Finetuning of Large Language Models
von: Aggarwal, Divyanshu, et al.
Veröffentlicht: (2024) -
Efficient Training of Language Models with Compact and Consistent Next Token Distributions
von: Sathe, Ashutosh, et al.
Veröffentlicht: (2024) -
A Unified Framework and Dataset for Assessing Societal Bias in Vision-Language Models
von: Sathe, Ashutosh, et al.
Veröffentlicht: (2024) -
MAFIA: Multi-Adapter Fused Inclusive LanguAge Models
von: Jain, Prachi, et al.
Veröffentlicht: (2024)