Top-b: Entropic Regulation of Relative Probability Bands in Autoregressive Language Processes
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Halder, Deepon, Dabre, Raj |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Scripts Through Time: A Survey of the Evolving Role of Transliteration in NLP
von: Jayakumar, Thanmay, et al.
Veröffentlicht: (2026)
von: Jayakumar, Thanmay, et al.
Veröffentlicht: (2026)
CycleDistill: Bootstrapping Machine Translation using LLMs with Cyclical Distillation
von: Halder, Deepon, et al.
Veröffentlicht: (2025)
von: Halder, Deepon, et al.
Veröffentlicht: (2025)
RiddleBench: A New Generative Reasoning Benchmark for LLMs
von: Halder, Deepon, et al.
Veröffentlicht: (2025)
von: Halder, Deepon, et al.
Veröffentlicht: (2025)
Multilingual TinyStories: A Synthetic Combinatorial Corpus of Indic Children's Stories for Training Small Language Models
von: Halder, Deepon, et al.
Veröffentlicht: (2026)
von: Halder, Deepon, et al.
Veröffentlicht: (2026)
Are Language Models Agnostic to Linguistically Grounded Perturbations? A Case Study of Indic Languages
von: Ghosh, Poulami, et al.
Veröffentlicht: (2024)
von: Ghosh, Poulami, et al.
Veröffentlicht: (2024)
Pretraining Language Models Using Translationese
von: Doshi, Meet, et al.
Veröffentlicht: (2024)
von: Doshi, Meet, et al.
Veröffentlicht: (2024)
Natural Language Processing for Dialects of a Language: A Survey
von: Joshi, Aditya, et al.
Veröffentlicht: (2024)
von: Joshi, Aditya, et al.
Veröffentlicht: (2024)
How effective is Multi-source pivoting for Translation of Low Resource Indian Languages?
von: Gaikwad, Pranav, et al.
Veröffentlicht: (2024)
von: Gaikwad, Pranav, et al.
Veröffentlicht: (2024)
IndicRAGSuite: Large-Scale Datasets and a Benchmark for Indian Language RAG Systems
von: Prasanjith, Pasunuti, et al.
Veröffentlicht: (2025)
von: Prasanjith, Pasunuti, et al.
Veröffentlicht: (2025)
An Empirical Study of In-context Learning in LLMs for Machine Translation
von: Chitale, Pranjal A., et al.
Veröffentlicht: (2024)
von: Chitale, Pranjal A., et al.
Veröffentlicht: (2024)
Mark My Words: A Robust Multilingual Model for Punctuation in Text and Speech Transcripts
von: Pulipaka, Sidharth, et al.
Veröffentlicht: (2025)
von: Pulipaka, Sidharth, et al.
Veröffentlicht: (2025)
A Morphology-Based Investigation of Positional Encodings
von: Ghosh, Poulami, et al.
Veröffentlicht: (2024)
von: Ghosh, Poulami, et al.
Veröffentlicht: (2024)
Pralekha: Cross-Lingual Document Alignment for Indic Languages
von: Suryanarayanan, Sanjay, et al.
Veröffentlicht: (2024)
von: Suryanarayanan, Sanjay, et al.
Veröffentlicht: (2024)
The Reasoning Lingua Franca: A Double-Edged Sword for Multilingual AI
von: Saji, Alan, et al.
Veröffentlicht: (2025)
von: Saji, Alan, et al.
Veröffentlicht: (2025)
IndicIFEval: A Benchmark for Verifiable Instruction-Following Evaluation in 14 Indic Languages
von: Jayakumar, Thanmay, et al.
Veröffentlicht: (2026)
von: Jayakumar, Thanmay, et al.
Veröffentlicht: (2026)
Probability Distributions Computed by Autoregressive Transformers
von: Yang, Andy, et al.
Veröffentlicht: (2025)
von: Yang, Andy, et al.
Veröffentlicht: (2025)
PrahokBART: A Pre-trained Sequence-to-Sequence Model for Khmer Natural Language Generation
von: Kaing, Hour, et al.
Veröffentlicht: (2025)
von: Kaing, Hour, et al.
Veröffentlicht: (2025)
How Good is Zero-Shot MT Evaluation for Low Resource Indian Languages?
von: Singh, Anushka, et al.
Veröffentlicht: (2024)
von: Singh, Anushka, et al.
Veröffentlicht: (2024)
An Empirical Comparison of Vocabulary Expansion and Initialization Approaches for Language Models
von: Mundra, Nandini, et al.
Veröffentlicht: (2024)
von: Mundra, Nandini, et al.
Veröffentlicht: (2024)
RomanSetu: Efficiently unlocking multilingual capabilities of Large Language Models via Romanization
von: Husain, Jaavid Aktar, et al.
Veröffentlicht: (2024)
von: Husain, Jaavid Aktar, et al.
Veröffentlicht: (2024)
Incremental Sentence Processing Mechanisms in Autoregressive Transformer Language Models
von: Hanna, Michael, et al.
Veröffentlicht: (2024)
von: Hanna, Michael, et al.
Veröffentlicht: (2024)
Context Dependence and Reliability in Autoregressive Language Models
von: Sengupta, Poushali, et al.
Veröffentlicht: (2026)
von: Sengupta, Poushali, et al.
Veröffentlicht: (2026)
PUB: A Pragmatics Understanding Benchmark for Assessing LLMs' Pragmatics Capabilities
von: Sravanthi, Settaluri Lakshmi, et al.
Veröffentlicht: (2024)
von: Sravanthi, Settaluri Lakshmi, et al.
Veröffentlicht: (2024)
Preferences of a Voice-First Nation: Large-Scale Pairwise Evaluation and Preference Analysis for TTS in Indian Languages
von: Anand, Srija, et al.
Veröffentlicht: (2026)
von: Anand, Srija, et al.
Veröffentlicht: (2026)
When Alignment Hurts: Decoupling Representational Spaces in Multilingual Models
von: Elshabrawy, Ahmed, et al.
Veröffentlicht: (2025)
von: Elshabrawy, Ahmed, et al.
Veröffentlicht: (2025)
Controlling Large Language Model Agents with Entropic Activation Steering
von: Rahn, Nate, et al.
Veröffentlicht: (2024)
von: Rahn, Nate, et al.
Veröffentlicht: (2024)
Cross-Lingual Auto Evaluation for Assessing Multilingual LLMs
von: Doddapaneni, Sumanth, et al.
Veröffentlicht: (2024)
von: Doddapaneni, Sumanth, et al.
Veröffentlicht: (2024)
TopK Language Models
von: Takahashi, Ryosuke, et al.
Veröffentlicht: (2025)
von: Takahashi, Ryosuke, et al.
Veröffentlicht: (2025)
Towards Building Large Scale Datasets and State-of-the-Art Automatic Speech Translation Systems for 14 Indian Languages
von: Sankar, Ashwin, et al.
Veröffentlicht: (2024)
von: Sankar, Ashwin, et al.
Veröffentlicht: (2024)
IteRABRe: Iterative Recovery-Aided Block Reduction
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2025)
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2025)
Limited-Resource Adapters Are Regularizers, Not Linguists
von: Fekete, Marcell, et al.
Veröffentlicht: (2025)
von: Fekete, Marcell, et al.
Veröffentlicht: (2025)
A Probability--Quality Trade-off in Aligned Language Models and its Relation to Sampling Adaptors
von: Tan, Naaman, et al.
Veröffentlicht: (2024)
von: Tan, Naaman, et al.
Veröffentlicht: (2024)
Causal Autoregressive Diffusion Language Model
von: Ruan, Junhao, et al.
Veröffentlicht: (2026)
von: Ruan, Junhao, et al.
Veröffentlicht: (2026)
Reversal Invariance in Autoregressive Language Models
von: Sahasrabudhe, Mihir
Veröffentlicht: (2025)
von: Sahasrabudhe, Mihir
Veröffentlicht: (2025)
TikZero: Zero-Shot Text-Guided Graphics Program Synthesis
von: Belouadi, Jonas, et al.
Veröffentlicht: (2025)
von: Belouadi, Jonas, et al.
Veröffentlicht: (2025)
Revisiting Knowledge Distillation for Autoregressive Language Models
von: Zhong, Qihuang, et al.
Veröffentlicht: (2024)
von: Zhong, Qihuang, et al.
Veröffentlicht: (2024)
Autoregressive Large Language Models are Computationally Universal
von: Schuurmans, Dale, et al.
Veröffentlicht: (2024)
von: Schuurmans, Dale, et al.
Veröffentlicht: (2024)
IndicLLMSuite: A Blueprint for Creating Pre-training and Fine-Tuning Datasets for Indian Languages
von: Khan, Mohammed Safi Ur Rahman, et al.
Veröffentlicht: (2024)
von: Khan, Mohammed Safi Ur Rahman, et al.
Veröffentlicht: (2024)
Combining Autoregressive and Autoencoder Language Models for Text Classification
von: Gonçalves, João
Veröffentlicht: (2024)
von: Gonçalves, João
Veröffentlicht: (2024)
OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models
von: Huang, Siming, et al.
Veröffentlicht: (2024)
von: Huang, Siming, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Scripts Through Time: A Survey of the Evolving Role of Transliteration in NLP
von: Jayakumar, Thanmay, et al.
Veröffentlicht: (2026) -
CycleDistill: Bootstrapping Machine Translation using LLMs with Cyclical Distillation
von: Halder, Deepon, et al.
Veröffentlicht: (2025) -
RiddleBench: A New Generative Reasoning Benchmark for LLMs
von: Halder, Deepon, et al.
Veröffentlicht: (2025) -
Multilingual TinyStories: A Synthetic Combinatorial Corpus of Indic Children's Stories for Training Small Language Models
von: Halder, Deepon, et al.
Veröffentlicht: (2026) -
Are Language Models Agnostic to Linguistically Grounded Perturbations? A Case Study of Indic Languages
von: Ghosh, Poulami, et al.
Veröffentlicht: (2024)