Test-Time Scaling with Repeated Sampling Improves Multilingual Text Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Gupta, Ashim, Srikumar, Vivek |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Found in Translation: Measuring Multilingual LLM Consistency as Simple as Translate then Evaluate
by: Gupta, Ashim, et al.
Published: (2025)
by: Gupta, Ashim, et al.
Published: (2025)
State Space Models are Strong Text Rerankers
by: Xu, Zhichao, et al.
Published: (2024)
by: Xu, Zhichao, et al.
Published: (2024)
Reinforcing Code Generation: Improving Text-to-SQL with Execution-Based Learning
by: Kulkarni, Atharv, et al.
Published: (2025)
by: Kulkarni, Atharv, et al.
Published: (2025)
An Empirical Investigation of Matrix Factorization Methods for Pre-trained Transformers
by: Gupta, Ashim, et al.
Published: (2024)
by: Gupta, Ashim, et al.
Published: (2024)
Whispers of Doubt Amidst Echoes of Triumph in NLP Robustness
by: Gupta, Ashim, et al.
Published: (2023)
by: Gupta, Ashim, et al.
Published: (2023)
Beyond Perplexity: Multi-dimensional Safety Evaluation of LLM Compression
by: Xu, Zhichao, et al.
Published: (2024)
by: Xu, Zhichao, et al.
Published: (2024)
Defragmenting Language Models: An Interpretability-based Approach for Vocabulary Expansion
by: Mehta, Maitrey, et al.
Published: (2026)
by: Mehta, Maitrey, et al.
Published: (2026)
Unequal Voices: How LLMs Construct Constrained Queer Narratives
by: Ghosal, Atreya, et al.
Published: (2025)
by: Ghosal, Atreya, et al.
Published: (2025)
InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis
by: Bentham, Oliver, et al.
Published: (2026)
by: Bentham, Oliver, et al.
Published: (2026)
LLM-Symbolic Integration for Robust Temporal Tabular Reasoning
by: Kulkarni, Atharv, et al.
Published: (2025)
by: Kulkarni, Atharv, et al.
Published: (2025)
Understanding the Logic of Direct Preference Alignment through Logic
by: Richardson, Kyle, et al.
Published: (2024)
by: Richardson, Kyle, et al.
Published: (2024)
Promptly Predicting Structures: The Return of Inference
by: Mehta, Maitrey, et al.
Published: (2024)
by: Mehta, Maitrey, et al.
Published: (2024)
Enhancing Question Answering on Charts Through Effective Pre-training Tasks
by: Gupta, Ashim, et al.
Published: (2024)
by: Gupta, Ashim, et al.
Published: (2024)
Multilingual Test-Time Scaling via Initial Thought Transfer
by: Bajpai, Prasoon, et al.
Published: (2025)
by: Bajpai, Prasoon, et al.
Published: (2025)
Map&Make: Schema Guided Text to Table Generation
by: Ahuja, Naman, et al.
Published: (2025)
by: Ahuja, Naman, et al.
Published: (2025)
In-Context Example Ordering Guided by Label Distributions
by: Xu, Zhichao, et al.
Published: (2024)
by: Xu, Zhichao, et al.
Published: (2024)
Improving Text-to-Image Generation with Input-Side Inference-Time Scaling
by: Chen, Ruibo, et al.
Published: (2025)
by: Chen, Ruibo, et al.
Published: (2025)
Leveraging LLM For Synchronizing Information Across Multilingual Tables
by: Khincha, Siddharth, et al.
Published: (2025)
by: Khincha, Siddharth, et al.
Published: (2025)
UNJOIN: Enhancing Multi-Table Text-to-SQL Generation via Schema Simplification
by: Ganesan, Poojah, et al.
Published: (2025)
by: Ganesan, Poojah, et al.
Published: (2025)
Distillation versus Contrastive Learning: How to Train Your Rerankers
by: Xu, Zhichao, et al.
Published: (2025)
by: Xu, Zhichao, et al.
Published: (2025)
Measuring and Improving Attentiveness to Partial Inputs with Counterfactuals
by: Elazar, Yanai, et al.
Published: (2023)
by: Elazar, Yanai, et al.
Published: (2023)
Accelerated Test-Time Scaling with Model-Free Speculative Sampling
by: Song, Woomin, et al.
Published: (2025)
by: Song, Woomin, et al.
Published: (2025)
RLShield: Practical Multi-Agent RL for Financial Cyber Defense with Attack-Surface MDPs and Real-Time Response Orchestration
by: Nayak, Srikumar
Published: (2026)
by: Nayak, Srikumar
Published: (2026)
ChartAttack: Testing the Vulnerability of LLMs to Malicious Prompting in Chart Generation
by: Ortiz-Barajas, Jesus-German, et al.
Published: (2026)
by: Ortiz-Barajas, Jesus-German, et al.
Published: (2026)
LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling
by: Zheng, Tong, et al.
Published: (2026)
by: Zheng, Tong, et al.
Published: (2026)
Adaptive Rectification Sampling for Test-Time Compute Scaling
by: Tan, Zhendong, et al.
Published: (2025)
by: Tan, Zhendong, et al.
Published: (2025)
Scale-free Characteristics of Multilingual Legal Texts and the Limitations of LLMs
by: Chen, Haoyang, et al.
Published: (2025)
by: Chen, Haoyang, et al.
Published: (2025)
MULTITuDE: Large-Scale Multilingual Machine-Generated Text Detection Benchmark
by: Macko, Dominik, et al.
Published: (2023)
by: Macko, Dominik, et al.
Published: (2023)
On the Role of Temperature Sampling in Test-Time Scaling
by: Wu, Yuheng, et al.
Published: (2025)
by: Wu, Yuheng, et al.
Published: (2025)
EEG-to-Text Translation: A Model for Deciphering Human Brain Activity
by: Murad, Saydul Akbar, et al.
Published: (2025)
by: Murad, Saydul Akbar, et al.
Published: (2025)
Agentar-Scale-SQL: Advancing Text-to-SQL through Orchestrated Test-Time Scaling
by: Wang, Pengfei, et al.
Published: (2025)
by: Wang, Pengfei, et al.
Published: (2025)
Authorship Obfuscation in Multilingual Machine-Generated Text Detection
by: Macko, Dominik, et al.
Published: (2024)
by: Macko, Dominik, et al.
Published: (2024)
HQFS: Hybrid Quantum Classical Financial Security with VQC Forecasting, QUBO Annealing, and Audit-Ready Post-Quantum Signing
by: Nayak, Srikumar
Published: (2026)
by: Nayak, Srikumar
Published: (2026)
Named Entity Recognition for Payment Data Using NLP
by: Nayak, Srikumar
Published: (2026)
by: Nayak, Srikumar
Published: (2026)
Multilingual Prompting for Improving LLM Generation Diversity
by: Wang, Qihan, et al.
Published: (2025)
by: Wang, Qihan, et al.
Published: (2025)
Iterative Deepening Sampling as Efficient Test-Time Scaling
by: Chen, Weizhe, et al.
Published: (2025)
by: Chen, Weizhe, et al.
Published: (2025)
Is That Your Final Answer? Test-Time Scaling Improves Selective Question Answering
by: Jurayj, William, et al.
Published: (2025)
by: Jurayj, William, et al.
Published: (2025)
mEdIT: Multilingual Text Editing via Instruction Tuning
by: Raheja, Vipul, et al.
Published: (2024)
by: Raheja, Vipul, et al.
Published: (2024)
PARIKSHA: A Large-Scale Investigation of Human-LLM Evaluator Agreement on Multilingual and Multi-Cultural Data
by: Watts, Ishaan, et al.
Published: (2024)
by: Watts, Ishaan, et al.
Published: (2024)
Test-Time Scaling with Reflective Generative Model
by: Wang, Zixiao, et al.
Published: (2025)
by: Wang, Zixiao, et al.
Published: (2025)
Similar Items
-
Found in Translation: Measuring Multilingual LLM Consistency as Simple as Translate then Evaluate
by: Gupta, Ashim, et al.
Published: (2025) -
State Space Models are Strong Text Rerankers
by: Xu, Zhichao, et al.
Published: (2024) -
Reinforcing Code Generation: Improving Text-to-SQL with Execution-Based Learning
by: Kulkarni, Atharv, et al.
Published: (2025) -
An Empirical Investigation of Matrix Factorization Methods for Pre-trained Transformers
by: Gupta, Ashim, et al.
Published: (2024) -
Whispers of Doubt Amidst Echoes of Triumph in NLP Robustness
by: Gupta, Ashim, et al.
Published: (2023)