QueST: Incentivizing LLMs to Generate Difficult Problems
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Hanxu, Zhang, Xingxing, Vamvas, Jannis, Sennrich, Rico, Wei, Furu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Source-primed Multi-turn Conversation Helps Large Language Models Translate Documents
by: Hu, Hanxu, et al.
Published: (2025)
by: Hu, Hanxu, et al.
Published: (2025)
Linear-time Minimum Bayes Risk Decoding with Reference Aggregation
by: Vamvas, Jannis, et al.
Published: (2024)
by: Vamvas, Jannis, et al.
Published: (2024)
SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents
by: Wastl, Michelle, et al.
Published: (2025)
by: Wastl, Michelle, et al.
Published: (2025)
Machine Translation Models are Zero-Shot Detectors of Translation Direction
by: Wastl, Michelle, et al.
Published: (2024)
by: Wastl, Michelle, et al.
Published: (2024)
Mitigating Hallucinations and Off-target Machine Translation with Source-Contrastive and Language-Contrastive Decoding
by: Sennrich, Rico, et al.
Published: (2023)
by: Sennrich, Rico, et al.
Published: (2023)
SwissBERT: The Multilingual Language Model for Switzerland
by: Vamvas, Jannis, et al.
Published: (2023)
by: Vamvas, Jannis, et al.
Published: (2023)
Modular Adaptation of Multilingual Encoders to Written Swiss German Dialect
by: Vamvas, Jannis, et al.
Published: (2024)
by: Vamvas, Jannis, et al.
Published: (2024)
Investigating Multi-Pivot Ensembling with Massively Multilingual Machine Translation Models
by: Mohammadshahi, Alireza, et al.
Published: (2023)
by: Mohammadshahi, Alireza, et al.
Published: (2023)
DeReason: A Difficulty-Aware Curriculum Improves Decoupled SFT-then-RL Training for General Reasoning
by: Hu, Hanxu, et al.
Published: (2026)
by: Hu, Hanxu, et al.
Published: (2026)
Leveraging In-Context Learning for Political Bias Testing of LLMs
by: Haller, Patrick, et al.
Published: (2025)
by: Haller, Patrick, et al.
Published: (2025)
20min-XD: A Comparable Corpus of Swiss News Articles
by: Wastl, Michelle, et al.
Published: (2025)
by: Wastl, Michelle, et al.
Published: (2025)
The Mediomatix Corpus: Parallel Data for Romansh Language Varieties via Comparable Schoolbooks
by: Hopton, Zachary, et al.
Published: (2025)
by: Hopton, Zachary, et al.
Published: (2025)
Translation Asymmetry in LLMs as a Data Augmentation Factor: A Case Study for 6 Romansh Language Varieties
by: Vamvas, Jannis, et al.
Published: (2026)
by: Vamvas, Jannis, et al.
Published: (2026)
Fine-tuning the SwissBERT Encoder Model for Embedding Sentences and Documents
by: Grosjean, Juri, et al.
Published: (2024)
by: Grosjean, Juri, et al.
Published: (2024)
Turning English-centric LLMs Into Polyglots: How Much Multilinguality Is Needed?
by: Kew, Tannon, et al.
Published: (2023)
by: Kew, Tannon, et al.
Published: (2023)
Robust Language Identification for Romansh Varieties
by: Model, Charlotte, et al.
Published: (2026)
by: Model, Charlotte, et al.
Published: (2026)
Measuring the Effect of Disfluency in Multilingual Knowledge Probing Benchmarks
by: Semenov, Kirill, et al.
Published: (2025)
by: Semenov, Kirill, et al.
Published: (2025)
RUMLEM: A Dictionary-Based Lemmatizer for Romansh
by: Fischer, Dominic P., et al.
Published: (2026)
by: Fischer, Dominic P., et al.
Published: (2026)
Examining Multilingual Embedding Models Cross-Lingually Through LLM-Generated Adversarial Examples
by: Michail, Andrianos, et al.
Published: (2025)
by: Michail, Andrianos, et al.
Published: (2025)
LLMs Encode How Difficult Problems Are
by: Lugoloobi, William, et al.
Published: (2025)
by: Lugoloobi, William, et al.
Published: (2025)
Evaluating Automatic Metrics with Incremental Machine Translation Systems
by: Wu, Guojun, et al.
Published: (2024)
by: Wu, Guojun, et al.
Published: (2024)
QueST: Persistent Queries as Semantic Monitors for Drift Suppression in Long-Horizon Tracking
by: Anand, Mayank, et al.
Published: (2026)
by: Anand, Mayank, et al.
Published: (2026)
MathScale: Scaling Instruction Tuning for Mathematical Reasoning
by: Tang, Zhengyang, et al.
Published: (2024)
by: Tang, Zhengyang, et al.
Published: (2024)
An Analysis of BPE Vocabulary Trimming in Neural Machine Translation
by: Cognetta, Marco, et al.
Published: (2024)
by: Cognetta, Marco, et al.
Published: (2024)
Bootstrap Your Own Context Length
by: Wang, Liang, et al.
Published: (2024)
by: Wang, Liang, et al.
Published: (2024)
Self-Boosting Large Language Models with Synthetic Preference Data
by: Dong, Qingxiu, et al.
Published: (2024)
by: Dong, Qingxiu, et al.
Published: (2024)
Conversational Lexicography: Querying Lexicographic Data on Knowledge Graphs with SPARQL through Natural Language
by: Sennrich, Kilian, et al.
Published: (2025)
by: Sennrich, Kilian, et al.
Published: (2025)
The Conception, Validation, and Reliability of the Questionnaire for Screen Time of Adolescents (QueST)
by: Knebel, Margarethe, et al.
Published: (2020)
by: Knebel, Margarethe, et al.
Published: (2020)
QueST: Self-Supervised Skill Abstractions for Learning Continuous Control
by: Mete, Atharva, et al.
Published: (2024)
by: Mete, Atharva, et al.
Published: (2024)
Generating Difficult-to-Translate Texts
by: Zouhar, Vilém, et al.
Published: (2025)
by: Zouhar, Vilém, et al.
Published: (2025)
Robust Native Language Identification through Agentic Decomposition
by: Uluslu, Ahmet Yavuz, et al.
Published: (2025)
by: Uluslu, Ahmet Yavuz, et al.
Published: (2025)
SAND-Math: Using LLMs to Generate Novel, Difficult and Useful Mathematics Questions and Answers
by: Manem, Chaitanya, et al.
Published: (2025)
by: Manem, Chaitanya, et al.
Published: (2025)
Expanding the WMT24++ Benchmark with Rumantsch Grischun, Sursilvan, Sutsilvan, Surmiran, Puter, and Vallader
by: Vamvas, Jannis, et al.
Published: (2025)
by: Vamvas, Jannis, et al.
Published: (2025)
Information Representation Fairness in Long-Document Embeddings: The Peculiar Interaction of Positional and Language Bias
by: Schuhmacher, Elias, et al.
Published: (2026)
by: Schuhmacher, Elias, et al.
Published: (2026)
Preference Optimization for Reasoning with Pseudo Feedback
by: Jiao, Fangkai, et al.
Published: (2024)
by: Jiao, Fangkai, et al.
Published: (2024)
CommonMorph: Participatory Morphological Documentation Platform
by: Mahmudi, Aso, et al.
Published: (2026)
by: Mahmudi, Aso, et al.
Published: (2026)
SignCLIP: Connecting Text and Sign Language by Contrastive Learning
by: Jiang, Zifan, et al.
Published: (2024)
by: Jiang, Zifan, et al.
Published: (2024)
BitNet a4.8: 4-bit Activations for 1-bit LLMs
by: Wang, Hongyu, et al.
Published: (2024)
by: Wang, Hongyu, et al.
Published: (2024)
BitNet v2: Native 4-bit Activations with Hadamard Transformation for 1-bit LLMs
by: Wang, Hongyu, et al.
Published: (2025)
by: Wang, Hongyu, et al.
Published: (2025)
BitNet b1.58 2B4T Technical Report
by: Ma, Shuming, et al.
Published: (2025)
by: Ma, Shuming, et al.
Published: (2025)
Similar Items
-
Source-primed Multi-turn Conversation Helps Large Language Models Translate Documents
by: Hu, Hanxu, et al.
Published: (2025) -
Linear-time Minimum Bayes Risk Decoding with Reference Aggregation
by: Vamvas, Jannis, et al.
Published: (2024) -
SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents
by: Wastl, Michelle, et al.
Published: (2025) -
Machine Translation Models are Zero-Shot Detectors of Translation Direction
by: Wastl, Michelle, et al.
Published: (2024) -
Mitigating Hallucinations and Off-target Machine Translation with Source-Contrastive and Language-Contrastive Decoding
by: Sennrich, Rico, et al.
Published: (2023)