Increasing the Difficulty of Automatically Generated Questions via Reinforcement Learning with Synthetic Preference
Fuente:
arXiv
Saved in:
| Main Authors: | Thorne, William, Robinson, Ambrose, Peng, Bohua, Lin, Chenghua, Maynard, Diana |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Large Language Models are Inconsistent and Biased Evaluators
by: Stureborg, Rickard, et al.
Published: (2024)
by: Stureborg, Rickard, et al.
Published: (2024)
Extracting Abstraction Dimensions by Identifying Syntax Pattern from Texts
by: Zhou, Jian, et al.
Published: (2025)
by: Zhou, Jian, et al.
Published: (2025)
Smotrom tvoja pa ander drogoj verden! Resurrecting Dead Pidgin with Generative Models: Russenorsk Case Study
by: Tikhonov, Alexey, et al.
Published: (2025)
by: Tikhonov, Alexey, et al.
Published: (2025)
Tailoring Vaccine Messaging with Common-Ground Opinions
by: Stureborg, Rickard, et al.
Published: (2024)
by: Stureborg, Rickard, et al.
Published: (2024)
TRAWL: Tensor Reduced and Approximated Weights for Large Language Models
by: Luo, Yiran, et al.
Published: (2024)
by: Luo, Yiran, et al.
Published: (2024)
MORQA: Benchmarking Evaluation Metrics for Medical Open-Ended Question Answering
by: Yim, Wen-wai, et al.
Published: (2025)
by: Yim, Wen-wai, et al.
Published: (2025)
Unveiling factors influencing judgment variation in Sentiment Analysis with Natural Language Processing and Statistics
by: Kellert, Olga, et al.
Published: (2024)
by: Kellert, Olga, et al.
Published: (2024)
AI in Investment Analysis: LLMs for Equity Stock Ratings
by: Papasotiriou, Kassiani, et al.
Published: (2024)
by: Papasotiriou, Kassiani, et al.
Published: (2024)
Uncovering Uncertainty in Transformer Inference
by: Brothers, Greyson, et al.
Published: (2024)
by: Brothers, Greyson, et al.
Published: (2024)
CogniLoad: A Synthetic Natural Language Reasoning Benchmark With Tunable Length, Intrinsic Difficulty, and Distractor Density
by: Kaiser, Daniel, et al.
Published: (2025)
by: Kaiser, Daniel, et al.
Published: (2025)
TableMoE: Neuro-Symbolic Routing for Structured Expert Reasoning in Multimodal Table Understanding
by: Zhang, Junwen, et al.
Published: (2025)
by: Zhang, Junwen, et al.
Published: (2025)
SIGMA: Scalable Spectral Insights for LLM Model Collapse
by: Gu, Yi, et al.
Published: (2026)
by: Gu, Yi, et al.
Published: (2026)
Atyaephyra at SemEval-2025 Task 4: Low-Rank Negative Preference Optimization
by: Bronec, Jan, et al.
Published: (2025)
by: Bronec, Jan, et al.
Published: (2025)
A Survey on Collaborating Small and Large Language Models for Performance, Cost-effectiveness, Cloud-edge Privacy, and Trustworthiness
by: Wang, Fali, et al.
Published: (2025)
by: Wang, Fali, et al.
Published: (2025)
Quantum NLP models on Natural Language Inference
by: Sun, Ling, et al.
Published: (2025)
by: Sun, Ling, et al.
Published: (2025)
Make Literature-Based Discovery Great Again through Reproducible Pipelines
by: Cestnik, Bojan, et al.
Published: (2025)
by: Cestnik, Bojan, et al.
Published: (2025)
Model Surgery: Modulating LLM's Behavior Via Simple Parameter Editing
by: Wang, Huanqian, et al.
Published: (2024)
by: Wang, Huanqian, et al.
Published: (2024)
Data and AI governance: Promoting equity, ethics, and fairness in large language models
by: Abhishek, Alok, et al.
Published: (2025)
by: Abhishek, Alok, et al.
Published: (2025)
Do Reasoning Models Enhance Embedding Models?
by: Chan, Wun Yu, et al.
Published: (2026)
by: Chan, Wun Yu, et al.
Published: (2026)
SHARP: Social Harm Analysis via Risk Profiles for Measuring Inequities in Large Language Models
by: Abhishek, Alok, et al.
Published: (2026)
by: Abhishek, Alok, et al.
Published: (2026)
Judgment2vec: Apply Graph Analytics to Searching and Recommendation of Similar Judgments
by: Shao, Hsuan-Lei
Published: (2024)
by: Shao, Hsuan-Lei
Published: (2024)
BEATS: Bias Evaluation and Assessment Test Suite for Large Language Models
by: Abhishek, Alok, et al.
Published: (2025)
by: Abhishek, Alok, et al.
Published: (2025)
Research on a hybrid LSTM-CNN-Attention model for text-based web content classification
by: Kuz, Mykola, et al.
Published: (2025)
by: Kuz, Mykola, et al.
Published: (2025)
Reasoning Promotes Robustness in Theory of Mind Tasks
by: de Haan, Ian B., et al.
Published: (2026)
by: de Haan, Ian B., et al.
Published: (2026)
Linguistic Collapse: Neural Collapse in (Large) Language Models
by: Wu, Robert, et al.
Published: (2024)
by: Wu, Robert, et al.
Published: (2024)
Rule Extraction in Machine Learning: Chat Incremental Pattern Constructor
by: Nwokocha, Caleb Princewill
Published: (2022)
by: Nwokocha, Caleb Princewill
Published: (2022)
The Association of Transformer-based Sentiment Analysis with Symptom Distress and Deterioration in Routine Psychotherapy Care
by: Faust, Douglas K., et al.
Published: (2026)
by: Faust, Douglas K., et al.
Published: (2026)
Making Every Head Count: Sparse Attention Without the Speed-Performance Trade-off
by: Zhao, Mingkuan, et al.
Published: (2025)
by: Zhao, Mingkuan, et al.
Published: (2025)
A Comparative Analysis of Noise Reduction Methods in Sentiment Analysis on Noisy Bangla Texts
by: Elahi, Kazi Toufique, et al.
Published: (2024)
by: Elahi, Kazi Toufique, et al.
Published: (2024)
Correctness is not Faithfulness in RAG Attributions
by: Wallat, Jonas, et al.
Published: (2024)
by: Wallat, Jonas, et al.
Published: (2024)
A transfer learning approach for automatic conflicts detection in software requirement sentence pairs based on dual encoders
by: Wang, Yizheng, et al.
Published: (2025)
by: Wang, Yizheng, et al.
Published: (2025)
Fine-Tuning Transformers: Vocabulary Transfer
by: Mosin, Vladislav, et al.
Published: (2021)
by: Mosin, Vladislav, et al.
Published: (2021)
A Comprehensive Survey of Small Language Models in the Era of Large Language Models: Techniques, Enhancements, Applications, Collaboration with LLMs, and Trustworthiness
by: Wang, Fali, et al.
Published: (2024)
by: Wang, Fali, et al.
Published: (2024)
Drama Engine: A Framework for Narrative Agents
by: Pichlmair, Martin, et al.
Published: (2024)
by: Pichlmair, Martin, et al.
Published: (2024)
CLEV: LLM-Based Evaluation Through Lightweight Efficient Voting for Free-Form Question-Answering
by: Badshah, Sher, et al.
Published: (2025)
by: Badshah, Sher, et al.
Published: (2025)
FANAL -- Financial Activity News Alerting Language Modeling Framework
by: Patel, Urjitkumar, et al.
Published: (2024)
by: Patel, Urjitkumar, et al.
Published: (2024)
Moral Semantics Survive Machine Translation: Cross-Lingual Evidence from Moral Foundations Corpora
by: Skorski, Maciej
Published: (2026)
by: Skorski, Maciej
Published: (2026)
GraphSkill: Documentation-Guided Hierarchical Retrieval-Augmented Coding for Complex Graph Reasoning
by: Wang, Fali, et al.
Published: (2026)
by: Wang, Fali, et al.
Published: (2026)
Machine Learning for Sentiment Analysis of Imported Food in Trinidad and Tobago
by: Daniels, Cassandra, et al.
Published: (2024)
by: Daniels, Cassandra, et al.
Published: (2024)
Recent Advances and Future Directions in Literature-Based Discovery
by: Kastrin, Andrej, et al.
Published: (2025)
by: Kastrin, Andrej, et al.
Published: (2025)
Similar Items
-
Large Language Models are Inconsistent and Biased Evaluators
by: Stureborg, Rickard, et al.
Published: (2024) -
Extracting Abstraction Dimensions by Identifying Syntax Pattern from Texts
by: Zhou, Jian, et al.
Published: (2025) -
Smotrom tvoja pa ander drogoj verden! Resurrecting Dead Pidgin with Generative Models: Russenorsk Case Study
by: Tikhonov, Alexey, et al.
Published: (2025) -
Tailoring Vaccine Messaging with Common-Ground Opinions
by: Stureborg, Rickard, et al.
Published: (2024) -
TRAWL: Tensor Reduced and Approximated Weights for Large Language Models
by: Luo, Yiran, et al.
Published: (2024)