An Overview of Large Language Models for Statisticians
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ji, Wenlong, Yuan, Weizhe, Getzen, Emily, Cho, Kyunghyun, Jordan, Michael I., Mei, Song, Weston, Jason E, Su, Weijie J., Xu, Jing, Zhang, Linjun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
System-Level Natural Language Feedback
von: Yuan, Weizhe, et al.
Veröffentlicht: (2023)
von: Yuan, Weizhe, et al.
Veröffentlicht: (2023)
Self-Rewarding Language Models
von: Yuan, Weizhe, et al.
Veröffentlicht: (2024)
von: Yuan, Weizhe, et al.
Veröffentlicht: (2024)
Following Length Constraints in Instructions
von: Yuan, Weizhe, et al.
Veröffentlicht: (2024)
von: Yuan, Weizhe, et al.
Veröffentlicht: (2024)
Iterative Reasoning Preference Optimization
von: Pang, Richard Yuanzhe, et al.
Veröffentlicht: (2024)
von: Pang, Richard Yuanzhe, et al.
Veröffentlicht: (2024)
Are Large Language Models Good Statisticians?
von: Zhu, Yizhang, et al.
Veröffentlicht: (2024)
von: Zhu, Yizhang, et al.
Veröffentlicht: (2024)
Leveraging Implicit Feedback from Deployment Data in Dialogue
von: Pang, Richard Yuanzhe, et al.
Veröffentlicht: (2023)
von: Pang, Richard Yuanzhe, et al.
Veröffentlicht: (2023)
NaturalReasoning: Reasoning in the Wild with 2.8M Challenging Questions
von: Yuan, Weizhe, et al.
Veröffentlicht: (2025)
von: Yuan, Weizhe, et al.
Veröffentlicht: (2025)
Are PPO-ed Language Models Hackable?
von: Anand, Suraj, et al.
Veröffentlicht: (2024)
von: Anand, Suraj, et al.
Veröffentlicht: (2024)
Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge
von: Wu, Tianhao, et al.
Veröffentlicht: (2024)
von: Wu, Tianhao, et al.
Veröffentlicht: (2024)
R.I.P.: Better Models by Survival of the Fittest Prompts
von: Yu, Ping, et al.
Veröffentlicht: (2025)
von: Yu, Ping, et al.
Veröffentlicht: (2025)
Efficient semantic uncertainty quantification in language models via diversity-steered sampling
von: Park, Ji Won, et al.
Veröffentlicht: (2025)
von: Park, Ji Won, et al.
Veröffentlicht: (2025)
First Tragedy, then Parse: History Repeats Itself in the New Era of Large Language Models
von: Saphra, Naomi, et al.
Veröffentlicht: (2023)
von: Saphra, Naomi, et al.
Veröffentlicht: (2023)
Thinking LLMs: General Instruction Following with Thought Generation
von: Wu, Tianhao, et al.
Veröffentlicht: (2024)
von: Wu, Tianhao, et al.
Veröffentlicht: (2024)
Reward Collapse in Aligning Large Language Models
von: Song, Ziang, et al.
Veröffentlicht: (2023)
von: Song, Ziang, et al.
Veröffentlicht: (2023)
Code-Switching In-Context Learning for Cross-Lingual Transfer of Large Language Models
von: Yoo, Haneul, et al.
Veröffentlicht: (2025)
von: Yoo, Haneul, et al.
Veröffentlicht: (2025)
LLMCRIT: Teaching Large Language Models to Use Criteria
von: Yuan, Weizhe, et al.
Veröffentlicht: (2024)
von: Yuan, Weizhe, et al.
Veröffentlicht: (2024)
CoT-Self-Instruct: Building high-quality synthetic prompts for reasoning and non-reasoning tasks
von: Yu, Ping, et al.
Veröffentlicht: (2025)
von: Yu, Ping, et al.
Veröffentlicht: (2025)
RESTRAIN: From Spurious Votes to Signals -- Self-Driven RL with Self-Penalization
von: Yu, Zhaoning, et al.
Veröffentlicht: (2025)
von: Yu, Zhaoning, et al.
Veröffentlicht: (2025)
StepWiser: Stepwise Generative Judges for Wiser Reasoning
von: Xiong, Wei, et al.
Veröffentlicht: (2025)
von: Xiong, Wei, et al.
Veröffentlicht: (2025)
Language Models as Causal Effect Generators
von: Bynum, Lucius E. J., et al.
Veröffentlicht: (2024)
von: Bynum, Lucius E. J., et al.
Veröffentlicht: (2024)
Training Large Language Models to Reason in a Continuous Latent Space
von: Hao, Shibo, et al.
Veröffentlicht: (2024)
von: Hao, Shibo, et al.
Veröffentlicht: (2024)
Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss
von: Xu, Jing, et al.
Veröffentlicht: (2023)
von: Xu, Jing, et al.
Veröffentlicht: (2023)
Distilling System 2 into System 1
von: Yu, Ping, et al.
Veröffentlicht: (2024)
von: Yu, Ping, et al.
Veröffentlicht: (2024)
On the Algorithmic Bias of Aligning Large Language Models with RLHF: Preference Collapse and Matching Regularization
von: Xiao, Jiancong, et al.
Veröffentlicht: (2024)
von: Xiao, Jiancong, et al.
Veröffentlicht: (2024)
Scaling Laws Are Unreliable for Downstream Tasks: A Reality Check
von: Lourie, Nicholas, et al.
Veröffentlicht: (2025)
von: Lourie, Nicholas, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Defense: Adaptive and Controllable Jailbreak Prevention for Large Language Models
von: Yang, Guangyu, et al.
Veröffentlicht: (2025)
von: Yang, Guangyu, et al.
Veröffentlicht: (2025)
Self-Consistency Preference Optimization
von: Prasad, Archiki, et al.
Veröffentlicht: (2024)
von: Prasad, Archiki, et al.
Veröffentlicht: (2024)
Generalization Measures for Zero-Shot Cross-Lingual Transfer
von: Bassi, Saksham, et al.
Veröffentlicht: (2024)
von: Bassi, Saksham, et al.
Veröffentlicht: (2024)
A Law of Next-Token Prediction in Large Language Models
von: He, Hangfeng, et al.
Veröffentlicht: (2024)
von: He, Hangfeng, et al.
Veröffentlicht: (2024)
The Geometry of Prompting: Unveiling Distinct Mechanisms of Task Adaptation in Language Models
von: Kirsanov, Artem, et al.
Veröffentlicht: (2025)
von: Kirsanov, Artem, et al.
Veröffentlicht: (2025)
Shared Heritage, Distinct Writing: Rethinking Resource Selection for East Asian Historical Documents
von: Song, Seyoung, et al.
Veröffentlicht: (2024)
von: Song, Seyoung, et al.
Veröffentlicht: (2024)
HERITAGE: An End-to-End Web Platform for Processing Korean Historical Documents in Hanja
von: Song, Seyoung, et al.
Veröffentlicht: (2025)
von: Song, Seyoung, et al.
Veröffentlicht: (2025)
Show Your Work with Confidence: Confidence Bands for Tuning Curves
von: Lourie, Nicholas, et al.
Veröffentlicht: (2023)
von: Lourie, Nicholas, et al.
Veröffentlicht: (2023)
Hyperparameter Loss Surfaces Are Simple Near their Optima
von: Lourie, Nicholas, et al.
Veröffentlicht: (2025)
von: Lourie, Nicholas, et al.
Veröffentlicht: (2025)
Bridging Offline and Online Reinforcement Learning for LLMs
von: Lanchantin, Jack, et al.
Veröffentlicht: (2025)
von: Lanchantin, Jack, et al.
Veröffentlicht: (2025)
Logic-of-Thought: Injecting Logic into Contexts for Full Reasoning in Large Language Models
von: Liu, Tongxuan, et al.
Veröffentlicht: (2024)
von: Liu, Tongxuan, et al.
Veröffentlicht: (2024)
SPICE: Self-Play In Corpus Environments Improves Reasoning
von: Liu, Bo, et al.
Veröffentlicht: (2025)
von: Liu, Bo, et al.
Veröffentlicht: (2025)
Branch-Solve-Merge Improves Large Language Model Evaluation and Generation
von: Saha, Swarnadeep, et al.
Veröffentlicht: (2023)
von: Saha, Swarnadeep, et al.
Veröffentlicht: (2023)
Aioli: A Unified Optimization Framework for Language Model Data Mixing
von: Chen, Mayee F., et al.
Veröffentlicht: (2024)
von: Chen, Mayee F., et al.
Veröffentlicht: (2024)
On the Relationship Between the Choice of Representation and In-Context Learning
von: Marinescu, Ioana, et al.
Veröffentlicht: (2025)
von: Marinescu, Ioana, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
System-Level Natural Language Feedback
von: Yuan, Weizhe, et al.
Veröffentlicht: (2023) -
Self-Rewarding Language Models
von: Yuan, Weizhe, et al.
Veröffentlicht: (2024) -
Following Length Constraints in Instructions
von: Yuan, Weizhe, et al.
Veröffentlicht: (2024) -
Iterative Reasoning Preference Optimization
von: Pang, Richard Yuanzhe, et al.
Veröffentlicht: (2024) -
Are Large Language Models Good Statisticians?
von: Zhu, Yizhang, et al.
Veröffentlicht: (2024)