Not-Just-Scaling Laws: Towards a Better Understanding of the Downstream Impact of Language Model Design Decisions
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Emmy, Bertsch, Amanda, Sutawika, Lintang, Tjuatja, Lindia, Fernandes, Patrick, Marinov, Lara, Chen, Michael, Singhal, Shreya, Lawrence, Carolin, Raghunathan, Aditi, Gashteovski, Kiril, Neubig, Graham |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
BehaviorBox: Automated Discovery of Fine-Grained Performance Differences Between Language Models
di: Tjuatja, Lindia, et al.
Pubblicazione: (2025)
di: Tjuatja, Lindia, et al.
Pubblicazione: (2025)
What Goes Into a LM Acceptability Judgment? Rethinking the Impact of Frequency and Length
di: Tjuatja, Lindia, et al.
Pubblicazione: (2024)
di: Tjuatja, Lindia, et al.
Pubblicazione: (2024)
Gained in Translation: Privileged Pairwise Judges Enhance Multilingual Reasoning
di: Sutawika, Lintang, et al.
Pubblicazione: (2026)
di: Sutawika, Lintang, et al.
Pubblicazione: (2026)
What do Language Models Learn and When? The Implicit Curriculum Hypothesis
di: Liu, Emmy, et al.
Pubblicazione: (2026)
di: Liu, Emmy, et al.
Pubblicazione: (2026)
Do LLMs exhibit human-like response biases? A case study in survey design
di: Tjuatja, Lindia, et al.
Pubblicazione: (2023)
di: Tjuatja, Lindia, et al.
Pubblicazione: (2023)
CMULAB: An Open-Source Framework for Training and Deployment of Natural Language Processing Models
di: Sheikh, Zaid, et al.
Pubblicazione: (2024)
di: Sheikh, Zaid, et al.
Pubblicazione: (2024)
Better Instruction-Following Through Minimum Bayes Risk
di: Wu, Ian, et al.
Pubblicazione: (2024)
di: Wu, Ian, et al.
Pubblicazione: (2024)
Compositional Steering of Large Language Models with Steering Tokens
di: Radevski, Gorjan, et al.
Pubblicazione: (2026)
di: Radevski, Gorjan, et al.
Pubblicazione: (2026)
GlossLM: A Massively Multilingual Corpus and Pretrained Model for Interlinear Glossed Text
di: Ginn, Michael, et al.
Pubblicazione: (2024)
di: Ginn, Michael, et al.
Pubblicazione: (2024)
Evaluating Language Models as Synthetic Data Generators
di: Kim, Seungone, et al.
Pubblicazione: (2024)
di: Kim, Seungone, et al.
Pubblicazione: (2024)
Massively Multilingual Joint Segmentation and Glossing
di: Ginn, Michael, et al.
Pubblicazione: (2026)
di: Ginn, Michael, et al.
Pubblicazione: (2026)
Scaling Evaluation-time Compute with Reasoning Models as Evaluators
di: Kim, Seungone, et al.
Pubblicazione: (2025)
di: Kim, Seungone, et al.
Pubblicazione: (2025)
MEDDxAgent: A Unified Modular Agent Framework for Explainable Automatic Differential Diagnosis
di: Rose, Daniel, et al.
Pubblicazione: (2025)
di: Rose, Daniel, et al.
Pubblicazione: (2025)
AgentQuest: A Modular Benchmark Framework to Measure Progress and Improve LLM Agents
di: Gioacchini, Luca, et al.
Pubblicazione: (2024)
di: Gioacchini, Luca, et al.
Pubblicazione: (2024)
An Incomplete Loop: Instruction Inference, Instruction Following, and In-context Learning in Language Models
di: Liu, Emmy, et al.
Pubblicazione: (2024)
di: Liu, Emmy, et al.
Pubblicazione: (2024)
Midtraining Bridges Pretraining and Posttraining Distributions
di: Liu, Emmy, et al.
Pubblicazione: (2025)
di: Liu, Emmy, et al.
Pubblicazione: (2025)
Wav2Gloss: Generating Interlinear Glossed Text from Speech
di: He, Taiqi, et al.
Pubblicazione: (2024)
di: He, Taiqi, et al.
Pubblicazione: (2024)
Repetition Improves Language Model Embeddings
di: Springer, Jacob Mitchell, et al.
Pubblicazione: (2024)
di: Springer, Jacob Mitchell, et al.
Pubblicazione: (2024)
Pangea: A Fully Open Multilingual Multimodal LLM for 39 Languages
di: Yue, Xiang, et al.
Pubblicazione: (2024)
di: Yue, Xiang, et al.
Pubblicazione: (2024)
Oolong: Evaluating Long Context Reasoning and Aggregation Capabilities
di: Bertsch, Amanda, et al.
Pubblicazione: (2025)
di: Bertsch, Amanda, et al.
Pubblicazione: (2025)
Efficient Many-Shot In-Context Learning with Dynamic Block-Sparse Attention
di: Xiao, Emily, et al.
Pubblicazione: (2025)
di: Xiao, Emily, et al.
Pubblicazione: (2025)
Multitask Learning Can Improve Worst-Group Outcomes
di: Kulkarni, Atharva, et al.
Pubblicazione: (2023)
di: Kulkarni, Atharva, et al.
Pubblicazione: (2023)
Leveraging Open Information Extraction for More Robust Domain Transfer of Event Trigger Detection
di: Dukić, David, et al.
Pubblicazione: (2023)
di: Dukić, David, et al.
Pubblicazione: (2023)
Synthetic Socratic Debates: Examining Persona Effects on Moral Decision and Persuasion Dynamics
di: Liu, Jiarui, et al.
Pubblicazione: (2025)
di: Liu, Jiarui, et al.
Pubblicazione: (2025)
Prompt-MII: Meta-Learning Instruction Induction for LLMs
di: Xiao, Emily, et al.
Pubblicazione: (2025)
di: Xiao, Emily, et al.
Pubblicazione: (2025)
CodeScout: An Effective Recipe for Reinforcement Learning of Code Search Agents
di: Sutawika, Lintang, et al.
Pubblicazione: (2026)
di: Sutawika, Lintang, et al.
Pubblicazione: (2026)
TextMineX: Data, Evaluation Framework and Ontology-guided LLM Pipeline for Humanitarian Mine Action
di: Zhou, Chenyue, et al.
Pubblicazione: (2025)
di: Zhou, Chenyue, et al.
Pubblicazione: (2025)
LightPAL: Lightweight Passage Retrieval for Open Domain Multi-Document Summarization
di: Enomoto, Masafumi, et al.
Pubblicazione: (2024)
di: Enomoto, Masafumi, et al.
Pubblicazione: (2024)
Robust Text Classification: Analyzing Prototype-Based Networks
di: Sourati, Zhivar, et al.
Pubblicazione: (2023)
di: Sourati, Zhivar, et al.
Pubblicazione: (2023)
In-Context Learning with Long-Context Models: An In-Depth Exploration
di: Bertsch, Amanda, et al.
Pubblicazione: (2024)
di: Bertsch, Amanda, et al.
Pubblicazione: (2024)
Divergences between Language Models and Human Brains
di: Zhou, Yuchen, et al.
Pubblicazione: (2023)
di: Zhou, Yuchen, et al.
Pubblicazione: (2023)
From Decoding to Meta-Generation: Inference-time Algorithms for Large Language Models
di: Welleck, Sean, et al.
Pubblicazione: (2024)
di: Welleck, Sean, et al.
Pubblicazione: (2024)
Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs
di: Zhong, Ziqian, et al.
Pubblicazione: (2025)
di: Zhong, Ziqian, et al.
Pubblicazione: (2025)
Overtrained Language Models Are Harder to Fine-Tune
di: Springer, Jacob Mitchell, et al.
Pubblicazione: (2025)
di: Springer, Jacob Mitchell, et al.
Pubblicazione: (2025)
Better Synthetic Data by Retrieving and Transforming Existing Datasets
di: Gandhi, Saumya, et al.
Pubblicazione: (2024)
di: Gandhi, Saumya, et al.
Pubblicazione: (2024)
Language Modeling with Editable External Knowledge
di: Li, Belinda Z., et al.
Pubblicazione: (2024)
di: Li, Belinda Z., et al.
Pubblicazione: (2024)
Self-Trained Verification for Training- and Test-Time Self-Improvement
di: Wu, Chen Henry, et al.
Pubblicazione: (2026)
di: Wu, Chen Henry, et al.
Pubblicazione: (2026)
Penetrating School Strata through Career Education. Program Evaluation.
di: Lindia, Albert, et al.
Pubblicazione: (1976)
di: Lindia, Albert, et al.
Pubblicazione: (1976)
CO 2 Electroreduction to CO Using Cu‐Supported NiO Catalyst: XPS Evidence of Redox Interaction Between Metal and Support
di: Akanksha Sharma, et al.
Pubblicazione: (2025)
di: Akanksha Sharma, et al.
Pubblicazione: (2025)
SELF-GUIDE: Better Task-Specific Instruction Following via Self-Synthetic Finetuning
di: Zhao, Chenyang, et al.
Pubblicazione: (2024)
di: Zhao, Chenyang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
BehaviorBox: Automated Discovery of Fine-Grained Performance Differences Between Language Models
di: Tjuatja, Lindia, et al.
Pubblicazione: (2025) -
What Goes Into a LM Acceptability Judgment? Rethinking the Impact of Frequency and Length
di: Tjuatja, Lindia, et al.
Pubblicazione: (2024) -
Gained in Translation: Privileged Pairwise Judges Enhance Multilingual Reasoning
di: Sutawika, Lintang, et al.
Pubblicazione: (2026) -
What do Language Models Learn and When? The Implicit Curriculum Hypothesis
di: Liu, Emmy, et al.
Pubblicazione: (2026) -
Do LLMs exhibit human-like response biases? A case study in survey design
di: Tjuatja, Lindia, et al.
Pubblicazione: (2023)