Saved in:
| Main Authors: | Johnson, Natasha, Bertsch, Amanda, Deal, Maria-Emil, Strubell, Emma |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2510.20926 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
To Build Our Future, We Must Know Our Past: Contextualizing Paradigm Shifts in Natural Language Processing
by: Gururaja, Sireesh, et al.
Published: (2023)
by: Gururaja, Sireesh, et al.
Published: (2023)
Computing the Formal and Institutional Boundaries of Contemporary Genre and Literary Fiction
by: Johnson, Natasha
Published: (2025)
by: Johnson, Natasha
Published: (2025)
Gradient Localization Improves Lifelong Pretraining of Language Models
by: Fernandez, Jared, et al.
Published: (2024)
by: Fernandez, Jared, et al.
Published: (2024)
A Taxonomy for Data Contamination in Large Language Models
by: Palavalli, Medha, et al.
Published: (2024)
by: Palavalli, Medha, et al.
Published: (2024)
Beyond Text: Characterizing Domain Expert Needs in Document Research
by: Gururaja, Sireesh, et al.
Published: (2025)
by: Gururaja, Sireesh, et al.
Published: (2025)
Oolong: Evaluating Long Context Reasoning and Aggregation Capabilities
by: Bertsch, Amanda, et al.
Published: (2025)
by: Bertsch, Amanda, et al.
Published: (2025)
In-Context Learning with Long-Context Models: An In-Depth Exploration
by: Bertsch, Amanda, et al.
Published: (2024)
by: Bertsch, Amanda, et al.
Published: (2024)
Stereotype or Personalization? User Identity Biases Chatbot Recommendations
by: Kantharuban, Anjali, et al.
Published: (2024)
by: Kantharuban, Anjali, et al.
Published: (2024)
FaNS: a Facet-based Narrative Similarity Metric
by: Akter, Mousumi, et al.
Published: (2023)
by: Akter, Mousumi, et al.
Published: (2023)
Energy and Carbon Considerations of Fine-Tuning BERT
by: Wang, Xiaorong, et al.
Published: (2023)
by: Wang, Xiaorong, et al.
Published: (2023)
Efficient Many-Shot In-Context Learning with Dynamic Block-Sparse Attention
by: Xiao, Emily, et al.
Published: (2025)
by: Xiao, Emily, et al.
Published: (2025)
Energy Considerations of Large Language Model Inference and Efficiency Optimizations
by: Fernandez, Jared, et al.
Published: (2025)
by: Fernandez, Jared, et al.
Published: (2025)
Kinetics: Rethinking Test-Time Scaling Laws
by: Sadhukhan, Ranajoy, et al.
Published: (2025)
by: Sadhukhan, Ranajoy, et al.
Published: (2025)
Just CHOP: Embarrassingly Simple LLM Compression
by: Jha, Ananya Harsh, et al.
Published: (2023)
by: Jha, Ananya Harsh, et al.
Published: (2023)
Multi-Facet Blending for Faceted Query-by-Example Retrieval
by: Do, Heejin, et al.
Published: (2024)
by: Do, Heejin, et al.
Published: (2024)
LCFO: Long Context and Long Form Output Dataset and Benchmarking
by: Costa-jussà, Marta R., et al.
Published: (2024)
by: Costa-jussà, Marta R., et al.
Published: (2024)
AlgoSimBench: Identifying Algorithmically Similar Problems for Competitive Programming
by: Li, Jierui, et al.
Published: (2025)
by: Li, Jierui, et al.
Published: (2025)
Prompt-MII: Meta-Learning Instruction Induction for LLMs
by: Xiao, Emily, et al.
Published: (2025)
by: Xiao, Emily, et al.
Published: (2025)
AboutMe: Using Self-Descriptions in Webpages to Document the Effects of English Pretraining Data Filters
by: Lucy, Li, et al.
Published: (2024)
by: Lucy, Li, et al.
Published: (2024)
Better Instruction-Following Through Minimum Bayes Risk
by: Wu, Ian, et al.
Published: (2024)
by: Wu, Ian, et al.
Published: (2024)
FictionalQA: A Dataset for Studying Memorization and Knowledge Acquisition
by: Kirchenbauer, John, et al.
Published: (2025)
by: Kirchenbauer, John, et al.
Published: (2025)
LFED: A Literary Fiction Evaluation Dataset for Large Language Models
by: Yu, Linhao, et al.
Published: (2024)
by: Yu, Linhao, et al.
Published: (2024)
Source-Aware Training Enables Knowledge Attribution in Language Models
by: Khalifa, Muhammad, et al.
Published: (2024)
by: Khalifa, Muhammad, et al.
Published: (2024)
SimSMoE: Solving Representational Collapse via Similarity Measure
by: Do, Giang, et al.
Published: (2024)
by: Do, Giang, et al.
Published: (2024)
Multi-Faceted Evaluation of Tool-Augmented Dialogue Systems
by: Hou, Zhaoyi Joey, et al.
Published: (2025)
by: Hou, Zhaoyi Joey, et al.
Published: (2025)
Multi-Facet Counterfactual Learning for Content Quality Evaluation
by: Zheng, Jiasheng, et al.
Published: (2024)
by: Zheng, Jiasheng, et al.
Published: (2024)
Scalable Data Ablation Approximations for Language Models through Modular Training and Merging
by: Na, Clara, et al.
Published: (2024)
by: Na, Clara, et al.
Published: (2024)
SimReg: Achieving Higher Performance in the Pretraining via Embedding Similarity Regularization
by: Sun, Yan, et al.
Published: (2026)
by: Sun, Yan, et al.
Published: (2026)
mPLM-Sim: Better Cross-Lingual Similarity and Transfer in Multilingual Pretrained Language Models
by: Lin, Peiqin, et al.
Published: (2023)
by: Lin, Peiqin, et al.
Published: (2023)
A Multilingual Similarity Dataset for News Article Frame
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
Rethinking Word Similarity: Semantic Similarity through Classification Confusion
by: Zhou, Kaitlyn, et al.
Published: (2025)
by: Zhou, Kaitlyn, et al.
Published: (2025)
PeeriScope: A Multi-Faceted Framework for Evaluating Peer Review Quality
by: Ebrahimi, Sajad, et al.
Published: (2026)
by: Ebrahimi, Sajad, et al.
Published: (2026)
PerSphere: A Comprehensive Framework for Multi-Faceted Perspective Retrieval and Summarization
by: Luo, Yun, et al.
Published: (2024)
by: Luo, Yun, et al.
Published: (2024)
Modelling Commonsense Commonalities with Multi-Facet Concept Embeddings
by: Kteich, Hanane, et al.
Published: (2024)
by: Kteich, Hanane, et al.
Published: (2024)
Think Like a Person Before Responding: A Multi-Faceted Evaluation of Persona-Guided LLMs for Countering Hate
by: Ngueajio, Mikel K., et al.
Published: (2025)
by: Ngueajio, Mikel K., et al.
Published: (2025)
Mapping the Political Discourse in the Brazilian Chamber of Deputies: A Multi-Faceted Computational Approach
by: Soriano, Flávio, et al.
Published: (2026)
by: Soriano, Flávio, et al.
Published: (2026)
Linguistically Conditioned Semantic Textual Similarity
by: Tu, Jingxuan, et al.
Published: (2024)
by: Tu, Jingxuan, et al.
Published: (2024)
Learning to Align Multi-Faceted Evaluation: A Unified and Robust Framework
by: Xu, Kaishuai, et al.
Published: (2025)
by: Xu, Kaishuai, et al.
Published: (2025)
Collage: Decomposable Rapid Prototyping for Information Extraction on Scientific PDFs
by: Gururaja, Sireesh, et al.
Published: (2024)
by: Gururaja, Sireesh, et al.
Published: (2024)
Semantic Similarity in Radiology Reports via LLMs and NER
by: Pearson, Beth, et al.
Published: (2025)
by: Pearson, Beth, et al.
Published: (2025)
Similar Items
-
To Build Our Future, We Must Know Our Past: Contextualizing Paradigm Shifts in Natural Language Processing
by: Gururaja, Sireesh, et al.
Published: (2023) -
Computing the Formal and Institutional Boundaries of Contemporary Genre and Literary Fiction
by: Johnson, Natasha
Published: (2025) -
Gradient Localization Improves Lifelong Pretraining of Language Models
by: Fernandez, Jared, et al.
Published: (2024) -
A Taxonomy for Data Contamination in Large Language Models
by: Palavalli, Medha, et al.
Published: (2024) -
Beyond Text: Characterizing Domain Expert Needs in Document Research
by: Gururaja, Sireesh, et al.
Published: (2025)