Do Generalisation Results Generalise?
Fuente:
arXiv
Saved in:
| Main Authors: | Boglioni, Matteo, Sgobbi, Andrea, Tavernini, Gabriel, Rita, Francesco, Mosbach, Marius, Pimentel, Tiago |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Role of Language Imbalance in Cross-lingual Generalisation: Insights from Cloned Language Experiments
by: Schäfer, Anton, et al.
Published: (2024)
by: Schäfer, Anton, et al.
Published: (2024)
Towards Generalising Neural Topical Representations
by: Yang, Xiaohao, et al.
Published: (2023)
by: Yang, Xiaohao, et al.
Published: (2023)
The Illusion of Superposition? A Principled Analysis of Latent Thinking in Language Models
by: Rizvi-Martel, Michael, et al.
Published: (2026)
by: Rizvi-Martel, Michael, et al.
Published: (2026)
Forecasting Downstream Performance of LLMs With Proxy Metrics
by: Patel, Arkil, et al.
Published: (2026)
by: Patel, Arkil, et al.
Published: (2026)
TRA: Better Length Generalisation with Threshold Relative Attention
by: Opper, Mattia, et al.
Published: (2025)
by: Opper, Mattia, et al.
Published: (2025)
Build the web for agents, not agents for the web
by: Lù, Xing Han, et al.
Published: (2025)
by: Lù, Xing Han, et al.
Published: (2025)
Rethinking the Evaluation of Alignment Methods: Insights into Diversity, Generalisation, and Safety
by: Janiak, Denis, et al.
Published: (2025)
by: Janiak, Denis, et al.
Published: (2025)
A Symbolic Framework for Evaluating Mathematical Reasoning and Generalisation with Transformers
by: Meadows, Jordan, et al.
Published: (2023)
by: Meadows, Jordan, et al.
Published: (2023)
Understanding the Effects of RLHF on LLM Generalisation and Diversity
by: Kirk, Robert, et al.
Published: (2023)
by: Kirk, Robert, et al.
Published: (2023)
Information Structure in Mappings: An Approach to Learning, Representation, and Generalisation
by: Conklin, Henry
Published: (2025)
by: Conklin, Henry
Published: (2025)
Convergence and Divergence of Language Models under Different Random Seeds
by: Fehlauer, Finlay, et al.
Published: (2025)
by: Fehlauer, Finlay, et al.
Published: (2025)
Value Drifts: Tracing Value Alignment During LLM Post-Training
by: Bhatia, Mehar, et al.
Published: (2025)
by: Bhatia, Mehar, et al.
Published: (2025)
A (More) Realistic Evaluation Setup for Generalisation of Community Models on Malicious Content Detection
by: Verhoeven, Ivo, et al.
Published: (2024)
by: Verhoeven, Ivo, et al.
Published: (2024)
Multi-Step Deductive Reasoning Over Natural Language: An Empirical Study on Out-of-Distribution Generalisation
by: Bao, Qiming, et al.
Published: (2022)
by: Bao, Qiming, et al.
Published: (2022)
What explains the success of cross-modal fine-tuning with ORCA?
by: García-de-Herreros, Paloma, et al.
Published: (2024)
by: García-de-Herreros, Paloma, et al.
Published: (2024)
Unsupervised Learning Approaches for Identifying ICU Patient Subgroups: Do Results Generalise?
by: Mayne, Harry, et al.
Published: (2024)
by: Mayne, Harry, et al.
Published: (2024)
Do LLMs Understand Romanian Driving Laws? A Study on Multimodal and Fine-Tuned Question Answering
by: Barbu, Eduard, et al.
Published: (2025)
by: Barbu, Eduard, et al.
Published: (2025)
Tokenisation via Convex Relaxations
by: Tempus, Jan, et al.
Published: (2026)
by: Tempus, Jan, et al.
Published: (2026)
Operationalising the Superficial Alignment Hypothesis via Task Complexity
by: Vergara-Browne, Tomás, et al.
Published: (2026)
by: Vergara-Browne, Tomás, et al.
Published: (2026)
Generalising E-prop to Deep Networks
by: Millidge, Beren
Published: (2025)
by: Millidge, Beren
Published: (2025)
Tokenisation over Bounded Alphabets is Hard
by: Kastreva, Violeta, et al.
Published: (2025)
by: Kastreva, Violeta, et al.
Published: (2025)
On the Effect of (Near) Duplicate Subwords in Language Modelling
by: Schäfer, Anton, et al.
Published: (2024)
by: Schäfer, Anton, et al.
Published: (2024)
Causal Estimation of Tokenisation Bias
by: Lesci, Pietro, et al.
Published: (2025)
by: Lesci, Pietro, et al.
Published: (2025)
A Culturally-Rich Romanian NLP Dataset from "Who Wants to Be a Millionaire?" Videos
by: Ganea, Alexandru-Gabriel, et al.
Published: (2025)
by: Ganea, Alexandru-Gabriel, et al.
Published: (2025)
Polyp Segmentation Generalisability of Pretrained Backbones
by: Sanderson, Edward, et al.
Published: (2024)
by: Sanderson, Edward, et al.
Published: (2024)
DoGE: Domain Reweighting with Generalization Estimation
by: Fan, Simin, et al.
Published: (2023)
by: Fan, Simin, et al.
Published: (2023)
EvIL: Evolution Strategies for Generalisable Imitation Learning
by: Sapora, Silvia, et al.
Published: (2024)
by: Sapora, Silvia, et al.
Published: (2024)
Why LLMs Cannot Think and How to Fix It
by: Jahrens, Marius, et al.
Published: (2025)
by: Jahrens, Marius, et al.
Published: (2025)
Demo: Statistically Significant Results On Biases and Errors of LLMs Do Not Guarantee Generalizable Results
by: Liu, Jonathan, et al.
Published: (2025)
by: Liu, Jonathan, et al.
Published: (2025)
Not All Data Are Unlearned Equally
by: Krishnan, Aravind, et al.
Published: (2025)
by: Krishnan, Aravind, et al.
Published: (2025)
Flatness Improves Backbone Generalisation in Few-shot Classification
by: Li, Rui, et al.
Published: (2024)
by: Li, Rui, et al.
Published: (2024)
Generalised Diffusion Probabilistic Scale-Spaces
by: Peter, Pascal
Published: (2023)
by: Peter, Pascal
Published: (2023)
Symmetry and Generalisation in Machine Learning
by: Elesedy, Hayder
Published: (2025)
by: Elesedy, Hayder
Published: (2025)
Symmetry and Generalisation in Neural Approximations of Renormalisation Transformations
by: Ashworth, Cassidy, et al.
Published: (2025)
by: Ashworth, Cassidy, et al.
Published: (2025)
Challenging Assumptions in Learning Generic Text Style Embeddings
by: Ostheimer, Phil, et al.
Published: (2025)
by: Ostheimer, Phil, et al.
Published: (2025)
Frequency and Generalisation of Periodic Activation Functions in Reinforcement Learning
by: Mavor-Parker, Augustine N., et al.
Published: (2024)
by: Mavor-Parker, Augustine N., et al.
Published: (2024)
Generalised Medical Phrase Grounding
by: Zhang, Wenjun, et al.
Published: (2025)
by: Zhang, Wenjun, et al.
Published: (2025)
The Generalised Kernel Covariance Measure
by: Bergen, Luca, et al.
Published: (2026)
by: Bergen, Luca, et al.
Published: (2026)
Unifying Linear-Time Attention via Latent Probabilistic Modelling
by: Dolga, Rares, et al.
Published: (2024)
by: Dolga, Rares, et al.
Published: (2024)
Federated Learning with Nonvacuous Generalisation Bounds
by: Jobic, Pierre, et al.
Published: (2023)
by: Jobic, Pierre, et al.
Published: (2023)
Similar Items
-
The Role of Language Imbalance in Cross-lingual Generalisation: Insights from Cloned Language Experiments
by: Schäfer, Anton, et al.
Published: (2024) -
Towards Generalising Neural Topical Representations
by: Yang, Xiaohao, et al.
Published: (2023) -
The Illusion of Superposition? A Principled Analysis of Latent Thinking in Language Models
by: Rizvi-Martel, Michael, et al.
Published: (2026) -
Forecasting Downstream Performance of LLMs With Proxy Metrics
by: Patel, Arkil, et al.
Published: (2026) -
TRA: Better Length Generalisation with Threshold Relative Attention
by: Opper, Mattia, et al.
Published: (2025)