Likelihood scoring for continuations of mathematical text: a self-supervised benchmark with tests for shortcut vulnerabilities
Fuente:
arXiv
Salvato in:
| Autore principale: | Ranard, Daniel |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Cropping outperforms dropout as an augmentation strategy for self-supervised training of text embeddings
di: González-Márquez, Rita, et al.
Pubblicazione: (2025)
di: González-Márquez, Rita, et al.
Pubblicazione: (2025)
BabySLM: language-acquisition-friendly benchmark of self-supervised spoken language models
di: Lavechin, Marvin, et al.
Pubblicazione: (2023)
di: Lavechin, Marvin, et al.
Pubblicazione: (2023)
A theoretical framework for self-supervised contrastive learning for continuous dependent data
di: Marusov, Alexander, et al.
Pubblicazione: (2025)
di: Marusov, Alexander, et al.
Pubblicazione: (2025)
eyeballvul: a future-proof benchmark for vulnerability detection in the wild
di: Chauvin, Timothee
Pubblicazione: (2024)
di: Chauvin, Timothee
Pubblicazione: (2024)
Do regularization methods for shortcut mitigation work as intended?
di: Hong, Haoyang, et al.
Pubblicazione: (2025)
di: Hong, Haoyang, et al.
Pubblicazione: (2025)
RobocupGym: A challenging continuous control benchmark in Robocup
di: Beukman, Michael, et al.
Pubblicazione: (2024)
di: Beukman, Michael, et al.
Pubblicazione: (2024)
Deployment of AI-Assisted Interventions: Capacity Constraints and Noisy Compliance
di: Chan, Carri W., et al.
Pubblicazione: (2026)
di: Chan, Carri W., et al.
Pubblicazione: (2026)
Contrastive Policy Gradient: Aligning LLMs on sequence-level scores in a supervised-friendly fashion
di: Flet-Berliac, Yannis, et al.
Pubblicazione: (2024)
di: Flet-Berliac, Yannis, et al.
Pubblicazione: (2024)
AlleNoise: large-scale text classification benchmark dataset with real-world label noise
di: Rączkowska, Alicja, et al.
Pubblicazione: (2024)
di: Rączkowska, Alicja, et al.
Pubblicazione: (2024)
Noise-robust zero-shot text-to-speech synthesis conditioned on self-supervised speech-representation model with adapters
di: Fujita, Kenichi, et al.
Pubblicazione: (2024)
di: Fujita, Kenichi, et al.
Pubblicazione: (2024)
Distributionally robust self-supervised learning for tabular data
di: Ghosh, Shantanu, et al.
Pubblicazione: (2024)
di: Ghosh, Shantanu, et al.
Pubblicazione: (2024)
Inpainting physics: self-supervised learning for context-driven fluid simulation
di: Weidner, Jonas, et al.
Pubblicazione: (2026)
di: Weidner, Jonas, et al.
Pubblicazione: (2026)
Semiparametric KSD test: unifying score and distance-based approaches for goodness-of-fit testing
di: Huang, Zhihan, et al.
Pubblicazione: (2025)
di: Huang, Zhihan, et al.
Pubblicazione: (2025)
Understanding the limitations of self-supervised learning for tabular anomaly detection
di: Mai, Kimberly T., et al.
Pubblicazione: (2023)
di: Mai, Kimberly T., et al.
Pubblicazione: (2023)
Inverse Entropic Optimal Transport Solves Semi-supervised Learning via Data Likelihood Maximization
di: Persiianov, Mikhail, et al.
Pubblicazione: (2024)
di: Persiianov, Mikhail, et al.
Pubblicazione: (2024)
Dynamic programming by polymorphic semiring algebraic shortcut fusion
di: Little, Max A., et al.
Pubblicazione: (2021)
di: Little, Max A., et al.
Pubblicazione: (2021)
On the social bias of speech self-supervised models
di: Lin, Yi-Cheng, et al.
Pubblicazione: (2024)
di: Lin, Yi-Cheng, et al.
Pubblicazione: (2024)
Information theoretic underpinning of self-supervised learning by clustering
di: Kittler, Josef, et al.
Pubblicazione: (2026)
di: Kittler, Josef, et al.
Pubblicazione: (2026)
RetrySQL: text-to-SQL training with retry data for self-correcting query generation
di: Rączkowska, Alicja, et al.
Pubblicazione: (2025)
di: Rączkowska, Alicja, et al.
Pubblicazione: (2025)
Math Takes Two: A test for emergent mathematical reasoning in communication
di: Cooper, Michael, et al.
Pubblicazione: (2026)
di: Cooper, Michael, et al.
Pubblicazione: (2026)
Systematic comparison of semi-supervised and self-supervised learning for medical image classification
di: Huang, Zhe, et al.
Pubblicazione: (2023)
di: Huang, Zhe, et al.
Pubblicazione: (2023)
Joint Embedding Predictive Architecture for self-supervised pretraining on polymer molecular graphs
di: Piccoli, Francesco, et al.
Pubblicazione: (2025)
di: Piccoli, Francesco, et al.
Pubblicazione: (2025)
There is more to graphs than meets the eye: Learning universal features with self-supervision
di: Das, Laya, et al.
Pubblicazione: (2023)
di: Das, Laya, et al.
Pubblicazione: (2023)
Graph self-supervised learning based on frequency corruption
di: Li, Haojie, et al.
Pubblicazione: (2026)
di: Li, Haojie, et al.
Pubblicazione: (2026)
Deployment-complete benchmarking
di: Mansouri, El Mustapha, et al.
Pubblicazione: (2026)
di: Mansouri, El Mustapha, et al.
Pubblicazione: (2026)
Visual concept ranking uncovers medical shortcuts used by large multimodal models
di: Janizek, Joseph D., et al.
Pubblicazione: (2026)
di: Janizek, Joseph D., et al.
Pubblicazione: (2026)
Probing self-attention in self-supervised speech models for cross-linguistic differences
di: Gopinath, Sai, et al.
Pubblicazione: (2024)
di: Gopinath, Sai, et al.
Pubblicazione: (2024)
Normalizing self-supervised learning for provably reliable Change Point Detection
di: Bazarova, Alexandra, et al.
Pubblicazione: (2024)
di: Bazarova, Alexandra, et al.
Pubblicazione: (2024)
Tempo estimation as fully self-supervised binary classification
di: Henkel, Florian, et al.
Pubblicazione: (2024)
di: Henkel, Florian, et al.
Pubblicazione: (2024)
Scalable Inference-Time Annealing with Surrogate Likelihood Estimators
di: Peñaherrera, Daniel, et al.
Pubblicazione: (2026)
di: Peñaherrera, Daniel, et al.
Pubblicazione: (2026)
Rethinking Test-time Likelihood: The Likelihood Path Principle and Its Application to OOD Detection
di: Huang, Sicong, et al.
Pubblicazione: (2024)
di: Huang, Sicong, et al.
Pubblicazione: (2024)
Featuremetric benchmarking: Quantum computer benchmarks based on circuit features
di: Proctor, Timothy, et al.
Pubblicazione: (2025)
di: Proctor, Timothy, et al.
Pubblicazione: (2025)
Competing for pixels: a self-play algorithm for weakly-supervised segmentation
di: Saeed, Shaheer U., et al.
Pubblicazione: (2024)
di: Saeed, Shaheer U., et al.
Pubblicazione: (2024)
Neural Likelihood Surfaces for Spatial Processes with Computationally Intensive or Intractable Likelihoods
di: Walchessen, Julia, et al.
Pubblicazione: (2023)
di: Walchessen, Julia, et al.
Pubblicazione: (2023)
Spectral Self-supervised Feature Selection
di: Segal, Daniel, et al.
Pubblicazione: (2024)
di: Segal, Daniel, et al.
Pubblicazione: (2024)
Maximum Likelihood Reinforcement Learning
di: Tajwar, Fahim, et al.
Pubblicazione: (2026)
di: Tajwar, Fahim, et al.
Pubblicazione: (2026)
Momentum Particle Maximum Likelihood
di: Lim, Jen Ning, et al.
Pubblicazione: (2023)
di: Lim, Jen Ning, et al.
Pubblicazione: (2023)
Likelihood-Free Variational Autoencoders
di: Xu, Chen, et al.
Pubblicazione: (2025)
di: Xu, Chen, et al.
Pubblicazione: (2025)
FRAPPE: $\underline{\text{F}}$ast $\underline{\text{Ra}}$nk $\underline{\text{App}}$roximation with $\underline{\text{E}}$xplainable Features for Tensors
di: Shiao, William, et al.
Pubblicazione: (2022)
di: Shiao, William, et al.
Pubblicazione: (2022)
Comparison of self-supervised in-domain and supervised out-domain transfer learning for bird species recognition
di: Ghaffari, Houtan, et al.
Pubblicazione: (2024)
di: Ghaffari, Houtan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Cropping outperforms dropout as an augmentation strategy for self-supervised training of text embeddings
di: González-Márquez, Rita, et al.
Pubblicazione: (2025) -
BabySLM: language-acquisition-friendly benchmark of self-supervised spoken language models
di: Lavechin, Marvin, et al.
Pubblicazione: (2023) -
A theoretical framework for self-supervised contrastive learning for continuous dependent data
di: Marusov, Alexander, et al.
Pubblicazione: (2025) -
eyeballvul: a future-proof benchmark for vulnerability detection in the wild
di: Chauvin, Timothee
Pubblicazione: (2024) -
Do regularization methods for shortcut mitigation work as intended?
di: Hong, Haoyang, et al.
Pubblicazione: (2025)