Guardado en:
| Autores principales: | Gideoni, Yonatan, Risi, Sebastian, Gal, Yarin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.16805 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Do Multilingual LLMs Think In English?
por: Schut, Lisa, et al.
Publicado: (2025)
por: Schut, Lisa, et al.
Publicado: (2025)
Stabilizing Policy Gradients for Sample-Efficient Reinforcement Learning in LLM Reasoning
por: Melo, Luckeciano C., et al.
Publicado: (2025)
por: Melo, Luckeciano C., et al.
Publicado: (2025)
Temporal-Difference Variational Continual Learning
por: Melo, Luckeciano C., et al.
Publicado: (2024)
por: Melo, Luckeciano C., et al.
Publicado: (2024)
In-Context Learning Learns Label Relationships but Is Not Conventional Learning
por: Kossen, Jannik, et al.
Publicado: (2023)
por: Kossen, Jannik, et al.
Publicado: (2023)
Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities
por: Nikitin, Alexander, et al.
Publicado: (2024)
por: Nikitin, Alexander, et al.
Publicado: (2024)
Evaluating & Reducing Deceptive Dialogue From Language Models with Multi-turn RL
por: Abdulhai, Marwa, et al.
Publicado: (2025)
por: Abdulhai, Marwa, et al.
Publicado: (2025)
Semantic Entropy Probes: Robust and Cheap Hallucination Detection in LLMs
por: Kossen, Jannik, et al.
Publicado: (2024)
por: Kossen, Jannik, et al.
Publicado: (2024)
Challenges and Considerations in the Evaluation of Bayesian Causal Discovery
por: Mamaghan, Amir Mohammad Karimi, et al.
Publicado: (2024)
por: Mamaghan, Amir Mohammad Karimi, et al.
Publicado: (2024)
TextCAVs: Debugging vision models using text
por: Nicolson, Angus, et al.
Publicado: (2024)
por: Nicolson, Angus, et al.
Publicado: (2024)
Leveraging Deep Learning for Physical Model Bias of Global Air Quality Estimates
por: Doerksen, Kelsey, et al.
Publicado: (2025)
por: Doerksen, Kelsey, et al.
Publicado: (2025)
Uncertainty Quantification for Surface Ozone Emulators using Deep Learning
por: Doerksen, Kelsey, et al.
Publicado: (2025)
por: Doerksen, Kelsey, et al.
Publicado: (2025)
Continuous Thought Machines
por: Darlow, Luke, et al.
Publicado: (2025)
por: Darlow, Luke, et al.
Publicado: (2025)
GPG: A Simple and Strong Reinforcement Learning Baseline for Model Reasoning
por: Chu, Xiangxiang, et al.
Publicado: (2025)
por: Chu, Xiangxiang, et al.
Publicado: (2025)
Deep Minds and Shallow Probes
por: Lee, Su Hyeong, et al.
Publicado: (2026)
por: Lee, Su Hyeong, et al.
Publicado: (2026)
Non-invasive Neural Decoding in Source Reconstructed Brain Space
por: Gideoni, Yonatan, et al.
Publicado: (2024)
por: Gideoni, Yonatan, et al.
Publicado: (2024)
No Mean Feat: Simple, Strong Baselines for Context Compression
por: Feldman, Yair, et al.
Publicado: (2025)
por: Feldman, Yair, et al.
Publicado: (2025)
Reasoning Introduces New Poisoning Attacks Yet Makes Them More Complicated
por: Foerster, Hanna, et al.
Publicado: (2025)
por: Foerster, Hanna, et al.
Publicado: (2025)
MolMix: A Simple Yet Effective Baseline for Multimodal Molecular Representation Learning
por: Manolache, Andrei, et al.
Publicado: (2024)
por: Manolache, Andrei, et al.
Publicado: (2024)
Iterative Deployment Improves Planning Skills in LLMs
por: Corrêa, Augusto B., et al.
Publicado: (2025)
por: Corrêa, Augusto B., et al.
Publicado: (2025)
SPA: A Simple but Tough-to-Beat Baseline for Knowledge Injection
por: Tang, Kexian, et al.
Publicado: (2026)
por: Tang, Kexian, et al.
Publicado: (2026)
Hedging Is Not All You Need: A Simple Baseline for Online Learning Under Haphazard Inputs
por: Buckchash, Himanshu, et al.
Publicado: (2024)
por: Buckchash, Himanshu, et al.
Publicado: (2024)
Probabilistic Modeling of Latent Agentic Substructures in Deep Neural Networks
por: Lee, Su Hyeong, et al.
Publicado: (2025)
por: Lee, Su Hyeong, et al.
Publicado: (2025)
When Does Structure Matter in Continual Learning? Dimensionality Controls When Modularity Shapes Representational Geometry
por: Korte, Kathrin, et al.
Publicado: (2026)
por: Korte, Kathrin, et al.
Publicado: (2026)
Deep Ignorance: Filtering Pretraining Data Builds Tamper-Resistant Safeguards into Open-Weight LLMs
por: O'Brien, Kyle, et al.
Publicado: (2025)
por: O'Brien, Kyle, et al.
Publicado: (2025)
A Simple Baseline for Stable and Plastic Neural Networks
por: Künzel, Étienne, et al.
Publicado: (2025)
por: Künzel, Étienne, et al.
Publicado: (2025)
An Embarrassingly Simple Baseline for Imbalanced Semi-Supervised Learning
por: Chen, Hao, et al.
Publicado: (2022)
por: Chen, Hao, et al.
Publicado: (2022)
Explaining Explainability: Recommendations for Effective Use of Concept Activation Vectors
por: Nicolson, Angus, et al.
Publicado: (2024)
por: Nicolson, Angus, et al.
Publicado: (2024)
Existing Large Language Model Unlearning Evaluations Are Inconclusive
por: Feng, Zhili, et al.
Publicado: (2025)
por: Feng, Zhili, et al.
Publicado: (2025)
LayerShuffle: Enhancing Robustness in Vision Transformers by Randomizing Layer Execution Order
por: Freiberger, Matthias, et al.
Publicado: (2024)
por: Freiberger, Matthias, et al.
Publicado: (2024)
A Simple, Solid, and Reproducible Baseline for Bridge Bidding AI
por: Kita, Haruka, et al.
Publicado: (2024)
por: Kita, Haruka, et al.
Publicado: (2024)
Is there Value in Reinforcement Learning?
por: Fox, Lior, et al.
Publicado: (2025)
por: Fox, Lior, et al.
Publicado: (2025)
On the Expressive Power of Sparse Geometric MPNNs
por: Sverdlov, Yonatan, et al.
Publicado: (2024)
por: Sverdlov, Yonatan, et al.
Publicado: (2024)
Structurally Flexible Neural Networks: Evolving the Building Blocks for General Agents
por: Pedersen, Joachim Winther, et al.
Publicado: (2024)
por: Pedersen, Joachim Winther, et al.
Publicado: (2024)
The Curse of Recursion: Training on Generated Data Makes Models Forget
por: Shumailov, Ilia, et al.
Publicado: (2023)
por: Shumailov, Ilia, et al.
Publicado: (2023)
More Test-Time Compute Can Hurt: Overestimation Bias in LLM Beam Search
por: Dalal, Gal, et al.
Publicado: (2026)
por: Dalal, Gal, et al.
Publicado: (2026)
Assessing Image Quality Using a Simple Generative Representation
por: Raviv, Simon, et al.
Publicado: (2024)
por: Raviv, Simon, et al.
Publicado: (2024)
Improved Distribution Estimation in $\ell_\infty$
por: Cohen, Doron, et al.
Publicado: (2026)
por: Cohen, Doron, et al.
Publicado: (2026)
Revisiting Multi-Permutation Equivariance through the Lens of Irreducible Representations
por: Sverdlov, Yonatan, et al.
Publicado: (2024)
por: Sverdlov, Yonatan, et al.
Publicado: (2024)
Competition-Aware CPC Forecasting with Near-Market Coverage
por: Frey, Sebastian, et al.
Publicado: (2026)
por: Frey, Sebastian, et al.
Publicado: (2026)
Policy Gradient with Tree Expansion
por: Dalal, Gal, et al.
Publicado: (2023)
por: Dalal, Gal, et al.
Publicado: (2023)
Ejemplares similares
-
Do Multilingual LLMs Think In English?
por: Schut, Lisa, et al.
Publicado: (2025) -
Stabilizing Policy Gradients for Sample-Efficient Reinforcement Learning in LLM Reasoning
por: Melo, Luckeciano C., et al.
Publicado: (2025) -
Temporal-Difference Variational Continual Learning
por: Melo, Luckeciano C., et al.
Publicado: (2024) -
In-Context Learning Learns Label Relationships but Is Not Conventional Learning
por: Kossen, Jannik, et al.
Publicado: (2023) -
Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities
por: Nikitin, Alexander, et al.
Publicado: (2024)