Evaluating the relationship between regularity and learnability in recursive numeral systems using Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Silvi, Andrea, Prasertsom, Ponrawee, Culbertson, Jennifer, Dubhashi, Devdatt, Johansson, Moa, Smith, Kenny |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Recursive numeral systems are highly regular and easy to process
von: Prasertsom, Ponrawee, et al.
Veröffentlicht: (2025)
von: Prasertsom, Ponrawee, et al.
Veröffentlicht: (2025)
Learning Efficient Recursive Numeral Systems via Reinforcement Learning
von: Silvi, Andrea, et al.
Veröffentlicht: (2024)
von: Silvi, Andrea, et al.
Veröffentlicht: (2024)
PACE: Procedural Abstractions for Communicating Efficiently
von: Thomas, Jonathan D., et al.
Veröffentlicht: (2024)
von: Thomas, Jonathan D., et al.
Veröffentlicht: (2024)
Learning Approximate and Exact Numeral Systems via Reinforcement Learning
von: Carlsson, Emil, et al.
Veröffentlicht: (2021)
von: Carlsson, Emil, et al.
Veröffentlicht: (2021)
Cultural evolution via iterated learning and communication explains efficient color naming systems
von: Carlsson, Emil, et al.
Veröffentlicht: (2023)
von: Carlsson, Emil, et al.
Veröffentlicht: (2023)
Active Preference Learning for Ordering Items In- and Out-of-sample
von: Bergström, Herman, et al.
Veröffentlicht: (2024)
von: Bergström, Herman, et al.
Veröffentlicht: (2024)
Pure Exploration in Bandits with Linear Constraints
von: Carlsson, Emil, et al.
Veröffentlicht: (2023)
von: Carlsson, Emil, et al.
Veröffentlicht: (2023)
FlashHead: Efficient Drop-In Replacement for the Classification Head in Language Model Inference
von: Tranheden, Wilhelm, et al.
Veröffentlicht: (2026)
von: Tranheden, Wilhelm, et al.
Veröffentlicht: (2026)
Learning Contextual Runtime Monitors for Safe AI-Based Autonomy
von: Luque-Cerpa, Alejandro, et al.
Veröffentlicht: (2026)
von: Luque-Cerpa, Alejandro, et al.
Veröffentlicht: (2026)
Representations as Language: An Information-Theoretic Framework for Interpretability
von: Conklin, Henry, et al.
Veröffentlicht: (2024)
von: Conklin, Henry, et al.
Veröffentlicht: (2024)
DeepMLF: Multimodal language model with learnable tokens for deep fusion in sentiment analysis
von: Georgiou, Efthymios, et al.
Veröffentlicht: (2025)
von: Georgiou, Efthymios, et al.
Veröffentlicht: (2025)
Vocabulary shapes cross-lingual variation of word-order learnability in language models
von: Martins, Jonas Mayer, et al.
Veröffentlicht: (2026)
von: Martins, Jonas Mayer, et al.
Veröffentlicht: (2026)
Reasoning in Transformers -- Mitigating Spurious Correlations and Reasoning Shortcuts
von: Enström, Daniel, et al.
Veröffentlicht: (2024)
von: Enström, Daniel, et al.
Veröffentlicht: (2024)
Fact Recall, Heuristics or Pure Guesswork? Precise Interpretations of Language Models for Fact Completion
von: Saynova, Denitsa, et al.
Veröffentlicht: (2024)
von: Saynova, Denitsa, et al.
Veröffentlicht: (2024)
Twitch: Learning Abstractions for Equational Theorem Proving
von: Axelrod, Guy, et al.
Veröffentlicht: (2026)
von: Axelrod, Guy, et al.
Veröffentlicht: (2026)
The learnability of natural concepts
von: Igor Douven
Veröffentlicht: (2024)
von: Igor Douven
Veröffentlicht: (2024)
Variational Quantum Optimization with Continuous Bandits
von: Wanner, Marc, et al.
Veröffentlicht: (2025)
von: Wanner, Marc, et al.
Veröffentlicht: (2025)
Specify What? Enhancing Neural Specification Synthesis by Symbolic Methods
von: Granberry, George, et al.
Veröffentlicht: (2024)
von: Granberry, George, et al.
Veröffentlicht: (2024)
Predicting Ground State Properties: Constant Sample Complexity and Deep Learning Algorithms
von: Wanner, Marc, et al.
Veröffentlicht: (2024)
von: Wanner, Marc, et al.
Veröffentlicht: (2024)
Benchmarking Debiasing Methods for LLM-based Parameter Estimates
von: de Pieuchon, Nicolas Audinet, et al.
Veröffentlicht: (2025)
von: de Pieuchon, Nicolas Audinet, et al.
Veröffentlicht: (2025)
What Happens to a Dataset Transformed by a Projection-based Concept Removal Method?
von: Johansson, Richard
Veröffentlicht: (2024)
von: Johansson, Richard
Veröffentlicht: (2024)
Identifying Non-Replicable Social Science Studies with Language Models
von: Saynova, Denitsa, et al.
Veröffentlicht: (2025)
von: Saynova, Denitsa, et al.
Veröffentlicht: (2025)
On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference
von: Ren, Siyu, et al.
Veröffentlicht: (2024)
von: Ren, Siyu, et al.
Veröffentlicht: (2024)
Sign of the Times: Evaluating the use of Large Language Models for Idiomaticity Detection
von: Phelps, Dylan, et al.
Veröffentlicht: (2024)
von: Phelps, Dylan, et al.
Veröffentlicht: (2024)
On the Military Applications of Large Language Models
von: Johansson, Satu, et al.
Veröffentlicht: (2025)
von: Johansson, Satu, et al.
Veröffentlicht: (2025)
Encode, Think, Decode: Scaling test-time reasoning with recursive latent thoughts
von: Koishekenov, Yeskendir, et al.
Veröffentlicht: (2025)
von: Koishekenov, Yeskendir, et al.
Veröffentlicht: (2025)
T2I-Eval-R1: Reinforcement Learning-Driven Reasoning for Interpretable Text-to-Image Evaluation
von: Ma, Zi-Ao, et al.
Veröffentlicht: (2025)
von: Ma, Zi-Ao, et al.
Veröffentlicht: (2025)
A Comprehensive Survey on Reinforcement Learning-based Agentic Search: Foundations, Roles, Optimizations, Evaluations, and Applications
von: Lin, Minhua, et al.
Veröffentlicht: (2025)
von: Lin, Minhua, et al.
Veröffentlicht: (2025)
On the Hidden Objective Biases of Group-based Reinforcement Learning
von: Fontana, Aleksandar, et al.
Veröffentlicht: (2026)
von: Fontana, Aleksandar, et al.
Veröffentlicht: (2026)
From Problem-Solving to Teaching Problem-Solving: Aligning LLMs with Pedagogy using Reinforcement Learning
von: Dinucu-Jianu, David, et al.
Veröffentlicht: (2025)
von: Dinucu-Jianu, David, et al.
Veröffentlicht: (2025)
Backdoor defense, learnability and obfuscation
von: Christiano, Paul, et al.
Veröffentlicht: (2024)
von: Christiano, Paul, et al.
Veröffentlicht: (2024)
Differentiable Evolutionary Reinforcement Learning
von: Cheng, Sitao, et al.
Veröffentlicht: (2025)
von: Cheng, Sitao, et al.
Veröffentlicht: (2025)
Bridging the Evaluation Gap: Leveraging Large Language Models for Topic Model Evaluation
von: Tan, Zhiyin, et al.
Veröffentlicht: (2025)
von: Tan, Zhiyin, et al.
Veröffentlicht: (2025)
Demonstration Selection for In-Context Learning via Reinforcement Learning
von: Wang, Xubin, et al.
Veröffentlicht: (2024)
von: Wang, Xubin, et al.
Veröffentlicht: (2024)
Can Large Language Models (or Humans) Disentangle Text?
von: de Pieuchon, Nicolas Audinet, et al.
Veröffentlicht: (2024)
von: de Pieuchon, Nicolas Audinet, et al.
Veröffentlicht: (2024)
PRL: Prompts from Reinforcement Learning
von: Batorski, Paweł, et al.
Veröffentlicht: (2025)
von: Batorski, Paweł, et al.
Veröffentlicht: (2025)
Discrete Prompt Compression with Reinforcement Learning
von: Jung, Hoyoun, et al.
Veröffentlicht: (2023)
von: Jung, Hoyoun, et al.
Veröffentlicht: (2023)
Generating Leakage-Free Benchmarks for Robust RAG Evaluation
von: Liu, Jiayi, et al.
Veröffentlicht: (2026)
von: Liu, Jiayi, et al.
Veröffentlicht: (2026)
A Cognitive Evaluation Benchmark of Image Reasoning and Description for Large Vision-Language Models
von: Song, Xiujie, et al.
Veröffentlicht: (2024)
von: Song, Xiujie, et al.
Veröffentlicht: (2024)
Automated Evaluation of Classroom Instructional Support with LLMs and BoWs: Connecting Global Predictions to Specific Feedback
von: Whitehill, Jacob, et al.
Veröffentlicht: (2023)
von: Whitehill, Jacob, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Recursive numeral systems are highly regular and easy to process
von: Prasertsom, Ponrawee, et al.
Veröffentlicht: (2025) -
Learning Efficient Recursive Numeral Systems via Reinforcement Learning
von: Silvi, Andrea, et al.
Veröffentlicht: (2024) -
PACE: Procedural Abstractions for Communicating Efficiently
von: Thomas, Jonathan D., et al.
Veröffentlicht: (2024) -
Learning Approximate and Exact Numeral Systems via Reinforcement Learning
von: Carlsson, Emil, et al.
Veröffentlicht: (2021) -
Cultural evolution via iterated learning and communication explains efficient color naming systems
von: Carlsson, Emil, et al.
Veröffentlicht: (2023)