Programming by Examples Meets Historical Linguistics: A Large Language Model Based Approach to Sound Law Induction
Fuente:
arXiv
Saved in:
| Main Authors: | Naik, Atharva, Agrawal, Darsh, Sng, Hong, Marr, Clayton, Zhang, Kexun, Robinson, Nathaniel R, Chang, Kalvin, Byrnes, Rebecca, Mysore, Aravind, Rose, Carolyn, Mortensen, David R |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can Large Language Models Code Like a Linguist?: A Case Study in Low Resource Sound Law Induction
by: Naik, Atharva, et al.
Published: (2024)
by: Naik, Atharva, et al.
Published: (2024)
PBEBench: A Multi-Step Programming by Examples Reasoning Benchmark inspired by Historical Linguistics
by: Naik, Atharva, et al.
Published: (2025)
by: Naik, Atharva, et al.
Published: (2025)
MetaLint: Easy-to-Hard Generalization for Code Linting
by: Naik, Atharva, et al.
Published: (2025)
by: Naik, Atharva, et al.
Published: (2025)
ReaComp: Compiling LLM Reasoning into Symbolic Solvers for Efficient Program Synthesis
by: Naik, Atharva, et al.
Published: (2026)
by: Naik, Atharva, et al.
Published: (2026)
Automating Sound Change Prediction for Phylogenetic Inference: A Tukanoan Case Study
by: Chang, Kalvin, et al.
Published: (2024)
by: Chang, Kalvin, et al.
Published: (2024)
CRScore++: Reinforcement Learning with Verifiable Tool and AI Feedback for Code Review
by: Kapadnis, Manav Nitin, et al.
Published: (2025)
by: Kapadnis, Manav Nitin, et al.
Published: (2025)
Data Augmentation for Code Translation with Comparable Corpora and Multiple References
by: Xie, Yiqing, et al.
Published: (2023)
by: Xie, Yiqing, et al.
Published: (2023)
An Empirical Study on Strong-Weak Model Collaboration for Repo-level Code Generation
by: Gandhi, Shubham, et al.
Published: (2025)
by: Gandhi, Shubham, et al.
Published: (2025)
CRScore: Grounding Automated Evaluation of Code Review Comments in Code Claims and Smells
by: Naik, Atharva, et al.
Published: (2024)
by: Naik, Atharva, et al.
Published: (2024)
PWESuite: Phonetic Word Embeddings and Tasks They Facilitate
by: Zouhar, Vilém, et al.
Published: (2023)
by: Zouhar, Vilém, et al.
Published: (2023)
ChartEditBench: Evaluating Grounded Multi-Turn Chart Editing in Multimodal Language Models
by: Kapadnis, Manav Nitin, et al.
Published: (2026)
by: Kapadnis, Manav Nitin, et al.
Published: (2026)
On the Limitations of Embedding Based Methods for Measuring Functional Correctness for Code Generation
by: Naik, Atharva
Published: (2024)
by: Naik, Atharva
Published: (2024)
Phonotactic Complexity across Dialects
by: Shim, Ryan Soh-Eun, et al.
Published: (2024)
by: Shim, Ryan Soh-Eun, et al.
Published: (2024)
Morpheme Induction for Emergent Language
by: Boldt, Brendon, et al.
Published: (2025)
by: Boldt, Brendon, et al.
Published: (2025)
Contribution of $π^0$ Exchange in Elastic Muon-Proton Scattering
by: Naik, Atharva, et al.
Published: (2024)
by: Naik, Atharva, et al.
Published: (2024)
PragWorld: A Benchmark Evaluating LLMs' Local World Model under Minimal Linguistic Alterations and Conversational Dynamics
by: Vashistha, Sachin, et al.
Published: (2025)
by: Vashistha, Sachin, et al.
Published: (2025)
Stationary Power-Law Solutions of Kinetic-Alfvénic Turbulence
by: Shen, Kexun, et al.
Published: (2025)
by: Shen, Kexun, et al.
Published: (2025)
Leveraging Allophony in Self-Supervised Speech Models for Atypical Pronunciation Assessment
by: Choi, Kwanghee, et al.
Published: (2025)
by: Choi, Kwanghee, et al.
Published: (2025)
kalvinroberts/RelayModelCode: Relay Model Code
by: Kalvin Roberts
Published: (2026)
by: Kalvin Roberts
Published: (2026)
Narrating Causal Graphs with Large Language Models
by: Phatak, Atharva, et al.
Published: (2024)
by: Phatak, Atharva, et al.
Published: (2024)
Hire a Linguist!: Learning Endangered Languages with In-Context Linguistic Descriptions
by: Zhang, Kexun, et al.
Published: (2024)
by: Zhang, Kexun, et al.
Published: (2024)
Scaling Self-Supervised Speech Models Uncovers Deep Linguistic Relationships: Evidence from the Pacific Cluster
by: Kim, Minu, et al.
Published: (2026)
by: Kim, Minu, et al.
Published: (2026)
Limited-Resource Adapters Are Regularizers, Not Linguists
by: Fekete, Marcell, et al.
Published: (2025)
by: Fekete, Marcell, et al.
Published: (2025)
LaMP: When Large Language Models Meet Personalization
by: Salemi, Alireza, et al.
Published: (2023)
by: Salemi, Alireza, et al.
Published: (2023)
Providing tailored reflection instructions in collaborative learning using large language models
by: Atharva Naik, et al.
Published: (2024)
by: Atharva Naik, et al.
Published: (2024)
Dynamics of number entropy for free fermionic systems in presence of defects and stochastic processes
by: Naik, Atharva, et al.
Published: (2025)
by: Naik, Atharva, et al.
Published: (2025)
Prediction-Based Markov Violation Scores for Detecting Non-Markovian Observations in Reinforcement Learning
by: Mysore, Naveen
Published: (2026)
by: Mysore, Naveen
Published: (2026)
Quantifying First-Order Markov Violations in Noisy Reinforcement Learning: A Causal Discovery Approach
by: Mysore, Naveen
Published: (2025)
by: Mysore, Naveen
Published: (2025)
Temporal Functional Circuits: From Spline Plots to Faithful Explanations in KAN Forecasting
by: Mysore, Naveen
Published: (2026)
by: Mysore, Naveen
Published: (2026)
DecompKAN: Decomposed Patch-KAN for Long-Term Time Series Forecasting
by: Mysore, Naveen
Published: (2026)
by: Mysore, Naveen
Published: (2026)
DialUp! Modeling the Language Continuum by Adapting Models to Dialects and Dialects to Models
by: Bafna, Niyati, et al.
Published: (2025)
by: Bafna, Niyati, et al.
Published: (2025)
Error Estimation for Adaptive Mesh Refinement in Droplet Simulations
by: Nathawani, Darsh, et al.
Published: (2025)
by: Nathawani, Darsh, et al.
Published: (2025)
A one-dimensional mathematical model for shear-induced droplet formation in co-flowing fluids
by: Nathawani, Darsh, et al.
Published: (2023)
by: Nathawani, Darsh, et al.
Published: (2023)
Linguistic Laws Meet Protein Sequences: A Comparative Analysis of Subword Tokenization Methods
by: Suyunu, Burak, et al.
Published: (2024)
by: Suyunu, Burak, et al.
Published: (2024)
Linguistic Structure Induction from Language Models
by: Momen, Omar
Published: (2024)
by: Momen, Omar
Published: (2024)
The Routledge Handbook of Historical Linguistics
Published: (2021)
Published: (2021)
Self-supervised Speech Representations Still Struggle with African American Vernacular English
by: Chang, Kalvin, et al.
Published: (2024)
by: Chang, Kalvin, et al.
Published: (2024)
PRiSM: Benchmarking Phone Realization in Speech Models
by: Bharadwaj, Shikhar, et al.
Published: (2026)
by: Bharadwaj, Shikhar, et al.
Published: (2026)
Harnessing Linguistic Dissimilarity for Language Generalization on Unseen Low-Resource Varieties
by: Kim, Jinju, et al.
Published: (2026)
by: Kim, Jinju, et al.
Published: (2026)
Generating Situated Reflection Triggers about Alternative Solution Paths: A Case Study of Generative AI for Computer-Supported Collaborative Learning
by: Naik, Atharva, et al.
Published: (2024)
by: Naik, Atharva, et al.
Published: (2024)
Similar Items
-
Can Large Language Models Code Like a Linguist?: A Case Study in Low Resource Sound Law Induction
by: Naik, Atharva, et al.
Published: (2024) -
PBEBench: A Multi-Step Programming by Examples Reasoning Benchmark inspired by Historical Linguistics
by: Naik, Atharva, et al.
Published: (2025) -
MetaLint: Easy-to-Hard Generalization for Code Linting
by: Naik, Atharva, et al.
Published: (2025) -
ReaComp: Compiling LLM Reasoning into Symbolic Solvers for Efficient Program Synthesis
by: Naik, Atharva, et al.
Published: (2026) -
Automating Sound Change Prediction for Phylogenetic Inference: A Tukanoan Case Study
by: Chang, Kalvin, et al.
Published: (2024)