Heterogeneity in Formal Linguistic Competence of Language Models: Is Data the Real Bottleneck?
Fuente:
arXiv
Salvato in:
| Autori principali: | Renduchintala, H S V N S Kowndinya, Bhatia, Sumit |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SMART: Submodular Data Mixture Strategy for Instruction Tuning
di: Renduchintala, H S V N S Kowndinya, et al.
Pubblicazione: (2024)
di: Renduchintala, H S V N S Kowndinya, et al.
Pubblicazione: (2024)
POSIX: A Prompt Sensitivity Index For Large Language Models
di: Chatterjee, Anwoy, et al.
Pubblicazione: (2024)
di: Chatterjee, Anwoy, et al.
Pubblicazione: (2024)
On the Effect of Instruction Tuning Loss on Generalization
di: Chatterjee, Anwoy, et al.
Pubblicazione: (2025)
di: Chatterjee, Anwoy, et al.
Pubblicazione: (2025)
Tied-Lora: Enhancing parameter efficiency of LoRA with weight tying
di: Renduchintala, Adithya, et al.
Pubblicazione: (2023)
di: Renduchintala, Adithya, et al.
Pubblicazione: (2023)
Between Circuits and Chomsky: Pre-pretraining on Formal Languages Imparts Linguistic Biases
di: Hu, Michael Y., et al.
Pubblicazione: (2025)
di: Hu, Michael Y., et al.
Pubblicazione: (2025)
Language Bottleneck Models for Qualitative Knowledge State Modeling
di: Berthon, Antonin, et al.
Pubblicazione: (2025)
di: Berthon, Antonin, et al.
Pubblicazione: (2025)
Consistency Is the Key: Detecting Hallucinations in LLM Generated Text By Checking Inconsistencies About Key Facts
di: Gupta, Raavi, et al.
Pubblicazione: (2025)
di: Gupta, Raavi, et al.
Pubblicazione: (2025)
Perceptions of Linguistic Uncertainty by Language Models and Humans
di: Belem, Catarina G, et al.
Pubblicazione: (2024)
di: Belem, Catarina G, et al.
Pubblicazione: (2024)
Linguistic Blind Spots of Large Language Models
di: Cheng, Jiali, et al.
Pubblicazione: (2025)
di: Cheng, Jiali, et al.
Pubblicazione: (2025)
Policy Learning with a Language Bottleneck
di: Srivastava, Megha, et al.
Pubblicazione: (2024)
di: Srivastava, Megha, et al.
Pubblicazione: (2024)
Are Small Language Models Ready to Compete with Large Language Models for Practical Applications?
di: Sinha, Neelabh, et al.
Pubblicazione: (2024)
di: Sinha, Neelabh, et al.
Pubblicazione: (2024)
Uncovering Competency Gaps in Large Language Models and Their Benchmarks
di: Bohacek, Maty, et al.
Pubblicazione: (2025)
di: Bohacek, Maty, et al.
Pubblicazione: (2025)
Regurgitative Training: The Value of Real Data in Training Large Language Models
di: Zhang, Jinghui, et al.
Pubblicazione: (2024)
di: Zhang, Jinghui, et al.
Pubblicazione: (2024)
LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety
di: Yang, Junxiao, et al.
Pubblicazione: (2026)
di: Yang, Junxiao, et al.
Pubblicazione: (2026)
ABBEL: LLM Agents Acting through Belief Bottlenecks Expressed in Language
di: Lidayan, Aly, et al.
Pubblicazione: (2025)
di: Lidayan, Aly, et al.
Pubblicazione: (2025)
Mixture of Heterogeneous Grouped Experts for Language Modeling
di: Ma, Zhicheng, et al.
Pubblicazione: (2026)
di: Ma, Zhicheng, et al.
Pubblicazione: (2026)
Exploring RL-based LLM Training for Formal Language Tasks with Programmed Rewards
di: Padula, Alexander G., et al.
Pubblicazione: (2024)
di: Padula, Alexander G., et al.
Pubblicazione: (2024)
Towards Real-world Human Behavior Simulation: Benchmarking Large Language Models on Long-horizon, Cross-scenario, Heterogeneous Behavior Traces
di: Chen, Jiawei, et al.
Pubblicazione: (2026)
di: Chen, Jiawei, et al.
Pubblicazione: (2026)
FVEL: Interactive Formal Verification Environment with Large Language Models via Theorem Proving
di: Lin, Xiaohan, et al.
Pubblicazione: (2024)
di: Lin, Xiaohan, et al.
Pubblicazione: (2024)
Concurrent Linguistic Error Detection (CLED): a New Methodology for Error Detection in Large Language Models
di: Zhu, Jinhua, et al.
Pubblicazione: (2024)
di: Zhu, Jinhua, et al.
Pubblicazione: (2024)
Experience Sharing in Mutual Reinforcement Learning for Heterogeneous Language Models
di: Liu, Xiaoze, et al.
Pubblicazione: (2026)
di: Liu, Xiaoze, et al.
Pubblicazione: (2026)
BIPEFT: Budget-Guided Iterative Search for Parameter Efficient Fine-Tuning of Large Pretrained Language Models
di: Chang, Aofei, et al.
Pubblicazione: (2024)
di: Chang, Aofei, et al.
Pubblicazione: (2024)
IndicSentEval: How Effectively do Multilingual Transformer Models encode Linguistic Properties for Indic Languages?
di: Aravapalli, Akhilesh, et al.
Pubblicazione: (2024)
di: Aravapalli, Akhilesh, et al.
Pubblicazione: (2024)
AraLingBench A Human-Annotated Benchmark for Evaluating Arabic Linguistic Capabilities of Large Language Models
di: Zbeeb, Mohammad, et al.
Pubblicazione: (2025)
di: Zbeeb, Mohammad, et al.
Pubblicazione: (2025)
Towards a Theoretical Understanding of Synthetic Data in LLM Post-Training: A Reverse-Bottleneck Perspective
di: Gan, Zeyu, et al.
Pubblicazione: (2024)
di: Gan, Zeyu, et al.
Pubblicazione: (2024)
Artificial Conversations, Real Results: Fostering Language Detection with Synthetic Data
di: Mohammadi, Fatemeh, et al.
Pubblicazione: (2025)
di: Mohammadi, Fatemeh, et al.
Pubblicazione: (2025)
Safe: Enhancing Mathematical Reasoning in Large Language Models via Retrospective Step-aware Formal Verification
di: Liu, Chengwu, et al.
Pubblicazione: (2025)
di: Liu, Chengwu, et al.
Pubblicazione: (2025)
Marco-o1 v2: Towards Widening The Distillation Bottleneck for Reasoning Models
di: Yin, Huifeng, et al.
Pubblicazione: (2025)
di: Yin, Huifeng, et al.
Pubblicazione: (2025)
FormalProofBench: Can Models Write Graduate Level Math Proofs That Are Formally Verified?
di: Ravi, Nikil, et al.
Pubblicazione: (2026)
di: Ravi, Nikil, et al.
Pubblicazione: (2026)
Deception Detection from Linguistic and Physiological Data Streams Using Bimodal Convolutional Neural Networks
di: Li, Panfeng, et al.
Pubblicazione: (2023)
di: Li, Panfeng, et al.
Pubblicazione: (2023)
Pooling Attention: Evaluating Pretrained Transformer Embeddings for Deception Classification
di: Mamtani, Sumit, et al.
Pubblicazione: (2025)
di: Mamtani, Sumit, et al.
Pubblicazione: (2025)
Efficient Real-time Refinement of Language Model Text Generation
di: Ko, Joonho, et al.
Pubblicazione: (2025)
di: Ko, Joonho, et al.
Pubblicazione: (2025)
Formal-LLM: Integrating Formal Language and Natural Language for Controllable LLM-based Agents
di: Li, Zelong, et al.
Pubblicazione: (2024)
di: Li, Zelong, et al.
Pubblicazione: (2024)
LexC-Gen: Generating Data for Extremely Low-Resource Languages with Large Language Models and Bilingual Lexicons
di: Yong, Zheng-Xin, et al.
Pubblicazione: (2024)
di: Yong, Zheng-Xin, et al.
Pubblicazione: (2024)
Lean-ing on Quality: How High-Quality Data Beats Diverse Multilingual Data in AutoFormalization
di: Chan, Willy, et al.
Pubblicazione: (2025)
di: Chan, Willy, et al.
Pubblicazione: (2025)
Linguistic Calibration of Long-Form Generations
di: Band, Neil, et al.
Pubblicazione: (2024)
di: Band, Neil, et al.
Pubblicazione: (2024)
Sectoral Coupling in Linguistic State Space
di: Dumbrava, Sebastian
Pubblicazione: (2025)
di: Dumbrava, Sebastian
Pubblicazione: (2025)
Scaling Data-Constrained Language Models
di: Muennighoff, Niklas, et al.
Pubblicazione: (2023)
di: Muennighoff, Niklas, et al.
Pubblicazione: (2023)
RWKU: Benchmarking Real-World Knowledge Unlearning for Large Language Models
di: Jin, Zhuoran, et al.
Pubblicazione: (2024)
di: Jin, Zhuoran, et al.
Pubblicazione: (2024)
RealCritic: Towards Effectiveness-Driven Evaluation of Language Model Critiques
di: Tang, Zhengyang, et al.
Pubblicazione: (2025)
di: Tang, Zhengyang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
SMART: Submodular Data Mixture Strategy for Instruction Tuning
di: Renduchintala, H S V N S Kowndinya, et al.
Pubblicazione: (2024) -
POSIX: A Prompt Sensitivity Index For Large Language Models
di: Chatterjee, Anwoy, et al.
Pubblicazione: (2024) -
On the Effect of Instruction Tuning Loss on Generalization
di: Chatterjee, Anwoy, et al.
Pubblicazione: (2025) -
Tied-Lora: Enhancing parameter efficiency of LoRA with weight tying
di: Renduchintala, Adithya, et al.
Pubblicazione: (2023) -
Between Circuits and Chomsky: Pre-pretraining on Formal Languages Imparts Linguistic Biases
di: Hu, Michael Y., et al.
Pubblicazione: (2025)