Learning and communication pressures in neural networks: Lessons from emergent communication
Fuente:
arXiv
Saved in:
| Main Authors: | Galke, Lukas, Raviv, Limor |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
What makes a language easy to deep-learn? Deep neural networks and humans similarly benefit from compositional structure
by: Galke, Lukas, et al.
Published: (2023)
by: Galke, Lukas, et al.
Published: (2023)
Tokenization and Morphology in Multilingual Language Models: A Comparative Analysis of mT5 and ByT5
by: Dang, Thao Anh, et al.
Published: (2024)
by: Dang, Thao Anh, et al.
Published: (2024)
Not Everything That Counts Can Be Counted: A Case for Safe Qualitative AI
by: Beltoft, Stine, et al.
Published: (2025)
by: Beltoft, Stine, et al.
Published: (2025)
Training Language Models to Use Prolog as a Tool
by: Mellgren, Niklas, et al.
Published: (2025)
by: Mellgren, Niklas, et al.
Published: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
by: Ashuach, Tomer, et al.
Published: (2025)
by: Ashuach, Tomer, et al.
Published: (2025)
Low-resource neural machine translation with morphological modeling
by: Nzeyimana, Antoine
Published: (2024)
by: Nzeyimana, Antoine
Published: (2024)
The Provenance Gap in Clinical AI: Evidence-Traceable Temporal Knowledge Graphs for Rare Disease Reasoning
by: Ahmed, Md Shamim, et al.
Published: (2026)
by: Ahmed, Md Shamim, et al.
Published: (2026)
Calibrated Confidence Estimation for Tabular Question Answering
by: Voss, Lukas
Published: (2026)
by: Voss, Lukas
Published: (2026)
Select or Project? Evaluating Lower-dimensional Vectors for LLM Training Data Explanations
by: Hinterleitner, Lukas, et al.
Published: (2026)
by: Hinterleitner, Lukas, et al.
Published: (2026)
ChronoMedKG: A Temporally-Grounded Biomedical Knowledge Graph and Benchmark for Clinical Reasoning
by: Ahmed, Md Shamim, et al.
Published: (2026)
by: Ahmed, Md Shamim, et al.
Published: (2026)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
by: Saji, Alan, et al.
Published: (2025)
by: Saji, Alan, et al.
Published: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
by: Peters, Sydney, et al.
Published: (2025)
by: Peters, Sydney, et al.
Published: (2025)
Learning to Generate Structured Output with Schema Reinforcement Learning
by: Lu, Yaxi, et al.
Published: (2025)
by: Lu, Yaxi, et al.
Published: (2025)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
by: Collado-Montañez, Jaime, et al.
Published: (2025)
by: Collado-Montañez, Jaime, et al.
Published: (2025)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
by: Smădu, Răzvan-Alexandru, et al.
Published: (2025)
by: Smădu, Răzvan-Alexandru, et al.
Published: (2025)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
by: Oketunji, Abiodun Finbarrs
Published: (2023)
by: Oketunji, Abiodun Finbarrs
Published: (2023)
Learning Translations via Matrix Completion
by: Wijaya, Derry, et al.
Published: (2024)
by: Wijaya, Derry, et al.
Published: (2024)
Surprise Calibration for Better In-Context Learning
by: Tan, Zhihang, et al.
Published: (2025)
by: Tan, Zhihang, et al.
Published: (2025)
AI Can Learn Scientific Taste
by: Tong, Jingqi, et al.
Published: (2026)
by: Tong, Jingqi, et al.
Published: (2026)
Active Few-Shot Learning for Text Classification
by: Ahmadnia, Saeed, et al.
Published: (2025)
by: Ahmadnia, Saeed, et al.
Published: (2025)
Combining Denoising Autoencoders with Contrastive Learning to fine-tune Transformer Models
by: Lopez-Avila, Alejo, et al.
Published: (2024)
by: Lopez-Avila, Alejo, et al.
Published: (2024)
Efficient Few-shot Learning for Multi-label Classification of Scientific Documents with Many Classes
by: Schopf, Tim, et al.
Published: (2024)
by: Schopf, Tim, et al.
Published: (2024)
Lacuna Language Learning: Leveraging RNNs for Ranked Text Completion in Digitized Coptic Manuscripts
by: Levine, Lauren, et al.
Published: (2024)
by: Levine, Lauren, et al.
Published: (2024)
MIX : a Multi-task Learning Approach to Solve Open-Domain Question Answering
by: Chaybouti, Sofian, et al.
Published: (2020)
by: Chaybouti, Sofian, et al.
Published: (2020)
Have Multimodal Large Language Models (MLLMs) Really Learned to Tell the Time on Analog Clocks?
by: Fu, Tairan, et al.
Published: (2025)
by: Fu, Tairan, et al.
Published: (2025)
SITA: Learning Speaker-Invariant and Tone-Aware Speech Representations for Low-Resource Tonal Languages
by: Xu, Tianyi, et al.
Published: (2026)
by: Xu, Tianyi, et al.
Published: (2026)
Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents
by: Wang, Renxi, et al.
Published: (2024)
by: Wang, Renxi, et al.
Published: (2024)
Arabic Hate Speech Identification and Masking in Social Media using Deep Learning Models and Pre-trained Models Fine-tuning
by: Doghmash, Salam Thabet, et al.
Published: (2025)
by: Doghmash, Salam Thabet, et al.
Published: (2025)
Influence-driven Curriculum Learning for Pre-training on Limited Data
by: Schoenegger, Loris, et al.
Published: (2025)
by: Schoenegger, Loris, et al.
Published: (2025)
Graphemic Normalization of the Perso-Arabic Script
by: Doctor, Raiomond, et al.
Published: (2022)
by: Doctor, Raiomond, et al.
Published: (2022)
Beyond Arabic: Software for Perso-Arabic Script Manipulation
by: Gutkin, Alexander, et al.
Published: (2023)
by: Gutkin, Alexander, et al.
Published: (2023)
Distinguishing Ignorance from Error in LLM Hallucinations
by: Simhi, Adi, et al.
Published: (2024)
by: Simhi, Adi, et al.
Published: (2024)
VotIE: Information Extraction from Meeting Minutes
by: Evans, José Pedro, et al.
Published: (2026)
by: Evans, José Pedro, et al.
Published: (2026)
Synthia: Scalable Grounded Persona Generation from Social Media Data
by: Rahimzadeh, Vahid, et al.
Published: (2025)
by: Rahimzadeh, Vahid, et al.
Published: (2025)
A systematic review of relation extraction task since the emergence of Transformers
by: Celian, Ringwald, et al.
Published: (2025)
by: Celian, Ringwald, et al.
Published: (2025)
Extracting Structured Insights from Financial News: An Augmented LLM Driven Approach
by: Dolphin, Rian, et al.
Published: (2024)
by: Dolphin, Rian, et al.
Published: (2024)
Algorithm for Semantic Network Generation from Texts of Low Resource Languages Such as Kiswahili
by: Wanjawa, Barack Wamkaya, et al.
Published: (2025)
by: Wanjawa, Barack Wamkaya, et al.
Published: (2025)
Eyla: Toward an Identity-Anchored LLM Architecture with Integrated Biological Priors -- Vision, Implementation Attempt, and Lessons from AI-Assisted Development
by: Aditto, Arif
Published: (2026)
by: Aditto, Arif
Published: (2026)
RuOpinionNE-2024: Extraction of Opinion Tuples from Russian News Texts
by: Loukachevitch, Natalia, et al.
Published: (2025)
by: Loukachevitch, Natalia, et al.
Published: (2025)
KinyaColBERT: A Lexically Grounded Retrieval Model for Low-Resource Retrieval-Augmented Generation
by: Nzeyimana, Antoine, et al.
Published: (2025)
by: Nzeyimana, Antoine, et al.
Published: (2025)
Similar Items
-
What makes a language easy to deep-learn? Deep neural networks and humans similarly benefit from compositional structure
by: Galke, Lukas, et al.
Published: (2023) -
Tokenization and Morphology in Multilingual Language Models: A Comparative Analysis of mT5 and ByT5
by: Dang, Thao Anh, et al.
Published: (2024) -
Not Everything That Counts Can Be Counted: A Case for Safe Qualitative AI
by: Beltoft, Stine, et al.
Published: (2025) -
Training Language Models to Use Prolog as a Tool
by: Mellgren, Niklas, et al.
Published: (2025) -
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
by: Ashuach, Tomer, et al.
Published: (2025)