Understanding the role of FFNs in driving multilingual behaviour in LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Bhattacharya, Sunit, Bojar, Ondřej |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Multimodal Shannon Game with Images
di: Zouhar, Vilém, et al.
Pubblicazione: (2023)
di: Zouhar, Vilém, et al.
Pubblicazione: (2023)
Finetuning LLMs for EvaCun 2025 token prediction shared task
di: Jon, Josef, et al.
Pubblicazione: (2025)
di: Jon, Josef, et al.
Pubblicazione: (2025)
End-to-end Automatic Speech Recognition and Speech Translation: Integration of Speech Foundational Models and LLMs
di: Luu, Nam, et al.
Pubblicazione: (2025)
di: Luu, Nam, et al.
Pubblicazione: (2025)
Overview of the Sensemaking Task at the ELOQUENT 2025 Lab: LLMs as Teachers, Students and Evaluators
di: Šindelář, Pavel, et al.
Pubblicazione: (2025)
di: Šindelář, Pavel, et al.
Pubblicazione: (2025)
Prompting LLMs: Length Control for Isometric Machine Translation
di: Javorský, Dávid, et al.
Pubblicazione: (2025)
di: Javorský, Dávid, et al.
Pubblicazione: (2025)
Quality and Quantity of Machine Translation References for Automatic Metrics
di: Zouhar, Vilém, et al.
Pubblicazione: (2024)
di: Zouhar, Vilém, et al.
Pubblicazione: (2024)
Intrinsic vs. Extrinsic Evaluation of Czech Sentence Embeddings: Semantic Relevance Doesn't Help with MT Evaluation
di: Barančíková, Petra, et al.
Pubblicazione: (2025)
di: Barančíková, Petra, et al.
Pubblicazione: (2025)
Continuous Rating as Reliable Human Evaluation of Simultaneous Speech Translation
di: Javorský, Dávid, et al.
Pubblicazione: (2022)
di: Javorský, Dávid, et al.
Pubblicazione: (2022)
MockConf: A Student Interpretation Dataset: Analysis, Word- and Span-level Alignment and Baselines
di: Javorský, Dávid, et al.
Pubblicazione: (2025)
di: Javorský, Dávid, et al.
Pubblicazione: (2025)
ParCzech4Speech: A New Speech Corpus Derived from Czech Parliamentary Data
di: Stankov, Vladislav, et al.
Pubblicazione: (2025)
di: Stankov, Vladislav, et al.
Pubblicazione: (2025)
Long-Form End-to-End Speech Translation via Latent Alignment Segmentation
di: Polák, Peter, et al.
Pubblicazione: (2023)
di: Polák, Peter, et al.
Pubblicazione: (2023)
Evaluating Optimal Reference Translations
di: Zouhar, Vilém, et al.
Pubblicazione: (2023)
di: Zouhar, Vilém, et al.
Pubblicazione: (2023)
Better Late Than Never: Meta-Evaluation of Latency Metrics for Simultaneous Speech-to-Text Translation
di: Polák, Peter, et al.
Pubblicazione: (2025)
di: Polák, Peter, et al.
Pubblicazione: (2025)
Corpus of Cross-lingual Dialogues with Minutes and Detection of Misunderstandings
di: Čechovič, Marko, et al.
Pubblicazione: (2025)
di: Čechovič, Marko, et al.
Pubblicazione: (2025)
Training data generation for context-dependent rubric-based short answer grading
di: Šindelář, Pavel, et al.
Pubblicazione: (2026)
di: Šindelář, Pavel, et al.
Pubblicazione: (2026)
Findings of the Third Automatic Minuting (AutoMin) Challenge
di: Shinde, Kartik, et al.
Pubblicazione: (2025)
di: Shinde, Kartik, et al.
Pubblicazione: (2025)
SRS-Stories: Vocabulary-constrained multilingual story generation for language learning
di: Kamzela, Wiktor, et al.
Pubblicazione: (2025)
di: Kamzela, Wiktor, et al.
Pubblicazione: (2025)
ParamMute: Suppressing Knowledge-Critical FFNs for Faithful Retrieval-Augmented Generation
di: Huang, Pengcheng, et al.
Pubblicazione: (2025)
di: Huang, Pengcheng, et al.
Pubblicazione: (2025)
How "Real" is Your Real-Time Simultaneous Speech-to-Text Translation System?
di: Papi, Sara, et al.
Pubblicazione: (2024)
di: Papi, Sara, et al.
Pubblicazione: (2024)
Understanding the effects of language-specific class imbalance in multilingual fine-tuning
di: Jung, Vincent, et al.
Pubblicazione: (2024)
di: Jung, Vincent, et al.
Pubblicazione: (2024)
Scalable multilingual PII annotation for responsible AI in LLMs
di: Meena, Bharti, et al.
Pubblicazione: (2025)
di: Meena, Bharti, et al.
Pubblicazione: (2025)
AnimatedLLM: Explaining LLMs with Interactive Visualizations
di: Kasner, Zdeněk, et al.
Pubblicazione: (2025)
di: Kasner, Zdeněk, et al.
Pubblicazione: (2025)
Investigating the interaction of linguistic and mathematical reasoning in language models using multilingual number puzzles
di: Bhattacharya, Antara Raaghavi, et al.
Pubblicazione: (2025)
di: Bhattacharya, Antara Raaghavi, et al.
Pubblicazione: (2025)
Beyond Traditional Benchmarks: Analyzing Behaviors of Open LLMs on Data-to-Text Generation
di: Kasner, Zdeněk, et al.
Pubblicazione: (2024)
di: Kasner, Zdeněk, et al.
Pubblicazione: (2024)
Halluverse-M^3: A multitask multilingual benchmark for hallucination in LLMs
di: Abdaljalil, Samir, et al.
Pubblicazione: (2026)
di: Abdaljalil, Samir, et al.
Pubblicazione: (2026)
LLM Compression: How Far Can We Go in Balancing Size and Performance?
di: Sk, Sahil, et al.
Pubblicazione: (2025)
di: Sk, Sahil, et al.
Pubblicazione: (2025)
MultiLoKo: a multilingual local knowledge benchmark for LLMs spanning 31 languages
di: Hupkes, Dieuwke, et al.
Pubblicazione: (2025)
di: Hupkes, Dieuwke, et al.
Pubblicazione: (2025)
LLMs as Span Annotators: A Comparative Study of LLMs and Humans
di: Kasner, Zdeněk, et al.
Pubblicazione: (2025)
di: Kasner, Zdeněk, et al.
Pubblicazione: (2025)
Reasoning Gets Harder for LLMs Inside A Dialogue
di: Kartáč, Ivan, et al.
Pubblicazione: (2026)
di: Kartáč, Ivan, et al.
Pubblicazione: (2026)
OpeNLGauge: An Explainable Metric for NLG Evaluation with Open-Weights LLMs
di: Kartáč, Ivan, et al.
Pubblicazione: (2025)
di: Kartáč, Ivan, et al.
Pubblicazione: (2025)
Inference-Time Structural Reasoning for Compositional Vision-Language Understanding
di: Bhattacharya, Amartya
Pubblicazione: (2026)
di: Bhattacharya, Amartya
Pubblicazione: (2026)
Morphosyntactic probing of multilingual BERT models
di: Acs, Judit, et al.
Pubblicazione: (2023)
di: Acs, Judit, et al.
Pubblicazione: (2023)
Can LLMs reason over extended multilingual contexts? Towards long-context evaluation beyond retrieval and haystacks
di: Hengle, Amey, et al.
Pubblicazione: (2025)
di: Hengle, Amey, et al.
Pubblicazione: (2025)
Evaluating the IWSLT2023 Speech Translation Tasks: Human Annotations, Automatic Metrics, and Segmentation
di: Sperber, Matthias, et al.
Pubblicazione: (2024)
di: Sperber, Matthias, et al.
Pubblicazione: (2024)
MEEDAV: A Synchronous Web Viewer for EEG, Eye-Tracking and Speech Data
di: Pijálek, Jan, et al.
Pubblicazione: (2026)
di: Pijálek, Jan, et al.
Pubblicazione: (2026)
Preliminary WMT24 Ranking of General MT Systems and LLMs
di: Kocmi, Tom, et al.
Pubblicazione: (2024)
di: Kocmi, Tom, et al.
Pubblicazione: (2024)
Continuous sentiment scores for literary and multilingual contexts
di: Lyngbaek, Laurits, et al.
Pubblicazione: (2025)
di: Lyngbaek, Laurits, et al.
Pubblicazione: (2025)
A Gold Standard Dataset and Evaluation Framework for Depression Detection and Explanation in Social Media using LLMs
di: Bolegave, Prajval, et al.
Pubblicazione: (2025)
di: Bolegave, Prajval, et al.
Pubblicazione: (2025)
Retrieval-augmented generation in multilingual settings
di: Chirkova, Nadezhda, et al.
Pubblicazione: (2024)
di: Chirkova, Nadezhda, et al.
Pubblicazione: (2024)
Quantitative Assessment of Intersectional Empathetic Bias and Understanding
di: Formanek, Vojtech, et al.
Pubblicazione: (2024)
di: Formanek, Vojtech, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Multimodal Shannon Game with Images
di: Zouhar, Vilém, et al.
Pubblicazione: (2023) -
Finetuning LLMs for EvaCun 2025 token prediction shared task
di: Jon, Josef, et al.
Pubblicazione: (2025) -
End-to-end Automatic Speech Recognition and Speech Translation: Integration of Speech Foundational Models and LLMs
di: Luu, Nam, et al.
Pubblicazione: (2025) -
Overview of the Sensemaking Task at the ELOQUENT 2025 Lab: LLMs as Teachers, Students and Evaluators
di: Šindelář, Pavel, et al.
Pubblicazione: (2025) -
Prompting LLMs: Length Control for Isometric Machine Translation
di: Javorský, Dávid, et al.
Pubblicazione: (2025)