Statement-Tuning Enables Efficient Cross-lingual Generalization in Encoder-only Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Elshabrawy, Ahmed, Nguyen, Thanh-Nhi, Kang, Yeeun, Feng, Lihan, Jain, Annant, Shaikh, Faadil Abdullah, Mansurov, Jonibek, Imam, Mohamed Fazli Mohamed, Ortiz-Barajas, Jesus-German, Chevi, Rendi, Aji, Alham Fikri |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Daisy-TTS: Simulating Wider Spectrum of Emotions via Prosody Embedding Decomposition
por: Chevi, Rendi, et al.
Publicado: (2024)
por: Chevi, Rendi, et al.
Publicado: (2024)
Data Laundering: Artificially Boosting Benchmark Results through Knowledge Distillation
por: Mansurov, Jonibek, et al.
Publicado: (2024)
por: Mansurov, Jonibek, et al.
Publicado: (2024)
Enabling Natural Zero-Shot Prompting on Encoder Models via Statement-Tuning
por: Elshabrawy, Ahmed, et al.
Publicado: (2024)
por: Elshabrawy, Ahmed, et al.
Publicado: (2024)
How Individual Traits and Language Styles Shape Preferences In Open-ended User-LLM Interaction: A Preliminary Study
por: Chevi, Rendi, et al.
Publicado: (2025)
por: Chevi, Rendi, et al.
Publicado: (2025)
Can Multimodal LLMs do Visual Temporal Understanding and Reasoning? The answer is No!
por: Imam, Mohamed Fazli, et al.
Publicado: (2025)
por: Imam, Mohamed Fazli, et al.
Publicado: (2025)
CLIP meets DINO for Tuning Zero-Shot Classifier using Unlabeled Image Collections
por: Imam, Mohamed Fazli, et al.
Publicado: (2024)
por: Imam, Mohamed Fazli, et al.
Publicado: (2024)
LoraxBench: A Multitask, Multilingual Benchmark Suite for 20 Indonesian Languages
por: Aji, Alham Fikri, et al.
Publicado: (2025)
por: Aji, Alham Fikri, et al.
Publicado: (2025)
Improving Low-Resource Machine Translation via Round-Trip Reinforcement Learning
por: Attia, Ahmed, et al.
Publicado: (2026)
por: Attia, Ahmed, et al.
Publicado: (2026)
When Alignment Hurts: Decoupling Representational Spaces in Multilingual Models
por: Elshabrawy, Ahmed, et al.
Publicado: (2025)
por: Elshabrawy, Ahmed, et al.
Publicado: (2025)
M4: Multi-generator, Multi-domain, and Multi-lingual Black-Box Machine-Generated Text Detection
por: Wang, Yuxia, et al.
Publicado: (2023)
por: Wang, Yuxia, et al.
Publicado: (2023)
Balanced Multi-Factor In-Context Learning for Multilingual Large Language Models
por: Kaneko, Masahiro, et al.
Publicado: (2025)
por: Kaneko, Masahiro, et al.
Publicado: (2025)
Sense Representations Are Inducible Interfaces
por: Cruz, Jan Christian Blaise, et al.
Publicado: (2026)
por: Cruz, Jan Christian Blaise, et al.
Publicado: (2026)
LLM Olympiad: Why Model Evaluation Needs a Sealed Exam
por: Cruz, Jan Christian Blaise, et al.
Publicado: (2026)
por: Cruz, Jan Christian Blaise, et al.
Publicado: (2026)
Extracting General-use Transformers for Low-resource Languages via Knowledge Distillation
por: Cruz, Jan Christian Blaise, et al.
Publicado: (2025)
por: Cruz, Jan Christian Blaise, et al.
Publicado: (2025)
Beyond Probabilities: Unveiling the Misalignment in Evaluating Large Language Models
por: Lyu, Chenyang, et al.
Publicado: (2024)
por: Lyu, Chenyang, et al.
Publicado: (2024)
MOMENTS: A Comprehensive Multimodal Benchmark for Theory of Mind
por: Villa-Cueva, Emilio, et al.
Publicado: (2025)
por: Villa-Cueva, Emilio, et al.
Publicado: (2025)
Language-Specific Latent Process Hinders Cross-Lingual Performance
por: Lim, Zheng Wei, et al.
Publicado: (2025)
por: Lim, Zheng Wei, et al.
Publicado: (2025)
The Privileged Students: On the Value of Initialization in Multilingual Knowledge Distillation
por: Wibowo, Haryo Akbarianto, et al.
Publicado: (2024)
por: Wibowo, Haryo Akbarianto, et al.
Publicado: (2024)
Crosslingual Reasoning through Test-Time Scaling
por: Yong, Zheng-Xin, et al.
Publicado: (2025)
por: Yong, Zheng-Xin, et al.
Publicado: (2025)
Efficient and Interpretable Grammatical Error Correction with Mixture of Experts
por: Qorib, Muhammad Reza, et al.
Publicado: (2024)
por: Qorib, Muhammad Reza, et al.
Publicado: (2024)
Multilinguality as Sense Adaptation
por: Cruz, Jan Christian Blaise, et al.
Publicado: (2026)
por: Cruz, Jan Christian Blaise, et al.
Publicado: (2026)
TextGames: Learning to Self-Play Text-Based Puzzle Games via Language Model Reasoning
por: Hudi, Frederikus, et al.
Publicado: (2025)
por: Hudi, Frederikus, et al.
Publicado: (2025)
Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation
por: Nur'aini, Khumaisa, et al.
Publicado: (2026)
por: Nur'aini, Khumaisa, et al.
Publicado: (2026)
Predicting the Order of Upcoming Tokens Improves Language Modeling
por: Zuhri, Zayd M. K., et al.
Publicado: (2025)
por: Zuhri, Zayd M. K., et al.
Publicado: (2025)
Multicultural Spyfall: Assessing LLMs through Dynamic Multilingual Social Deduction Game
por: Wibowo, Haryo Akbarianto, et al.
Publicado: (2026)
por: Wibowo, Haryo Akbarianto, et al.
Publicado: (2026)
Softpick: No Attention Sink, No Massive Activations with Rectified Softmax
por: Zuhri, Zayd M. K., et al.
Publicado: (2025)
por: Zuhri, Zayd M. K., et al.
Publicado: (2025)
KazMMLU: Evaluating Language Models on Kazakh, Russian, and Regional Knowledge of Kazakhstan
por: Togmanov, Mukhammed, et al.
Publicado: (2025)
por: Togmanov, Mukhammed, et al.
Publicado: (2025)
M4GT-Bench: Evaluation Benchmark for Black-Box Machine-Generated Text Detection
por: Wang, Yuxia, et al.
Publicado: (2024)
por: Wang, Yuxia, et al.
Publicado: (2024)
From Surveys to Narratives: Rethinking Cultural Value Adaptation in LLMs
por: Adilazuarda, Muhammad Farid, et al.
Publicado: (2025)
por: Adilazuarda, Muhammad Farid, et al.
Publicado: (2025)
LaMini-LM: A Diverse Herd of Distilled Models from Large-Scale Instructions
por: Wu, Minghao, et al.
Publicado: (2023)
por: Wu, Minghao, et al.
Publicado: (2023)
SEA-SafeguardBench: Evaluating AI Safety in SEA Languages and Cultures
por: Tasawong, Panuthep, et al.
Publicado: (2025)
por: Tasawong, Panuthep, et al.
Publicado: (2025)
SEA-Guard: Culturally Grounded Multilingual Safeguard for Southeast Asia
por: Tasawong, Panuthep, et al.
Publicado: (2026)
por: Tasawong, Panuthep, et al.
Publicado: (2026)
MLKV: Multi-Layer Key-Value Heads for Memory Efficient Transformer Decoding
por: Zuhri, Zayd Muhammad Kawakibi, et al.
Publicado: (2024)
por: Zuhri, Zayd Muhammad Kawakibi, et al.
Publicado: (2024)
SemEval-2024 Task 8: Multidomain, Multimodel and Multilingual Machine-Generated Text Detection
por: Wang, Yuxia, et al.
Publicado: (2024)
por: Wang, Yuxia, et al.
Publicado: (2024)
LinguAlchemy: Fusing Typological and Geographical Elements for Unseen Language Generalization
por: Adilazuarda, Muhammad Farid, et al.
Publicado: (2024)
por: Adilazuarda, Muhammad Farid, et al.
Publicado: (2024)
LinguDistill: Recovering Linguistic Ability in Vision-Language Models via Selective Cross-Modal Distillation
por: Irawan, Patrick Amadeus, et al.
Publicado: (2026)
por: Irawan, Patrick Amadeus, et al.
Publicado: (2026)
QLESS: A Quantized Approach for Data Valuation and Selection in Large Language Model Fine-Tuning
por: Ananta, Moses, et al.
Publicado: (2025)
por: Ananta, Moses, et al.
Publicado: (2025)
Cultural Conditioning or Placebo? On the Effectiveness of Socio-Demographic Prompting
por: Mukherjee, Sagnik, et al.
Publicado: (2024)
por: Mukherjee, Sagnik, et al.
Publicado: (2024)
IteRABRe: Iterative Recovery-Aided Block Reduction
por: Wibowo, Haryo Akbarianto, et al.
Publicado: (2025)
por: Wibowo, Haryo Akbarianto, et al.
Publicado: (2025)
Conservación en la frontera : La conservación transfronteriza ser una parte importante de los esfuerzos para la conservación de los bosques tropicales en el siglo XXI
por: Imam Bakarr, Mohamed
Publicado: (1998)
por: Imam Bakarr, Mohamed
Publicado: (1998)
Ejemplares similares
-
Daisy-TTS: Simulating Wider Spectrum of Emotions via Prosody Embedding Decomposition
por: Chevi, Rendi, et al.
Publicado: (2024) -
Data Laundering: Artificially Boosting Benchmark Results through Knowledge Distillation
por: Mansurov, Jonibek, et al.
Publicado: (2024) -
Enabling Natural Zero-Shot Prompting on Encoder Models via Statement-Tuning
por: Elshabrawy, Ahmed, et al.
Publicado: (2024) -
How Individual Traits and Language Styles Shape Preferences In Open-ended User-LLM Interaction: A Preliminary Study
por: Chevi, Rendi, et al.
Publicado: (2025) -
Can Multimodal LLMs do Visual Temporal Understanding and Reasoning? The answer is No!
por: Imam, Mohamed Fazli, et al.
Publicado: (2025)