NALA_MAINZ at BLP-2025 Task 2: A Multi-agent Approach for Bangla Instruction to Python Code Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Saadi, Hossain Shaikh, Alam, Faria, Sanz-Guerrero, Mario, Bui, Minh Duc, Mager, Manuel, von der Wense, Katharina |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
JGU Mainz's Submission to the WMT25 Shared Task on LLMs with Limited Resources for Slavic Languages: MT and QA
por: Saadi, Hossain Shaikh, et al.
Publicado: (2025)
por: Saadi, Hossain Shaikh, et al.
Publicado: (2025)
Mind the Gap: A Closer Look at Tokenization for Multiple-Choice Question Answering with LLMs
por: Sanz-Guerrero, Mario, et al.
Publicado: (2025)
por: Sanz-Guerrero, Mario, et al.
Publicado: (2025)
From If-Statements to ML Pipelines: Revisiting Bias in Code-Generation
por: Bui, Minh Duc, et al.
Publicado: (2026)
por: Bui, Minh Duc, et al.
Publicado: (2026)
The Trade-off between Performance, Efficiency, and Fairness in Adapter Modules for Text Classification
por: Bui, Minh Duc, et al.
Publicado: (2024)
por: Bui, Minh Duc, et al.
Publicado: (2024)
Meenz bleibt Meenz, but Large Language Models Do Not Speak Its Dialect
por: Bui, Minh Duc, et al.
Publicado: (2026)
por: Bui, Minh Duc, et al.
Publicado: (2026)
Multi3Hate: Multimodal, Multilingual, and Multicultural Hate Speech Detection with Vision-Language Models
por: Bui, Minh Duc, et al.
Publicado: (2024)
por: Bui, Minh Duc, et al.
Publicado: (2024)
PyBangla at BLP-2025 Task 2: Enhancing Bangla-to-Python Code Generation with Iterative Self-Correction and Multilingual Agents
por: Islam, Jahidul, et al.
Publicado: (2025)
por: Islam, Jahidul, et al.
Publicado: (2025)
Mitigating Label Length Bias in Large Language Models
por: Sanz-Guerrero, Mario, et al.
Publicado: (2025)
por: Sanz-Guerrero, Mario, et al.
Publicado: (2025)
Corrective In-Context Learning: Evaluating Self-Correction in Large Language Models
por: Sanz-Guerrero, Mario, et al.
Publicado: (2025)
por: Sanz-Guerrero, Mario, et al.
Publicado: (2025)
Retriv at BLP-2025 Task 2: Test-Driven Feedback-Guided Framework for Bangla-to-Python Code Generation
por: Asib, K M Nafi, et al.
Publicado: (2025)
por: Asib, K M Nafi, et al.
Publicado: (2025)
Knowledge Distillation vs. Pretraining from Scratch under a Fixed (Computation) Budget
por: Bui, Minh Duc, et al.
Publicado: (2024)
por: Bui, Minh Duc, et al.
Publicado: (2024)
Large Language Models Discriminate Against Speakers of German Dialects
por: Bui, Minh Duc, et al.
Publicado: (2025)
por: Bui, Minh Duc, et al.
Publicado: (2025)
Asking Again and Again: Exploring LLM Robustness to Repeated Questions
por: Shaier, Sagi, et al.
Publicado: (2024)
por: Shaier, Sagi, et al.
Publicado: (2024)
On Generalization across Measurement Systems: LLMs Entail More Test-Time Compute for Underrepresented Cultures
por: Bui, Minh Duc, et al.
Publicado: (2025)
por: Bui, Minh Duc, et al.
Publicado: (2025)
Greater accessibility can amplify discrimination in generative AI
por: Holtermann, Carolin, et al.
Publicado: (2026)
por: Holtermann, Carolin, et al.
Publicado: (2026)
BLP-2023 Task 2: Sentiment Analysis
por: Hasan, Md. Arid, et al.
Publicado: (2023)
por: Hasan, Md. Arid, et al.
Publicado: (2023)
Improving Low-Resource Morphological Inflection via Self-Supervised Objectives
por: Wiemerslage, Adam, et al.
Publicado: (2025)
por: Wiemerslage, Adam, et al.
Publicado: (2025)
Retriv at BLP-2025 Task 1: A Transformer Ensemble and Multi-Task Learning Approach for Bangla Hate Speech Identification
por: Saha, Sourav, et al.
Publicado: (2025)
por: Saha, Sourav, et al.
Publicado: (2025)
Interdisciplinary Research in Conversation: A Case Study in Computational Morphology for Language Documentation
por: Rice, Enora, et al.
Publicado: (2025)
por: Rice, Enora, et al.
Publicado: (2025)
Model-Based Ranking of Source Languages for Zero-Shot Cross-Lingual Transfer
por: Ebrahimi, Abteen, et al.
Publicado: (2025)
por: Ebrahimi, Abteen, et al.
Publicado: (2025)
Implicitly Aligning Humans and Autonomous Agents through Shared Task Abstractions
por: Aroca-Ouellette, Stéphane, et al.
Publicado: (2025)
por: Aroca-Ouellette, Stéphane, et al.
Publicado: (2025)
NALA: an Effective and Interpretable Entity Alignment Method
por: Xu, Chuanhao, et al.
Publicado: (2024)
por: Xu, Chuanhao, et al.
Publicado: (2024)
Desiderata for the Context Use of Question Answering Systems
por: Shaier, Sagi, et al.
Publicado: (2024)
por: Shaier, Sagi, et al.
Publicado: (2024)
Who Are All The Stochastic Parrots Imitating? They Should Tell Us!
por: Shaier, Sagi, et al.
Publicado: (2023)
por: Shaier, Sagi, et al.
Publicado: (2023)
It Is Not About What You Say, It Is About How You Say It: A Surprisingly Simple Approach for Improving Reading Comprehension
por: Shaier, Sagi, et al.
Publicado: (2024)
por: Shaier, Sagi, et al.
Publicado: (2024)
CLIX: Cross-Lingual Explanations of Idiomatic Expressions
por: Gluck, Aaron, et al.
Publicado: (2025)
por: Gluck, Aaron, et al.
Publicado: (2025)
LLM-Based Multi-Task Bangla Hate Speech Detection: Type, Severity, and Target
por: Hasan, Md Arid, et al.
Publicado: (2025)
por: Hasan, Md Arid, et al.
Publicado: (2025)
Comparing Template-based and Template-free Language Model Probing
por: Shaier, Sagi, et al.
Publicado: (2024)
por: Shaier, Sagi, et al.
Publicado: (2024)
Gradient Masters at BLP-2025 Task 1: Advancing Low-Resource NLP for Bengali using Ensemble-Based Adversarial Training for Hate Speech Detection
por: Hoque, Syed Mohaiminul, et al.
Publicado: (2025)
por: Hoque, Syed Mohaiminul, et al.
Publicado: (2025)
How Good LLMs Are at Answering Bangla Medical Visual Questions? Dataset and Benchmarking
por: Ahmed, Rafid, et al.
Publicado: (2026)
por: Ahmed, Rafid, et al.
Publicado: (2026)
From Priest to Doctor: Domain Adaptation for Low-Resource Neural Machine Translation
por: Marashian, Ali, et al.
Publicado: (2024)
por: Marashian, Ali, et al.
Publicado: (2024)
Untangling the Influence of Typology, Data and Model Architecture on Ranking Transfer Languages for Cross-Lingual POS Tagging
por: Rice, Enora, et al.
Publicado: (2025)
por: Rice, Enora, et al.
Publicado: (2025)
TAMS: Translation-Assisted Morphological Segmentation
por: Rice, Enora, et al.
Publicado: (2024)
por: Rice, Enora, et al.
Publicado: (2024)
More Experts Than Galaxies: Conditionally-overlapping Experts With Biologically-Inspired Fixed Routing
por: Shaier, Sagi, et al.
Publicado: (2024)
por: Shaier, Sagi, et al.
Publicado: (2024)
Estimation of BLP models with high-dimensional controls
por: Jin, Hua
Publicado: (2026)
por: Jin, Hua
Publicado: (2026)
Measuring Contextual Informativeness in Child-Directed Text
por: Valentini, Maria, et al.
Publicado: (2024)
por: Valentini, Maria, et al.
Publicado: (2024)
Lost in the Middle, and In-Between: Enhancing Language Models' Ability to Reason Over Long Contexts in Multi-Hop QA
por: Baker, George Arthur, et al.
Publicado: (2024)
por: Baker, George Arthur, et al.
Publicado: (2024)
MALAMUTE: A Multilingual, Highly-granular, Template-free, Education-based Probing Dataset
por: Shaier, Sagi, et al.
Publicado: (2024)
por: Shaier, Sagi, et al.
Publicado: (2024)
BanglaForge: LLM Collaboration with Self-Refinement for Bangla Code Generation
por: Dihan, Mahir Labib, et al.
Publicado: (2025)
por: Dihan, Mahir Labib, et al.
Publicado: (2025)
Accelerating Bangla NLP Tasks with Automatic Mixed Precision: Resource-Efficient Training Preserving Model Efficacy
por: Opi, Md Mehrab Hossain, et al.
Publicado: (2025)
por: Opi, Md Mehrab Hossain, et al.
Publicado: (2025)
Ejemplares similares
-
JGU Mainz's Submission to the WMT25 Shared Task on LLMs with Limited Resources for Slavic Languages: MT and QA
por: Saadi, Hossain Shaikh, et al.
Publicado: (2025) -
Mind the Gap: A Closer Look at Tokenization for Multiple-Choice Question Answering with LLMs
por: Sanz-Guerrero, Mario, et al.
Publicado: (2025) -
From If-Statements to ML Pipelines: Revisiting Bias in Code-Generation
por: Bui, Minh Duc, et al.
Publicado: (2026) -
The Trade-off between Performance, Efficiency, and Fairness in Adapter Modules for Text Classification
por: Bui, Minh Duc, et al.
Publicado: (2024) -
Meenz bleibt Meenz, but Large Language Models Do Not Speak Its Dialect
por: Bui, Minh Duc, et al.
Publicado: (2026)