Do LLMs Understand Romanian Driving Laws? A Study on Multimodal and Fine-Tuned Question Answering
Fuente:
arXiv
Saved in:
| Main Authors: | Barbu, Eduard, Dumitran, Adrian Marius |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GRILE: A Benchmark for Grammar Reasoning and Explanation in Romanian LLMs
by: Dumitran, Adrian-Marius, et al.
Published: (2025)
by: Dumitran, Adrian-Marius, et al.
Published: (2025)
A Culturally-Rich Romanian NLP Dataset from "Who Wants to Be a Millionaire?" Videos
by: Ganea, Alexandru-Gabriel, et al.
Published: (2025)
by: Ganea, Alexandru-Gabriel, et al.
Published: (2025)
RoBiologyDataChoiceQA: A Romanian Dataset for improving Biology understanding of Large Language Models
by: Ghinea, Dragos-Dumitru, et al.
Published: (2025)
by: Ghinea, Dragos-Dumitru, et al.
Published: (2025)
Exploring Large Language Models for Translating Romanian Computational Problems into English
by: Dumitran, Adrian Marius, et al.
Published: (2025)
by: Dumitran, Adrian Marius, et al.
Published: (2025)
MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks
by: Marius, Dumitran Adrian, et al.
Published: (2025)
by: Marius, Dumitran Adrian, et al.
Published: (2025)
Do LLMs Understand Your Translations? Evaluating Paragraph-level MT with Question Answering
by: Fernandes, Patrick, et al.
Published: (2025)
by: Fernandes, Patrick, et al.
Published: (2025)
Fine-Tuning vs. RAG for Multi-Hop Question Answering with Novel Knowledge
by: Yang, Zhuoyi, et al.
Published: (2026)
by: Yang, Zhuoyi, et al.
Published: (2026)
On the Calibration of Multilingual Question Answering LLMs
by: Yang, Yahan, et al.
Published: (2023)
by: Yang, Yahan, et al.
Published: (2023)
A Large-Scale Benchmark for Evaluating Large Language Models on Medical Question Answering in Romanian
by: Rogoz, Ana-Cristina, et al.
Published: (2025)
by: Rogoz, Ana-Cristina, et al.
Published: (2025)
A Cross-Lingual Analysis of Bias in Large Language Models Using Romanian History
by: Cocu, Matei-Iulian, et al.
Published: (2025)
by: Cocu, Matei-Iulian, et al.
Published: (2025)
BacPrep: Lessons from Deploying an LLM-Based Bacalaureat Assessment Platform
by: Dumitran, Adrian-Marius, et al.
Published: (2025)
by: Dumitran, Adrian-Marius, et al.
Published: (2025)
Parameter Efficient Multimodal Instruction Tuning for Romanian Vision Language Models
by: Dima, George-Andrei, et al.
Published: (2025)
by: Dima, George-Andrei, et al.
Published: (2025)
Language Model Knowledge Distillation for Efficient Question Answering in Spanish
by: Bazaga, Adrián, et al.
Published: (2023)
by: Bazaga, Adrián, et al.
Published: (2023)
RoMathExam: A Longitudinal Dataset of Romanian Math Exams (1895-2025) with a Seven-Decade Core (1957-2025)
by: Cuclea, Luca-Ncolae, et al.
Published: (2026)
by: Cuclea, Luca-Ncolae, et al.
Published: (2026)
Fine-Tuning Large Language Models and Evaluating Retrieval Methods for Improved Question Answering on Building Codes
by: Aqib, Mohammad, et al.
Published: (2025)
by: Aqib, Mohammad, et al.
Published: (2025)
Towards Understanding Fine-Tuning Mechanisms of LLMs via Circuit Analysis
by: Wang, Xu, et al.
Published: (2025)
by: Wang, Xu, et al.
Published: (2025)
QEFT: Quantization for Efficient Fine-Tuning of LLMs
by: Lee, Changhun, et al.
Published: (2024)
by: Lee, Changhun, et al.
Published: (2024)
Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
by: Bakman, Yavuz, et al.
Published: (2025)
by: Bakman, Yavuz, et al.
Published: (2025)
Do LLMs Find Human Answers To Fact-Driven Questions Perplexing? A Case Study on Reddit
by: Seegmiller, Parker, et al.
Published: (2024)
by: Seegmiller, Parker, et al.
Published: (2024)
ALKAFI-LLAMA3: Fine-Tuning LLMs for Precise Legal Understanding in Palestine
by: Qasem, Rabee, et al.
Published: (2024)
by: Qasem, Rabee, et al.
Published: (2024)
Fine-tuning Multi-hop Question Answering with Hierarchical Graph Network
by: Xiong, Guanming
Published: (2020)
by: Xiong, Guanming
Published: (2020)
FRAMES-VQA: Benchmarking Fine-Tuning Robustness across Multi-Modal Shifts in Visual Question Answering
by: Huang, Chengyue, et al.
Published: (2025)
by: Huang, Chengyue, et al.
Published: (2025)
Simplify-This: A Comparative Analysis of Prompt-Based and Fine-Tuned LLMs
by: Cohen, Eilam, et al.
Published: (2026)
by: Cohen, Eilam, et al.
Published: (2026)
PocketLLM: Enabling On-Device Fine-Tuning for Personalized LLMs
by: Peng, Dan, et al.
Published: (2024)
by: Peng, Dan, et al.
Published: (2024)
Structured RAG for Answering Aggregative Questions
by: Koshorek, Omri, et al.
Published: (2025)
by: Koshorek, Omri, et al.
Published: (2025)
On Mechanistic Circuits for Extractive Question-Answering
by: Basu, Samyadeep, et al.
Published: (2025)
by: Basu, Samyadeep, et al.
Published: (2025)
ESQA: Event Sequences Question Answering
by: Abdullaeva, Irina, et al.
Published: (2024)
by: Abdullaeva, Irina, et al.
Published: (2024)
Retrieval Augmented Question Answering: When Should LLMs Admit Ignorance?
by: Wang, Dingmin, et al.
Published: (2025)
by: Wang, Dingmin, et al.
Published: (2025)
Towards Understanding the Fragility of Multilingual LLMs against Fine-Tuning Attacks
by: Poppi, Samuele, et al.
Published: (2024)
by: Poppi, Samuele, et al.
Published: (2024)
FoQA: A Faroese Question-Answering Dataset
by: Simonsen, Annika, et al.
Published: (2025)
by: Simonsen, Annika, et al.
Published: (2025)
SAFT: Structure-Aware Fine-Tuning of LLMs for AMR-to-Text Generation
by: Kamel, Rafiq, et al.
Published: (2025)
by: Kamel, Rafiq, et al.
Published: (2025)
L3Cube-IndicQuest: A Benchmark Question Answering Dataset for Evaluating Knowledge of LLMs in Indic Context
by: Rohera, Pritika, et al.
Published: (2024)
by: Rohera, Pritika, et al.
Published: (2024)
Retrieval-Augmented Generation for Domain-Specific Question Answering: A Case Study on Pittsburgh and CMU
by: Sun, Haojia, et al.
Published: (2024)
by: Sun, Haojia, et al.
Published: (2024)
Interpretable LLM-based Table Question Answering
by: Nguyen, Giang, et al.
Published: (2024)
by: Nguyen, Giang, et al.
Published: (2024)
Explainable Fact-checking through Question Answering
by: Yang, Jing, et al.
Published: (2021)
by: Yang, Jing, et al.
Published: (2021)
WikiMixQA: A Multimodal Benchmark for Question Answering over Tables and Charts
by: Foroutan, Negar, et al.
Published: (2025)
by: Foroutan, Negar, et al.
Published: (2025)
Comprehensive Modeling and Question Answering of Cancer Clinical Practice Guidelines using LLMs
by: Gupta, Bhumika, et al.
Published: (2025)
by: Gupta, Bhumika, et al.
Published: (2025)
A Novel Cartography-Based Curriculum Learning Method Applied on RoNLI: The First Romanian Natural Language Inference Corpus
by: Poesina, Eduard, et al.
Published: (2024)
by: Poesina, Eduard, et al.
Published: (2024)
Token Buncher: Shielding LLMs from Harmful Reinforcement Learning Fine-Tuning
by: Feng, Weitao, et al.
Published: (2025)
by: Feng, Weitao, et al.
Published: (2025)
OWLViz: An Open-World Benchmark for Visual Question Answering
by: Nguyen, Thuy, et al.
Published: (2025)
by: Nguyen, Thuy, et al.
Published: (2025)
Similar Items
-
GRILE: A Benchmark for Grammar Reasoning and Explanation in Romanian LLMs
by: Dumitran, Adrian-Marius, et al.
Published: (2025) -
A Culturally-Rich Romanian NLP Dataset from "Who Wants to Be a Millionaire?" Videos
by: Ganea, Alexandru-Gabriel, et al.
Published: (2025) -
RoBiologyDataChoiceQA: A Romanian Dataset for improving Biology understanding of Large Language Models
by: Ghinea, Dragos-Dumitru, et al.
Published: (2025) -
Exploring Large Language Models for Translating Romanian Computational Problems into English
by: Dumitran, Adrian Marius, et al.
Published: (2025) -
MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks
by: Marius, Dumitran Adrian, et al.
Published: (2025)