Human and Automatic Interpretation of Romanian Noun Compounds
Fuente:
arXiv
Salvato in:
| Autori principali: | Marinescu, Ioana, Fellbaum, Christiane |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Dependency Parsing Across the Resource Spectrum: Evaluating Architectures on High and Low-Resource Languages
di: Guan, Kevin, et al.
Pubblicazione: (2026)
di: Guan, Kevin, et al.
Pubblicazione: (2026)
Lugha-Llama: Adapting Large Language Models for African Languages
di: Buzaaba, Happy, et al.
Pubblicazione: (2025)
di: Buzaaba, Happy, et al.
Pubblicazione: (2025)
Automatic Replication of LLM Mistakes in Medical Conversations
di: Proniakin, Oleksii, et al.
Pubblicazione: (2025)
di: Proniakin, Oleksii, et al.
Pubblicazione: (2025)
Supernova Event Dataset: Interpreting Large Language Models' Personality through Critical Event Analysis
di: Agarwal, Pranav, et al.
Pubblicazione: (2025)
di: Agarwal, Pranav, et al.
Pubblicazione: (2025)
Interpreting LLM-as-a-Judge Policies via Verifiable Global Explanations
di: Gajcin, Jasmina, et al.
Pubblicazione: (2025)
di: Gajcin, Jasmina, et al.
Pubblicazione: (2025)
The Role of Handling Attributive Nouns in Improving Chinese-To-English Machine Translation
di: Wang, Lisa, et al.
Pubblicazione: (2024)
di: Wang, Lisa, et al.
Pubblicazione: (2024)
Evaluating Adjective-Noun Compositionality in LLMs: Functional vs Representational Perspectives
di: Dhar, Ruchira, et al.
Pubblicazione: (2026)
di: Dhar, Ruchira, et al.
Pubblicazione: (2026)
RoMath: A Mathematical Reasoning Benchmark in Romanian
di: Cosma, Adrian, et al.
Pubblicazione: (2024)
di: Cosma, Adrian, et al.
Pubblicazione: (2024)
Improving Legal Judgement Prediction in Romanian with Long Text Encoders
di: Masala, Mihai, et al.
Pubblicazione: (2024)
di: Masala, Mihai, et al.
Pubblicazione: (2024)
Information Flow Routes: Automatically Interpreting Language Models at Scale
di: Ferrando, Javier, et al.
Pubblicazione: (2024)
di: Ferrando, Javier, et al.
Pubblicazione: (2024)
A Cross-Lingual Analysis of Bias in Large Language Models Using Romanian History
di: Cocu, Matei-Iulian, et al.
Pubblicazione: (2025)
di: Cocu, Matei-Iulian, et al.
Pubblicazione: (2025)
AutoMetrics: Approximate Human Judgements with Automatically Generated Evaluators
di: Ryan, Michael J., et al.
Pubblicazione: (2025)
di: Ryan, Michael J., et al.
Pubblicazione: (2025)
Large Language Models Badly Generalize across Option Length, Problem Types, and Irrelevant Noun Replacements
di: Zhao, Guangxiang, et al.
Pubblicazione: (2025)
di: Zhao, Guangxiang, et al.
Pubblicazione: (2025)
A LLM-Powered Automatic Grading Framework with Human-Level Guidelines Optimization
di: Chu, Yucheng, et al.
Pubblicazione: (2024)
di: Chu, Yucheng, et al.
Pubblicazione: (2024)
Automatic Analysis of Collaboration Through Human Conversational Data Resources: A Review
di: Yu, Yi, et al.
Pubblicazione: (2026)
di: Yu, Yi, et al.
Pubblicazione: (2026)
Comparing Human and Large Language Model Interpretation of Implicit Information
di: De Santis, Antonio, et al.
Pubblicazione: (2026)
di: De Santis, Antonio, et al.
Pubblicazione: (2026)
Don't Believe Everything You Read: Enhancing Summarization Interpretability through Automatic Identification of Hallucinations in Large Language Models
di: Vakharia, Priyesh, et al.
Pubblicazione: (2023)
di: Vakharia, Priyesh, et al.
Pubblicazione: (2023)
Improving Statistical Significance in Human Evaluation of Automatic Metrics via Soft Pairwise Accuracy
di: Thompson, Brian, et al.
Pubblicazione: (2024)
di: Thompson, Brian, et al.
Pubblicazione: (2024)
Summarization Metrics for Spanish and Basque: Do Automatic Scores and LLM-Judges Correlate with Humans?
di: Barnes, Jeremy, et al.
Pubblicazione: (2025)
di: Barnes, Jeremy, et al.
Pubblicazione: (2025)
What's In My Human Feedback? Learning Interpretable Descriptions of Preference Data
di: Movva, Rajiv, et al.
Pubblicazione: (2025)
di: Movva, Rajiv, et al.
Pubblicazione: (2025)
Parameter Efficient Multimodal Instruction Tuning for Romanian Vision Language Models
di: Dima, George-Andrei, et al.
Pubblicazione: (2025)
di: Dima, George-Andrei, et al.
Pubblicazione: (2025)
SIMBA UQ: Similarity-Based Aggregation for Uncertainty Quantification in Large Language Models
di: Bhattacharjya, Debarun, et al.
Pubblicazione: (2025)
di: Bhattacharjya, Debarun, et al.
Pubblicazione: (2025)
Automatic Control With Human-Like Reasoning: Exploring Language Model Embodied Air Traffic Agents
di: Andriuškevičius, Justas, et al.
Pubblicazione: (2024)
di: Andriuškevičius, Justas, et al.
Pubblicazione: (2024)
Differentiating Between Human-Written and AI-Generated Texts Using Automatically Extracted Linguistic Features
di: Georgiou, Georgios P.
Pubblicazione: (2024)
di: Georgiou, Georgios P.
Pubblicazione: (2024)
Building Large-Scale English-Romanian Literary Translation Resources with Open Models
di: Nadas, Mihai, et al.
Pubblicazione: (2025)
di: Nadas, Mihai, et al.
Pubblicazione: (2025)
MobEvolve: An Agentic Self-Evolving Heuristic System for Interpretable Human Mobility Generation
di: He, Junlin, et al.
Pubblicazione: (2026)
di: He, Junlin, et al.
Pubblicazione: (2026)
Are Bias Evaluation Methods Biased ?
di: Berrayana, Lina, et al.
Pubblicazione: (2025)
di: Berrayana, Lina, et al.
Pubblicazione: (2025)
PoPreRo: A New Dataset for Popularity Prediction of Romanian Reddit Posts
di: Rogoz, Ana-Cristina, et al.
Pubblicazione: (2024)
di: Rogoz, Ana-Cristina, et al.
Pubblicazione: (2024)
RoLegalGEC: Legal Domain Grammatical Error Detection and Correction Dataset for Romanian
di: Timpuriu, Mircea, et al.
Pubblicazione: (2026)
di: Timpuriu, Mircea, et al.
Pubblicazione: (2026)
FactReasoner: A Probabilistic Approach to Long-Form Factuality Assessment for Large Language Models
di: Marinescu, Radu, et al.
Pubblicazione: (2025)
di: Marinescu, Radu, et al.
Pubblicazione: (2025)
A Large-Scale Benchmark for Evaluating Large Language Models on Medical Question Answering in Romanian
di: Rogoz, Ana-Cristina, et al.
Pubblicazione: (2025)
di: Rogoz, Ana-Cristina, et al.
Pubblicazione: (2025)
RoMathExam: A Longitudinal Dataset of Romanian Math Exams (1895-2025) with a Seven-Decade Core (1957-2025)
di: Cuclea, Luca-Ncolae, et al.
Pubblicazione: (2026)
di: Cuclea, Luca-Ncolae, et al.
Pubblicazione: (2026)
Re-evaluating Automatic LLM System Ranking for Alignment with Human Preference
di: Gao, Mingqi, et al.
Pubblicazione: (2024)
di: Gao, Mingqi, et al.
Pubblicazione: (2024)
Automatic Summarization of Long Documents
di: Chhibbar, Naman, et al.
Pubblicazione: (2024)
di: Chhibbar, Naman, et al.
Pubblicazione: (2024)
Convergences and Divergences between Automatic Assessment and Human Evaluation: Insights from Comparing ChatGPT-Generated Translation and Neural Machine Translation
di: Jiang, Zhaokun, et al.
Pubblicazione: (2024)
di: Jiang, Zhaokun, et al.
Pubblicazione: (2024)
FactCorrector: A Graph-Inspired Approach to Long-Form Factuality Correction of Large Language Models
di: Carnerero-Cano, Javier, et al.
Pubblicazione: (2026)
di: Carnerero-Cano, Javier, et al.
Pubblicazione: (2026)
An Automatic Question Usability Evaluation Toolkit
di: Moore, Steven, et al.
Pubblicazione: (2024)
di: Moore, Steven, et al.
Pubblicazione: (2024)
Automatic Legal Writing Evaluation of LLMs
di: Pires, Ramon, et al.
Pubblicazione: (2025)
di: Pires, Ramon, et al.
Pubblicazione: (2025)
Robust Planning with Compound LLM Architectures: An LLM-Modulo Approach
di: Gundawar, Atharva, et al.
Pubblicazione: (2024)
di: Gundawar, Atharva, et al.
Pubblicazione: (2024)
A System for Automatic English Text Expansion
di: Méndez, Silvia García, et al.
Pubblicazione: (2024)
di: Méndez, Silvia García, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Dependency Parsing Across the Resource Spectrum: Evaluating Architectures on High and Low-Resource Languages
di: Guan, Kevin, et al.
Pubblicazione: (2026) -
Lugha-Llama: Adapting Large Language Models for African Languages
di: Buzaaba, Happy, et al.
Pubblicazione: (2025) -
Automatic Replication of LLM Mistakes in Medical Conversations
di: Proniakin, Oleksii, et al.
Pubblicazione: (2025) -
Supernova Event Dataset: Interpreting Large Language Models' Personality through Critical Event Analysis
di: Agarwal, Pranav, et al.
Pubblicazione: (2025) -
Interpreting LLM-as-a-Judge Policies via Verifiable Global Explanations
di: Gajcin, Jasmina, et al.
Pubblicazione: (2025)