LLM-based Atomic Propositions help weak extractors: Evaluation of a Propositioner for triplet extraction
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Pommeret, Luc, Gerald, Thomas, Paroubek, Patrick, Ghannay, Sahar, Servan, Christophe, Rosset, Sophie |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Leveraging Information Retrieval to Enhance Spoken Language Understanding Prompts in Few-Shot Learning
par: Lepagnol, Pierre, et autres
Publié: (2025)
par: Lepagnol, Pierre, et autres
Publié: (2025)
New Semantic Task for the French Spoken Language Understanding MEDIA Benchmark
par: Alavoine, Nadège, et autres
Publié: (2024)
par: Alavoine, Nadège, et autres
Publié: (2024)
mALBERT: Is a Compact Multilingual BERT Model Still Worth It?
par: Servan, Christophe, et autres
Publié: (2024)
par: Servan, Christophe, et autres
Publié: (2024)
Small Language Models are Good Too: An Empirical Study of Zero-Shot Classification
par: Lepagnol, Pierre, et autres
Publié: (2024)
par: Lepagnol, Pierre, et autres
Publié: (2024)
Evaluation of Clinical Trials Reporting Quality using Large Language Models
par: Laï-king, Mathieu, et autres
Publié: (2025)
par: Laï-king, Mathieu, et autres
Publié: (2025)
Chitchat as Interference: Adding User Backstories to Task-Oriented Dialogues
par: Stricker, Armand, et autres
Publié: (2024)
par: Stricker, Armand, et autres
Publié: (2024)
A Unified Approach to Emotion Detection and Task-Oriented Dialogue Modeling
par: Stricker, Armand, et autres
Publié: (2024)
par: Stricker, Armand, et autres
Publié: (2024)
Enhancing Task-Oriented Dialogues with Chitchat: a Comparative Study Based on Lexical Diversity and Divergence
par: Stricker, Armand, et autres
Publié: (2023)
par: Stricker, Armand, et autres
Publié: (2023)
THIVLVC: Retrieval Augmented Dependency Parsing for Latin
par: Pommeret, Luc, et autres
Publié: (2026)
par: Pommeret, Luc, et autres
Publié: (2026)
Pre-training data selection for biomedical domain adaptation using journal impact metrics
par: Laï-king, Mathieu, et autres
Publié: (2024)
par: Laï-king, Mathieu, et autres
Publié: (2024)
A dual task learning approach to fine-tune a multilingual semantic speech encoder for Spoken Language Understanding
par: Laperrière, Gaëlle, et autres
Publié: (2024)
par: Laperrière, Gaëlle, et autres
Publié: (2024)
A Diversity Diet for a Healthier Model: A Case Study of French ModernBERT
par: Estève, Louis, et autres
Publié: (2026)
par: Estève, Louis, et autres
Publié: (2026)
Semantic enrichment towards efficient speech representations
par: Laperrière, Gaëlle, et autres
Publié: (2023)
par: Laperrière, Gaëlle, et autres
Publié: (2023)
A Benchmark Evaluation of Clinical Named Entity Recognition in French
par: Bannour, Nesrine, et autres
Publié: (2024)
par: Bannour, Nesrine, et autres
Publié: (2024)
A semantically enhanced dual encoder for aspect sentiment triplet extraction
par: Jiang, Baoxing, et autres
Publié: (2023)
par: Jiang, Baoxing, et autres
Publié: (2023)
Collaborative and Proactive Management of Task-Oriented Conversations
par: Saedi, Arezoo, et autres
Publié: (2025)
par: Saedi, Arezoo, et autres
Publié: (2025)
The Hawthorne Effect in Reasoning Models: Evaluating and Steering Test Awareness
par: Abdelnabi, Sahar, et autres
Publié: (2025)
par: Abdelnabi, Sahar, et autres
Publié: (2025)
Scalable and Domain-General Abstractive Proposition Segmentation
par: Hosseini, Mohammad Javad, et autres
Publié: (2024)
par: Hosseini, Mohammad Javad, et autres
Publié: (2024)
Valuable Hallucinations: Realizable Non-realistic Propositions
par: Chen, Qiucheng, et autres
Publié: (2025)
par: Chen, Qiucheng, et autres
Publié: (2025)
Not Worth Mentioning? A Pilot Study on Salient Proposition Annotation
par: Zeldes, Amir, et autres
Publié: (2026)
par: Zeldes, Amir, et autres
Publié: (2026)
Rethinking Atomic Decomposition for LLM Judges: A Prompt-Controlled Study of Reference-Grounded QA Evaluation
par: Zhang, Xinran
Publié: (2026)
par: Zhang, Xinran
Publié: (2026)
Benchmarking foundation models as feature extractors for weakly-supervised computational pathology
par: Neidlinger, Peter, et autres
Publié: (2024)
par: Neidlinger, Peter, et autres
Publié: (2024)
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
par: Hashemi, Helia, et autres
Publié: (2024)
par: Hashemi, Helia, et autres
Publié: (2024)
Rosetta-PL: Propositional Logic as a Benchmark for Large Language Model Reasoning
par: Baek, Shaun, et autres
Publié: (2025)
par: Baek, Shaun, et autres
Publié: (2025)
A Navigational Approach for Comprehensive RAG via Traversal over Proposition Graphs
par: Delmas, Maxime, et autres
Publié: (2026)
par: Delmas, Maxime, et autres
Publié: (2026)
ESGReveal: An LLM-based approach for extracting structured data from ESG reports
par: Zou, Yi, et autres
Publié: (2023)
par: Zou, Yi, et autres
Publié: (2023)
Monitoring Latent World States in Language Models with Propositional Probes
par: Feng, Jiahai, et autres
Publié: (2024)
par: Feng, Jiahai, et autres
Publié: (2024)
A Dataset for Pharmacovigilance in German, French, and Japanese: Annotating Adverse Drug Reactions across Languages
par: Raithel, Lisa, et autres
Publié: (2024)
par: Raithel, Lisa, et autres
Publié: (2024)
Are LLM-based Evaluators Confusing NLG Quality Criteria?
par: Hu, Xinyu, et autres
Publié: (2024)
par: Hu, Xinyu, et autres
Publié: (2024)
LLM-based NLG Evaluation: Current Status and Challenges
par: Gao, Mingqi, et autres
Publié: (2024)
par: Gao, Mingqi, et autres
Publié: (2024)
Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory
par: Sun, Jingwei, et autres
Publié: (2026)
par: Sun, Jingwei, et autres
Publié: (2026)
We Politely Insist: Your LLM Must Learn the Persian Art of Taarof
par: Sadr, Nikta Gohari, et autres
Publié: (2025)
par: Sadr, Nikta Gohari, et autres
Publié: (2025)
Cooperation, Competition, and Maliciousness: LLM-Stakeholders Interactive Negotiation
par: Abdelnabi, Sahar, et autres
Publié: (2023)
par: Abdelnabi, Sahar, et autres
Publié: (2023)
From Directions to Cones: Exploring Multidimensional Representations of Propositional Facts in LLMs
par: Yu, Stanley, et autres
Publié: (2025)
par: Yu, Stanley, et autres
Publié: (2025)
Scalable Multi-phase Word Embedding Using Conjunctive Propositional Clauses
par: Kadhim, Ahmed K., et autres
Publié: (2025)
par: Kadhim, Ahmed K., et autres
Publié: (2025)
LP-Eval: Rubric and Dataset for Measuring the Quality of Legal Proposition Generation
par: Xu, Shanshan, et autres
Publié: (2026)
par: Xu, Shanshan, et autres
Publié: (2026)
PropRAG: Guiding Retrieval with Beam Search over Proposition Paths
par: Wang, Jingjin, et autres
Publié: (2025)
par: Wang, Jingjin, et autres
Publié: (2025)
Editing Arbitrary Propositions in LLMs without Subject Labels
par: Feigenbaum, Itai, et autres
Publié: (2024)
par: Feigenbaum, Itai, et autres
Publié: (2024)
Evaluation of Oncotimia: An LLM based system for supporting tumour boards
par: Lorenzo, Luis, et autres
Publié: (2026)
par: Lorenzo, Luis, et autres
Publié: (2026)
Models That Know How Evaluations Are Designed Score Safer
par: Deckenbach, Katharina, et autres
Publié: (2026)
par: Deckenbach, Katharina, et autres
Publié: (2026)
Documents similaires
-
Leveraging Information Retrieval to Enhance Spoken Language Understanding Prompts in Few-Shot Learning
par: Lepagnol, Pierre, et autres
Publié: (2025) -
New Semantic Task for the French Spoken Language Understanding MEDIA Benchmark
par: Alavoine, Nadège, et autres
Publié: (2024) -
mALBERT: Is a Compact Multilingual BERT Model Still Worth It?
par: Servan, Christophe, et autres
Publié: (2024) -
Small Language Models are Good Too: An Empirical Study of Zero-Shot Classification
par: Lepagnol, Pierre, et autres
Publié: (2024) -
Evaluation of Clinical Trials Reporting Quality using Large Language Models
par: Laï-king, Mathieu, et autres
Publié: (2025)