Synthetic Dataset for Evaluating Complex Compositional Knowledge for Natural Language Inference
Fuente:
arXiv
Guardado en:
| Autores principales: | Akoju, Sushma Anand, Vacareanu, Robert, Riaz, Haris, Blanco, Eduardo, Surdeanu, Mihai |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Best of Both Worlds: A Pliable and Generalizable Neuro-Symbolic Approach for Relation Classification
por: Vacareanu, Robert, et al.
Publicado: (2024)
por: Vacareanu, Robert, et al.
Publicado: (2024)
MorphNLI: A Stepwise Approach to Natural Language Inference Using Text Morphing
por: Negru, Vlad Andrei, et al.
Publicado: (2025)
por: Negru, Vlad Andrei, et al.
Publicado: (2025)
Towards Realistic Few-Shot Relation Extraction: A New Meta Dataset and Evaluation
por: Alam, Fahmida, et al.
Publicado: (2024)
por: Alam, Fahmida, et al.
Publicado: (2024)
Say Less, Mean More: Leveraging Pragmatics in Retrieval-Augmented Generation
por: Riaz, Haris, et al.
Publicado: (2025)
por: Riaz, Haris, et al.
Publicado: (2025)
From Words to Numbers: Your Large Language Model Is Secretly A Capable Regressor When Given In-Context Examples
por: Vacareanu, Robert, et al.
Publicado: (2024)
por: Vacareanu, Robert, et al.
Publicado: (2024)
ELLEN: Extremely Lightly Supervised Learning For Efficient Named Entity Recognition
por: Riaz, Haris, et al.
Publicado: (2024)
por: Riaz, Haris, et al.
Publicado: (2024)
Structured Semantic Information Helps Retrieve Better Examples for In-Context Learning Applied to Few-Shot Relation Extraction
por: Chakma, Aunabil, et al.
Publicado: (2026)
por: Chakma, Aunabil, et al.
Publicado: (2026)
Can LLMs Judge Debates? Evaluating Non-Linear Reasoning via Argumentation Theory Semantics
por: Sanayei, Reza, et al.
Publicado: (2025)
por: Sanayei, Reza, et al.
Publicado: (2025)
When and Where Did it Happen? An Encoder-Decoder Model to Identify Scenario Context
por: Noriega-Atala, Enrique, et al.
Publicado: (2024)
por: Noriega-Atala, Enrique, et al.
Publicado: (2024)
Fane at SemEval-2025 Task 10: Zero-Shot Entity Framing with Large Language Models
por: Fane, Enfa, et al.
Publicado: (2025)
por: Fane, Enfa, et al.
Publicado: (2025)
Data Contamination Quiz: A Tool to Detect and Estimate Contamination in Large Language Models
por: Golchin, Shahriar, et al.
Publicado: (2023)
por: Golchin, Shahriar, et al.
Publicado: (2023)
Time Travel in LLMs: Tracing Data Contamination in Large Language Models
por: Golchin, Shahriar, et al.
Publicado: (2023)
por: Golchin, Shahriar, et al.
Publicado: (2023)
Memorization in In-Context Learning
por: Golchin, Shahriar, et al.
Publicado: (2024)
por: Golchin, Shahriar, et al.
Publicado: (2024)
A Lightweight Explainable Guardrail for Prompt Safety
por: Islam, Md Asiful, et al.
Publicado: (2026)
por: Islam, Md Asiful, et al.
Publicado: (2026)
Bridging the Long-Tail Gap: Robust Retrieval-Augmented Relation Completion via Multi-Stage Paraphrase Infusion
por: Alam, Fahmida, et al.
Publicado: (2026)
por: Alam, Fahmida, et al.
Publicado: (2026)
Finding a Wolf in Sheep's Clothing: Combating Adversarial Text-To-Image Prompts with Text Summarization
por: Cooper, Portia, et al.
Publicado: (2024)
por: Cooper, Portia, et al.
Publicado: (2024)
MetaSynth: Meta-Prompting-Driven Agentic Scaffolds for Diverse Synthetic Data Generation
por: Riaz, Haris, et al.
Publicado: (2025)
por: Riaz, Haris, et al.
Publicado: (2025)
Evaluating Multilingual and Code-Switched Alignment in LLMs via Synthetic Natural Language Inference
por: Abdaljalil, Samir, et al.
Publicado: (2025)
por: Abdaljalil, Samir, et al.
Publicado: (2025)
Defining and Evaluating Decision and Composite Risk in Language Models Applied to Natural Language Inference
por: Shen, Ke, et al.
Publicado: (2024)
por: Shen, Ke, et al.
Publicado: (2024)
Peeking inside the Black-Box: Reinforcement Learning for Explainable and Accurate Relation Extraction
por: Guo, Xinyu, et al.
Publicado: (2025)
por: Guo, Xinyu, et al.
Publicado: (2025)
Let's CONFER: A Dataset for Evaluating Natural Language Inference Models on CONditional InFERence and Presupposition
por: Azin, Tara, et al.
Publicado: (2025)
por: Azin, Tara, et al.
Publicado: (2025)
Natural Language Inference Improves Compositionality in Vision-Language Models
por: Cascante-Bonilla, Paola, et al.
Publicado: (2024)
por: Cascante-Bonilla, Paola, et al.
Publicado: (2024)
Enhancing Transformer RNNs with Multiple Temporal Perspectives
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2024)
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2024)
BNLI: A Linguistically-Refined Bengali Dataset for Natural Language Inference
por: Haque, Farah Binta, et al.
Publicado: (2025)
por: Haque, Farah Binta, et al.
Publicado: (2025)
Synthetic Data Generation Using Large Language Models: Advances in Text and Code
por: Nadas, Mihai, et al.
Publicado: (2025)
por: Nadas, Mihai, et al.
Publicado: (2025)
MorphPiece : A Linguistic Tokenizer for Large Language Models
por: Jabbar, Haris
Publicado: (2023)
por: Jabbar, Haris
Publicado: (2023)
A New Benchmark Dataset and Mixture-of-Experts Language Models for Adversarial Natural Language Inference in Vietnamese
por: Van Huynh, Tin, et al.
Publicado: (2024)
por: Van Huynh, Tin, et al.
Publicado: (2024)
Filling the Gap: Is Commonsense Knowledge Generation useful for Natural Language Inference?
por: Jayaweera, Chathuri, et al.
Publicado: (2025)
por: Jayaweera, Chathuri, et al.
Publicado: (2025)
Multi-head attention debiasing and contrastive learning for mitigating Dataset Artifacts in Natural Language Inference
por: Sivakoti, Karthik
Publicado: (2024)
por: Sivakoti, Karthik
Publicado: (2024)
A Synthetic Dataset for Personal Attribute Inference
por: Yukhymenko, Hanna, et al.
Publicado: (2024)
por: Yukhymenko, Hanna, et al.
Publicado: (2024)
How Is LLM Reasoning Distracted by Irrelevant Context? An Analysis Using a Controlled Benchmark
por: Yang, Minglai, et al.
Publicado: (2025)
por: Yang, Minglai, et al.
Publicado: (2025)
CopySpec: Accelerating LLMs with Speculative Copy-and-Paste Without Compromising Quality
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2025)
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2025)
The Heap: A Contamination-Free Multilingual Code Dataset for Evaluating Large Language Models
por: Katzy, Jonathan, et al.
Publicado: (2025)
por: Katzy, Jonathan, et al.
Publicado: (2025)
AlignSAE: Concept-Aligned Sparse Autoencoders
por: Yang, Minglai, et al.
Publicado: (2025)
por: Yang, Minglai, et al.
Publicado: (2025)
On Reference (In-)Determinacy in Natural Language Inference
por: Chen, Sihao, et al.
Publicado: (2025)
por: Chen, Sihao, et al.
Publicado: (2025)
Quantum Natural Language Processing: A Comprehensive Review of Models, Methods, and Applications
por: Nausheen, Farha, et al.
Publicado: (2025)
por: Nausheen, Farha, et al.
Publicado: (2025)
Lost in Inference: Rediscovering the Role of Natural Language Inference for Large Language Models
por: Madaan, Lovish, et al.
Publicado: (2024)
por: Madaan, Lovish, et al.
Publicado: (2024)
Evaluating Large Language Models for Diacritic Restoration in Romanian Texts: A Comparative Study
por: Nadas, Mihai, et al.
Publicado: (2025)
por: Nadas, Mihai, et al.
Publicado: (2025)
Evaluating Gender Bias of Pre-trained Language Models in Natural Language Inference by Considering All Labels
por: Anantaprayoon, Panatchakorn, et al.
Publicado: (2023)
por: Anantaprayoon, Panatchakorn, et al.
Publicado: (2023)
Fine-Grained Natural Language Inference Based Faithfulness Evaluation for Diverse Summarisation Tasks
por: Zhang, Huajian, et al.
Publicado: (2024)
por: Zhang, Huajian, et al.
Publicado: (2024)
Ejemplares similares
-
Best of Both Worlds: A Pliable and Generalizable Neuro-Symbolic Approach for Relation Classification
por: Vacareanu, Robert, et al.
Publicado: (2024) -
MorphNLI: A Stepwise Approach to Natural Language Inference Using Text Morphing
por: Negru, Vlad Andrei, et al.
Publicado: (2025) -
Towards Realistic Few-Shot Relation Extraction: A New Meta Dataset and Evaluation
por: Alam, Fahmida, et al.
Publicado: (2024) -
Say Less, Mean More: Leveraging Pragmatics in Retrieval-Augmented Generation
por: Riaz, Haris, et al.
Publicado: (2025) -
From Words to Numbers: Your Large Language Model Is Secretly A Capable Regressor When Given In-Context Examples
por: Vacareanu, Robert, et al.
Publicado: (2024)