Evidence Absence Is Not Evidence Insufficiency: Diagnosing NEI Construction Artifacts in Fact Verification
Fuente:
arXiv
Guardado en:
| Autores principales: | Qiu, Jingxi, Han, Zeyu, Huang, Cheng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SURE-RAG: Sufficiency and Uncertainty-Aware Evidence Verification for Selective Retrieval-Augmented Generation
por: Qiu, Jingxi, et al.
Publicado: (2026)
por: Qiu, Jingxi, et al.
Publicado: (2026)
SemLink: A Semantic-Aware Automated Test Oracle for Hyperlink Verification using Siamese Sentence-BERT
por: Yang, Guan-Yan, et al.
Publicado: (2026)
por: Yang, Guan-Yan, et al.
Publicado: (2026)
Assessing the Ability of ChatGPT to Screen Articles for Systematic Reviews
por: Syriani, Eugene, et al.
Publicado: (2023)
por: Syriani, Eugene, et al.
Publicado: (2023)
Toward building next-generation Geocoding systems: a systematic review
por: Yin, Zhengcong, et al.
Publicado: (2025)
por: Yin, Zhengcong, et al.
Publicado: (2025)
CodeRAG: Finding Relevant and Necessary Knowledge for Retrieval-Augmented Repository-Level Code Completion
por: Zhang, Sheng, et al.
Publicado: (2025)
por: Zhang, Sheng, et al.
Publicado: (2025)
Iterative Self-Training for Code Generation via Reinforced Re-Ranking
por: Sorokin, Nikita, et al.
Publicado: (2025)
por: Sorokin, Nikita, et al.
Publicado: (2025)
Selective Shot Learning for Code Explanation
por: Bhattacharya, Paheli, et al.
Publicado: (2024)
por: Bhattacharya, Paheli, et al.
Publicado: (2024)
ProCQA: A Large-scale Community-based Programming Question Answering Dataset for Code Search
por: Li, Zehan, et al.
Publicado: (2024)
por: Li, Zehan, et al.
Publicado: (2024)
SaraCoder: Orchestrating Semantic and Structural Cues for Resource-Optimized Repository-Level Code Completion
por: Chen, Xiaohan, et al.
Publicado: (2025)
por: Chen, Xiaohan, et al.
Publicado: (2025)
CodeKGC: Code Language Model for Generative Knowledge Graph Construction
por: Bi, Zhen, et al.
Publicado: (2023)
por: Bi, Zhen, et al.
Publicado: (2023)
When "Better" Prompts Hurt: Evaluation-Driven Iteration for LLM Applications
por: Commey, Daniel
Publicado: (2026)
por: Commey, Daniel
Publicado: (2026)
Towards AI Evaluation in Domain-Specific RAG Systems: The AgriHubi Case Study
por: Hasan, Md. Toufique, et al.
Publicado: (2026)
por: Hasan, Md. Toufique, et al.
Publicado: (2026)
ORFuzz: Fuzzing the "Other Side" of LLM Safety -- Testing Over-Refusal
por: Zhang, Haonan, et al.
Publicado: (2025)
por: Zhang, Haonan, et al.
Publicado: (2025)
Studying and Recommending Information Highlighting in Stack Overflow Answers
por: Ahmed, Shahla Shaan, et al.
Publicado: (2024)
por: Ahmed, Shahla Shaan, et al.
Publicado: (2024)
Rewriting the Code: A Simple Method for Large Language Model Augmented Code Search
por: Li, Haochen, et al.
Publicado: (2024)
por: Li, Haochen, et al.
Publicado: (2024)
Domain-Specific Retrieval-Augmented Generation Using Vector Stores, Knowledge Graphs, and Tensor Factorization
por: Barron, Ryan C., et al.
Publicado: (2024)
por: Barron, Ryan C., et al.
Publicado: (2024)
LLM Agents Improve Semantic Code Search
por: Jain, Sarthak, et al.
Publicado: (2024)
por: Jain, Sarthak, et al.
Publicado: (2024)
cAST: Enhancing Code Retrieval-Augmented Generation with Structural Chunking via Abstract Syntax Tree
por: Zhang, Yilin, et al.
Publicado: (2025)
por: Zhang, Yilin, et al.
Publicado: (2025)
Embedding-based search in JetBrains IDEs
por: Abramov, Evgeny, et al.
Publicado: (2024)
por: Abramov, Evgeny, et al.
Publicado: (2024)
The Invisible Hand of AI Libraries Shaping Open Source Projects and Communities
por: Esposito, Matteo, et al.
Publicado: (2026)
por: Esposito, Matteo, et al.
Publicado: (2026)
Automating Database-Native Function Code Synthesis with LLMs
por: Zhou, Wei, et al.
Publicado: (2026)
por: Zhou, Wei, et al.
Publicado: (2026)
Credible, Unreliable or Leaked?: Evidence Verification for Enhanced Automated Fact-checking
por: Chrysidis, Zacharias, et al.
Publicado: (2024)
por: Chrysidis, Zacharias, et al.
Publicado: (2024)
GNN-Coder: Boosting Semantic Code Retrieval with Combined GNNs and Transformer
por: Ye, Yufan, et al.
Publicado: (2025)
por: Ye, Yufan, et al.
Publicado: (2025)
Automating Categorization of Scientific Texts with In-Context Learning and Prompt-Chaining in Large Language Models
por: Shahi, Gautam Kishore, et al.
Publicado: (2026)
por: Shahi, Gautam Kishore, et al.
Publicado: (2026)
SQuaD: The Software Quality Dataset
por: Robredo, Mikel, et al.
Publicado: (2025)
por: Robredo, Mikel, et al.
Publicado: (2025)
ReCode: Updating Code API Knowledge with Reinforcement Learning
por: Wu, Haoze, et al.
Publicado: (2025)
por: Wu, Haoze, et al.
Publicado: (2025)
MGS3: A Multi-Granularity Self-Supervised Code Search Framework
por: Li, Rui, et al.
Publicado: (2025)
por: Li, Rui, et al.
Publicado: (2025)
Descriptor: C++ Self-Admitted Technical Debt Dataset (CppSATD)
por: Pham, Phuoc, et al.
Publicado: (2025)
por: Pham, Phuoc, et al.
Publicado: (2025)
yProv4DV: Reproducible Data Visualization Scripts Out of the Box
por: Padovani, Gabriele, et al.
Publicado: (2026)
por: Padovani, Gabriele, et al.
Publicado: (2026)
What About Emotions? Guiding Fine-Grained Emotion Extraction from Mobile App Reviews
por: Motger, Quim, et al.
Publicado: (2025)
por: Motger, Quim, et al.
Publicado: (2025)
Evaluating LLM-Based Mobile App Recommendations: An Empirical Study
por: Motger, Quim, et al.
Publicado: (2025)
por: Motger, Quim, et al.
Publicado: (2025)
Incremental Analysis of Legacy Applications Using Knowledge Graphs for Application Modernization
por: Krishnan, Saravanan, et al.
Publicado: (2025)
por: Krishnan, Saravanan, et al.
Publicado: (2025)
FLOWER: Flow-Oriented Entity-Relationship Tool
por: Moskalev, Dmitry
Publicado: (2025)
por: Moskalev, Dmitry
Publicado: (2025)
Use as Directed? A Comparison of Software Tools Intended to Check Rigor and Transparency of Published Work
por: Eckmann, Peter, et al.
Publicado: (2025)
por: Eckmann, Peter, et al.
Publicado: (2025)
SBAN: A Framework & Multi-Dimensional Dataset for Large Language Model Pre-Training and Software Code Mining
por: Jelodar, Hamed, et al.
Publicado: (2025)
por: Jelodar, Hamed, et al.
Publicado: (2025)
Leveraging Graph-RAG and Prompt Engineering to Enhance LLM-Based Automated Requirement Traceability and Compliance Checks
por: Masoudifard, Arsalan, et al.
Publicado: (2024)
por: Masoudifard, Arsalan, et al.
Publicado: (2024)
An Empirical Study of Multi-Agent RAG for Real-World University Admissions Counseling
por: Nguyen-Duc, Anh, et al.
Publicado: (2025)
por: Nguyen-Duc, Anh, et al.
Publicado: (2025)
Optimizing Retrieval Augmented Generation for Object Constraint Language
por: Li, Kevin Chenhao, et al.
Publicado: (2025)
por: Li, Kevin Chenhao, et al.
Publicado: (2025)
Deep Code Search with Naming-Agnostic Contrastive Multi-View Learning
por: Feng, Jiadong, et al.
Publicado: (2024)
por: Feng, Jiadong, et al.
Publicado: (2024)
A Survey on Query-based API Recommendation
por: Wei, Moshi, et al.
Publicado: (2023)
por: Wei, Moshi, et al.
Publicado: (2023)
Ejemplares similares
-
SURE-RAG: Sufficiency and Uncertainty-Aware Evidence Verification for Selective Retrieval-Augmented Generation
por: Qiu, Jingxi, et al.
Publicado: (2026) -
SemLink: A Semantic-Aware Automated Test Oracle for Hyperlink Verification using Siamese Sentence-BERT
por: Yang, Guan-Yan, et al.
Publicado: (2026) -
Assessing the Ability of ChatGPT to Screen Articles for Systematic Reviews
por: Syriani, Eugene, et al.
Publicado: (2023) -
Toward building next-generation Geocoding systems: a systematic review
por: Yin, Zhengcong, et al.
Publicado: (2025) -
CodeRAG: Finding Relevant and Necessary Knowledge for Retrieval-Augmented Repository-Level Code Completion
por: Zhang, Sheng, et al.
Publicado: (2025)