Towards Robust Universal Information Extraction: Benchmark, Evaluation, and Solution
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhu, Jizhao, Shi, Akang, Li, Zixuan, Bai, Long, Jin, Xiaolong, Guo, Jiafeng, Cheng, Xueqi |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Towards Event Extraction with Massive Types: LLM-based Collaborative Annotation and Partitioning Extraction
par: Liu, Wenxuan, et autres
Publié: (2025)
par: Liu, Wenxuan, et autres
Publié: (2025)
KnowCoder-X: Boosting Multilingual Information Extraction via Code
par: Zuo, Yuxin, et autres
Publié: (2024)
par: Zuo, Yuxin, et autres
Publié: (2024)
Temporal Knowledge Graph Question Answering: A Survey
par: Su, Miao, et autres
Publié: (2024)
par: Su, Miao, et autres
Publié: (2024)
Mixture Policy based Multi-Hop Reasoning over N-tuple Temporal Knowledge Graphs
par: Hou, Zhongni, et autres
Publié: (2025)
par: Hou, Zhongni, et autres
Publié: (2025)
Evaluating Implicit Bias in Large Language Models by Attacking From a Psychometric Perspective
par: Wen, Yuchen, et autres
Publié: (2024)
par: Wen, Yuchen, et autres
Publié: (2024)
RouteRAG: Efficient Retrieval-Augmented Generation from Text and Graph via Reinforcement Learning
par: Guo, Yucan, et autres
Publié: (2025)
par: Guo, Yucan, et autres
Publié: (2025)
KnowCoder: Coding Structured Knowledge into LLMs for Universal Information Extraction
par: Li, Zixuan, et autres
Publié: (2024)
par: Li, Zixuan, et autres
Publié: (2024)
Bagging-Based Model Merging for Robust General Text Embeddings
par: Zhang, Hengran, et autres
Publié: (2026)
par: Zhang, Hengran, et autres
Publié: (2026)
Robust Neural Information Retrieval: An Adversarial and Out-of-distribution Perspective
par: Liu, Yu-An, et autres
Publié: (2024)
par: Liu, Yu-An, et autres
Publié: (2024)
G2S: A General-to-Specific Learning Framework for Temporal Knowledge Graph Forecasting with Large Language Models
par: Bai, Long, et autres
Publié: (2025)
par: Bai, Long, et autres
Publié: (2025)
Towards Knowledgeable Deep Research: Framework and Benchmark
par: Liu, Wenxuan, et autres
Publié: (2026)
par: Liu, Wenxuan, et autres
Publié: (2026)
Nested Event Extraction upon Pivot Element Recogniton
par: Ren, Weicheng, et autres
Publié: (2023)
par: Ren, Weicheng, et autres
Publié: (2023)
Estimating Commonsense Plausibility through Semantic Shifts
par: Cui, Wanqing, et autres
Publié: (2025)
par: Cui, Wanqing, et autres
Publié: (2025)
Class-Incremental Few-Shot Event Detection
par: Zhao, Kailin, et autres
Publié: (2024)
par: Zhao, Kailin, et autres
Publié: (2024)
An Iterative Utility Judgment Framework Inspired by Philosophical Relevance via LLMs
par: Zhang, Hengran, et autres
Publié: (2024)
par: Zhang, Hengran, et autres
Publié: (2024)
Beyond Dialogue Time: Temporal Semantic Memory for Personalized LLM Agents
par: Su, Miao, et autres
Publié: (2026)
par: Su, Miao, et autres
Publié: (2026)
KnowCoder-A1: Incentivizing Agentic Reasoning Capability with Outcome Supervision for KBQA
par: Chen, Zhuo, et autres
Publié: (2025)
par: Chen, Zhuo, et autres
Publié: (2025)
Look Globally and Reason: Two-stage Path Reasoning over Sparse Knowledge Graphs
par: Guan, Saiping, et autres
Publié: (2024)
par: Guan, Saiping, et autres
Publié: (2024)
A Survey of Link Prediction in N-ary Knowledge Graphs
par: Wei, Jiyao, et autres
Publié: (2025)
par: Wei, Jiyao, et autres
Publié: (2025)
An In-Context Schema Understanding Method for Knowledge Base Question Answering
par: Liu, Yantao, et autres
Publié: (2023)
par: Liu, Yantao, et autres
Publié: (2023)
QUITO-X: A New Perspective on Context Compression from the Information Bottleneck Theory
par: Wang, Yihang, et autres
Publié: (2024)
par: Wang, Yihang, et autres
Publié: (2024)
Few-shot Link Prediction on N-ary Facts
par: Wei, Jiyao, et autres
Publié: (2023)
par: Wei, Jiyao, et autres
Publié: (2023)
Preserving Knowledge Invariance: Rethinking Robustness Evaluation of Open Information Extraction
par: Qi, Ji, et autres
Publié: (2023)
par: Qi, Ji, et autres
Publié: (2023)
Utility-Focused LLM Annotation for Retrieval and Retrieval-Augmented Generation
par: Zhang, Hengran, et autres
Publié: (2025)
par: Zhang, Hengran, et autres
Publié: (2025)
Unleashing the Power of LLMs in Dense Retrieval with Query Likelihood Modeling
par: Zhang, Hengran, et autres
Publié: (2025)
par: Zhang, Hengran, et autres
Publié: (2025)
A Claim Decomposition Benchmark for Long-form Answer Verification
par: Zhang, Zhihao, et autres
Publié: (2024)
par: Zhang, Zhihao, et autres
Publié: (2024)
Foot-In-The-Door: A Multi-turn Jailbreak for LLMs
par: Weng, Zixuan, et autres
Publié: (2025)
par: Weng, Zixuan, et autres
Publié: (2025)
JudgeAgent: Beyond Static Benchmarks for Knowledge-Driven and Dynamic LLM Evaluation
par: Shi, Zhichao, et autres
Publié: (2025)
par: Shi, Zhichao, et autres
Publié: (2025)
LLM-Specific Utility: A New Perspective for Retrieval-Augmented Generation
par: Zhang, Hengran, et autres
Publié: (2025)
par: Zhang, Hengran, et autres
Publié: (2025)
Generating Leakage-Free Benchmarks for Robust RAG Evaluation
par: Liu, Jiayi, et autres
Publié: (2026)
par: Liu, Jiayi, et autres
Publié: (2026)
RefineX: Learning to Refine Pre-training Data at Scale from Expert-Guided Programs
par: Bi, Baolong, et autres
Publié: (2025)
par: Bi, Baolong, et autres
Publié: (2025)
Distilling a Small Utility-Based Passage Selector to Enhance Retrieval-Augmented Generation
par: Zhang, Hengran, et autres
Publié: (2025)
par: Zhang, Hengran, et autres
Publié: (2025)
Label Drop for Multi-Aspect Relation Modeling in Universal Information Extraction
par: Yang, Lu, et autres
Publié: (2025)
par: Yang, Lu, et autres
Publié: (2025)
STRUCTSENSE: A Task-Agnostic Agentic Framework for Structured Information Extraction with Human-In-The-Loop Evaluation and Benchmarking
par: Chhetri, Tek Raj, et autres
Publié: (2025)
par: Chhetri, Tek Raj, et autres
Publié: (2025)
CMMaTH: A Chinese Multi-modal Math Skill Evaluation Benchmark for Foundation Models
par: Li, Zhong-Zhi, et autres
Publié: (2024)
par: Li, Zhong-Zhi, et autres
Publié: (2024)
CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports
par: Zhang, Xiao Yu Cindy, et autres
Publié: (2025)
par: Zhang, Xiao Yu Cindy, et autres
Publié: (2025)
EduEval: A Hierarchical Cognitive Benchmark for Evaluating Large Language Models in Chinese Education
par: Ma, Guoqing, et autres
Publié: (2025)
par: Ma, Guoqing, et autres
Publié: (2025)
Benchmarking Large Language Models on CFLUE -- A Chinese Financial Language Understanding Evaluation Dataset
par: Zhu, Jie, et autres
Publié: (2024)
par: Zhu, Jie, et autres
Publié: (2024)
On Robustness and Reliability of Benchmark-Based Evaluation of LLMs
par: Lunardi, Riccardo, et autres
Publié: (2025)
par: Lunardi, Riccardo, et autres
Publié: (2025)
KnowCoder-V2: Deep Knowledge Analysis
par: Li, Zixuan, et autres
Publié: (2025)
par: Li, Zixuan, et autres
Publié: (2025)
Documents similaires
-
Towards Event Extraction with Massive Types: LLM-based Collaborative Annotation and Partitioning Extraction
par: Liu, Wenxuan, et autres
Publié: (2025) -
KnowCoder-X: Boosting Multilingual Information Extraction via Code
par: Zuo, Yuxin, et autres
Publié: (2024) -
Temporal Knowledge Graph Question Answering: A Survey
par: Su, Miao, et autres
Publié: (2024) -
Mixture Policy based Multi-Hop Reasoning over N-tuple Temporal Knowledge Graphs
par: Hou, Zhongni, et autres
Publié: (2025) -
Evaluating Implicit Bias in Large Language Models by Attacking From a Psychometric Perspective
par: Wen, Yuchen, et autres
Publié: (2024)