Rethinking Relation Extraction: Beyond Shortcuts to Generalization with a Debiased Benchmark
Fuente:
arXiv
Guardado en:
| Autores principales: | He, Liang, Chu, Yougang, Wu, Zhen, Zhang, Jianbing, Dai, Xinyu, Chen, Jiajun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MixRED: A Mix-lingual Relation Extraction Dataset
por: Kong, Lingxing, et al.
Publicado: (2024)
por: Kong, Lingxing, et al.
Publicado: (2024)
SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents
por: Cheng, Kanzhi, et al.
Publicado: (2024)
por: Cheng, Kanzhi, et al.
Publicado: (2024)
MAS-Bench: A Unified Benchmark for Shortcut-Augmented Hybrid Mobile GUI Agents
por: Zhao, Pengxiang, et al.
Publicado: (2025)
por: Zhao, Pengxiang, et al.
Publicado: (2025)
Persona-Aware Alignment Framework for Personalized Dialogue Generation
por: Li, Guanrong, et al.
Publicado: (2025)
por: Li, Guanrong, et al.
Publicado: (2025)
Spurious Correlations and Beyond: Understanding and Mitigating Shortcut Learning in SDOH Extraction with Large Language Models
por: Sakib, Fardin Ahsan, et al.
Publicado: (2025)
por: Sakib, Fardin Ahsan, et al.
Publicado: (2025)
General Phrase Debiaser: Debiasing Masked Language Models at a Multi-Token Level
por: Shi, Bingkang, et al.
Publicado: (2023)
por: Shi, Bingkang, et al.
Publicado: (2023)
GenRES: Rethinking Evaluation for Generative Relation Extraction in the Era of Large Language Models
por: Jiang, Pengcheng, et al.
Publicado: (2024)
por: Jiang, Pengcheng, et al.
Publicado: (2024)
Recognize Your Orchestrator: An Entropy Dynamics Perspective for LLM Multi-Agent Systems
por: Zhu, Junze, et al.
Publicado: (2026)
por: Zhu, Junze, et al.
Publicado: (2026)
From Projection to Prediction: Beyond Logits for Scalable Language Models
por: Dong, Jianbing, et al.
Publicado: (2025)
por: Dong, Jianbing, et al.
Publicado: (2025)
Causal Evidence for Attention Head Imbalance in Modality Conflict Hallucination
por: Jiang, Jinrui, et al.
Publicado: (2026)
por: Jiang, Jinrui, et al.
Publicado: (2026)
DeepShop: A Benchmark for Deep Research Shopping Agents
por: Lyu, Yougang, et al.
Publicado: (2025)
por: Lyu, Yougang, et al.
Publicado: (2025)
Counterfactual Language Reasoning for Explainable Recommendation Systems
por: Li, Guanrong, et al.
Publicado: (2025)
por: Li, Guanrong, et al.
Publicado: (2025)
Relative-Absolute Fusion: Rethinking Feature Extraction in Image-Based Iterative Method Selection for Solving Sparse Linear Systems
por: Zhang, Kaiqi, et al.
Publicado: (2025)
por: Zhang, Kaiqi, et al.
Publicado: (2025)
Revisiting, Benchmarking and Understanding Unsupervised Graph Domain Adaptation
por: Liu, Meihan, et al.
Publicado: (2024)
por: Liu, Meihan, et al.
Publicado: (2024)
Rethinking Generative Recommender Tokenizer: Recsys-Native Encoding and Semantic Quantization Beyond LLMs
por: Liang, Yu, et al.
Publicado: (2026)
por: Liang, Yu, et al.
Publicado: (2026)
Rethinking Propagation for Unsupervised Graph Domain Adaptation
por: Liu, Meihan, et al.
Publicado: (2024)
por: Liu, Meihan, et al.
Publicado: (2024)
Rethinking the Soft Conflict Pseudo Boolean Constraint on MaxSAT Local Search Solvers
por: Zheng, Jiongzhi, et al.
Publicado: (2024)
por: Zheng, Jiongzhi, et al.
Publicado: (2024)
Contextual Multi-Objective Optimization: Rethinking Objectives in Frontier AI Systems
por: Zhou, Jie, et al.
Publicado: (2026)
por: Zhou, Jie, et al.
Publicado: (2026)
Empirical Analysis of Dialogue Relation Extraction with Large Language Models
por: Li, Guozheng, et al.
Publicado: (2024)
por: Li, Guozheng, et al.
Publicado: (2024)
Debiased Offline Representation Learning for Fast Online Adaptation in Non-stationary Dynamics
por: Zhang, Xinyu, et al.
Publicado: (2024)
por: Zhang, Xinyu, et al.
Publicado: (2024)
Retrieving, Rethinking and Revising: The Chain-of-Verification Can Improve Retrieval Augmented Generation
por: He, Bolei, et al.
Publicado: (2024)
por: He, Bolei, et al.
Publicado: (2024)
Audio-centric Video Understanding Benchmark without Text Shortcut
por: Yang, Yudong, et al.
Publicado: (2025)
por: Yang, Yudong, et al.
Publicado: (2025)
A Neuro-Symbolic Benchmark Suite for Concept Quality and Reasoning Shortcuts
por: Bortolotti, Samuele, et al.
Publicado: (2024)
por: Bortolotti, Samuele, et al.
Publicado: (2024)
Debiased Multimodal Understanding for Human Language Sequences
por: Xu, Zhi, et al.
Publicado: (2024)
por: Xu, Zhi, et al.
Publicado: (2024)
ShortcutsBench: A Large-Scale Real-world Benchmark for API-based Agents
por: Shen, Haiyang, et al.
Publicado: (2024)
por: Shen, Haiyang, et al.
Publicado: (2024)
Recall, Retrieve and Reason: Towards Better In-Context Relation Extraction
por: Li, Guozheng, et al.
Publicado: (2024)
por: Li, Guozheng, et al.
Publicado: (2024)
Beyond Low-rank Decomposition: A Shortcut Approach for Efficient On-Device Learning
por: Nguyen, Le-Trung, et al.
Publicado: (2025)
por: Nguyen, Le-Trung, et al.
Publicado: (2025)
Out-of-distribution Evidence-aware Fake News Detection via Dual Adversarial Debiasing
por: Liu, Qiang, et al.
Publicado: (2023)
por: Liu, Qiang, et al.
Publicado: (2023)
A Causal Adjustment Module for Debiasing Scene Graph Generation
por: Liu, Li, et al.
Publicado: (2025)
por: Liu, Li, et al.
Publicado: (2025)
Personalized Generative Models for Contextual Debiasing
por: Liang, Xinran, et al.
Publicado: (2026)
por: Liang, Xinran, et al.
Publicado: (2026)
Debiasing Kernel-Based Generative Models
por: Qin, Tian, et al.
Publicado: (2025)
por: Qin, Tian, et al.
Publicado: (2025)
UDA: Unsupervised Debiasing Alignment for Pair-wise LLM-as-a-Judge
por: Zhang, Yang, et al.
Publicado: (2025)
por: Zhang, Yang, et al.
Publicado: (2025)
An Embarrassingly Simple Graph Heuristic Reveals Shortcut-Solvable Benchmarks for Sequential Recommendation
por: Han, Haoyu, et al.
Publicado: (2026)
por: Han, Haoyu, et al.
Publicado: (2026)
Customized Retrieval-Augmented Generation with LLM for Debiasing Recommendation Unlearning
por: Zhang, Haichao, et al.
Publicado: (2025)
por: Zhang, Haichao, et al.
Publicado: (2025)
Dual Debiasing for Noisy In-Context Learning for Text Generation
por: Liang, Siqi, et al.
Publicado: (2025)
por: Liang, Siqi, et al.
Publicado: (2025)
Unlocking Instructive In-Context Learning with Tabular Prompting for Relational Triple Extraction
por: Li, Guozheng, et al.
Publicado: (2024)
por: Li, Guozheng, et al.
Publicado: (2024)
Beyond Spurious Signals: Debiasing Multimodal Large Language Models via Counterfactual Inference and Adaptive Expert Routing
por: Wu, Zichen, et al.
Publicado: (2025)
por: Wu, Zichen, et al.
Publicado: (2025)
Rethinking Weak Supervision in Anomaly Detection: A Comprehensive Benchmark
por: Yao, Xu, et al.
Publicado: (2026)
por: Yao, Xu, et al.
Publicado: (2026)
SenseMath: Do LLMs Have Number Sense? Evaluating Shortcut Use, Judgment, and Generation
por: Zhuang, Haomin, et al.
Publicado: (2026)
por: Zhuang, Haomin, et al.
Publicado: (2026)
Graph-Augmented Relation Extraction Model with LLMs-Generated Support Document
por: Dong, Vicky, et al.
Publicado: (2024)
por: Dong, Vicky, et al.
Publicado: (2024)
Ejemplares similares
-
MixRED: A Mix-lingual Relation Extraction Dataset
por: Kong, Lingxing, et al.
Publicado: (2024) -
SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents
por: Cheng, Kanzhi, et al.
Publicado: (2024) -
MAS-Bench: A Unified Benchmark for Shortcut-Augmented Hybrid Mobile GUI Agents
por: Zhao, Pengxiang, et al.
Publicado: (2025) -
Persona-Aware Alignment Framework for Personalized Dialogue Generation
por: Li, Guanrong, et al.
Publicado: (2025) -
Spurious Correlations and Beyond: Understanding and Mitigating Shortcut Learning in SDOH Extraction with Large Language Models
por: Sakib, Fardin Ahsan, et al.
Publicado: (2025)