Weak Reward Model Transforms Generative Models into Robust Causal Event Extraction Systems
Fuente:
arXiv
Saved in:
| Main Authors: | da Silva, Italo Luis, Yan, Hanqi, Gui, Lin, He, Yulan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GraphMind: Interactive Novelty Assessment System for Accelerating Scientific Discovery
by: da Silva, Italo Luis, et al.
Published: (2025)
by: da Silva, Italo Luis, et al.
Published: (2025)
Addressing Order Sensitivity of In-Context Demonstration Examples in Causal Language Models
by: Xiang, Yanzheng, et al.
Published: (2024)
by: Xiang, Yanzheng, et al.
Published: (2024)
Soft Reasoning: Navigating Solution Spaces in Large Language Models through Controlled Embedding Exploration
by: Zhu, Qinglin, et al.
Published: (2025)
by: Zhu, Qinglin, et al.
Published: (2025)
Mirror: A Multiple-perspective Self-Reflection Method for Knowledge-rich Reasoning
by: Yan, Hanqi, et al.
Published: (2024)
by: Yan, Hanqi, et al.
Published: (2024)
Explainable Recommender with Geometric Information Bottleneck
by: Yan, Hanqi, et al.
Published: (2023)
by: Yan, Hanqi, et al.
Published: (2023)
The Mystery of In-Context Learning: A Comprehensive Survey on Interpretation and Analysis
by: Zhou, Yuxiang, et al.
Published: (2023)
by: Zhou, Yuxiang, et al.
Published: (2023)
Beyond Perplexity: Let the Reader Select Retrieval Summaries via Spectrum Projection Score
by: Hu, Zhanghao, et al.
Published: (2025)
by: Hu, Zhanghao, et al.
Published: (2025)
Counterfactual Generation with Identifiability Guarantees
by: Yan, Hanqi, et al.
Published: (2024)
by: Yan, Hanqi, et al.
Published: (2024)
SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers
by: Xiang, Yanzheng, et al.
Published: (2025)
by: Xiang, Yanzheng, et al.
Published: (2025)
Encourage or Inhibit Monosemanticity? Revisit Monosemanticity from a Feature Decorrelation Perspective
by: Yan, Hanqi, et al.
Published: (2024)
by: Yan, Hanqi, et al.
Published: (2024)
Beyond Prompting: An Efficient Embedding Framework for Open-Domain Question Answering
by: Hu, Zhanghao, et al.
Published: (2025)
by: Hu, Zhanghao, et al.
Published: (2025)
Are NLP Models Good at Tracing Thoughts: An Overview of Narrative Understanding
by: Zhu, Lixing, et al.
Published: (2023)
by: Zhu, Lixing, et al.
Published: (2023)
Beyond RAG for Agent Memory: Retrieval by Decoupling and Aggregation
by: Hu, Zhanghao, et al.
Published: (2026)
by: Hu, Zhanghao, et al.
Published: (2026)
Towards Unified Task Embeddings Across Multiple Models: Bridging the Gap for Prompt-Based Large Language Models and Beyond
by: Wang, Xinyu, et al.
Published: (2024)
by: Wang, Xinyu, et al.
Published: (2024)
Leveraging ChatGPT in Pharmacovigilance Event Extraction: An Empirical Study
by: Sun, Zhaoyue, et al.
Published: (2024)
by: Sun, Zhaoyue, et al.
Published: (2024)
Multi-Layer Ranking with Large Language Models for News Source Recommendation
by: Zhang, Wenjia, et al.
Published: (2024)
by: Zhang, Wenjia, et al.
Published: (2024)
Cascading Large Language Models for Salient Event Graph Generation
by: Tan, Xingwei, et al.
Published: (2024)
by: Tan, Xingwei, et al.
Published: (2024)
When Thinking Backfires: Mechanistic Insights Into Reasoning-Induced Misalignment
by: Yan, Hanqi, et al.
Published: (2025)
by: Yan, Hanqi, et al.
Published: (2025)
A Survey of Automatic Hallucination Evaluation on Natural Language Generation
by: Qi, Siya, et al.
Published: (2024)
by: Qi, Siya, et al.
Published: (2024)
Weak-to-Strong Preference Optimization: Stealing Reward from Weak Aligned Model
by: Zhu, Wenhong, et al.
Published: (2024)
by: Zhu, Wenhong, et al.
Published: (2024)
PECAN: LLM-Guided Dynamic Progress Control with Attention-Guided Hierarchical Weighted Graph for Long-Document QA
by: Wang, Xinyu, et al.
Published: (2024)
by: Wang, Xinyu, et al.
Published: (2024)
CODI: Compressing Chain-of-Thought into Continuous Space via Self-Distillation
by: Shen, Zhenyi, et al.
Published: (2025)
by: Shen, Zhenyi, et al.
Published: (2025)
Fix the Structural Bottleneck: Context Compression via Explicit Information Transmission
by: Ye, Jiangnan, et al.
Published: (2026)
by: Ye, Jiangnan, et al.
Published: (2026)
Reward Modeling with Weak Supervision for Language Models
by: Hauptvogel, Ben, et al.
Published: (2024)
by: Hauptvogel, Ben, et al.
Published: (2024)
Causal Prompting: Debiasing Large Language Model Prompting based on Front-Door Adjustment
by: Zhang, Congzhi, et al.
Published: (2024)
by: Zhang, Congzhi, et al.
Published: (2024)
Length Generalization of Causal Transformers without Position Encoding
by: Wang, Jie, et al.
Published: (2024)
by: Wang, Jie, et al.
Published: (2024)
Extracting Event Temporal Relations via Hyperbolic Geometry
by: Tan, Xingwei, et al.
Published: (2021)
by: Tan, Xingwei, et al.
Published: (2021)
Two Heads Are Better Than One: Dual-Model Verbal Reflection at Inference-Time
by: Li, Jiazheng, et al.
Published: (2025)
by: Li, Jiazheng, et al.
Published: (2025)
Set-Aligning Framework for Auto-Regressive Event Temporal Graph Generation
by: Tan, Xingwei, et al.
Published: (2024)
by: Tan, Xingwei, et al.
Published: (2024)
A Structure-aware Generative Model for Biomedical Event Extraction
by: Yuan, Haohan, et al.
Published: (2024)
by: Yuan, Haohan, et al.
Published: (2024)
SymbolicThought: Integrating Language Models and Symbolic Reasoning for Consistent and Interpretable Human Relationship Understanding
by: Zhao, Runcong, et al.
Published: (2025)
by: Zhao, Runcong, et al.
Published: (2025)
Event Extraction in Large Language Model
by: Li, Bobo, et al.
Published: (2025)
by: Li, Bobo, et al.
Published: (2025)
Robust Reward Modeling for Large Language Models via Causal Decomposition
by: Lu, Yunsheng, et al.
Published: (2026)
by: Lu, Yunsheng, et al.
Published: (2026)
reWordBench: Benchmarking and Improving the Robustness of Reward Models with Transformed Inputs
by: Wu, Zhaofeng, et al.
Published: (2025)
by: Wu, Zhaofeng, et al.
Published: (2025)
Large Language Models Fall Short: Understanding Complex Relationships in Detective Narratives
by: Zhao, Runcong, et al.
Published: (2024)
by: Zhao, Runcong, et al.
Published: (2024)
RRM: Robust Reward Model Training Mitigates Reward Hacking
by: Liu, Tianqi, et al.
Published: (2024)
by: Liu, Tianqi, et al.
Published: (2024)
Stop the Flip-Flop: Context-Preserving Verification for Fast Revocable Diffusion Decoding
by: Xiang, Yanzheng, et al.
Published: (2026)
by: Xiang, Yanzheng, et al.
Published: (2026)
Reward Models Can Improve Themselves: Reward-Guided Adversarial Failure Mode Discovery for Robust Reward Modeling
by: Pathmanathan, Pankayaraj, et al.
Published: (2025)
by: Pathmanathan, Pankayaraj, et al.
Published: (2025)
PLAYER*: Enhancing LLM-based Multi-Agent Communication and Interaction in Murder Mystery Games
by: Zhu, Qinglin, et al.
Published: (2024)
by: Zhu, Qinglin, et al.
Published: (2024)
GRAM: A Generative Foundation Reward Model for Reward Generalization
by: Wang, Chenglong, et al.
Published: (2025)
by: Wang, Chenglong, et al.
Published: (2025)
Similar Items
-
GraphMind: Interactive Novelty Assessment System for Accelerating Scientific Discovery
by: da Silva, Italo Luis, et al.
Published: (2025) -
Addressing Order Sensitivity of In-Context Demonstration Examples in Causal Language Models
by: Xiang, Yanzheng, et al.
Published: (2024) -
Soft Reasoning: Navigating Solution Spaces in Large Language Models through Controlled Embedding Exploration
by: Zhu, Qinglin, et al.
Published: (2025) -
Mirror: A Multiple-perspective Self-Reflection Method for Knowledge-rich Reasoning
by: Yan, Hanqi, et al.
Published: (2024) -
Explainable Recommender with Geometric Information Bottleneck
by: Yan, Hanqi, et al.
Published: (2023)