ACCESS : A Benchmark for Abstract Causal Event Discovery and Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Vo, Vy, Qu, Lizhen, Feng, Tao, Hua, Yuncheng, Kang, Xiaoxi, Fan, Songhai, Dwyer, Tim, Soon, Lay-Ki, Haffari, Gholamreza |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Assistive Large Language Model Agents for Socially-Aware Negotiation Dialogues
by: Hua, Yuncheng, et al.
Published: (2024)
by: Hua, Yuncheng, et al.
Published: (2024)
CausalScore: An Automatic Reference-Free Metric for Assessing Response Relevance in Open-Domain Dialogue Systems
by: Feng, Tao, et al.
Published: (2024)
by: Feng, Tao, et al.
Published: (2024)
On the Reliability of Large Language Models for Causal Discovery
by: Feng, Tao, et al.
Published: (2024)
by: Feng, Tao, et al.
Published: (2024)
IRIS: An Iterative and Integrated Framework for Verifiable Causal Discovery in the Absence of Tabular Data
by: Feng, Tao, et al.
Published: (2025)
by: Feng, Tao, et al.
Published: (2025)
Causal Discovery Inspired Unsupervised Domain Adaptation for Emotion-Cause Pair Extraction
by: Hua, Yuncheng, et al.
Published: (2024)
by: Hua, Yuncheng, et al.
Published: (2024)
IMO: Greedy Layer-Wise Sparse Representation Learning for Out-of-Distribution Text Classification with Pre-trained Models
by: Feng, Tao, et al.
Published: (2024)
by: Feng, Tao, et al.
Published: (2024)
Automating IRAC Analysis in Malaysian Contract Law using a Semi-Structured Knowledge Base
by: Kang, Xiaoxi, et al.
Published: (2024)
by: Kang, Xiaoxi, et al.
Published: (2024)
RENOVI: A Benchmark Towards Remediating Norm Violations in Socio-Cultural Conversations
by: Zhan, Haolan, et al.
Published: (2024)
by: Zhan, Haolan, et al.
Published: (2024)
SCAR: Data Selection via Style Consistency-Aware Response Ranking for Efficient Instruction-Tuning of Large Language Models
by: Li, Zhuang, et al.
Published: (2024)
by: Li, Zhuang, et al.
Published: (2024)
RIDE: Enhancing Large Language Model Alignment through Restyled In-Context Learning Demonstration Exemplars
by: Hua, Yuncheng, et al.
Published: (2025)
by: Hua, Yuncheng, et al.
Published: (2025)
Towards Probing Speech-Specific Risks in Large Multimodal Models: A Taxonomy, Benchmark, and Insights
by: Yang, Hao, et al.
Published: (2024)
by: Yang, Hao, et al.
Published: (2024)
Can ChatGPT Perform Reasoning Using the IRAC Method in Analyzing Legal Scenarios Like a Lawyer?
by: Kang, Xiaoxi, et al.
Published: (2023)
by: Kang, Xiaoxi, et al.
Published: (2023)
NAP^2: A Benchmark for Naturalness and Privacy-Preserving Text Rewriting by Learning from Human
by: Huang, Shuo, et al.
Published: (2024)
by: Huang, Shuo, et al.
Published: (2024)
Zero-Shot Privacy-Aware Text Rewriting via Iterative Tree Search
by: Huang, Shuo, et al.
Published: (2025)
by: Huang, Shuo, et al.
Published: (2025)
Jigsaw Puzzles: Splitting Harmful Questions to Jailbreak Large Language Models
by: Yang, Hao, et al.
Published: (2024)
by: Yang, Hao, et al.
Published: (2024)
Audio Is the Achilles' Heel: Red Teaming Audio Large Multimodal Models
by: Yang, Hao, et al.
Published: (2024)
by: Yang, Hao, et al.
Published: (2024)
Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models
by: Yang, Hao, et al.
Published: (2025)
by: Yang, Hao, et al.
Published: (2025)
LePREC: Reasoning as Classification over Structured Factors for Assessing Relevance of Legal Issues
by: Wang, Fanyu, et al.
Published: (2026)
by: Wang, Fanyu, et al.
Published: (2026)
Evidence-based Distributional Alignment for Large Language Models
by: Pham, Viet-Thanh, et al.
Published: (2026)
by: Pham, Viet-Thanh, et al.
Published: (2026)
The Best of Both Worlds: Bridging Quality and Diversity in Data Selection with Bipartite Graph
by: Wu, Minghao, et al.
Published: (2024)
by: Wu, Minghao, et al.
Published: (2024)
Mixture-of-Skills: Learning to Optimize Data Usage for Fine-Tuning Large Language Models
by: Wu, Minghao, et al.
Published: (2024)
by: Wu, Minghao, et al.
Published: (2024)
Physics-Grounded Motion Forecasting via Equation Discovery for Trajectory-Guided Image-to-Video Generation
by: Feng, Tao, et al.
Published: (2025)
by: Feng, Tao, et al.
Published: (2025)
LiveCultureBench: a Multi-Agent, Multi-Cultural Benchmark for Large Language Models in Dynamic Social Simulations
by: Pham, Viet-Thanh, et al.
Published: (2026)
by: Pham, Viet-Thanh, et al.
Published: (2026)
Importance-Aware Data Augmentation for Document-Level Neural Machine Translation
by: Wu, Minghao, et al.
Published: (2024)
by: Wu, Minghao, et al.
Published: (2024)
Unbiased Sliced Wasserstein Kernels for High-Quality Audio Captioning
by: Luong, Manh, et al.
Published: (2025)
by: Luong, Manh, et al.
Published: (2025)
Let's Negotiate! A Survey of Negotiation Dialogue Systems
by: Zhan, Haolan, et al.
Published: (2024)
by: Zhan, Haolan, et al.
Published: (2024)
Adapting Large Language Models for Document-Level Machine Translation
by: Wu, Minghao, et al.
Published: (2024)
by: Wu, Minghao, et al.
Published: (2024)
Resurfacing Paralinguistic Awareness in Large Audio Language Models
by: Yang, Hao, et al.
Published: (2026)
by: Yang, Hao, et al.
Published: (2026)
SADAS: A Dialogue Assistant System Towards Remediating Norm Violations in Bilingual Socio-Cultural Conversations
by: Hua, Yuncheng, et al.
Published: (2024)
by: Hua, Yuncheng, et al.
Published: (2024)
'Finance Wizard' at the FinLLM Challenge Task: Financial Text Summarization
by: Lee, Meisin, et al.
Published: (2024)
by: Lee, Meisin, et al.
Published: (2024)
Mini-BEHAVIOR-Gran: Revealing U-Shaped Effects of Instruction Granularity on Language-Guided Embodied Agents
by: Huang, Sukai, et al.
Published: (2026)
by: Huang, Sukai, et al.
Published: (2026)
Ordering-based Causal Discovery via Generalized Score Matching
by: Vo, Vy, et al.
Published: (2026)
by: Vo, Vy, et al.
Published: (2026)
Reasoning on Graphs: Faithful and Interpretable Large Language Model Reasoning
by: Luo, Linhao, et al.
Published: (2023)
by: Luo, Linhao, et al.
Published: (2023)
A Directed Graph Model and Experimental Framework for Design and Study of Time-Dependent Text Visualisation
by: Fan, Songhai, et al.
Published: (2026)
by: Fan, Songhai, et al.
Published: (2026)
AIPO: Learning to Reason from Active Interaction
by: Liu, Junnan, et al.
Published: (2026)
by: Liu, Junnan, et al.
Published: (2026)
Towards Inference-time Scaling for Continuous Space Reasoning
by: Wang, Minghan, et al.
Published: (2025)
by: Wang, Minghan, et al.
Published: (2025)
SituatedThinker: Grounding LLM Reasoning with Real-World through Situated Thinking
by: Liu, Junnan, et al.
Published: (2025)
by: Liu, Junnan, et al.
Published: (2025)
Towards Event Extraction from Speech with Contextual Clues
by: Kang, Jingqi, et al.
Published: (2024)
by: Kang, Jingqi, et al.
Published: (2024)
Scalable Frame-based Construction of Sociocultural NormBases for Socially-Aware Dialogues
by: Qu, Shilin, et al.
Published: (2024)
by: Qu, Shilin, et al.
Published: (2024)
Multi-Layer Scheduling for MoE-Based LLM Reasoning
by: Sun, Yifan, et al.
Published: (2026)
by: Sun, Yifan, et al.
Published: (2026)
Similar Items
-
Assistive Large Language Model Agents for Socially-Aware Negotiation Dialogues
by: Hua, Yuncheng, et al.
Published: (2024) -
CausalScore: An Automatic Reference-Free Metric for Assessing Response Relevance in Open-Domain Dialogue Systems
by: Feng, Tao, et al.
Published: (2024) -
On the Reliability of Large Language Models for Causal Discovery
by: Feng, Tao, et al.
Published: (2024) -
IRIS: An Iterative and Integrated Framework for Verifiable Causal Discovery in the Absence of Tabular Data
by: Feng, Tao, et al.
Published: (2025) -
Causal Discovery Inspired Unsupervised Domain Adaptation for Emotion-Cause Pair Extraction
by: Hua, Yuncheng, et al.
Published: (2024)