EconCausal: A Context-Aware Economic Reasoning Benchmark for Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Donggyu, Yun, Hyeok, Cha, Meeyoung, Park, Sungwon, Park, Sangyoon, Kim, Jihee |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Ideological Bias in LLMs' Economic Causal Reasoning
von: Lee, Donggyu, et al.
Veröffentlicht: (2026)
von: Lee, Donggyu, et al.
Veröffentlicht: (2026)
Adversarial Style Augmentation via Large Language Model for Robust Fake News Detection
von: Park, Sungwon, et al.
Veröffentlicht: (2024)
von: Park, Sungwon, et al.
Veröffentlicht: (2024)
GeoSEE: Regional Socio-Economic Estimation With a Large Language Model
von: Han, Sungwon, et al.
Veröffentlicht: (2024)
von: Han, Sungwon, et al.
Veröffentlicht: (2024)
Generalizable Disaster Damage Assessment via Change Detection with Vision Foundation Model
von: Ahn, Kyeongjin, et al.
Veröffentlicht: (2024)
von: Ahn, Kyeongjin, et al.
Veröffentlicht: (2024)
Dropouts in Confidence: Moral Uncertainty in Human-LLM Alignment
von: Kwon, Jea, et al.
Veröffentlicht: (2025)
von: Kwon, Jea, et al.
Veröffentlicht: (2025)
In-Context Examples Suppress Scientific Knowledge Recall in LLMs
von: Jang, Chaemin, et al.
Veröffentlicht: (2026)
von: Jang, Chaemin, et al.
Veröffentlicht: (2026)
ERA: Evidence-based Reliability Alignment for Honest Retrieval-Augmented Generation
von: Shin, Sunguk, et al.
Veröffentlicht: (2026)
von: Shin, Sunguk, et al.
Veröffentlicht: (2026)
How Training Data Shapes the Use of Parametric and In-Context Knowledge in Language Models
von: Kim, Minsung, et al.
Veröffentlicht: (2025)
von: Kim, Minsung, et al.
Veröffentlicht: (2025)
GeoReg: Weight-Constrained Few-Shot Regression for Socio-Economic Estimation using LLM
von: Ahn, Kyeongjin, et al.
Veröffentlicht: (2025)
von: Ahn, Kyeongjin, et al.
Veröffentlicht: (2025)
Generalizable Slum Detection from Satellite Imagery with Mixture-of-Experts
von: Lee, Sumin, et al.
Veröffentlicht: (2025)
von: Lee, Sumin, et al.
Veröffentlicht: (2025)
GECKO: Generative Language Model for English, Code and Korean
von: Oh, Sungwoo, et al.
Veröffentlicht: (2024)
von: Oh, Sungwoo, et al.
Veröffentlicht: (2024)
KITE: A Benchmark for Evaluating Korean Instruction-Following Abilities in Large Language Models
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
Representing the Under-Represented: Cultural and Core Capability Benchmarks for Developing Thai Large Language Models
von: Kim, Dahyun, et al.
Veröffentlicht: (2024)
von: Kim, Dahyun, et al.
Veröffentlicht: (2024)
EconEvals: Benchmarks and Litmus Tests for Economic Decision-Making by LLM Agents
von: Fish, Sara, et al.
Veröffentlicht: (2025)
von: Fish, Sara, et al.
Veröffentlicht: (2025)
X-Teaming Evolutionary M2S: Automated Discovery of Multi-turn to Single-turn Jailbreak Templates
von: Kim, Hyunjun, et al.
Veröffentlicht: (2025)
von: Kim, Hyunjun, et al.
Veröffentlicht: (2025)
SAAS: Solving Ability Amplification Strategy for Enhanced Mathematical Reasoning in Large Language Models
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2024)
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2024)
FeRG-LLM : Feature Engineering by Reason Generation Large Language Models
von: Ko, Jeonghyun, et al.
Veröffentlicht: (2025)
von: Ko, Jeonghyun, et al.
Veröffentlicht: (2025)
Open Ko-LLM Leaderboard: Evaluating Large Language Models in Korean with Ko-H5 Benchmark
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
InvThink: Premortem Reasoning for Safer Language Models
von: Kim, Yubin, et al.
Veröffentlicht: (2025)
von: Kim, Yubin, et al.
Veröffentlicht: (2025)
Cross-lingual Collapse: How Language-Centric Foundation Models Shape Reasoning in Large Language Models
von: Park, Cheonbok, et al.
Veröffentlicht: (2025)
von: Park, Cheonbok, et al.
Veröffentlicht: (2025)
EconProver: Towards More Economical Test-Time Scaling for Automated Theorem Proving
von: Li, Mukai, et al.
Veröffentlicht: (2025)
von: Li, Mukai, et al.
Veröffentlicht: (2025)
Improving Multi-hop Logical Reasoning in Knowledge Graphs with Context-Aware Query Representation Learning
von: Kim, Jeonghoon, et al.
Veröffentlicht: (2024)
von: Kim, Jeonghoon, et al.
Veröffentlicht: (2024)
CALRK-Bench: Evaluating Context-Aware Legal Reasoning in Korean Law
von: Jung, JiHyeok, et al.
Veröffentlicht: (2026)
von: Jung, JiHyeok, et al.
Veröffentlicht: (2026)
Dataverse: Open-Source ETL (Extract, Transform, Load) Pipeline for Large Language Models
von: Park, Hyunbyung, et al.
Veröffentlicht: (2024)
von: Park, Hyunbyung, et al.
Veröffentlicht: (2024)
M2S: Multi-turn to Single-turn jailbreak in Red Teaming for LLMs
von: Ha, Junwoo, et al.
Veröffentlicht: (2025)
von: Ha, Junwoo, et al.
Veröffentlicht: (2025)
FLEX: A Benchmark for Evaluating Robustness of Fairness in Large Language Models
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
CausalVLBench: Benchmarking Visual Causal Reasoning in Large Vision-Language Models
von: Komanduri, Aneesh, et al.
Veröffentlicht: (2025)
von: Komanduri, Aneesh, et al.
Veröffentlicht: (2025)
Ko-MuSR: A Multistep Soft Reasoning Benchmark for LLMs Capable of Understanding Korean
von: Park, Chanwoo, et al.
Veröffentlicht: (2025)
von: Park, Chanwoo, et al.
Veröffentlicht: (2025)
Context-Robust Knowledge Editing for Language Models
von: Park, Haewon, et al.
Veröffentlicht: (2025)
von: Park, Haewon, et al.
Veröffentlicht: (2025)
Finding Answers in Thought Matters: Revisiting Evaluation on Large Language Models with Reasoning
von: Jo, Hwiyeol, et al.
Veröffentlicht: (2025)
von: Jo, Hwiyeol, et al.
Veröffentlicht: (2025)
NoisyCausal: A Benchmark for Evaluating Causal Reasoning Under Structured Noise
von: Xu, Zhi, et al.
Veröffentlicht: (2026)
von: Xu, Zhi, et al.
Veröffentlicht: (2026)
EconNLI: Evaluating Large Language Models on Economics Reasoning
von: Guo, Yue, et al.
Veröffentlicht: (2024)
von: Guo, Yue, et al.
Veröffentlicht: (2024)
EXAONE 4.0: Unified Large Language Models Integrating Non-reasoning and Reasoning Modes
von: Bae, Kyunghoon, et al.
Veröffentlicht: (2025)
von: Bae, Kyunghoon, et al.
Veröffentlicht: (2025)
CASE-Bench: Context-Aware SafEty Benchmark for Large Language Models
von: Sun, Guangzhi, et al.
Veröffentlicht: (2025)
von: Sun, Guangzhi, et al.
Veröffentlicht: (2025)
Enhancing Robustness of Retrieval-Augmented Language Models with In-Context Learning
von: Park, Seong-Il, et al.
Veröffentlicht: (2024)
von: Park, Seong-Il, et al.
Veröffentlicht: (2024)
Large Language Models Are Better Logical Fallacy Reasoners with Counterargument, Explanation, and Goal-Aware Prompt Formulation
von: Jeong, Jiwon, et al.
Veröffentlicht: (2025)
von: Jeong, Jiwon, et al.
Veröffentlicht: (2025)
Does "Reasoning" with Large Language Models Improve Recognizing, Generating, and Reframing Unhelpful Thoughts?
von: Qi, Yilin, et al.
Veröffentlicht: (2025)
von: Qi, Yilin, et al.
Veröffentlicht: (2025)
EXAONE Deep: Reasoning Enhanced Language Models
von: Bae, Kyunghoon, et al.
Veröffentlicht: (2025)
von: Bae, Kyunghoon, et al.
Veröffentlicht: (2025)
Tabular Feature Discovery With Reasoning Type Exploration
von: Han, Sungwon, et al.
Veröffentlicht: (2025)
von: Han, Sungwon, et al.
Veröffentlicht: (2025)
InfoCausalQA:Can Models Perform Non-explicit Causal Reasoning Based on Infographic?
von: Ka, Keummin, et al.
Veröffentlicht: (2025)
von: Ka, Keummin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Ideological Bias in LLMs' Economic Causal Reasoning
von: Lee, Donggyu, et al.
Veröffentlicht: (2026) -
Adversarial Style Augmentation via Large Language Model for Robust Fake News Detection
von: Park, Sungwon, et al.
Veröffentlicht: (2024) -
GeoSEE: Regional Socio-Economic Estimation With a Large Language Model
von: Han, Sungwon, et al.
Veröffentlicht: (2024) -
Generalizable Disaster Damage Assessment via Change Detection with Vision Foundation Model
von: Ahn, Kyeongjin, et al.
Veröffentlicht: (2024) -
Dropouts in Confidence: Moral Uncertainty in Human-LLM Alignment
von: Kwon, Jea, et al.
Veröffentlicht: (2025)