Advancing ESG Intelligence: An Expert-level Agent and Comprehensive Benchmark for Sustainable Finance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Yilei, Zhang, Wentao, Xiao, Lei, Zheng, Yandan, Liu, Mengpu, Lim, Wei Yang Bryan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AgentOrchestra: Orchestrating Multi-Agent Intelligence with the Tool-Environment-Agent(TEA) Protocol
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
STORM: A Spatio-Temporal Factor Model Based on Dual Vector Quantized Variational Autoencoders for Financial Trading
von: Zhao, Yilei, et al.
Veröffentlicht: (2024)
von: Zhao, Yilei, et al.
Veröffentlicht: (2024)
Towards Competent AI for Fundamental Analysis in Finance: A Benchmark Dataset and Evaluation
von: Wu, Zonghan, et al.
Veröffentlicht: (2025)
von: Wu, Zonghan, et al.
Veröffentlicht: (2025)
Agent Manufacturing: Foundation-Model Agents as First-Class Industrial Entities
von: Zhang, Yilei
Veröffentlicht: (2026)
von: Zhang, Yilei
Veröffentlicht: (2026)
VoiceAgentEval: A Dual-Dimensional Benchmark for Expert-Level Intelligent Voice-Agent Evaluation of Xbench's Professional-Aligned Series
von: Xu, Pengyu, et al.
Veröffentlicht: (2025)
von: Xu, Pengyu, et al.
Veröffentlicht: (2025)
Ready Jurist One: Benchmarking Language Agents for Legal Intelligence in Dynamic Environments
von: Jia, Zheng, et al.
Veröffentlicht: (2025)
von: Jia, Zheng, et al.
Veröffentlicht: (2025)
EvoCodeBench: A Human-Performance Benchmark for Self-Evolving LLM-Driven Coding Systems
von: Zhang, Wentao, et al.
Veröffentlicht: (2026)
von: Zhang, Wentao, et al.
Veröffentlicht: (2026)
COSINT-Agent: A Knowledge-Driven Multimodal Agent for Chinese Open Source Intelligence
von: Li, Wentao, et al.
Veröffentlicht: (2025)
von: Li, Wentao, et al.
Veröffentlicht: (2025)
Empowering Sustainable Finance with Artificial Intelligence: A Framework for Responsible Implementation
von: Pavlidis, Georgios
Veröffentlicht: (2025)
von: Pavlidis, Georgios
Veröffentlicht: (2025)
DAComp: Benchmarking Data Agents across the Full Data Intelligence Lifecycle
von: Lei, Fangyu, et al.
Veröffentlicht: (2025)
von: Lei, Fangyu, et al.
Veröffentlicht: (2025)
ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence
von: Ma, Menghe, et al.
Veröffentlicht: (2026)
von: Ma, Menghe, et al.
Veröffentlicht: (2026)
ESGenius: Benchmarking LLMs on Environmental, Social, and Governance (ESG) and Sustainability Knowledge
von: He, Chaoyue, et al.
Veröffentlicht: (2025)
von: He, Chaoyue, et al.
Veröffentlicht: (2025)
Is Your VLM for Autonomous Driving Safety-Ready? A Comprehensive Benchmark for Evaluating External and In-Cabin Risks
von: Meng, Xianhui, et al.
Veröffentlicht: (2025)
von: Meng, Xianhui, et al.
Veröffentlicht: (2025)
AlphaForgeBench: Benchmarking End-to-End Trading Strategy Design with Large Language Models
von: Zhang, Wentao, et al.
Veröffentlicht: (2026)
von: Zhang, Wentao, et al.
Veröffentlicht: (2026)
GeoAgentBench: A Dynamic Execution Benchmark for Tool-Augmented Agents in Spatial Analysis
von: Yu, Bo, et al.
Veröffentlicht: (2026)
von: Yu, Bo, et al.
Veröffentlicht: (2026)
FinWorld: An All-in-One Open-Source Platform for End-to-End Financial AI Research and Deployment
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
FraudBench: A Multimodal Benchmark for Detecting AI-Generated Fraudulent Refund Evidence
von: Yan, Xinyu, et al.
Veröffentlicht: (2026)
von: Yan, Xinyu, et al.
Veröffentlicht: (2026)
EHRStruct: A Comprehensive Benchmark Framework for Evaluating Large Language Models on Structured Electronic Health Record Tasks
von: Yang, Xiao, et al.
Veröffentlicht: (2025)
von: Yang, Xiao, et al.
Veröffentlicht: (2025)
AI Agents for Sustainable SMEs: A Green ESG Assessment Framework
von: Trinh, Viet, et al.
Veröffentlicht: (2026)
von: Trinh, Viet, et al.
Veröffentlicht: (2026)
Recent Advances in Multi-modal 3D Intelligence: A Comprehensive Survey and Evaluation
von: Lei, Yinjie, et al.
Veröffentlicht: (2023)
von: Lei, Yinjie, et al.
Veröffentlicht: (2023)
ELAIPBench: A Benchmark for Expert-Level Artificial Intelligence Paper Understanding
von: Dai, Xinbang, et al.
Veröffentlicht: (2025)
von: Dai, Xinbang, et al.
Veröffentlicht: (2025)
ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices
von: Kong, Dezhi, et al.
Veröffentlicht: (2026)
von: Kong, Dezhi, et al.
Veröffentlicht: (2026)
Integrating ESG and AI: A Comprehensive Responsible AI Assessment Framework
von: Lee, Sung Une, et al.
Veröffentlicht: (2024)
von: Lee, Sung Une, et al.
Veröffentlicht: (2024)
ESG-Bench: Benchmarking Long-Context ESG Reports for Hallucination Mitigation
von: Sun, Siqi, et al.
Veröffentlicht: (2026)
von: Sun, Siqi, et al.
Veröffentlicht: (2026)
DataCross: A Unified Benchmark and Agent Framework for Cross-Modal Heterogeneous Data Analysis
von: Qi, Ruyi, et al.
Veröffentlicht: (2026)
von: Qi, Ruyi, et al.
Veröffentlicht: (2026)
MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs
von: Yuan, Jiakang, et al.
Veröffentlicht: (2025)
von: Yuan, Jiakang, et al.
Veröffentlicht: (2025)
ELABORATION: A Comprehensive Benchmark on Human-LLM Competitive Programming
von: Yang, Xinwei, et al.
Veröffentlicht: (2025)
von: Yang, Xinwei, et al.
Veröffentlicht: (2025)
SPA-Bench: A Comprehensive Benchmark for SmartPhone Agent Evaluation
von: Chen, Jingxuan, et al.
Veröffentlicht: (2024)
von: Chen, Jingxuan, et al.
Veröffentlicht: (2024)
Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values
von: Dong, Haonan, et al.
Veröffentlicht: (2026)
von: Dong, Haonan, et al.
Veröffentlicht: (2026)
Data and System Perspectives of Sustainable Artificial Intelligence
von: Xie, Tao, et al.
Veröffentlicht: (2025)
von: Xie, Tao, et al.
Veröffentlicht: (2025)
FORTIS: Benchmarking Over-Privilege in Agent Skills
von: Li, Shawn, et al.
Veröffentlicht: (2026)
von: Li, Shawn, et al.
Veröffentlicht: (2026)
TestAgent: An Adaptive and Intelligent Expert for Human Assessment
von: Yu, Junhao, et al.
Veröffentlicht: (2025)
von: Yu, Junhao, et al.
Veröffentlicht: (2025)
MobileBench-OL: A Comprehensive Chinese Benchmark for Evaluating Mobile GUI Agents in Real-World Environment
von: Wu, Qinzhuo, et al.
Veröffentlicht: (2026)
von: Wu, Qinzhuo, et al.
Veröffentlicht: (2026)
Bench-CoE: a Framework for Collaboration of Experts from Benchmark
von: Wang, Yuanshuai, et al.
Veröffentlicht: (2024)
von: Wang, Yuanshuai, et al.
Veröffentlicht: (2024)
General-Purpose Aerial Intelligent Agents Empowered by Large Language Models
von: Zhao, Ji, et al.
Veröffentlicht: (2025)
von: Zhao, Ji, et al.
Veröffentlicht: (2025)
Terra Nova: A Comprehensive Challenge Environment for Intelligent Agents
von: McInroe, Trevor
Veröffentlicht: (2025)
von: McInroe, Trevor
Veröffentlicht: (2025)
GUI Testing Arena: A Unified Benchmark for Advancing Autonomous GUI Testing Agent
von: Zhao, Kangjia, et al.
Veröffentlicht: (2024)
von: Zhao, Kangjia, et al.
Veröffentlicht: (2024)
TravelEval: A Comprehensive Benchmarking Framework for Evaluating LLM-Powered Travel Planning Agents
von: Chen, Weiyi, et al.
Veröffentlicht: (2026)
von: Chen, Weiyi, et al.
Veröffentlicht: (2026)
A Multimodal Foundation Agent for Financial Trading: Tool-Augmented, Diversified, and Generalist
von: Zhang, Wentao, et al.
Veröffentlicht: (2024)
von: Zhang, Wentao, et al.
Veröffentlicht: (2024)
Intelligent Computing: The Latest Advances, Challenges and Future
von: Zhu, Shiqiang, et al.
Veröffentlicht: (2022)
von: Zhu, Shiqiang, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
AgentOrchestra: Orchestrating Multi-Agent Intelligence with the Tool-Environment-Agent(TEA) Protocol
von: Zhang, Wentao, et al.
Veröffentlicht: (2025) -
STORM: A Spatio-Temporal Factor Model Based on Dual Vector Quantized Variational Autoencoders for Financial Trading
von: Zhao, Yilei, et al.
Veröffentlicht: (2024) -
Towards Competent AI for Fundamental Analysis in Finance: A Benchmark Dataset and Evaluation
von: Wu, Zonghan, et al.
Veröffentlicht: (2025) -
Agent Manufacturing: Foundation-Model Agents as First-Class Industrial Entities
von: Zhang, Yilei
Veröffentlicht: (2026) -
VoiceAgentEval: A Dual-Dimensional Benchmark for Expert-Level Intelligent Voice-Agent Evaluation of Xbench's Professional-Aligned Series
von: Xu, Pengyu, et al.
Veröffentlicht: (2025)