Measuring Hong Kong Massive Multi-Task Language Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cao, Chuxue, Zhu, Zhenghao, Zhu, Junqi, Lu, Guoying, Peng, Siyu, Dai, Juntao, Shi, Weijie, Han, Sirui, Guo, Yike |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HKGAI-V1: Towards Regional Sovereign Large Language Model for Hong Kong
von: Han, Sirui, et al.
Veröffentlicht: (2025)
von: Han, Sirui, et al.
Veröffentlicht: (2025)
SafeLawBench: Towards Safe Alignment of Large Language Models
von: Cao, Chuxue, et al.
Veröffentlicht: (2025)
von: Cao, Chuxue, et al.
Veröffentlicht: (2025)
MedInsightBench: Evaluating Medical Analytics Agents Through Multi-Step Insight Discovery in Multimodal Medical Data
von: Zhu, Zhenghao, et al.
Veröffentlicht: (2025)
von: Zhu, Zhenghao, et al.
Veröffentlicht: (2025)
Perception, Understanding and Reasoning, A Multimodal Benchmark for Video Fake News Detection
von: Yakun, Cui, et al.
Veröffentlicht: (2025)
von: Yakun, Cui, et al.
Veröffentlicht: (2025)
Towards Advanced Mathematical Reasoning for LLMs via First-Order Logic Theorem Proving
von: Cao, Chuxue, et al.
Veröffentlicht: (2025)
von: Cao, Chuxue, et al.
Veröffentlicht: (2025)
LRAS: Advanced Legal Reasoning with Agentic Search
von: Zhou, Yujin, et al.
Veröffentlicht: (2026)
von: Zhou, Yujin, et al.
Veröffentlicht: (2026)
SafeMT: Multi-turn Safety for Multimodal Language Models
von: Zhu, Han, et al.
Veröffentlicht: (2025)
von: Zhu, Han, et al.
Veröffentlicht: (2025)
Benchmarking Multi-National Value Alignment for Large Language Models
von: Shi, Weijie, et al.
Veröffentlicht: (2025)
von: Shi, Weijie, et al.
Veröffentlicht: (2025)
MMFCTUB: Multi-Modal Financial Credit Table Understanding Benchmark
von: Yakun, Cui, et al.
Veröffentlicht: (2026)
von: Yakun, Cui, et al.
Veröffentlicht: (2026)
InsightEval: An Expert-Curated Benchmark for Assessing Insight Discovery in LLM-Driven Data Agents
von: Zhu, Zhenghao, et al.
Veröffentlicht: (2025)
von: Zhu, Zhenghao, et al.
Veröffentlicht: (2025)
Hong Kong
von: Edwards, Mike
Veröffentlicht: (1997)
von: Edwards, Mike
Veröffentlicht: (1997)
Hong Kong
Veröffentlicht: (2007)
Veröffentlicht: (2007)
LegalReasoner: Step-wised Verification-Correction for Legal Judgment Reasoning
von: Shi, Weijie, et al.
Veröffentlicht: (2025)
von: Shi, Weijie, et al.
Veröffentlicht: (2025)
Pushing the Boundaries of Natural Reasoning: Interleaved Bonus from Formal-Logic Verification
von: Cao, Chuxue, et al.
Veröffentlicht: (2026)
von: Cao, Chuxue, et al.
Veröffentlicht: (2026)
Hong Kong 2012
Veröffentlicht: (2012)
Veröffentlicht: (2012)
Rumpus in Hong Kong
Veröffentlicht: (1997)
Veröffentlicht: (1997)
Orchid Flora of Hong Kong II: Rediscovery of Zeuxine membranacea in Hong Kong
von: Kit Shan Wong, et al.
Veröffentlicht: (2025)
von: Kit Shan Wong, et al.
Veröffentlicht: (2025)
Global Hong Kong: Post‐2019 migration and the new Hong Kong diaspora
von: Gordon Mathews
Veröffentlicht: (2026)
von: Gordon Mathews
Veröffentlicht: (2026)
30th Anniversary of the Department of Chemistry, City University of Hong Kong
von: Zonglong Zhu, et al.
Veröffentlicht: (2024)
von: Zonglong Zhu, et al.
Veröffentlicht: (2024)
Letter from Hong Kong: Vaccination trends for respiratory viral infections in Hong Kong
von: Ken Ka Pang Chan, et al.
Veröffentlicht: (2024)
von: Ken Ka Pang Chan, et al.
Veröffentlicht: (2024)
ThinkPatterns-21k: A Systematic Study on the Impact of Thinking Patterns in LLMs
von: Wen, Pengcheng, et al.
Veröffentlicht: (2025)
von: Wen, Pengcheng, et al.
Veröffentlicht: (2025)
Religion and Official Politics in Contemporary China, Hong Kong, and Taiwan
von: Ting Guo
Veröffentlicht: (2026)
von: Ting Guo
Veröffentlicht: (2026)
Hong Kong Journal of Radiology
Veröffentlicht: (2021)
Veröffentlicht: (2021)
Hong Kong Medical Journal
Veröffentlicht: (2004)
Veröffentlicht: (2004)
Hong Kong Physiotherapy Journal
Veröffentlicht: (2016)
Veröffentlicht: (2016)
Hong Kong, democracy denied
Veröffentlicht: (2008)
Veröffentlicht: (2008)
Hong Kong. China's gamble
Veröffentlicht: (1996)
Veröffentlicht: (1996)
Hong Kong countdown to 1997
Veröffentlicht: (1997)
Veröffentlicht: (1997)
Ambulatory labour in Hong Kong
von: Fengxuan Xue, et al.
Veröffentlicht: (1980)
von: Fengxuan Xue, et al.
Veröffentlicht: (1980)
Labour in Hong Kong in 1934
Veröffentlicht: (1936)
Veröffentlicht: (1936)
Automate Strategy Finding with LLM in Quant Investment
von: Kou, Zhizhuo, et al.
Veröffentlicht: (2024)
von: Kou, Zhizhuo, et al.
Veröffentlicht: (2024)
MMTU: A Massive Multi-Task Table Understanding and Reasoning Benchmark
von: Xing, Junjie, et al.
Veröffentlicht: (2025)
von: Xing, Junjie, et al.
Veröffentlicht: (2025)
Can Post-Training Transform LLMs into Causal Reasoners?
von: Chen, Junqi, et al.
Veröffentlicht: (2026)
von: Chen, Junqi, et al.
Veröffentlicht: (2026)
When Slower Isn't Truer: Inverse Scaling Law of Truthfulness in Multimodal Reasoning
von: Fang, Sitong, et al.
Veröffentlicht: (2025)
von: Fang, Sitong, et al.
Veröffentlicht: (2025)
Unlocking Data Value in Finance: A Study on Distillation and Difficulty-Aware Training
von: Cao, Chuxue, et al.
Veröffentlicht: (2026)
von: Cao, Chuxue, et al.
Veröffentlicht: (2026)
A Multi‐Scale Approach to Assessing Positioning Accuracy in High‐Density Urban Environments of Hong Kong
von: Tianyu Li, et al.
Veröffentlicht: (2026)
von: Tianyu Li, et al.
Veröffentlicht: (2026)
Hong Kong Journal of Occupational Therapy
Veröffentlicht: (2016)
Veröffentlicht: (2016)
Hong Kong Journal of Emergency Medicine
Veröffentlicht: (2020)
Veröffentlicht: (2020)
Contesting Education and Identity in Hong Kong
von: Jackson, Liz
Veröffentlicht: (2025)
von: Jackson, Liz
Veröffentlicht: (2025)
Hong Kong Engineering Geological Model
von: Zhao, Zening, et al.
Veröffentlicht: (2026)
von: Zhao, Zening, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
HKGAI-V1: Towards Regional Sovereign Large Language Model for Hong Kong
von: Han, Sirui, et al.
Veröffentlicht: (2025) -
SafeLawBench: Towards Safe Alignment of Large Language Models
von: Cao, Chuxue, et al.
Veröffentlicht: (2025) -
MedInsightBench: Evaluating Medical Analytics Agents Through Multi-Step Insight Discovery in Multimodal Medical Data
von: Zhu, Zhenghao, et al.
Veröffentlicht: (2025) -
Perception, Understanding and Reasoning, A Multimodal Benchmark for Video Fake News Detection
von: Yakun, Cui, et al.
Veröffentlicht: (2025) -
Towards Advanced Mathematical Reasoning for LLMs via First-Order Logic Theorem Proving
von: Cao, Chuxue, et al.
Veröffentlicht: (2025)