Exploring the Capability Boundaries of LLMs in Mastering of Chinese Chouxiang Language
Fuente:
arXiv
Guardado en:
| Autores principales: | Lin, Dianqing, Lan, Tian, Zhu, Jiali, Li, Jiang, Chen, Wei, Liu, Xu, Aruukhan, Su, Xiangdong, Hou, Hongxu, Gao, Guanglai |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry
por: Li, Jiang, et al.
Publicado: (2026)
por: Li, Jiang, et al.
Publicado: (2026)
McBE: A Multi-task Chinese Bias Evaluation Benchmark for Large Language Models
por: Lan, Tian, et al.
Publicado: (2025)
por: Lan, Tian, et al.
Publicado: (2025)
Mitigating Heterogeneity among Factor Tensors via Lie Group Manifolds for Tensor Decomposition Based Temporal Knowledge Graph Embedding
por: Li, Jiang, et al.
Publicado: (2024)
por: Li, Jiang, et al.
Publicado: (2024)
TransERR: Translation-based Knowledge Graph Embedding via Efficient Relation Rotation
por: Li, Jiang, et al.
Publicado: (2023)
por: Li, Jiang, et al.
Publicado: (2023)
RepCali: High Efficient Fine-tuning Via Representation Calibration in Latent Space for Pre-trained Language Models
por: Zhang, Fujun, et al.
Publicado: (2025)
por: Zhang, Fujun, et al.
Publicado: (2025)
Explore the Reasoning Capability of LLMs in the Chess Testbed
por: Wang, Shu, et al.
Publicado: (2024)
por: Wang, Shu, et al.
Publicado: (2024)
Mitigating Biases in Language Models via Bias Unlearning
por: Liu, Dianqing, et al.
Publicado: (2025)
por: Liu, Dianqing, et al.
Publicado: (2025)
FoundaBench: Evaluating Chinese Fundamental Knowledge Capabilities of Large Language Models
por: Li, Wei, et al.
Publicado: (2024)
por: Li, Wei, et al.
Publicado: (2024)
Can LLMs Act as Historians? Evaluating Historical Research Capabilities of LLMs via the Chinese Imperial Examination
por: Gao, Lirong, et al.
Publicado: (2026)
por: Gao, Lirong, et al.
Publicado: (2026)
Olapa-MCoT: Enhancing the Chinese Mathematical Reasoning Capability of LLMs
por: Zhu, Shaojie, et al.
Publicado: (2023)
por: Zhu, Shaojie, et al.
Publicado: (2023)
Leveraging Importance Sampling to Detach Alignment Modules from Large Language Models
por: Liu, Yi, et al.
Publicado: (2025)
por: Liu, Yi, et al.
Publicado: (2025)
Training-Inference Consistent Segmented Execution for Long-Context LLMs
por: Shang, Xianpeng, et al.
Publicado: (2026)
por: Shang, Xianpeng, et al.
Publicado: (2026)
How Chinese are Chinese Language Models? The Puzzling Lack of Language Policy in China's LLMs
por: Wen-Yi, Andrea W, et al.
Publicado: (2024)
por: Wen-Yi, Andrea W, et al.
Publicado: (2024)
LPO: Discovering Missed Peephole Optimizations with Large Language Models
por: Xu, Zhenyang, et al.
Publicado: (2025)
por: Xu, Zhenyang, et al.
Publicado: (2025)
CNSL-bench: Benchmarking the Sign Language Understanding Capabilities of MLLMs on Chinese National Sign Language
por: Zhao, Rui, et al.
Publicado: (2026)
por: Zhao, Rui, et al.
Publicado: (2026)
Think in Sentences: Explicit Sentence Boundaries Enhance Language Model's Capabilities
por: Liu, Zhichen, et al.
Publicado: (2026)
por: Liu, Zhichen, et al.
Publicado: (2026)
OpenEval: Benchmarking Chinese LLMs across Capability, Alignment and Safety
por: Liu, Chuang, et al.
Publicado: (2024)
por: Liu, Chuang, et al.
Publicado: (2024)
Evaluating the Generation Capabilities of Large Chinese Language Models
por: Zeng, Hui, et al.
Publicado: (2023)
por: Zeng, Hui, et al.
Publicado: (2023)
Analysis of Indic Language Capabilities in LLMs
por: Vaidya, Aatman, et al.
Publicado: (2025)
por: Vaidya, Aatman, et al.
Publicado: (2025)
STEM: Efficient Relative Capability Evaluation of LLMs through Structured Transition Samples
por: Hu, Haiquan, et al.
Publicado: (2025)
por: Hu, Haiquan, et al.
Publicado: (2025)
Is LLM an Overconfident Judge? Unveiling the Capabilities of LLMs in Detecting Offensive Language with Annotation Disagreement
por: Lu, Junyu, et al.
Publicado: (2025)
por: Lu, Junyu, et al.
Publicado: (2025)
CNNSum: Exploring Long-Context Summarization with Large Language Models in Chinese Novels
por: Wei, Lingxiao, et al.
Publicado: (2024)
por: Wei, Lingxiao, et al.
Publicado: (2024)
L^2GC:Lorentzian Linear Graph Convolutional Networks for Node Classification
por: Liang, Qiuyu, et al.
Publicado: (2024)
por: Liang, Qiuyu, et al.
Publicado: (2024)
Evaluating LLMs on Chinese Idiom Translation
por: Yang, Cai, et al.
Publicado: (2025)
por: Yang, Cai, et al.
Publicado: (2025)
Reasoning Under Uncertainty: Exploring Probabilistic Reasoning Capabilities of LLMs
por: Pournemat, Mobina, et al.
Publicado: (2025)
por: Pournemat, Mobina, et al.
Publicado: (2025)
Unifying Dual-Space Embedding for Entity Alignment via Contrastive Learning
por: Wang, Cunda, et al.
Publicado: (2024)
por: Wang, Cunda, et al.
Publicado: (2024)
Chinese Word Boundary Recovery through Character Alignment Projection
por: Wang, Lusha, et al.
Publicado: (2026)
por: Wang, Lusha, et al.
Publicado: (2026)
Are the LLMs Capable of Maintaining at Least the Language Genus?
por: Mitrović, Sandra, et al.
Publicado: (2025)
por: Mitrović, Sandra, et al.
Publicado: (2025)
CSRP: Chain-of-Thought Reasoning for Chinese Text Correction via Reinforcement Learning with Efficiency-Aware Rewards
por: Tian, Wei, et al.
Publicado: (2026)
por: Tian, Wei, et al.
Publicado: (2026)
ZeroSearch: Incentivize the Search Capability of LLMs without Searching
por: Sun, Hao, et al.
Publicado: (2025)
por: Sun, Hao, et al.
Publicado: (2025)
SCAN: Structured Capability Assessment and Navigation for LLMs
por: Wang, Zongqi, et al.
Publicado: (2025)
por: Wang, Zongqi, et al.
Publicado: (2025)
Analyzing LLMs' Knowledge Boundary Cognition Across Languages Through the Lens of Internal Representations
por: Xiao, Chenghao, et al.
Publicado: (2025)
por: Xiao, Chenghao, et al.
Publicado: (2025)
Exploring the Capabilities of ChatGPT in Ancient Chinese Translation and Person Name Recognition
por: Si, Shijing, et al.
Publicado: (2023)
por: Si, Shijing, et al.
Publicado: (2023)
RL-PLUS: Countering Capability Boundary Collapse of LLMs in Reinforcement Learning with Hybrid-policy Optimization
por: Dong, Yihong, et al.
Publicado: (2025)
por: Dong, Yihong, et al.
Publicado: (2025)
From Hallucination to Structure Snowballing: The Alignment Tax of Constrained Decoding in LLM Reflection
por: Zhou, Hongxu
Publicado: (2026)
por: Zhou, Hongxu
Publicado: (2026)
Benchmarking Chinese Commonsense Reasoning of LLMs: From Chinese-Specifics to Reasoning-Memorization Correlations
por: Sun, Jiaxing, et al.
Publicado: (2024)
por: Sun, Jiaxing, et al.
Publicado: (2024)
Leveraging Large Language Models for Generalizing Peephole Optimizations
por: Liao, Chunhao, et al.
Publicado: (2026)
por: Liao, Chunhao, et al.
Publicado: (2026)
PlanGenLLMs: A Modern Survey of LLM Planning Capabilities
por: Wei, Hui, et al.
Publicado: (2025)
por: Wei, Hui, et al.
Publicado: (2025)
AC-EVAL: Evaluating Ancient Chinese Language Understanding in Large Language Models
por: Wei, Yuting, et al.
Publicado: (2024)
por: Wei, Yuting, et al.
Publicado: (2024)
Confidence v.s. Critique: A Decomposition of Self-Correction Capability for LLMs
por: Yang, Zhe, et al.
Publicado: (2024)
por: Yang, Zhe, et al.
Publicado: (2024)
Ejemplares similares
-
Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry
por: Li, Jiang, et al.
Publicado: (2026) -
McBE: A Multi-task Chinese Bias Evaluation Benchmark for Large Language Models
por: Lan, Tian, et al.
Publicado: (2025) -
Mitigating Heterogeneity among Factor Tensors via Lie Group Manifolds for Tensor Decomposition Based Temporal Knowledge Graph Embedding
por: Li, Jiang, et al.
Publicado: (2024) -
TransERR: Translation-based Knowledge Graph Embedding via Efficient Relation Rotation
por: Li, Jiang, et al.
Publicado: (2023) -
RepCali: High Efficient Fine-tuning Via Representation Calibration in Latent Space for Pre-trained Language Models
por: Zhang, Fujun, et al.
Publicado: (2025)