Unveiling Linguistic Regions in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Zhihao, Zhao, Jun, Zhang, Qi, Gui, Tao, Huang, Xuanjing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLaMA Beyond English: An Empirical Study on Language Capability Transfer
von: Zhao, Jun, et al.
Veröffentlicht: (2024)
von: Zhao, Jun, et al.
Veröffentlicht: (2024)
ToolSword: Unveiling Safety Issues of Large Language Models in Tool Learning Across Three Stages
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
RoCoIns: Enhancing Robustness of Large Language Models through Code-Style Instructions
von: Zhang, Yuansen, et al.
Veröffentlicht: (2024)
von: Zhang, Yuansen, et al.
Veröffentlicht: (2024)
Self-Demos: Eliciting Out-of-Demonstration Generalizability in Large Language Models
von: He, Wei, et al.
Veröffentlicht: (2024)
von: He, Wei, et al.
Veröffentlicht: (2024)
LLM-DA: Data Augmentation via Large Language Models for Few-Shot Named Entity Recognition
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
Exploring the Compositional Deficiency of Large Language Models in Mathematical Reasoning
von: Zhao, Jun, et al.
Veröffentlicht: (2024)
von: Zhao, Jun, et al.
Veröffentlicht: (2024)
Are Large Language Models Good Prompt Optimizers?
von: Ma, Ruotian, et al.
Veröffentlicht: (2024)
von: Ma, Ruotian, et al.
Veröffentlicht: (2024)
Navigating the OverKill in Large Language Models
von: Shi, Chenyu, et al.
Veröffentlicht: (2024)
von: Shi, Chenyu, et al.
Veröffentlicht: (2024)
Self-Polish: Enhance Reasoning in Large Language Models via Problem Refinement
von: Xi, Zhiheng, et al.
Veröffentlicht: (2023)
von: Xi, Zhiheng, et al.
Veröffentlicht: (2023)
CodeChameleon: Personalized Encryption Framework for Jailbreaking Large Language Models
von: Lv, Huijie, et al.
Veröffentlicht: (2024)
von: Lv, Huijie, et al.
Veröffentlicht: (2024)
Emergent Structured Representations Support Flexible In-Context Inference in Large Language Models
von: Xu, Ningyu, et al.
Veröffentlicht: (2026)
von: Xu, Ningyu, et al.
Veröffentlicht: (2026)
LongAgent: Scaling Language Models to 128k Context through Multi-Agent Collaboration
von: Zhao, Jun, et al.
Veröffentlicht: (2024)
von: Zhao, Jun, et al.
Veröffentlicht: (2024)
On the Tip of the Tongue: Analyzing Conceptual Representation in Large Language Models with Reverse-Dictionary Probe
von: Xu, Ningyu, et al.
Veröffentlicht: (2024)
von: Xu, Ningyu, et al.
Veröffentlicht: (2024)
RoTBench: A Multi-Level Benchmark for Evaluating the Robustness of Large Language Models in Tool Learning
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
LongHeads: Multi-Head Attention is Secretly a Long Context Processor
von: Lu, Yi, et al.
Veröffentlicht: (2024)
von: Lu, Yi, et al.
Veröffentlicht: (2024)
Domain Generalization via Causal Adjustment for Cross-Domain Sentiment Analysis
von: Wang, Siyin, et al.
Veröffentlicht: (2024)
von: Wang, Siyin, et al.
Veröffentlicht: (2024)
Advancing Translation Preference Modeling with RLHF: A Step Towards Cost-Effective Solution
von: Xu, Nuo, et al.
Veröffentlicht: (2024)
von: Xu, Nuo, et al.
Veröffentlicht: (2024)
CCTU: A Benchmark for Tool Use under Complex Constraints
von: Ye, Junjie, et al.
Veröffentlicht: (2026)
von: Ye, Junjie, et al.
Veröffentlicht: (2026)
Unveiling the Truth and Facilitating Change: Towards Agent-based Large-scale Social Movement Simulation
von: Mou, Xinyi, et al.
Veröffentlicht: (2024)
von: Mou, Xinyi, et al.
Veröffentlicht: (2024)
Compression Hacking: A Supplementary Perspective on Informatics Properties of Language Models from Geometric Distortion
von: Zang, Jianxiang, et al.
Veröffentlicht: (2025)
von: Zang, Jianxiang, et al.
Veröffentlicht: (2025)
ToolEyes: Fine-Grained Evaluation for Tool Learning Capabilities of Large Language Models in Real-world Scenarios
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
LLMEval-Fair: A Large-Scale Longitudinal Study on Robust and Fair Evaluation of Large Language Models
von: Zhang, Ming, et al.
Veröffentlicht: (2025)
von: Zhang, Ming, et al.
Veröffentlicht: (2025)
Inverse-Q*: Token Level Reinforcement Learning for Aligning Large Language Models Without Preference Data
von: Xia, Han, et al.
Veröffentlicht: (2024)
von: Xia, Han, et al.
Veröffentlicht: (2024)
Enhancing the Capability and Robustness of Large Language Models through Reinforcement Learning-Driven Query Refinement
von: Wang, Xiaohua, et al.
Veröffentlicht: (2024)
von: Wang, Xiaohua, et al.
Veröffentlicht: (2024)
Feedback-Driven Tool-Use Improvements in Large Language Models via Automated Build Environments
von: Ye, Junjie, et al.
Veröffentlicht: (2025)
von: Ye, Junjie, et al.
Veröffentlicht: (2025)
LoRAMoE: Alleviate World Knowledge Forgetting in Large Language Models via MoE-Style Plugin
von: Dou, Shihan, et al.
Veröffentlicht: (2023)
von: Dou, Shihan, et al.
Veröffentlicht: (2023)
Advancing Block Diffusion Language Models for Test-Time Scaling
von: Lu, Yi, et al.
Veröffentlicht: (2026)
von: Lu, Yi, et al.
Veröffentlicht: (2026)
Length Generalization of Causal Transformers without Position Encoding
von: Wang, Jie, et al.
Veröffentlicht: (2024)
von: Wang, Jie, et al.
Veröffentlicht: (2024)
TL-Training: A Task-Feature-Based Framework for Training Large Language Models in Tool Use
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
von: Ye, Junjie, et al.
Veröffentlicht: (2024)
Subspace Defense: Discarding Adversarial Perturbations by Learning a Subspace for Clean Signals
von: Zheng, Rui, et al.
Veröffentlicht: (2024)
von: Zheng, Rui, et al.
Veröffentlicht: (2024)
Error Classification of Large Language Models on Math Word Problems: A Dynamically Adaptive Framework
von: Sun, Yuhong, et al.
Veröffentlicht: (2025)
von: Sun, Yuhong, et al.
Veröffentlicht: (2025)
ToolHop: A Query-Driven Benchmark for Evaluating Large Language Models in Multi-Hop Tool Use
von: Ye, Junjie, et al.
Veröffentlicht: (2025)
von: Ye, Junjie, et al.
Veröffentlicht: (2025)
Improving RL Exploration for LLM Reasoning through Retrospective Replay
von: Dou, Shihan, et al.
Veröffentlicht: (2025)
von: Dou, Shihan, et al.
Veröffentlicht: (2025)
What's Wrong with Your Code Generated by Large Language Models? An Extensive Study
von: Dou, Shihan, et al.
Veröffentlicht: (2024)
von: Dou, Shihan, et al.
Veröffentlicht: (2024)
Distill Visual Chart Reasoning Ability from LLMs to MLLMs
von: He, Wei, et al.
Veröffentlicht: (2024)
von: He, Wei, et al.
Veröffentlicht: (2024)
Unveiling the Misuse Potential of Base Large Language Models via In-Context Learning
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
Beyond Boundaries: Learning a Universal Entity Taxonomy across Datasets and Languages for Open Named Entity Recognition
von: Yang, Yuming, et al.
Veröffentlicht: (2024)
von: Yang, Yuming, et al.
Veröffentlicht: (2024)
Effective Length Extrapolation via Dimension-Wise Positional Embeddings Manipulation
von: Lu, Yi, et al.
Veröffentlicht: (2025)
von: Lu, Yi, et al.
Veröffentlicht: (2025)
Intrinsic Model Weaknesses: How Priming Attacks Unveil Vulnerabilities in Large Language Models
von: Huang, Yuyi, et al.
Veröffentlicht: (2025)
von: Huang, Yuyi, et al.
Veröffentlicht: (2025)
MulDimIF: A Multi-Dimensional Constraint Framework for Evaluating and Improving Instruction Following in Large Language Models
von: Ye, Junjie, et al.
Veröffentlicht: (2025)
von: Ye, Junjie, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LLaMA Beyond English: An Empirical Study on Language Capability Transfer
von: Zhao, Jun, et al.
Veröffentlicht: (2024) -
ToolSword: Unveiling Safety Issues of Large Language Models in Tool Learning Across Three Stages
von: Ye, Junjie, et al.
Veröffentlicht: (2024) -
RoCoIns: Enhancing Robustness of Large Language Models through Code-Style Instructions
von: Zhang, Yuansen, et al.
Veröffentlicht: (2024) -
Self-Demos: Eliciting Out-of-Demonstration Generalizability in Large Language Models
von: He, Wei, et al.
Veröffentlicht: (2024) -
LLM-DA: Data Augmentation via Large Language Models for Few-Shot Named Entity Recognition
von: Ye, Junjie, et al.
Veröffentlicht: (2024)