Gespeichert in:
| Hauptverfasser: | Zhu, Hongyin, Peng, Hao, Lyu, Zhiheng, Hou, Lei, Li, Juanzi, Xiao, Jinghui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2021
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2109.01048 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Pre-training Distillation for Large Language Models: A Design Space Exploration
von: Peng, Hao, et al.
Veröffentlicht: (2024)
von: Peng, Hao, et al.
Veröffentlicht: (2024)
AtomR: Atomic Operator-Empowered Large Language Models for Heterogeneous Knowledge Reasoning
von: Xin, Amy, et al.
Veröffentlicht: (2024)
von: Xin, Amy, et al.
Veröffentlicht: (2024)
R-Eval: A Unified Toolkit for Evaluating Domain Knowledge of Retrieval Augmented Large Language Models
von: Tu, Shangqing, et al.
Veröffentlicht: (2024)
von: Tu, Shangqing, et al.
Veröffentlicht: (2024)
Adaptive Few-shot Prompting for Machine Translation with Pre-trained Language Models
von: Tang, Lei, et al.
Veröffentlicht: (2025)
von: Tang, Lei, et al.
Veröffentlicht: (2025)
ADELIE: Aligning Large Language Models on Information Extraction
von: Qi, Yunjia, et al.
Veröffentlicht: (2024)
von: Qi, Yunjia, et al.
Veröffentlicht: (2024)
MRCEval: A Comprehensive, Challenging and Accessible Machine Reading Comprehension Benchmark
von: Ma, Shengkun, et al.
Veröffentlicht: (2025)
von: Ma, Shengkun, et al.
Veröffentlicht: (2025)
Constraint Back-translation Improves Complex Instruction Following of Large Language Models
von: Qi, Yunjia, et al.
Veröffentlicht: (2024)
von: Qi, Yunjia, et al.
Veröffentlicht: (2024)
Challenges and Responses in the Practice of Large Language Models
von: Zhu, Hongyin
Veröffentlicht: (2024)
von: Zhu, Hongyin
Veröffentlicht: (2024)
Architectural Foundations for the Large Language Model Infrastructures
von: Zhu, Hongyin
Veröffentlicht: (2024)
von: Zhu, Hongyin
Veröffentlicht: (2024)
Unifying Ontology Construction and Semantic Alignment for Deterministic Enterprise Reasoning at Scale
von: Zhu, Hongyin
Veröffentlicht: (2026)
von: Zhu, Hongyin
Veröffentlicht: (2026)
OpenEP: Open-Ended Future Event Prediction
von: Guan, Yong, et al.
Veröffentlicht: (2024)
von: Guan, Yong, et al.
Veröffentlicht: (2024)
Event-level Knowledge Editing
von: Peng, Hao, et al.
Veröffentlicht: (2024)
von: Peng, Hao, et al.
Veröffentlicht: (2024)
KB-Plugin: A Plug-and-play Framework for Large Language Models to Induce Programs over Low-resourced Knowledge Bases
von: Zhang, Jiajie, et al.
Veröffentlicht: (2024)
von: Zhang, Jiajie, et al.
Veröffentlicht: (2024)
Domain Pre-training Impact on Representations
von: Gonzalez-Gutierrez, Cesar, et al.
Veröffentlicht: (2025)
von: Gonzalez-Gutierrez, Cesar, et al.
Veröffentlicht: (2025)
Data Mixing Agent: Learning to Re-weight Domains for Continual Pre-training
von: Yang, Kailai, et al.
Veröffentlicht: (2025)
von: Yang, Kailai, et al.
Veröffentlicht: (2025)
Boundary-Guided Policy Optimization for Memory-efficient RL of Diffusion Large Language Models
von: Lin, Nianyi, et al.
Veröffentlicht: (2025)
von: Lin, Nianyi, et al.
Veröffentlicht: (2025)
WildReward: Learning Reward Models from In-the-Wild Human Interactions
von: Peng, Hao, et al.
Veröffentlicht: (2026)
von: Peng, Hao, et al.
Veröffentlicht: (2026)
How Proficient Are Large Language Models in Formal Languages? An In-Depth Insight for Knowledge Base Question Answering
von: Liu, Jinxin, et al.
Veröffentlicht: (2024)
von: Liu, Jinxin, et al.
Veröffentlicht: (2024)
Untangle the KNOT: Interweaving Conflicting Knowledge and Reasoning Skills in Large Language Models
von: Liu, Yantao, et al.
Veröffentlicht: (2024)
von: Liu, Yantao, et al.
Veröffentlicht: (2024)
AGENTIF: Benchmarking Instruction Following of Large Language Models in Agentic Scenarios
von: Qi, Yunjia, et al.
Veröffentlicht: (2025)
von: Qi, Yunjia, et al.
Veröffentlicht: (2025)
Construct, Align, and Reason: Large Ontology Models for Enterprise Knowledge Management
von: Zhang, Yao, et al.
Veröffentlicht: (2026)
von: Zhang, Yao, et al.
Veröffentlicht: (2026)
Knowledge-to-Jailbreak: Investigating Knowledge-driven Jailbreaking Attacks for Large Language Models
von: Tu, Shangqing, et al.
Veröffentlicht: (2024)
von: Tu, Shangqing, et al.
Veröffentlicht: (2024)
Climate Change from Large Language Models
von: Zhu, Hongyin, et al.
Veröffentlicht: (2023)
von: Zhu, Hongyin, et al.
Veröffentlicht: (2023)
RM-Bench: Benchmarking Reward Models of Language Models with Subtlety and Style
von: Liu, Yantao, et al.
Veröffentlicht: (2024)
von: Liu, Yantao, et al.
Veröffentlicht: (2024)
Evaluating Generative Language Models in Information Extraction as Subjective Question Correction
von: Fan, Yuchen, et al.
Veröffentlicht: (2024)
von: Fan, Yuchen, et al.
Veröffentlicht: (2024)
StoryAlign: Evaluating and Training Reward Models for Story Generation
von: Xia, Haotian, et al.
Veröffentlicht: (2026)
von: Xia, Haotian, et al.
Veröffentlicht: (2026)
VerIF: Verification Engineering for Reinforcement Learning in Instruction Following
von: Peng, Hao, et al.
Veröffentlicht: (2025)
von: Peng, Hao, et al.
Veröffentlicht: (2025)
Guiding LLM Post-training Data Engineering with Model Internals from Sparse Autoencoders
von: Jing, Yi, et al.
Veröffentlicht: (2026)
von: Jing, Yi, et al.
Veröffentlicht: (2026)
Agentic Reward Modeling: Integrating Human Preferences with Verifiable Correctness Signals for Reliable Reward Systems
von: Peng, Hao, et al.
Veröffentlicht: (2025)
von: Peng, Hao, et al.
Veröffentlicht: (2025)
Pre-trained Language Model with Prompts for Temporal Knowledge Graph Completion
von: Xu, Wenjie, et al.
Veröffentlicht: (2023)
von: Xu, Wenjie, et al.
Veröffentlicht: (2023)
LLMAEL: Large Language Models are Good Context Augmenters for Entity Linking
von: Xin, Amy, et al.
Veröffentlicht: (2024)
von: Xin, Amy, et al.
Veröffentlicht: (2024)
WaterBench: Towards Holistic Evaluation of Watermarks for Large Language Models
von: Tu, Shangqing, et al.
Veröffentlicht: (2023)
von: Tu, Shangqing, et al.
Veröffentlicht: (2023)
KoCo: Conditioning Language Model Pre-training on Knowledge Coordinates
von: Li, Yudong, et al.
Veröffentlicht: (2026)
von: Li, Yudong, et al.
Veröffentlicht: (2026)
MAVEN-Fact: A Large-scale Event Factuality Detection Dataset
von: Li, Chunyang, et al.
Veröffentlicht: (2024)
von: Li, Chunyang, et al.
Veröffentlicht: (2024)
Reranking Passages with Coarse-to-Fine Neural Retriever Enhanced by List-Context Information
von: Zhu, Hongyin
Veröffentlicht: (2023)
von: Zhu, Hongyin
Veröffentlicht: (2023)
EventSum: A Large-Scale Event-Centric Summarization Dataset for Chinese Multi-News Documents
von: Zhu, Mengna, et al.
Veröffentlicht: (2024)
von: Zhu, Mengna, et al.
Veröffentlicht: (2024)
Knowledge Graphs and Pre-trained Language Models enhanced Representation Learning for Conversational Recommender Systems
von: Qiu, Zhangchi, et al.
Veröffentlicht: (2023)
von: Qiu, Zhangchi, et al.
Veröffentlicht: (2023)
Pre-training Limited Memory Language Models with Internal and External Knowledge
von: Zhao, Linxi, et al.
Veröffentlicht: (2025)
von: Zhao, Linxi, et al.
Veröffentlicht: (2025)
Velocitune: A Velocity-based Dynamic Domain Reweighting Method for Continual Pre-training
von: Luo, Zheheng, et al.
Veröffentlicht: (2024)
von: Luo, Zheheng, et al.
Veröffentlicht: (2024)
Industry Risk Assessment via Hierarchical Financial Data Using Stock Market Sentiment Indicators
von: Zhu, Hongyin
Veröffentlicht: (2023)
von: Zhu, Hongyin
Veröffentlicht: (2023)
Ähnliche Einträge
-
Pre-training Distillation for Large Language Models: A Design Space Exploration
von: Peng, Hao, et al.
Veröffentlicht: (2024) -
AtomR: Atomic Operator-Empowered Large Language Models for Heterogeneous Knowledge Reasoning
von: Xin, Amy, et al.
Veröffentlicht: (2024) -
R-Eval: A Unified Toolkit for Evaluating Domain Knowledge of Retrieval Augmented Large Language Models
von: Tu, Shangqing, et al.
Veröffentlicht: (2024) -
Adaptive Few-shot Prompting for Machine Translation with Pre-trained Language Models
von: Tang, Lei, et al.
Veröffentlicht: (2025) -
ADELIE: Aligning Large Language Models on Information Extraction
von: Qi, Yunjia, et al.
Veröffentlicht: (2024)