SciInstruct: a Self-Reflective Instruction Annotated Dataset for Training Scientific Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Dan, Hu, Ziniu, Zhoubian, Sining, Du, Zhengxiao, Yang, Kaiyu, Wang, Zihan, Yue, Yisong, Dong, Yuxiao, Tang, Jie |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search
par: Zhang, Dan, et autres
Publié: (2024)
par: Zhang, Dan, et autres
Publié: (2024)
ReST-RL: Achieving Accurate Code Reasoning of LLMs with Optimized Self-Training and Decoding
par: Zhoubian, Sining, et autres
Publié: (2025)
par: Zhoubian, Sining, et autres
Publié: (2025)
DataSciBench: An LLM Agent Benchmark for Data Science
par: Zhang, Dan, et autres
Publié: (2025)
par: Zhang, Dan, et autres
Publié: (2025)
Rock Classification Based on Residual Networks
par: Zhoubian, Sining, et autres
Publié: (2024)
par: Zhoubian, Sining, et autres
Publié: (2024)
Understanding Emergent Abilities of Language Models from the Loss Perspective
par: Du, Zhengxiao, et autres
Publié: (2024)
par: Du, Zhengxiao, et autres
Publié: (2024)
TDRM: Smooth Reward Models with Temporal Difference for LLM RL and Inference
par: Zhang, Dan, et autres
Publié: (2025)
par: Zhang, Dan, et autres
Publié: (2025)
Self-Control of LLM Behaviors by Compressing Suffix Gradient into Prefix Controller
par: Cai, Min, et autres
Publié: (2024)
par: Cai, Min, et autres
Publié: (2024)
TreeRL: LLM Reinforcement Learning with On-Policy Tree Search
par: Hou, Zhenyu, et autres
Publié: (2025)
par: Hou, Zhenyu, et autres
Publié: (2025)
Self-Evolving Visual Concept Library using Vision-Language Critics
par: Sehgal, Atharva, et autres
Publié: (2025)
par: Sehgal, Atharva, et autres
Publié: (2025)
Scaling Speech-Text Pre-training with Synthetic Interleaved Data
par: Zeng, Aohan, et autres
Publié: (2024)
par: Zeng, Aohan, et autres
Publié: (2024)
SciIF: Benchmarking Scientific Instruction Following Towards Rigorous Scientific Intelligence
par: Su, Encheng, et autres
Publié: (2026)
par: Su, Encheng, et autres
Publié: (2026)
ChatGLM-Math: Improving Math Problem-Solving in Large Language Models with a Self-Critique Pipeline
par: Xu, Yifan, et autres
Publié: (2024)
par: Xu, Yifan, et autres
Publié: (2024)
InstructEngine: Instruction-driven Text-to-Image Alignment
par: Lu, Xingyu, et autres
Publié: (2025)
par: Lu, Xingyu, et autres
Publié: (2025)
Instruct-of-Reflection: Enhancing Large Language Models Iterative Reflection Capabilities via Dynamic-Meta Instruction
par: Liu, Liping, et autres
Publié: (2025)
par: Liu, Liping, et autres
Publié: (2025)
GLM-4-Voice: Towards Intelligent and Human-Like End-to-End Spoken Chatbot
par: Zeng, Aohan, et autres
Publié: (2024)
par: Zeng, Aohan, et autres
Publié: (2024)
Strategist: Self-improvement of LLM Decision Making via Bi-Level Tree Search
par: Light, Jonathan, et autres
Publié: (2024)
par: Light, Jonathan, et autres
Publié: (2024)
VisScience: An Extensive Benchmark for Evaluating K12 Educational Multi-modal Scientific Reasoning
par: Jiang, Zhihuan, et autres
Publié: (2024)
par: Jiang, Zhihuan, et autres
Publié: (2024)
ComplexFuncBench: Exploring Multi-Step and Constrained Function Calling under Long-Context Scenario
par: Zhong, Lucen, et autres
Publié: (2025)
par: Zhong, Lucen, et autres
Publié: (2025)
InverseCoder: Self-improving Instruction-Tuned Code LLMs with Inverse-Instruct
par: Wu, Yutong, et autres
Publié: (2024)
par: Wu, Yutong, et autres
Publié: (2024)
Repurposing Annotation Guidelines to Instruct LLM Annotators: A Case Study
par: Kim, Kon Woo, et autres
Publié: (2025)
par: Kim, Kon Woo, et autres
Publié: (2025)
AutoRE: Document-Level Relation Extraction with Large Language Models
par: Xue, Lilong, et autres
Publié: (2024)
par: Xue, Lilong, et autres
Publié: (2024)
Does RLHF Scale? Exploring the Impacts From Data, Model, and Method
par: Hou, Zhenyu, et autres
Publié: (2024)
par: Hou, Zhenyu, et autres
Publié: (2024)
SPaR: Self-Play with Tree-Search Refinement to Improve Instruction-Following in Large Language Models
par: Cheng, Jiale, et autres
Publié: (2024)
par: Cheng, Jiale, et autres
Publié: (2024)
Extensive Self-Contrast Enables Feedback-Free Language Model Alignment
par: Liu, Xiao, et autres
Publié: (2024)
par: Liu, Xiao, et autres
Publié: (2024)
InstructPart: Task-Oriented Part Segmentation with Instruction Reasoning
par: Wan, Zifu, et autres
Publié: (2025)
par: Wan, Zifu, et autres
Publié: (2025)
LuxInstruct: A Cross-Lingual Instruction Tuning Dataset For Luxembourgish
par: Philippy, Fred, et autres
Publié: (2025)
par: Philippy, Fred, et autres
Publié: (2025)
InstructIE: A Bilingual Instruction-based Information Extraction Dataset
par: Gui, Honghao, et autres
Publié: (2023)
par: Gui, Honghao, et autres
Publié: (2023)
SciBench: Evaluating College-Level Scientific Problem-Solving Abilities of Large Language Models
par: Wang, Xiaoxuan, et autres
Publié: (2023)
par: Wang, Xiaoxuan, et autres
Publié: (2023)
Dataset Distillation for Offline Reinforcement Learning
par: Light, Jonathan, et autres
Publié: (2024)
par: Light, Jonathan, et autres
Publié: (2024)
Instructing LLMs to Negotiate using Reinforcement Learning with Verifiable Rewards
par: Liu, Shuze Daniel, et autres
Publié: (2026)
par: Liu, Shuze Daniel, et autres
Publié: (2026)
SWE-Dev: Building Software Engineering Agents with Training and Inference Scaling
par: Wang, Haoran, et autres
Publié: (2025)
par: Wang, Haoran, et autres
Publié: (2025)
SciEvent: Benchmarking Multi-domain Scientific Event Extraction
par: Dong, Bofu, et autres
Publié: (2025)
par: Dong, Bofu, et autres
Publié: (2025)
InstructAttribute: Fine-grained Object Attributes editing with Instruction
par: Yin, Xingxi, et autres
Publié: (2025)
par: Yin, Xingxi, et autres
Publié: (2025)
InstructSAM: Segment Any Instance with Any Instructions
par: Yuan, Yuqian, et autres
Publié: (2026)
par: Yuan, Yuqian, et autres
Publié: (2026)
SearchInstruct: Enhancing Domain Adaptation via Retrieval-Based Instruction Dataset Creation
par: Barati, Iman, et autres
Publié: (2025)
par: Barati, Iman, et autres
Publié: (2025)
InstructLR: A Scalable Approach to Create Instruction Dataset for Under-Resourced Languages
par: Keita, Mamadou K., et autres
Publié: (2025)
par: Keita, Mamadou K., et autres
Publié: (2025)
OpenCodeInstruct: A Large-scale Instruction Tuning Dataset for Code LLMs
par: Ahmad, Wasi Uddin, et autres
Publié: (2025)
par: Ahmad, Wasi Uddin, et autres
Publié: (2025)
InstructDoc: A Dataset for Zero-Shot Generalization of Visual Document Understanding with Instructions
par: Tanaka, Ryota, et autres
Publié: (2024)
par: Tanaka, Ryota, et autres
Publié: (2024)
Learning to Instruct for Visual Instruction Tuning
par: Zhou, Zhihan, et autres
Publié: (2025)
par: Zhou, Zhihan, et autres
Publié: (2025)
SciER: An Entity and Relation Extraction Dataset for Datasets, Methods, and Tasks in Scientific Documents
par: Zhang, Qi, et autres
Publié: (2024)
par: Zhang, Qi, et autres
Publié: (2024)
Documents similaires
-
ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search
par: Zhang, Dan, et autres
Publié: (2024) -
ReST-RL: Achieving Accurate Code Reasoning of LLMs with Optimized Self-Training and Decoding
par: Zhoubian, Sining, et autres
Publié: (2025) -
DataSciBench: An LLM Agent Benchmark for Data Science
par: Zhang, Dan, et autres
Publié: (2025) -
Rock Classification Based on Residual Networks
par: Zhoubian, Sining, et autres
Publié: (2024) -
Understanding Emergent Abilities of Language Models from the Loss Perspective
par: Du, Zhengxiao, et autres
Publié: (2024)