CRAFT: Extracting and Tuning Cultural Instructions from the Wild
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Bin, Lin, Geyu, Liu, Zhengyuan, Wei, Chengwei, Chen, Nancy F. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Resilience of Large Language Models for Noisy Instructions
by: Wang, Bin, et al.
Published: (2024)
by: Wang, Bin, et al.
Published: (2024)
CrossIn: An Efficient Instruction Tuning Approach for Cross-Lingual Knowledge Alignment
by: Lin, Geyu, et al.
Published: (2024)
by: Lin, Geyu, et al.
Published: (2024)
Personality-aware Student Simulation for Conversational Intelligent Tutoring Systems
by: Liu, Zhengyuan, et al.
Published: (2024)
by: Liu, Zhengyuan, et al.
Published: (2024)
CoinMath: Harnessing the Power of Coding Instruction for Math LLMs
by: Wei, Chengwei, et al.
Published: (2024)
by: Wei, Chengwei, et al.
Published: (2024)
Instructive Dialogue Summarization with Query Aggregations
by: Wang, Bin, et al.
Published: (2023)
by: Wang, Bin, et al.
Published: (2023)
Learning Planning-based Reasoning by Trajectories Collection and Process Reward Synthesizing
by: Jiao, Fangkai, et al.
Published: (2024)
by: Jiao, Fangkai, et al.
Published: (2024)
Scaffolding Language Learning via Multi-modal Tutoring Systems with Pedagogical Instructions
by: Liu, Zhengyuan, et al.
Published: (2024)
by: Liu, Zhengyuan, et al.
Published: (2024)
AudioBench: A Universal Benchmark for Audio Large Language Models
by: Wang, Bin, et al.
Published: (2024)
by: Wang, Bin, et al.
Published: (2024)
Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems
by: Wei, Chengwei, et al.
Published: (2025)
by: Wei, Chengwei, et al.
Published: (2025)
SeaEval for Multilingual Foundation Models: From Cross-Lingual Alignment to Cultural Reasoning
by: Wang, Bin, et al.
Published: (2023)
by: Wang, Bin, et al.
Published: (2023)
IFEval-Audio: Benchmarking Instruction-Following Capability in Audio-based Large Language Models
by: Gao, Yiming, et al.
Published: (2025)
by: Gao, Yiming, et al.
Published: (2025)
Instruction Following without Instruction Tuning
by: Hewitt, John, et al.
Published: (2024)
by: Hewitt, John, et al.
Published: (2024)
WildLong: Synthesizing Realistic Long-Context Instruction Data at Scale
by: Li, Jiaxi, et al.
Published: (2025)
by: Li, Jiaxi, et al.
Published: (2025)
Not All Documents Are What You Need for Extracting Instruction Tuning Data
by: Zhang, Chi, et al.
Published: (2025)
by: Zhang, Chi, et al.
Published: (2025)
IterSelectTune: An Iterative Training Framework for Efficient Instruction-Tuning Data Selection
by: Song, Jielin, et al.
Published: (2024)
by: Song, Jielin, et al.
Published: (2024)
A Unified Causal View of Instruction Tuning
by: Chen, Lu, et al.
Published: (2024)
by: Chen, Lu, et al.
Published: (2024)
WildIFEval: Instruction Following in the Wild
by: Lior, Gili, et al.
Published: (2025)
by: Lior, Gili, et al.
Published: (2025)
Instruction Tuning With Loss Over Instructions
by: Shi, Zhengyan, et al.
Published: (2024)
by: Shi, Zhengyan, et al.
Published: (2024)
SingaKids: A Multilingual Multimodal Dialogic Tutor for Language Learning
by: Liu, Zhengyuan, et al.
Published: (2025)
by: Liu, Zhengyuan, et al.
Published: (2025)
CRAFT: Customizing LLMs by Creating and Retrieving from Specialized Toolsets
by: Yuan, Lifan, et al.
Published: (2023)
by: Yuan, Lifan, et al.
Published: (2023)
Reinforcing Compositional Retrieval: Retrieving Step-by-Step for Composing Informative Contexts
by: Long, Quanyu, et al.
Published: (2025)
by: Long, Quanyu, et al.
Published: (2025)
CRAFT: Cultural Russian-Oriented Dataset Adaptation for Focused Text-to-Image Generation
by: Vasilev, Viacheslav, et al.
Published: (2025)
by: Vasilev, Viacheslav, et al.
Published: (2025)
GIFT: Guided Fine-Tuning and Transfer for Enhancing Instruction-Tuned Language Models
by: Ruan, Zhiwen, et al.
Published: (2026)
by: Ruan, Zhiwen, et al.
Published: (2026)
MoWE-Audio: Multitask AudioLLMs with Mixture of Weak Encoders
by: Zhang, Wenyu, et al.
Published: (2024)
by: Zhang, Wenyu, et al.
Published: (2024)
Minimal Clips, Maximum Salience: Long Video Summarization via Key Moment Extraction
by: Pennec, Galann, et al.
Published: (2025)
by: Pennec, Galann, et al.
Published: (2025)
Integrating Video and Text: A Balanced Approach to Multimodal Summary Generation and Evaluation
by: Pennec, Galann, et al.
Published: (2025)
by: Pennec, Galann, et al.
Published: (2025)
InfoDensity: Rewarding Information-Dense Traces for Efficient Reasoning
by: Wei, Chengwei, et al.
Published: (2026)
by: Wei, Chengwei, et al.
Published: (2026)
Exploring Self-supervised Logic-enhanced Training for Large Language Models
by: Jiao, Fangkai, et al.
Published: (2023)
by: Jiao, Fangkai, et al.
Published: (2023)
Multi-Agent Collaboration for Multilingual Code Instruction Tuning
by: Yang, Jian, et al.
Published: (2025)
by: Yang, Jian, et al.
Published: (2025)
Beyond IID: Optimizing Instruction Learning from the Perspective of Instruction Interaction and Dependency
by: Zhao, Hanyu, et al.
Published: (2024)
by: Zhao, Hanyu, et al.
Published: (2024)
Persuasion Dynamics in LLMs: Investigating Robustness and Adaptability in Knowledge and Safety with DuET-PD
by: Tan, Bryan Chen Zhengyu, et al.
Published: (2025)
by: Tan, Bryan Chen Zhengyu, et al.
Published: (2025)
UltraIF: Advancing Instruction Following from the Wild
by: An, Kaikai, et al.
Published: (2025)
by: An, Kaikai, et al.
Published: (2025)
Instruction Tuning on Public Government and Cultural Data for Low-Resource Language: a Case Study in Kazakh
by: Laiyk, Nurkhan, et al.
Published: (2025)
by: Laiyk, Nurkhan, et al.
Published: (2025)
MIDB: Multilingual Instruction Data Booster for Enhancing Cultural Equality in Multilingual Instruction Synthesis
by: Liu, Yilun, et al.
Published: (2025)
by: Liu, Yilun, et al.
Published: (2025)
CRAFT: Training-Free Cascaded Retrieval for Tabular QA
by: Singh, Adarsh, et al.
Published: (2025)
by: Singh, Adarsh, et al.
Published: (2025)
Instruction-Tuning Data Synthesis from Scratch via Web Reconstruction
by: Jiang, Yuxin, et al.
Published: (2025)
by: Jiang, Yuxin, et al.
Published: (2025)
JsonTuning: Towards Generalizable, Robust, and Controllable Instruction Tuning
by: Gao, Chang, et al.
Published: (2023)
by: Gao, Chang, et al.
Published: (2023)
DnA-Eval: Enhancing Large Language Model Evaluation through Decomposition and Aggregation
by: Li, Minzhi, et al.
Published: (2024)
by: Li, Minzhi, et al.
Published: (2024)
COGENT: A Curriculum-oriented Framework for Generating Grade-appropriate Educational Content
by: Liu, Zhengyuan, et al.
Published: (2025)
by: Liu, Zhengyuan, et al.
Published: (2025)
Synthetic Data (Almost) from Scratch: Generalized Instruction Tuning for Language Models
by: Li, Haoran, et al.
Published: (2024)
by: Li, Haoran, et al.
Published: (2024)
Similar Items
-
Resilience of Large Language Models for Noisy Instructions
by: Wang, Bin, et al.
Published: (2024) -
CrossIn: An Efficient Instruction Tuning Approach for Cross-Lingual Knowledge Alignment
by: Lin, Geyu, et al.
Published: (2024) -
Personality-aware Student Simulation for Conversational Intelligent Tutoring Systems
by: Liu, Zhengyuan, et al.
Published: (2024) -
CoinMath: Harnessing the Power of Coding Instruction for Math LLMs
by: Wei, Chengwei, et al.
Published: (2024) -
Instructive Dialogue Summarization with Query Aggregations
by: Wang, Bin, et al.
Published: (2023)