MAmmoTH2: Scaling Instructions from the Web
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yue, Xiang, Zheng, Tuney, Zhang, Ge, Chen, Wenhu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MAmmoTH-VL: Eliciting Multimodal Reasoning with Instruction Tuning at Scale
von: Guo, Jarvis, et al.
Veröffentlicht: (2024)
von: Guo, Jarvis, et al.
Veröffentlicht: (2024)
VisualWebInstruct: Scaling up Multimodal Instruction Data through Web Search
von: Jia, Yiming, et al.
Veröffentlicht: (2025)
von: Jia, Yiming, et al.
Veröffentlicht: (2025)
LongIns: A Challenging Long-context Instruction-based Exam for LLMs
von: Gavin, Shawn, et al.
Veröffentlicht: (2024)
von: Gavin, Shawn, et al.
Veröffentlicht: (2024)
Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
Long-context LLMs Struggle with Long In-context Learning
von: Li, Tianle, et al.
Veröffentlicht: (2024)
von: Li, Tianle, et al.
Veröffentlicht: (2024)
StructLM: Towards Building Generalist Models for Structured Knowledge Grounding
von: Zhuang, Alex, et al.
Veröffentlicht: (2024)
von: Zhuang, Alex, et al.
Veröffentlicht: (2024)
OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement
von: Zheng, Tianyu, et al.
Veröffentlicht: (2024)
von: Zheng, Tianyu, et al.
Veröffentlicht: (2024)
MagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing
von: Zhang, Kai, et al.
Veröffentlicht: (2023)
von: Zhang, Kai, et al.
Veröffentlicht: (2023)
Does Math Reasoning Improve General LLM Capabilities? Understanding Transferability of LLM Reasoning
von: Huan, Maggie, et al.
Veröffentlicht: (2025)
von: Huan, Maggie, et al.
Veröffentlicht: (2025)
VisCoder: Fine-Tuning LLMs for Executable Python Visualization Code Generation
von: Ni, Yuansheng, et al.
Veröffentlicht: (2025)
von: Ni, Yuansheng, et al.
Veröffentlicht: (2025)
VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation
von: Ku, Max, et al.
Veröffentlicht: (2023)
von: Ku, Max, et al.
Veröffentlicht: (2023)
General-Reasoner: Advancing LLM Reasoning Across All Domains
von: Ma, Xueguang, et al.
Veröffentlicht: (2025)
von: Ma, Xueguang, et al.
Veröffentlicht: (2025)
WebExplorer: Explore and Evolve for Training Long-Horizon Web Agents
von: Liu, Junteng, et al.
Veröffentlicht: (2025)
von: Liu, Junteng, et al.
Veröffentlicht: (2025)
BrowserAgent: Building Web Agents with Human-Inspired Web Browsing Actions
von: Yu, Tao, et al.
Veröffentlicht: (2025)
von: Yu, Tao, et al.
Veröffentlicht: (2025)
OphIn-500K: Curating Web-Scale Visual Instructions for Scaling Ophthalmic Multimodal Large Language Models
von: Dong, Xuanzhao, et al.
Veröffentlicht: (2026)
von: Dong, Xuanzhao, et al.
Veröffentlicht: (2026)
MANTIS: Interleaved Multi-Image Instruction Tuning
von: Jiang, Dongfu, et al.
Veröffentlicht: (2024)
von: Jiang, Dongfu, et al.
Veröffentlicht: (2024)
PixelWorld: How Far Are We from Perceiving Everything as Pixels?
von: Lyu, Zhiheng, et al.
Veröffentlicht: (2025)
von: Lyu, Zhiheng, et al.
Veröffentlicht: (2025)
MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark
von: Yue, Xiang, et al.
Veröffentlicht: (2024)
von: Yue, Xiang, et al.
Veröffentlicht: (2024)
EditReward: A Human-Aligned Reward Model for Instruction-Guided Image Editing
von: Wu, Keming, et al.
Veröffentlicht: (2025)
von: Wu, Keming, et al.
Veröffentlicht: (2025)
E^2-LLM: Efficient and Extreme Length Extension of Large Language Models
von: Liu, Jiaheng, et al.
Veröffentlicht: (2024)
von: Liu, Jiaheng, et al.
Veröffentlicht: (2024)
TIGERScore: Towards Building Explainable Metric for All Text Generation Tasks
von: Jiang, Dongfu, et al.
Veröffentlicht: (2023)
von: Jiang, Dongfu, et al.
Veröffentlicht: (2023)
On the Multi-turn Instruction Following for Conversational Web Agents
von: Deng, Yang, et al.
Veröffentlicht: (2024)
von: Deng, Yang, et al.
Veröffentlicht: (2024)
Instruction-Tuning Data Synthesis from Scratch via Web Reconstruction
von: Jiang, Yuxin, et al.
Veröffentlicht: (2025)
von: Jiang, Yuxin, et al.
Veröffentlicht: (2025)
SCRIBES: Web-Scale Script-Based Semi-Structured Data Extraction with Reinforcement Learning
von: Liu, Shicheng, et al.
Veröffentlicht: (2025)
von: Liu, Shicheng, et al.
Veröffentlicht: (2025)
ScholarCopilot: Training Large Language Models for Academic Writing with Accurate Citations
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
LongRAG: Enhancing Retrieval-Augmented Generation with Long-context LLMs
von: Jiang, Ziyan, et al.
Veröffentlicht: (2024)
von: Jiang, Ziyan, et al.
Veröffentlicht: (2024)
AutoKaggle: A Multi-Agent Framework for Autonomous Data Science Competitions
von: Li, Ziming, et al.
Veröffentlicht: (2024)
von: Li, Ziming, et al.
Veröffentlicht: (2024)
Augmenting Black-box LLMs with Medical Textbooks for Biomedical Question Answering
von: Wang, Yubo, et al.
Veröffentlicht: (2023)
von: Wang, Yubo, et al.
Veröffentlicht: (2023)
VisCoder2: Building Multi-Language Visualization Coding Agents
von: Ni, Yuansheng, et al.
Veröffentlicht: (2025)
von: Ni, Yuansheng, et al.
Veröffentlicht: (2025)
Reconstructive Visual Instruction Tuning
von: Wang, Haochen, et al.
Veröffentlicht: (2024)
von: Wang, Haochen, et al.
Veröffentlicht: (2024)
Critique-Coder: Enhancing Coder Models by Critique Reinforcement Learning
von: Ruan, Chi, et al.
Veröffentlicht: (2025)
von: Ruan, Chi, et al.
Veröffentlicht: (2025)
Large-Scale Data Selection for Instruction Tuning
von: Ivison, Hamish, et al.
Veröffentlicht: (2025)
von: Ivison, Hamish, et al.
Veröffentlicht: (2025)
Long Context Alignment with Short Instructions and Synthesized Positions
von: Wu, Wenhao, et al.
Veröffentlicht: (2024)
von: Wu, Wenhao, et al.
Veröffentlicht: (2024)
RefuteBench 2.0 -- Agentic Benchmark for Dynamic Evaluation of LLM Responses to Refutation Instruction
von: Yan, Jianhao, et al.
Veröffentlicht: (2025)
von: Yan, Jianhao, et al.
Veröffentlicht: (2025)
WebWeaver: Structuring Web-Scale Evidence with Dynamic Outlines for Open-Ended Deep Research
von: Li, Zijian, et al.
Veröffentlicht: (2025)
von: Li, Zijian, et al.
Veröffentlicht: (2025)
FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale
von: Patel, Ajay, et al.
Veröffentlicht: (2026)
von: Patel, Ajay, et al.
Veröffentlicht: (2026)
Instruction Anchor: Dissecting the Mechanistic Dynamics of Modality Arbitration
von: Zhang, Yu, et al.
Veröffentlicht: (2026)
von: Zhang, Yu, et al.
Veröffentlicht: (2026)
WEPO: Web Element Preference Optimization for LLM-based Web Navigation
von: Liu, Jiarun, et al.
Veröffentlicht: (2024)
von: Liu, Jiarun, et al.
Veröffentlicht: (2024)
The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale
von: Penedo, Guilherme, et al.
Veröffentlicht: (2024)
von: Penedo, Guilherme, et al.
Veröffentlicht: (2024)
WEBSERV: A Full-Stack and RL-Ready Web Environment for Training Web Agents at Scale
von: Lu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Lu, Yuxuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MAmmoTH-VL: Eliciting Multimodal Reasoning with Instruction Tuning at Scale
von: Guo, Jarvis, et al.
Veröffentlicht: (2024) -
VisualWebInstruct: Scaling up Multimodal Instruction Data through Web Search
von: Jia, Yiming, et al.
Veröffentlicht: (2025) -
LongIns: A Challenging Long-context Instruction-based Exam for LLMs
von: Gavin, Shawn, et al.
Veröffentlicht: (2024) -
Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate
von: Wang, Yubo, et al.
Veröffentlicht: (2025) -
Long-context LLMs Struggle with Long In-context Learning
von: Li, Tianle, et al.
Veröffentlicht: (2024)