RAFT: Adapting Language Model to Domain Specific RAG
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Tianjun, Patil, Shishir G., Jain, Naman, Shen, Sheng, Zaharia, Matei, Stoica, Ion, Gonzalez, Joseph E. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Specifications: The missing link to making the development of LLM systems an engineering discipline
von: Stoica, Ion, et al.
Veröffentlicht: (2024)
von: Stoica, Ion, et al.
Veröffentlicht: (2024)
BARE: Leveraging Base Language Models for Few-Shot Synthetic Data Generation
von: Zhu, Alan, et al.
Veröffentlicht: (2025)
von: Zhu, Alan, et al.
Veröffentlicht: (2025)
DeepScholar-Bench: A Live Benchmark and Automated Evaluation for Generative Research Synthesis
von: Patel, Liana, et al.
Veröffentlicht: (2025)
von: Patel, Liana, et al.
Veröffentlicht: (2025)
GoEX: Perspectives and Designs Towards a Runtime for Autonomous LLM Applications
von: Patil, Shishir G., et al.
Veröffentlicht: (2024)
von: Patil, Shishir G., et al.
Veröffentlicht: (2024)
The Price Reversal Phenomenon: When Cheaper Reasoning Models Cost More
von: Chen, Lingjiao, et al.
Veröffentlicht: (2026)
von: Chen, Lingjiao, et al.
Veröffentlicht: (2026)
Optimizing Model Selection for Compound AI Systems
von: Chen, Lingjiao, et al.
Veröffentlicht: (2025)
von: Chen, Lingjiao, et al.
Veröffentlicht: (2025)
RAG over Thinking Traces Can Improve Reasoning Tasks
von: Arabzadeh, Negar, et al.
Veröffentlicht: (2026)
von: Arabzadeh, Negar, et al.
Veröffentlicht: (2026)
HashAttention: Semantic Sparsity for Faster Inference
von: Desai, Aditya, et al.
Veröffentlicht: (2024)
von: Desai, Aditya, et al.
Veröffentlicht: (2024)
Are More LLM Calls All You Need? Towards Scaling Laws of Compound Inference Systems
von: Chen, Lingjiao, et al.
Veröffentlicht: (2024)
von: Chen, Lingjiao, et al.
Veröffentlicht: (2024)
LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!
von: Li, Dacheng, et al.
Veröffentlicht: (2025)
von: Li, Dacheng, et al.
Veröffentlicht: (2025)
GSO: Challenging Software Optimization Tasks for Evaluating SWE-Agents
von: Shetty, Manish, et al.
Veröffentlicht: (2025)
von: Shetty, Manish, et al.
Veröffentlicht: (2025)
Post-Training Sparse Attention with Double Sparsity
von: Yang, Shuo, et al.
Veröffentlicht: (2024)
von: Yang, Shuo, et al.
Veröffentlicht: (2024)
How to Train Your Advisor: Steering Black-Box LLMs with Advisor Models
von: Asawa, Parth, et al.
Veröffentlicht: (2025)
von: Asawa, Parth, et al.
Veröffentlicht: (2025)
Recursive Introspection: Teaching Language Model Agents How to Self-Improve
von: Qu, Yuxiao, et al.
Veröffentlicht: (2024)
von: Qu, Yuxiao, et al.
Veröffentlicht: (2024)
Networks of Networks: Complexity Class Principles Applied to Compound AI Systems Design
von: Davis, Jared Quincy, et al.
Veröffentlicht: (2024)
von: Davis, Jared Quincy, et al.
Veröffentlicht: (2024)
DSPy Assertions: Computational Constraints for Self-Refining Language Model Pipelines
von: Singhvi, Arnav, et al.
Veröffentlicht: (2023)
von: Singhvi, Arnav, et al.
Veröffentlicht: (2023)
MemGPT: Towards LLMs as Operating Systems
von: Packer, Charles, et al.
Veröffentlicht: (2023)
von: Packer, Charles, et al.
Veröffentlicht: (2023)
OR-Bench: An Over-Refusal Benchmark for Large Language Models
von: Cui, Justin, et al.
Veröffentlicht: (2024)
von: Cui, Justin, et al.
Veröffentlicht: (2024)
Chain-of-Rank: Enhancing Large Language Models for Domain-Specific RAG in Edge Device
von: Lee, Juntae, et al.
Veröffentlicht: (2025)
von: Lee, Juntae, et al.
Veröffentlicht: (2025)
LLoCO: Learning Long Contexts Offline
von: Tan, Sijun, et al.
Veröffentlicht: (2024)
von: Tan, Sijun, et al.
Veröffentlicht: (2024)
Tabular Embedding Model (TEM): Finetuning Embedding Models For Tabular RAG Applications
von: Khanna, Sujit, et al.
Veröffentlicht: (2024)
von: Khanna, Sujit, et al.
Veröffentlicht: (2024)
SGLang: Efficient Execution of Structured Language Model Programs
von: Zheng, Lianmin, et al.
Veröffentlicht: (2023)
von: Zheng, Lianmin, et al.
Veröffentlicht: (2023)
Reasoning Models Can Be Effective Without Thinking
von: Ma, Wenjie, et al.
Veröffentlicht: (2025)
von: Ma, Wenjie, et al.
Veröffentlicht: (2025)
DS SERVE: A Framework for Efficient and Scalable Neural Retrieval
von: Liu, Jinjian, et al.
Veröffentlicht: (2025)
von: Liu, Jinjian, et al.
Veröffentlicht: (2025)
MoE-Lightning: High-Throughput MoE Inference on Memory-constrained GPUs
von: Cao, Shiyi, et al.
Veröffentlicht: (2024)
von: Cao, Shiyi, et al.
Veröffentlicht: (2024)
Domain-Specific Data Generation Framework for RAG Adaptation
von: Tian, Chris Xing, et al.
Veröffentlicht: (2025)
von: Tian, Chris Xing, et al.
Veröffentlicht: (2025)
DO-RAG: A Domain-Specific QA Framework Using Knowledge Graph-Enhanced Retrieval-Augmented Generation
von: Opoku, David Osei, et al.
Veröffentlicht: (2025)
von: Opoku, David Osei, et al.
Veröffentlicht: (2025)
vAttention: Verified Sparse Attention
von: Desai, Aditya, et al.
Veröffentlicht: (2025)
von: Desai, Aditya, et al.
Veröffentlicht: (2025)
Sleep-time Compute: Beyond Inference Scaling at Test-time
von: Lin, Kevin, et al.
Veröffentlicht: (2025)
von: Lin, Kevin, et al.
Veröffentlicht: (2025)
ARES: An Automated Evaluation Framework for Retrieval-Augmented Generation Systems
von: Saad-Falcon, Jon, et al.
Veröffentlicht: (2023)
von: Saad-Falcon, Jon, et al.
Veröffentlicht: (2023)
Optimizing Instructions and Demonstrations for Multi-Stage Language Model Programs
von: Opsahl-Ong, Krista, et al.
Veröffentlicht: (2024)
von: Opsahl-Ong, Krista, et al.
Veröffentlicht: (2024)
AdaEvolve: Adaptive LLM Driven Zeroth-Order Optimization
von: Cemri, Mert, et al.
Veröffentlicht: (2026)
von: Cemri, Mert, et al.
Veröffentlicht: (2026)
Methodology of Adapting Large English Language Models for Specific Cultural Contexts
von: Zhang, Wenjing, et al.
Veröffentlicht: (2024)
von: Zhang, Wenjing, et al.
Veröffentlicht: (2024)
RAGalyst: Automated Human-Aligned Agentic Evaluation for Domain-Specific RAG
von: Gao, Joshua, et al.
Veröffentlicht: (2025)
von: Gao, Joshua, et al.
Veröffentlicht: (2025)
AI-Driven Research for Databases
von: Cheng, Audrey, et al.
Veröffentlicht: (2026)
von: Cheng, Audrey, et al.
Veröffentlicht: (2026)
Adaptive-RAG: Learning to Adapt Retrieval-Augmented Large Language Models through Question Complexity
von: Jeong, Soyeong, et al.
Veröffentlicht: (2024)
von: Jeong, Soyeong, et al.
Veröffentlicht: (2024)
$L^*LM$: Learning Automata from Examples using Natural Language Oracles
von: Vazquez-Chanlatte, Marcell, et al.
Veröffentlicht: (2024)
von: Vazquez-Chanlatte, Marcell, et al.
Veröffentlicht: (2024)
SimRAG: Self-Improving Retrieval-Augmented Generation for Adapting Large Language Models to Specialized Domains
von: Xu, Ran, et al.
Veröffentlicht: (2024)
von: Xu, Ran, et al.
Veröffentlicht: (2024)
Semantic Operators: A Declarative Model for Rich, AI-based Data Processing
von: Patel, Liana, et al.
Veröffentlicht: (2024)
von: Patel, Liana, et al.
Veröffentlicht: (2024)
Cross-Domain Content Generation with Domain-Specific Small Language Models
von: Maloo, Ankit, et al.
Veröffentlicht: (2024)
von: Maloo, Ankit, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Specifications: The missing link to making the development of LLM systems an engineering discipline
von: Stoica, Ion, et al.
Veröffentlicht: (2024) -
BARE: Leveraging Base Language Models for Few-Shot Synthetic Data Generation
von: Zhu, Alan, et al.
Veröffentlicht: (2025) -
DeepScholar-Bench: A Live Benchmark and Automated Evaluation for Generative Research Synthesis
von: Patel, Liana, et al.
Veröffentlicht: (2025) -
GoEX: Perspectives and Designs Towards a Runtime for Autonomous LLM Applications
von: Patil, Shishir G., et al.
Veröffentlicht: (2024) -
The Price Reversal Phenomenon: When Cheaper Reasoning Models Cost More
von: Chen, Lingjiao, et al.
Veröffentlicht: (2026)