AlpaGasus: Training A Better Alpaca with Fewer Data
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Lichang, Li, Shiyang, Yan, Jun, Wang, Hai, Gunaratna, Kalpa, Yadav, Vikas, Tang, Zheng, Srinivasan, Vijay, Zhou, Tianyi, Huang, Heng, Jin, Hongxia |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Backdooring Instruction-Tuned Large Language Models with Virtual Prompt Injection
di: Yan, Jun, et al.
Pubblicazione: (2023)
di: Yan, Jun, et al.
Pubblicazione: (2023)
Instruction-following Evaluation through Verbalizer Manipulation
di: Li, Shiyang, et al.
Pubblicazione: (2023)
di: Li, Shiyang, et al.
Pubblicazione: (2023)
PathFinder: MCTS and LLM Feedback-based Path Selection for Multi-Hop Question Answering
di: Maram, Durga Prasad, et al.
Pubblicazione: (2025)
di: Maram, Durga Prasad, et al.
Pubblicazione: (2025)
Paraphrase and Aggregate with Large Language Models for Minimizing Intent Classification Errors
di: Yadav, Vikas, et al.
Pubblicazione: (2024)
di: Yadav, Vikas, et al.
Pubblicazione: (2024)
Explicit Diversity Conditions for Effective Question Answer Generation with Large Language Models
di: Yadav, Vikas, et al.
Pubblicazione: (2024)
di: Yadav, Vikas, et al.
Pubblicazione: (2024)
AlpaCare:Instruction-tuned Large Language Models for Medical Application
di: Zhang, Xinlu, et al.
Pubblicazione: (2023)
di: Zhang, Xinlu, et al.
Pubblicazione: (2023)
Ranking Free RAG: Replacing Re-ranking with Selection in RAG for Sensitive Domains
di: Saxena, Yash, et al.
Pubblicazione: (2025)
di: Saxena, Yash, et al.
Pubblicazione: (2025)
IMRNNs: An Efficient Method for Interpretable Dense Retrieval via Embedding Modulation
di: Saxena, Yash, et al.
Pubblicazione: (2026)
di: Saxena, Yash, et al.
Pubblicazione: (2026)
Few and Fewer: Learning Better from Few Examples Using Fewer Base Classes
di: Lafargue, Raphael, et al.
Pubblicazione: (2024)
di: Lafargue, Raphael, et al.
Pubblicazione: (2024)
Towards Efficient CoT Distillation: Self-Guided Rationale Selector for Better Performance with Fewer Rationales
di: Yan, Jianzhi, et al.
Pubblicazione: (2025)
di: Yan, Jianzhi, et al.
Pubblicazione: (2025)
Depth-supervised NeRF: Fewer Views and Faster Training for Free
di: Deng, Kangle, et al.
Pubblicazione: (2021)
di: Deng, Kangle, et al.
Pubblicazione: (2021)
$p1$: Better Prompt Optimization with Fewer Prompts
di: Gao, Zhaolin, et al.
Pubblicazione: (2026)
di: Gao, Zhaolin, et al.
Pubblicazione: (2026)
Partial Channel Network: Compute Fewer, Perform Better
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
ReGATE: Learning Faster and Better with Fewer Tokens in MLLMs
di: Li, Chaoyu, et al.
Pubblicazione: (2025)
di: Li, Chaoyu, et al.
Pubblicazione: (2025)
Fewer, but Better: On the Benefits of Surfactant‐Free Colloidal Syntheses of Nanomaterials
di: Márton Varga, et al.
Pubblicazione: (2025)
di: Márton Varga, et al.
Pubblicazione: (2025)
Dynamic Noise Preference Optimization: Self-Improvement of Large Language Models with Self-Synthetic Data
di: Yang, Haoyan, et al.
Pubblicazione: (2025)
di: Yang, Haoyan, et al.
Pubblicazione: (2025)
GLOBAL SUPPORT FOR AL QAEDA AND OSAMA BIN LADEN: AN INCREASE OR DECREASE?
di: Rohan Gunaratna
Pubblicazione: (2011)
di: Rohan Gunaratna
Pubblicazione: (2011)
UNDERSTANDING THE CHALLENGE OF IDEOLOGICAL EXTREMISM
di: Rohan Gunaratna
Pubblicazione: (2008)
di: Rohan Gunaratna
Pubblicazione: (2008)
GLOBAL THREAT FORECAST
di: Rohan Gunaratna
Pubblicazione: (2017)
di: Rohan Gunaratna
Pubblicazione: (2017)
THE BATTLEFIELD OF THE MIND: REHABILITATING MUSLIM TERRORISTS
di: Rohan Gunaratna
Pubblicazione: (2009)
di: Rohan Gunaratna
Pubblicazione: (2009)
Terrorist threat in 2014
di: Rohan Gunaratna
Pubblicazione: (2014)
di: Rohan Gunaratna
Pubblicazione: (2014)
MUMBAI INVESTIGATION: THE OPERATIVES, MASTERMINDS AND ENDURING THREAT
di: Rohan Gunaratna
Pubblicazione: (2009)
di: Rohan Gunaratna
Pubblicazione: (2009)
Communities Defeat Terrorism: Post-9/11 Community Engagement Strategies
di: Rohan Gunaratna
Pubblicazione: (2011)
di: Rohan Gunaratna
Pubblicazione: (2011)
The Year of Living Dangerously: Threat Assessment 2006
di: Rohan Gunaratna
Pubblicazione: (2006)
di: Rohan Gunaratna
Pubblicazione: (2006)
GLOBAL THREAT ASSESSMENT 2009
di: Rohan Gunaratna
Pubblicazione: (2009)
di: Rohan Gunaratna
Pubblicazione: (2009)
A NEW THREAT LANDSCAPE IN 2015
di: Rohan Gunaratna
Pubblicazione: (2015)
di: Rohan Gunaratna
Pubblicazione: (2015)
GLOBAL TERRORISM IN 2016
di: Rohan Gunaratna
Pubblicazione: (2016)
di: Rohan Gunaratna
Pubblicazione: (2016)
Caracterización fisonómico-estructural de vegetación serrana (Alpa Corral-Córdoba-Argentina)
di: S. Suárez
Pubblicazione: (1997)
di: S. Suárez
Pubblicazione: (1997)
Monolingual or Multilingual Instruction Tuning: Which Makes a Better Alpaca
di: Chen, Pinzhen, et al.
Pubblicazione: (2023)
di: Chen, Pinzhen, et al.
Pubblicazione: (2023)
Fewer May Be Better: Enhancing Offline Reinforcement Learning with Reduced Dataset
di: Yang, Yiqin, et al.
Pubblicazione: (2025)
di: Yang, Yiqin, et al.
Pubblicazione: (2025)
XNet v2: Fewer Limitations, Better Results and Greater Universality
di: Zhou, Yanfeng, et al.
Pubblicazione: (2024)
di: Zhou, Yanfeng, et al.
Pubblicazione: (2024)
From Lists to Emojis: How Format Bias Affects Model Alignment
di: Zhang, Xuanchang, et al.
Pubblicazione: (2024)
di: Zhang, Xuanchang, et al.
Pubblicazione: (2024)
AlpaPICO: Extraction of PICO Frames from Clinical Trial Documents Using LLMs
di: Ghosh, Madhusudan, et al.
Pubblicazione: (2024)
di: Ghosh, Madhusudan, et al.
Pubblicazione: (2024)
FQGA-single: Towards Fewer Training Epochs and Fewer Model Parameters for Image-to-Image Translation Tasks
di: Yang, Cho
Pubblicazione: (2024)
di: Yang, Cho
Pubblicazione: (2024)
Can LLMs Speak For Diverse People? Tuning LLMs via Debate to Generate Controllable Controversial Statements
di: Li, Ming, et al.
Pubblicazione: (2024)
di: Li, Ming, et al.
Pubblicazione: (2024)
$p$-groups with small number of character degrees and their normal subgroups
di: Talukdar, Nabajit, et al.
Pubblicazione: (2024)
di: Talukdar, Nabajit, et al.
Pubblicazione: (2024)
Some results on GVZ-groups with two character degrees
di: Talukdar, Nabajit, et al.
Pubblicazione: (2024)
di: Talukdar, Nabajit, et al.
Pubblicazione: (2024)
GVZ-groups with two character degrees
di: Talukdar, Nabajit, et al.
Pubblicazione: (2024)
di: Talukdar, Nabajit, et al.
Pubblicazione: (2024)
Kolmogorov Arnold Networks (KANs) for Imbalanced Data -- An Empirical Perspective
di: Yadav, Pankaj, et al.
Pubblicazione: (2025)
di: Yadav, Pankaj, et al.
Pubblicazione: (2025)
MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data
di: Han, Tianyu, et al.
Pubblicazione: (2023)
di: Han, Tianyu, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Backdooring Instruction-Tuned Large Language Models with Virtual Prompt Injection
di: Yan, Jun, et al.
Pubblicazione: (2023) -
Instruction-following Evaluation through Verbalizer Manipulation
di: Li, Shiyang, et al.
Pubblicazione: (2023) -
PathFinder: MCTS and LLM Feedback-based Path Selection for Multi-Hop Question Answering
di: Maram, Durga Prasad, et al.
Pubblicazione: (2025) -
Paraphrase and Aggregate with Large Language Models for Minimizing Intent Classification Errors
di: Yadav, Vikas, et al.
Pubblicazione: (2024) -
Explicit Diversity Conditions for Effective Question Answer Generation with Large Language Models
di: Yadav, Vikas, et al.
Pubblicazione: (2024)