LML-DAP: Language Model Learning a Dataset for Data-Augmented Prediction
Fuente:
arXiv
Saved in:
| Main Author: | Vadlapati, Praneeth |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AutoPureData: Automated Filtering of Undesirable Web Data to Update LLM Knowledge
by: Vadlapati, Praneeth
Published: (2024)
by: Vadlapati, Praneeth
Published: (2024)
Large Language Model Augmented Exercise Retrieval for Personalized Language Learning
by: Xu, Austin, et al.
Published: (2024)
by: Xu, Austin, et al.
Published: (2024)
Augmenting Lateral Thinking in Language Models with Humor and Riddle Data for the BRAINTEASER Task
by: Ghashami, Mina, et al.
Published: (2024)
by: Ghashami, Mina, et al.
Published: (2024)
NyayaAnumana & INLegalLlama: The Largest Indian Legal Judgment Prediction Dataset and Specialized Language Model for Enhanced Decision Analysis
by: Nigam, Shubham Kumar, et al.
Published: (2024)
by: Nigam, Shubham Kumar, et al.
Published: (2024)
Soft Prompt Tuning for Augmenting Dense Retrieval with Large Language Models
by: Peng, Zhiyuan, et al.
Published: (2023)
by: Peng, Zhiyuan, et al.
Published: (2023)
Knowledge-Augmented Large Language Models for Personalized Contextual Query Suggestion
by: Baek, Jinheon, et al.
Published: (2023)
by: Baek, Jinheon, et al.
Published: (2023)
Rethinking Legal Judgement Prediction in a Realistic Scenario in the Era of Large Language Models
by: Nigam, Shubham Kumar, et al.
Published: (2024)
by: Nigam, Shubham Kumar, et al.
Published: (2024)
SimRAG: Self-Improving Retrieval-Augmented Generation for Adapting Large Language Models to Specialized Domains
by: Xu, Ran, et al.
Published: (2024)
by: Xu, Ran, et al.
Published: (2024)
OAEI-LLM-T: A TBox Benchmark Dataset for Understanding Large Language Model Hallucinations in Ontology Matching
by: Qiang, Zhangcheng, et al.
Published: (2025)
by: Qiang, Zhangcheng, et al.
Published: (2025)
REGEN: A Dataset and Benchmarks with Natural Language Critiques and Narratives
by: Su, Kun, et al.
Published: (2025)
by: Su, Kun, et al.
Published: (2025)
Self-Augmented In-Context Learning for Unsupervised Word Translation
by: Li, Yaoyiran, et al.
Published: (2024)
by: Li, Yaoyiran, et al.
Published: (2024)
Towards Robust Evaluation: A Comprehensive Taxonomy of Datasets and Metrics for Open Domain Question Answering in the Era of Large Language Models
by: Srivastava, Akchay, et al.
Published: (2024)
by: Srivastava, Akchay, et al.
Published: (2024)
Explainable Lane Change Prediction for Near-Crash Scenarios Using Knowledge Graph Embeddings and Retrieval Augmented Generation
by: Manzour, M., et al.
Published: (2025)
by: Manzour, M., et al.
Published: (2025)
AgentMove: A Large Language Model based Agentic Framework for Zero-shot Next Location Prediction
by: Feng, Jie, et al.
Published: (2024)
by: Feng, Jie, et al.
Published: (2024)
Improving Crash Data Quality with Large Language Models: Evidence from Secondary Crash Narratives in Kentucky
by: Zhang, Xu, et al.
Published: (2025)
by: Zhang, Xu, et al.
Published: (2025)
GRAM: Generative Retrieval Augmented Matching of Data Schemas in the Context of Data Security
by: Liu, Xuanqing, et al.
Published: (2024)
by: Liu, Xuanqing, et al.
Published: (2024)
IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languages
by: Nigam, Shubham Kumar, et al.
Published: (2026)
by: Nigam, Shubham Kumar, et al.
Published: (2026)
Understanding Survey Paper Taxonomy about Large Language Models via Graph Representation Learning
by: Zhuang, Jun, et al.
Published: (2024)
by: Zhuang, Jun, et al.
Published: (2024)
Let Multimodal Embedders Learn When to Augment Query via Adaptive Query Augmentation
by: Kim, Wongyu, et al.
Published: (2025)
by: Kim, Wongyu, et al.
Published: (2025)
News Recommendation with Category Description by a Large Language Model
by: Yada, Yuki, et al.
Published: (2024)
by: Yada, Yuki, et al.
Published: (2024)
CoRAG: Collaborative Retrieval-Augmented Generation
by: Muhamed, Aashiq, et al.
Published: (2025)
by: Muhamed, Aashiq, et al.
Published: (2025)
Scaling Retrieval-Based Language Models with a Trillion-Token Datastore
by: Shao, Rulin, et al.
Published: (2024)
by: Shao, Rulin, et al.
Published: (2024)
Distilling a Small Utility-Based Passage Selector to Enhance Retrieval-Augmented Generation
by: Zhang, Hengran, et al.
Published: (2025)
by: Zhang, Hengran, et al.
Published: (2025)
Familiarity-Aware Evidence Compression for Retrieval-Augmented Generation
by: Jung, Dongwon, et al.
Published: (2024)
by: Jung, Dongwon, et al.
Published: (2024)
Retrieval Augmented Generation for Domain-specific Question Answering
by: Sharma, Sanat, et al.
Published: (2024)
by: Sharma, Sanat, et al.
Published: (2024)
Optimization of Retrieval-Augmented Generation Context with Outlier Detection
by: Bulgakov, Vitaly
Published: (2024)
by: Bulgakov, Vitaly
Published: (2024)
From Data to Behavior: Predicting Unintended Model Behaviors Before Training
by: Wang, Mengru, et al.
Published: (2026)
by: Wang, Mengru, et al.
Published: (2026)
The Factuality of Large Language Models in the Legal Domain
by: Hamdani, Rajaa El, et al.
Published: (2024)
by: Hamdani, Rajaa El, et al.
Published: (2024)
User Embedding Model for Personalized Language Prompting
by: Doddapaneni, Sumanth, et al.
Published: (2024)
by: Doddapaneni, Sumanth, et al.
Published: (2024)
Approaching Human-Level Forecasting with Language Models
by: Halawi, Danny, et al.
Published: (2024)
by: Halawi, Danny, et al.
Published: (2024)
Policy-Gradient Training of Language Models for Ranking
by: Gao, Ge, et al.
Published: (2023)
by: Gao, Ge, et al.
Published: (2023)
On Bilingual Lexicon Induction with Large Language Models
by: Li, Yaoyiran, et al.
Published: (2023)
by: Li, Yaoyiran, et al.
Published: (2023)
Reducing hallucination in structured outputs via Retrieval-Augmented Generation
by: Béchard, Patrice, et al.
Published: (2024)
by: Béchard, Patrice, et al.
Published: (2024)
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation
by: Dong, Guanting, et al.
Published: (2024)
by: Dong, Guanting, et al.
Published: (2024)
InstructRetro: Instruction Tuning post Retrieval-Augmented Pretraining
by: Wang, Boxin, et al.
Published: (2023)
by: Wang, Boxin, et al.
Published: (2023)
Enhancing Retrieval-Augmented Generation with Entity Linking for Educational Platforms
by: Granata, Francesco, et al.
Published: (2025)
by: Granata, Francesco, et al.
Published: (2025)
SynCABEL: Synthetic Contextualized Augmentation for Biomedical Entity Linking
by: Remaki, Adam, et al.
Published: (2026)
by: Remaki, Adam, et al.
Published: (2026)
Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation
by: Zhang, Chenghao, et al.
Published: (2025)
by: Zhang, Chenghao, et al.
Published: (2025)
BookWorm: A Dataset for Character Description and Analysis
by: Papoudakis, Argyrios, et al.
Published: (2024)
by: Papoudakis, Argyrios, et al.
Published: (2024)
Graph-based Confidence Calibration for Large Language Models
by: Li, Yukun, et al.
Published: (2024)
by: Li, Yukun, et al.
Published: (2024)
Similar Items
-
AutoPureData: Automated Filtering of Undesirable Web Data to Update LLM Knowledge
by: Vadlapati, Praneeth
Published: (2024) -
Large Language Model Augmented Exercise Retrieval for Personalized Language Learning
by: Xu, Austin, et al.
Published: (2024) -
Augmenting Lateral Thinking in Language Models with Humor and Riddle Data for the BRAINTEASER Task
by: Ghashami, Mina, et al.
Published: (2024) -
NyayaAnumana & INLegalLlama: The Largest Indian Legal Judgment Prediction Dataset and Specialized Language Model for Enhanced Decision Analysis
by: Nigam, Shubham Kumar, et al.
Published: (2024) -
Soft Prompt Tuning for Augmenting Dense Retrieval with Large Language Models
by: Peng, Zhiyuan, et al.
Published: (2023)