LongLaMP: A Benchmark for Personalized Long-form Text Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kumar, Ishita, Viswanathan, Snigdha, Yerra, Sushrita, Salemi, Alireza, Rossi, Ryan A., Dernoncourt, Franck, Deilamsalehy, Hanieh, Chen, Xiang, Zhang, Ruiyi, Agarwal, Shubham, Lipka, Nedim, Van Nguyen, Chien, Nguyen, Thien Huu, Zamani, Hamed |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LaMP-QA: A Benchmark for Personalized Long-form Question Answering
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
Taipan: Efficient and Expressive State Space Language Models with Selective Attention
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2024)
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2024)
Comprehensive and Practical Evaluation of Retrieval-Augmented Generation Systems for Medical Question Answering
von: Ngo, Nghia Trung, et al.
Veröffentlicht: (2024)
von: Ngo, Nghia Trung, et al.
Veröffentlicht: (2024)
ExPerT: Effective and Explainable Evaluation of Personalized Long-Form Text Generation
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
Identifying Speakers in Dialogue Transcripts: A Text-based Approach Using Pretrained Language Models
von: Nguyen, Minh, et al.
Veröffentlicht: (2024)
von: Nguyen, Minh, et al.
Veröffentlicht: (2024)
A Multi-LLM Debiasing Framework
von: Owens, Deonna M., et al.
Veröffentlicht: (2024)
von: Owens, Deonna M., et al.
Veröffentlicht: (2024)
LaMP: When Large Language Models Meet Personalization
von: Salemi, Alireza, et al.
Veröffentlicht: (2023)
von: Salemi, Alireza, et al.
Veröffentlicht: (2023)
Lizard: An Efficient Linearization Framework for Large Language Models
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2025)
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2025)
mSCoRe: a $M$ultilingual and Scalable Benchmark for $S$kill-based $Co$mmonsense $Re$asoning
von: Ngo, Nghia Trung, et al.
Veröffentlicht: (2025)
von: Ngo, Nghia Trung, et al.
Veröffentlicht: (2025)
Blind to the Human Touch: Overlap Bias in LLM-Based Summary Evaluation
von: Fang, Jiangnan, et al.
Veröffentlicht: (2026)
von: Fang, Jiangnan, et al.
Veröffentlicht: (2026)
Towards a Search Engine for Machines: Unified Ranking for Multiple Retrieval-Augmented Large Language Models
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
Learning to Rank for Multiple Retrieval-Augmented Models through Iterative Utility Maximization
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
Learning from Natural Language Feedback for Personalized Question Answering
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
Improving User Privacy in Personalized Generation: Client-Side Retrieval-Augmented Modification of Server-Side Generated Speculations
von: Salemi, Alireza, et al.
Veröffentlicht: (2026)
von: Salemi, Alireza, et al.
Veröffentlicht: (2026)
Comparing Retrieval-Augmentation and Parameter-Efficient Fine-Tuning for Privacy-Preserving Personalization of Large Language Models
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
Evaluating Retrieval Quality in Retrieval-Augmented Generation
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
Trust but Verify: Introducing DAVinCI -- A Framework for Dual Attribution and Verification in Claim Inference for Language Models
von: Rawte, Vipula, et al.
Veröffentlicht: (2026)
von: Rawte, Vipula, et al.
Veröffentlicht: (2026)
NoLiMa: Long-Context Evaluation Beyond Literal Matching
von: Modarressi, Ali, et al.
Veröffentlicht: (2025)
von: Modarressi, Ali, et al.
Veröffentlicht: (2025)
Document Attribution: Examining Citation Relationships using Large Language Models
von: Rawte, Vipula, et al.
Veröffentlicht: (2025)
von: Rawte, Vipula, et al.
Veröffentlicht: (2025)
ULLME: A Unified Framework for Large Language Model Embeddings with Generation-Augmented Learning
von: Man, Hieu, et al.
Veröffentlicht: (2024)
von: Man, Hieu, et al.
Veröffentlicht: (2024)
Orthrus: Memory-Efficient Parallel Token Generation via Dual-View Diffusion
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2026)
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2026)
OATS: Opinion Aspect Target Sentiment Quadruple Extraction Dataset for Aspect-Based Sentiment Analysis
von: Chebolu, Siva Uday Sampreeth, et al.
Veröffentlicht: (2023)
von: Chebolu, Siva Uday Sampreeth, et al.
Veröffentlicht: (2023)
ROAST: Review-level Opinion Aspect Sentiment Target Joint Detection for ABSA
von: Chebolu, Siva Uday Sampreeth, et al.
Veröffentlicht: (2024)
von: Chebolu, Siva Uday Sampreeth, et al.
Veröffentlicht: (2024)
Beyond Factual Accuracy: Evaluating Coverage of Diverse Factual Information in Long-form Text Generation
von: Samarinas, Chris, et al.
Veröffentlicht: (2025)
von: Samarinas, Chris, et al.
Veröffentlicht: (2025)
Learning to Reason for Multi-Step Retrieval of Personal Context in Personalized Question Answering
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2026)
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2026)
CIIR@LiveRAG 2025: Optimizing Multi-Agent Retrieval Augmented Generation through Self-Training
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
Plan-and-Refine: Diverse and Comprehensive Retrieval-Augmented Generation
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
Optimization Methods for Personalizing Large Language Models through Retrieval Augmentation
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
KaPQA: Knowledge-Augmented Product Question-Answering
von: Eppalapally, Swetha, et al.
Veröffentlicht: (2024)
von: Eppalapally, Swetha, et al.
Veröffentlicht: (2024)
Towards Enhancing Coherence in Extractive Summarization: Dataset and Experiments with LLMs
von: Parmar, Mihir, et al.
Veröffentlicht: (2024)
von: Parmar, Mihir, et al.
Veröffentlicht: (2024)
Beyond a Million Tokens: Benchmarking and Enhancing Long-Term Memory in LLMs
von: Tavakoli, Mohammad, et al.
Veröffentlicht: (2025)
von: Tavakoli, Mohammad, et al.
Veröffentlicht: (2025)
Explainable Disentangled Representation Learning for Generalizable Authorship Attribution in the Era of Generative AI
von: Man, Hieu, et al.
Veröffentlicht: (2026)
von: Man, Hieu, et al.
Veröffentlicht: (2026)
Multi-LLM Text Summarization
von: Fang, Jiangnan, et al.
Veröffentlicht: (2024)
von: Fang, Jiangnan, et al.
Veröffentlicht: (2024)
Evaluation of Agents under Simulated AI Marketplace Dynamics
von: Kim, To Eun, et al.
Veröffentlicht: (2026)
von: Kim, To Eun, et al.
Veröffentlicht: (2026)
Critic-R: Improving Agentic Search using Instruction-tuned Retrievers with Natural Language Introspective Feedback
von: Alam, Md Zarif Ul, et al.
Veröffentlicht: (2026)
von: Alam, Md Zarif Ul, et al.
Veröffentlicht: (2026)
ChartLens: Fine-grained Visual Attribution in Charts
von: Suri, Manan, et al.
Veröffentlicht: (2025)
von: Suri, Manan, et al.
Veröffentlicht: (2025)
Structured Uncertainty guided Clarification for LLM Agents
von: Suri, Manan, et al.
Veröffentlicht: (2025)
von: Suri, Manan, et al.
Veröffentlicht: (2025)
Self-Debiasing Large Language Models: Zero-Shot Recognition and Reduction of Stereotypes
von: Gallegos, Isabel O., et al.
Veröffentlicht: (2024)
von: Gallegos, Isabel O., et al.
Veröffentlicht: (2024)
VipAct: Visual-Perception Enhancement via Specialized VLM Agent Collaboration and Tool-use
von: Zhang, Zhehao, et al.
Veröffentlicht: (2024)
von: Zhang, Zhehao, et al.
Veröffentlicht: (2024)
A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality
von: Elmoghany, Mohamed, et al.
Veröffentlicht: (2025)
von: Elmoghany, Mohamed, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LaMP-QA: A Benchmark for Personalized Long-form Question Answering
von: Salemi, Alireza, et al.
Veröffentlicht: (2025) -
Taipan: Efficient and Expressive State Space Language Models with Selective Attention
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2024) -
Comprehensive and Practical Evaluation of Retrieval-Augmented Generation Systems for Medical Question Answering
von: Ngo, Nghia Trung, et al.
Veröffentlicht: (2024) -
ExPerT: Effective and Explainable Evaluation of Personalized Long-Form Text Generation
von: Salemi, Alireza, et al.
Veröffentlicht: (2025) -
Identifying Speakers in Dialogue Transcripts: A Text-based Approach Using Pretrained Language Models
von: Nguyen, Minh, et al.
Veröffentlicht: (2024)