LongRecall: A Structured Approach for Robust Recall Evaluation in Long-Form Text
Fuente:
arXiv
Salvato in:
| Autori principali: | Ardestani, MohamamdJavad, Kamalloo, Ehsan, Rafiei, Davood |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Hindsight is 20/20: Building Agent Memory that Retains, Recalls, and Reflects
di: Latimer, Chris, et al.
Pubblicazione: (2025)
di: Latimer, Chris, et al.
Pubblicazione: (2025)
LFOSum: Summarizing Long-form Opinions with Large Language Models
di: Nayeem, Mir Tafseer, et al.
Pubblicazione: (2024)
di: Nayeem, Mir Tafseer, et al.
Pubblicazione: (2024)
OpinioRAG: Towards Generating User-Centric Opinion Highlights from Large-scale Online Reviews
di: Nayeem, Mir Tafseer, et al.
Pubblicazione: (2025)
di: Nayeem, Mir Tafseer, et al.
Pubblicazione: (2025)
Recall: Empowering Multimodal Embedding for Edge Devices
di: Cai, Dongqi, et al.
Pubblicazione: (2024)
di: Cai, Dongqi, et al.
Pubblicazione: (2024)
PRISM: Agentic Retrieval with LLMs for Multi-Hop Question Answering
di: Nahid, Md Mahadi Hasan, et al.
Pubblicazione: (2025)
di: Nahid, Md Mahadi Hasan, et al.
Pubblicazione: (2025)
LongKey: Keyphrase Extraction for Long Documents
di: Alves, Jeovane Honorio, et al.
Pubblicazione: (2024)
di: Alves, Jeovane Honorio, et al.
Pubblicazione: (2024)
ExPerT: Effective and Explainable Evaluation of Personalized Long-Form Text Generation
di: Salemi, Alireza, et al.
Pubblicazione: (2025)
di: Salemi, Alireza, et al.
Pubblicazione: (2025)
NormTab: Improving Symbolic Reasoning in LLMs Through Tabular Data Normalization
di: Nahid, Md Mahadi Hasan, et al.
Pubblicazione: (2024)
di: Nahid, Md Mahadi Hasan, et al.
Pubblicazione: (2024)
PluriHopRAG: Exhaustive, Recall-Sensitive QA Through Corpus-Specific Document Structure Learning
di: Sveistrys, Mykolas, et al.
Pubblicazione: (2025)
di: Sveistrys, Mykolas, et al.
Pubblicazione: (2025)
Retrieval meets Long Context Large Language Models
di: Xu, Peng, et al.
Pubblicazione: (2023)
di: Xu, Peng, et al.
Pubblicazione: (2023)
Reasoning-Enhanced Self-Training for Long-Form Personalized Text Generation
di: Salemi, Alireza, et al.
Pubblicazione: (2025)
di: Salemi, Alireza, et al.
Pubblicazione: (2025)
Large Language Models are Learnable Planners for Long-Term Recommendation
di: Shi, Wentao, et al.
Pubblicazione: (2024)
di: Shi, Wentao, et al.
Pubblicazione: (2024)
Microstructures and Accuracy of Graph Recall by Large Language Models
di: Wang, Yanbang, et al.
Pubblicazione: (2024)
di: Wang, Yanbang, et al.
Pubblicazione: (2024)
Eliciting In-context Retrieval and Reasoning for Long-context Large Language Models
di: Qiu, Yifu, et al.
Pubblicazione: (2025)
di: Qiu, Yifu, et al.
Pubblicazione: (2025)
Can't Remember Details in Long Documents? You Need Some R&R
di: Agrawal, Devanshu, et al.
Pubblicazione: (2024)
di: Agrawal, Devanshu, et al.
Pubblicazione: (2024)
StructMem: Structured Memory for Long-Horizon Behavior in LLMs
di: Xu, Buqiang, et al.
Pubblicazione: (2026)
di: Xu, Buqiang, et al.
Pubblicazione: (2026)
ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities
di: Xu, Peng, et al.
Pubblicazione: (2024)
di: Xu, Peng, et al.
Pubblicazione: (2024)
DuoLens: A Framework for Robust Detection of Machine-Generated Multilingual Text and Code
di: Agrawal, Shriyansh, et al.
Pubblicazione: (2025)
di: Agrawal, Shriyansh, et al.
Pubblicazione: (2025)
RAIR: A Rule-Aware Benchmark Uniting Challenging Long-Tail and Visual Salience Subset for E-commerce Relevance Assessment
di: Lu, Chenji, et al.
Pubblicazione: (2025)
di: Lu, Chenji, et al.
Pubblicazione: (2025)
Is Implicit Knowledge Enough for LLMs? A RAG Approach for Tree-based Structures
di: Gupte, Mihir, et al.
Pubblicazione: (2025)
di: Gupte, Mihir, et al.
Pubblicazione: (2025)
Inside CORE-KG: Evaluating Structured Prompting and Coreference Resolution for Knowledge Graphs
di: Meher, Dipak, et al.
Pubblicazione: (2025)
di: Meher, Dipak, et al.
Pubblicazione: (2025)
A Systematic Evaluation of LLM Strategies for Mental Health Text Analysis: Fine-tuning vs. Prompt Engineering vs. RAG
di: Kermani, Arshia, et al.
Pubblicazione: (2025)
di: Kermani, Arshia, et al.
Pubblicazione: (2025)
Structured Legal Document Generation in India: A Model-Agnostic Wrapper Approach with VidhikDastaavej
di: Nigam, Shubham Kumar, et al.
Pubblicazione: (2025)
di: Nigam, Shubham Kumar, et al.
Pubblicazione: (2025)
Towards Robust Evaluation: A Comprehensive Taxonomy of Datasets and Metrics for Open Domain Question Answering in the Era of Large Language Models
di: Srivastava, Akchay, et al.
Pubblicazione: (2024)
di: Srivastava, Akchay, et al.
Pubblicazione: (2024)
FullRecall: A Semantic Search-Based Ranking Approach for Maximizing Recall in Patent Retrieval
di: Ali, Amna, et al.
Pubblicazione: (2025)
di: Ali, Amna, et al.
Pubblicazione: (2025)
RUST-BENCH: Benchmarking LLM Reasoning on Unstructured Text within Structured Tables
di: Abhyankar, Nikhil, et al.
Pubblicazione: (2025)
di: Abhyankar, Nikhil, et al.
Pubblicazione: (2025)
Recall Them All: Retrieval-Augmented Language Models for Long Object List Extraction from Long Documents
di: Singhania, Sneha, et al.
Pubblicazione: (2024)
di: Singhania, Sneha, et al.
Pubblicazione: (2024)
MUST-RAG: MUSical Text Question Answering with Retrieval Augmented Generation
di: Kwon, Daeyong, et al.
Pubblicazione: (2025)
di: Kwon, Daeyong, et al.
Pubblicazione: (2025)
KGGen: Extracting Knowledge Graphs from Plain Text with Language Models
di: Mo, Belinda, et al.
Pubblicazione: (2025)
di: Mo, Belinda, et al.
Pubblicazione: (2025)
Deep-Reporter: Deep Research for Grounded Multimodal Long-Form Generation
di: Ye, Fangda, et al.
Pubblicazione: (2026)
di: Ye, Fangda, et al.
Pubblicazione: (2026)
Towards a Robust Retrieval-Based Summarization System
di: Liu, Shengjie, et al.
Pubblicazione: (2024)
di: Liu, Shengjie, et al.
Pubblicazione: (2024)
UniGLM: Training One Unified Language Model for Text-Attributed Graph Embedding
di: Fang, Yi, et al.
Pubblicazione: (2024)
di: Fang, Yi, et al.
Pubblicazione: (2024)
SemanticZip: A Pilot Framework for Lossy Text Compression with LLMs as Semantic Decompressors
di: Trukhina, Natalia, et al.
Pubblicazione: (2026)
di: Trukhina, Natalia, et al.
Pubblicazione: (2026)
Robust Neural Information Retrieval: An Adversarial and Out-of-distribution Perspective
di: Liu, Yu-An, et al.
Pubblicazione: (2024)
di: Liu, Yu-An, et al.
Pubblicazione: (2024)
MixLM: High-Throughput and Effective LLM Ranking via Text-Embedding Mix-Interaction
di: Li, Guoyao, et al.
Pubblicazione: (2025)
di: Li, Guoyao, et al.
Pubblicazione: (2025)
When LLMs are Unfit Use FastFit: Fast and Effective Text Classification with Many Classes
di: Yehudai, Asaf, et al.
Pubblicazione: (2024)
di: Yehudai, Asaf, et al.
Pubblicazione: (2024)
Q-PEFT: Query-dependent Parameter Efficient Fine-tuning for Text Reranking with Large Language Models
di: Peng, Zhiyuan, et al.
Pubblicazione: (2024)
di: Peng, Zhiyuan, et al.
Pubblicazione: (2024)
Approaching Human-Level Forecasting with Language Models
di: Halawi, Danny, et al.
Pubblicazione: (2024)
di: Halawi, Danny, et al.
Pubblicazione: (2024)
Semantic Recall for Vector Search
di: Kuffo, Leonardo, et al.
Pubblicazione: (2026)
di: Kuffo, Leonardo, et al.
Pubblicazione: (2026)
STRUM-LLM: Attributed and Structured Contrastive Summarization
di: Gunel, Beliz, et al.
Pubblicazione: (2024)
di: Gunel, Beliz, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Hindsight is 20/20: Building Agent Memory that Retains, Recalls, and Reflects
di: Latimer, Chris, et al.
Pubblicazione: (2025) -
LFOSum: Summarizing Long-form Opinions with Large Language Models
di: Nayeem, Mir Tafseer, et al.
Pubblicazione: (2024) -
OpinioRAG: Towards Generating User-Centric Opinion Highlights from Large-scale Online Reviews
di: Nayeem, Mir Tafseer, et al.
Pubblicazione: (2025) -
Recall: Empowering Multimodal Embedding for Edge Devices
di: Cai, Dongqi, et al.
Pubblicazione: (2024) -
PRISM: Agentic Retrieval with LLMs for Multi-Hop Question Answering
di: Nahid, Md Mahadi Hasan, et al.
Pubblicazione: (2025)