Long document summarization using page specific target text alignment and distilling page importance
Fuente:
arXiv
Guardado en:
| Autores principales: | Devi, Pushpa, Agrawal, Ayush, Dubey, Ashutosh, Chowdary, C. Ravindranath |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SlideSpawn: An Automatic Slides Generation System for Research Publications
por: Kumar, Keshav, et al.
Publicado: (2024)
por: Kumar, Keshav, et al.
Publicado: (2024)
Loops On Retrieval Augmented Generation (LoRAG)
por: Thakur, Ayush, et al.
Publicado: (2024)
por: Thakur, Ayush, et al.
Publicado: (2024)
Abstractive summarization from Audio Transcription
por: Derkach, Ilia
Publicado: (2024)
por: Derkach, Ilia
Publicado: (2024)
Is text normalization relevant for classifying medieval charters?
por: Atzenhofer-Baumgartner, Florian, et al.
Publicado: (2024)
por: Atzenhofer-Baumgartner, Florian, et al.
Publicado: (2024)
Benchmarking pre-trained text embedding models in aligning built asset information
por: Shahinmoghadam, Mehrzad, et al.
Publicado: (2024)
por: Shahinmoghadam, Mehrzad, et al.
Publicado: (2024)
$\text{R}^2\text{R}$: A Route-to-Rerank Post-Training Framework for Multi-Domain Decoder-Only Rerankers
por: Wang, Xinyu, et al.
Publicado: (2025)
por: Wang, Xinyu, et al.
Publicado: (2025)
Facilitating phenotyping from clinical texts: the medkit library
por: Neuraz, Antoine, et al.
Publicado: (2024)
por: Neuraz, Antoine, et al.
Publicado: (2024)
How important is Recall for Measuring Retrieval Quality?
por: Schwartz, Shelly, et al.
Publicado: (2025)
por: Schwartz, Shelly, et al.
Publicado: (2025)
Is Architectural Complexity Overrated? Competitive and Interpretable Knowledge Graph Completion with RelatE
por: Chakraborty, Abhijit, et al.
Publicado: (2025)
por: Chakraborty, Abhijit, et al.
Publicado: (2025)
Evaluation of the phi-3-mini SLM for identification of texts related to medicine, health, and sports injuries
por: Brogly, Chris, et al.
Publicado: (2025)
por: Brogly, Chris, et al.
Publicado: (2025)
Retrieving Contextual Information for Long-Form Question Answering using Weak Supervision
por: Christmann, Philipp, et al.
Publicado: (2024)
por: Christmann, Philipp, et al.
Publicado: (2024)
Causality extraction from medical text using Large Language Models (LLMs)
por: Gopalakrishnan, Seethalakshmi, et al.
Publicado: (2024)
por: Gopalakrishnan, Seethalakshmi, et al.
Publicado: (2024)
Introducing Super RAGs in Mistral 8x7B-v1
por: Thakur, Ayush, et al.
Publicado: (2024)
por: Thakur, Ayush, et al.
Publicado: (2024)
MRGSEM-Sum: An Unsupervised Multi-document Summarization Framework based on Multi-Relational Graphs and Structural Entropy Minimization
por: Zhang, Yongbing, et al.
Publicado: (2025)
por: Zhang, Yongbing, et al.
Publicado: (2025)
Deep Learning Based Named Entity Recognition Models for Recipes
por: Goel, Mansi, et al.
Publicado: (2024)
por: Goel, Mansi, et al.
Publicado: (2024)
REARANK: Reasoning Re-ranking Agent via Reinforcement Learning
por: Zhang, Le, et al.
Publicado: (2025)
por: Zhang, Le, et al.
Publicado: (2025)
OneSearch-V2: The Latent Reasoning Enhanced Self-distillation Generative Search Framework
por: Chen, Ben, et al.
Publicado: (2026)
por: Chen, Ben, et al.
Publicado: (2026)
EIoU-EMC: A Novel Loss for Domain-specific Nested Entity Recognition
por: Zhang, Jian, et al.
Publicado: (2025)
por: Zhang, Jian, et al.
Publicado: (2025)
DREditor: An Time-efficient Approach for Building a Domain-specific Dense Retrieval Model
por: Huang, Chen, et al.
Publicado: (2024)
por: Huang, Chen, et al.
Publicado: (2024)
Passage-specific Prompt Tuning for Passage Reranking in Question Answering with Large Language Models
por: Wu, Xuyang, et al.
Publicado: (2024)
por: Wu, Xuyang, et al.
Publicado: (2024)
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation
por: Wang, Shuting, et al.
Publicado: (2024)
por: Wang, Shuting, et al.
Publicado: (2024)
FsPONER: Few-shot Prompt Optimization for Named Entity Recognition in Domain-specific Scenarios
por: Tang, Yongjian, et al.
Publicado: (2024)
por: Tang, Yongjian, et al.
Publicado: (2024)
OkraLong: A Flexible Retrieval-Augmented Framework for Long-Text Query Processing
por: Hui, Yulong, et al.
Publicado: (2025)
por: Hui, Yulong, et al.
Publicado: (2025)
SciNets: Graph-Constrained Multi-Hop Reasoning for Scientific Literature Synthesis
por: Dubey, Sauhard
Publicado: (2025)
por: Dubey, Sauhard
Publicado: (2025)
Long Dialog Summarization: An Analysis
por: Mullick, Ankan, et al.
Publicado: (2024)
por: Mullick, Ankan, et al.
Publicado: (2024)
Can't Remember Details in Long Documents? You Need Some R&R
por: Agrawal, Devanshu, et al.
Publicado: (2024)
por: Agrawal, Devanshu, et al.
Publicado: (2024)
Adaptive Data-Resilient Multi-Modal Hierarchical Multi-Label Book Genre Identification
por: Nareti, Utsav Kumar, et al.
Publicado: (2025)
por: Nareti, Utsav Kumar, et al.
Publicado: (2025)
Aug2Search: Enhancing Facebook Marketplace Search with LLM-Generated Synthetic Data Augmentation
por: Xi, Ruijie, et al.
Publicado: (2025)
por: Xi, Ruijie, et al.
Publicado: (2025)
NEAR$^2$: A Nested Embedding Approach to Efficient Product Retrieval and Ranking
por: Qian, Shenbin, et al.
Publicado: (2025)
por: Qian, Shenbin, et al.
Publicado: (2025)
Recall Them All: Retrieval-Augmented Language Models for Long Object List Extraction from Long Documents
por: Singhania, Sneha, et al.
Publicado: (2024)
por: Singhania, Sneha, et al.
Publicado: (2024)
Noise reduction in BERT NER models for clinical entity extraction
por: Jiwani, Kuldeep, et al.
Publicado: (2026)
por: Jiwani, Kuldeep, et al.
Publicado: (2026)
Leveraging Hierarchical Organization for Medical Multi-document Summarization
por: Hsu, Yi-Li, et al.
Publicado: (2025)
por: Hsu, Yi-Li, et al.
Publicado: (2025)
Multi-Relation Extraction in Entity Pairs using Global Context
por: Nilesh, et al.
Publicado: (2025)
por: Nilesh, et al.
Publicado: (2025)
Who's important? -- SUnSET: Synergistic Understanding of Stakeholder, Events and Time for Timeline Generation
por: Sim, Tiviatis, et al.
Publicado: (2025)
por: Sim, Tiviatis, et al.
Publicado: (2025)
Unstructured Evidence Attribution for Long Context Query Focused Summarization
por: Wright, Dustin, et al.
Publicado: (2025)
por: Wright, Dustin, et al.
Publicado: (2025)
Graph-Based Retriever Captures the Long Tail of Biomedical Knowledge
por: Delile, Julien, et al.
Publicado: (2024)
por: Delile, Julien, et al.
Publicado: (2024)
Writing in the Margins: Better Inference Pattern for Long Context Retrieval
por: Russak, Melisa, et al.
Publicado: (2024)
por: Russak, Melisa, et al.
Publicado: (2024)
Positional Bias in Long-Document Ranking: Impact, Assessment, and Mitigation
por: Boytsov, Leonid, et al.
Publicado: (2022)
por: Boytsov, Leonid, et al.
Publicado: (2022)
Millions of $\text{GeAR}$-s: Extending GraphRAG to Millions of Documents
por: Shen, Zhili, et al.
Publicado: (2025)
por: Shen, Zhili, et al.
Publicado: (2025)
Exploring Selective Retrieval-Augmentation for Long-Tail Legal Text Classification
por: Mao, Boheng
Publicado: (2025)
por: Mao, Boheng
Publicado: (2025)
Ejemplares similares
-
SlideSpawn: An Automatic Slides Generation System for Research Publications
por: Kumar, Keshav, et al.
Publicado: (2024) -
Loops On Retrieval Augmented Generation (LoRAG)
por: Thakur, Ayush, et al.
Publicado: (2024) -
Abstractive summarization from Audio Transcription
por: Derkach, Ilia
Publicado: (2024) -
Is text normalization relevant for classifying medieval charters?
por: Atzenhofer-Baumgartner, Florian, et al.
Publicado: (2024) -
Benchmarking pre-trained text embedding models in aligning built asset information
por: Shahinmoghadam, Mehrzad, et al.
Publicado: (2024)