Is text normalization relevant for classifying medieval charters?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Atzenhofer-Baumgartner, Florian, Kovács, Tamás |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
What Do Humanities Scholars Need? A User Model for Recommendation in Digital Archives
von: Atzenhofer-Baumgartner, Florian, et al.
Veröffentlicht: (2026)
von: Atzenhofer-Baumgartner, Florian, et al.
Veröffentlicht: (2026)
A Multistakeholder Approach to Value-Driven Co-Design of Recommender System Evaluation Metrics in Digital Archives
von: Atzenhofer-Baumgartner, Florian, et al.
Veröffentlicht: (2025)
von: Atzenhofer-Baumgartner, Florian, et al.
Veröffentlicht: (2025)
Specialized text classification: an approach to classifying Open Banking transactions
von: TA, Duc Tuyen, et al.
Veröffentlicht: (2025)
von: TA, Duc Tuyen, et al.
Veröffentlicht: (2025)
Value Identification in Multistakeholder Recommender Systems for Humanities and Historical Research: The Case of the Digital Archive Monasterium.net
von: Atzenhofer-Baumgartner, Florian, et al.
Veröffentlicht: (2024)
von: Atzenhofer-Baumgartner, Florian, et al.
Veröffentlicht: (2024)
Challenges in Implementing a Recommender System for Historical Research in the Humanities
von: Atzenhofer-Baumgartner, Florian, et al.
Veröffentlicht: (2024)
von: Atzenhofer-Baumgartner, Florian, et al.
Veröffentlicht: (2024)
$\text{R}^2\text{R}$: A Route-to-Rerank Post-Training Framework for Multi-Domain Decoder-Only Rerankers
von: Wang, Xinyu, et al.
Veröffentlicht: (2025)
von: Wang, Xinyu, et al.
Veröffentlicht: (2025)
Facilitating phenotyping from clinical texts: the medkit library
von: Neuraz, Antoine, et al.
Veröffentlicht: (2024)
von: Neuraz, Antoine, et al.
Veröffentlicht: (2024)
LLM-as-classifier: Semi-Supervised, Iterative Framework for Hierarchical Text Classification using Large Language Models
von: You, Doohee, et al.
Veröffentlicht: (2025)
von: You, Doohee, et al.
Veröffentlicht: (2025)
Long document summarization using page specific target text alignment and distilling page importance
von: Devi, Pushpa, et al.
Veröffentlicht: (2025)
von: Devi, Pushpa, et al.
Veröffentlicht: (2025)
Evaluation of the phi-3-mini SLM for identification of texts related to medicine, health, and sports injuries
von: Brogly, Chris, et al.
Veröffentlicht: (2025)
von: Brogly, Chris, et al.
Veröffentlicht: (2025)
PeerQA: A Scientific Question Answering Dataset from Peer Reviews
von: Baumgärtner, Tim, et al.
Veröffentlicht: (2025)
von: Baumgärtner, Tim, et al.
Veröffentlicht: (2025)
Conversational Exploratory Search of Scholarly Publications Using Knowledge Graphs
von: Schneider, Phillip, et al.
Veröffentlicht: (2024)
von: Schneider, Phillip, et al.
Veröffentlicht: (2024)
An Analysis of Datasets, Metrics and Models in Keyphrase Generation
von: Boudin, Florian, et al.
Veröffentlicht: (2025)
von: Boudin, Florian, et al.
Veröffentlicht: (2025)
Enhancing Answer Attribution for Faithful Text Generation with Large Language Models
von: Vladika, Juraj, et al.
Veröffentlicht: (2024)
von: Vladika, Juraj, et al.
Veröffentlicht: (2024)
WikiHint: A Human-Annotated Dataset for Hint Ranking and Generation
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
Self-Compositional Data Augmentation for Scientific Keyphrase Generation
von: Houbre, Mael, et al.
Veröffentlicht: (2024)
von: Houbre, Mael, et al.
Veröffentlicht: (2024)
Millions of $\text{GeAR}$-s: Extending GraphRAG to Millions of Documents
von: Shen, Zhili, et al.
Veröffentlicht: (2025)
von: Shen, Zhili, et al.
Veröffentlicht: (2025)
Causality extraction from medical text using Large Language Models (LLMs)
von: Gopalakrishnan, Seethalakshmi, et al.
Veröffentlicht: (2024)
von: Gopalakrishnan, Seethalakshmi, et al.
Veröffentlicht: (2024)
Classification performance and reproducibility of GPT-4 omni for information extraction from veterinary electronic health records
von: Wulcan, Judit M, et al.
Veröffentlicht: (2024)
von: Wulcan, Judit M, et al.
Veröffentlicht: (2024)
Evaluating Large Language Models in Semantic Parsing for Conversational Question Answering over Knowledge Graphs
von: Schneider, Phillip, et al.
Veröffentlicht: (2024)
von: Schneider, Phillip, et al.
Veröffentlicht: (2024)
Recent advances in text embedding: A Comprehensive Review of Top-Performing Methods on the MTEB Benchmark
von: Cao, Hongliu
Veröffentlicht: (2024)
von: Cao, Hongliu
Veröffentlicht: (2024)
Automated Generation of Research Workflows from Academic Papers: A Full-text Mining Framework
von: Zhang, Heng, et al.
Veröffentlicht: (2025)
von: Zhang, Heng, et al.
Veröffentlicht: (2025)
LMK > CLS: Landmark Pooling for Dense Embeddings
von: Doshi, Meet, et al.
Veröffentlicht: (2026)
von: Doshi, Meet, et al.
Veröffentlicht: (2026)
Comparing Knowledge Sources for Open-Domain Scientific Claim Verification
von: Vladika, Juraj, et al.
Veröffentlicht: (2024)
von: Vladika, Juraj, et al.
Veröffentlicht: (2024)
Improving Health Question Answering with Reliable and Time-Aware Evidence Retrieval
von: Vladika, Juraj, et al.
Veröffentlicht: (2024)
von: Vladika, Juraj, et al.
Veröffentlicht: (2024)
MedSEBA: Synthesizing Evidence-Based Answers Grounded in Evolving Medical Literature
von: Vladika, Juraj, et al.
Veröffentlicht: (2025)
von: Vladika, Juraj, et al.
Veröffentlicht: (2025)
Granite Embedding R2 Models
von: Awasthy, Parul, et al.
Veröffentlicht: (2025)
von: Awasthy, Parul, et al.
Veröffentlicht: (2025)
Survey: Understand the challenges of MachineLearning Experts using Named EntityRecognition Tools
von: Freund, Florian, et al.
Veröffentlicht: (2025)
von: Freund, Florian, et al.
Veröffentlicht: (2025)
Query Attribute Modeling: Improving search relevance with Semantic Search and Meta Data Filtering
von: Menon, Karthik, et al.
Veröffentlicht: (2025)
von: Menon, Karthik, et al.
Veröffentlicht: (2025)
Granite Embedding Models
von: Awasthy, Parul, et al.
Veröffentlicht: (2025)
von: Awasthy, Parul, et al.
Veröffentlicht: (2025)
Benchmarking pre-trained text embedding models in aligning built asset information
von: Shahinmoghadam, Mehrzad, et al.
Veröffentlicht: (2024)
von: Shahinmoghadam, Mehrzad, et al.
Veröffentlicht: (2024)
Unmasking Superspreaders: Data-Driven Approaches for Identifying and Comparing Key Influencers of Conspiracy Theories on X.com
von: Kramer, Florian, et al.
Veröffentlicht: (2026)
von: Kramer, Florian, et al.
Veröffentlicht: (2026)
From Topology to Retrieval: Decoding Embedding Spaces with Unified Signatures
von: Rottach, Florian, et al.
Veröffentlicht: (2025)
von: Rottach, Florian, et al.
Veröffentlicht: (2025)
CodeTaxo: Enhancing Taxonomy Expansion with Limited Examples via Code Language Prompts
von: Zeng, Qingkai, et al.
Veröffentlicht: (2024)
von: Zeng, Qingkai, et al.
Veröffentlicht: (2024)
Ask Optimal Questions: Aligning Large Language Models with Retriever's Preference in Conversation
von: Yoon, Chanwoong, et al.
Veröffentlicht: (2024)
von: Yoon, Chanwoong, et al.
Veröffentlicht: (2024)
Bootstrap Your Own Context Length
von: Wang, Liang, et al.
Veröffentlicht: (2024)
von: Wang, Liang, et al.
Veröffentlicht: (2024)
Bridging Personalization and Control in Scientific Personalized Search
von: Mysore, Sheshera, et al.
Veröffentlicht: (2024)
von: Mysore, Sheshera, et al.
Veröffentlicht: (2024)
Streamlining Systematic Reviews: A Novel Application of Large Language Models
von: Trad, Fouad, et al.
Veröffentlicht: (2024)
von: Trad, Fouad, et al.
Veröffentlicht: (2024)
Item-Language Model for Conversational Recommendation
von: Yang, Li, et al.
Veröffentlicht: (2024)
von: Yang, Li, et al.
Veröffentlicht: (2024)
A Text is Worth Several Tokens: Text Embedding from LLMs Secretly Aligns Well with The Key Tokens
von: Nie, Zhijie, et al.
Veröffentlicht: (2024)
von: Nie, Zhijie, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
What Do Humanities Scholars Need? A User Model for Recommendation in Digital Archives
von: Atzenhofer-Baumgartner, Florian, et al.
Veröffentlicht: (2026) -
A Multistakeholder Approach to Value-Driven Co-Design of Recommender System Evaluation Metrics in Digital Archives
von: Atzenhofer-Baumgartner, Florian, et al.
Veröffentlicht: (2025) -
Specialized text classification: an approach to classifying Open Banking transactions
von: TA, Duc Tuyen, et al.
Veröffentlicht: (2025) -
Value Identification in Multistakeholder Recommender Systems for Humanities and Historical Research: The Case of the Digital Archive Monasterium.net
von: Atzenhofer-Baumgartner, Florian, et al.
Veröffentlicht: (2024) -
Challenges in Implementing a Recommender System for Historical Research in the Humanities
von: Atzenhofer-Baumgartner, Florian, et al.
Veröffentlicht: (2024)