GPT-5 vs Other LLMs in Long Short-Context Performance
Fuente:
arXiv
Guardado en:
| Autores principales: | Esmi, Nima, Nezhad-Moghaddam, Maryam, Borhani, Fatemeh, Shahbahrami, Asadollah, Daemdoost, Amin, Gaydadjiev, Georgi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Motion Compensation for Real Time Ultrasound Scanning in Robotically Assisted Prostate Biopsy Procedures
por: Markulin, Matija, et al.
Publicado: (2026)
por: Markulin, Matija, et al.
Publicado: (2026)
BALI: Enhancing Biomedical Language Representations through Knowledge Graph and Language Model Alignment
por: Sakhovskiy, Andrey, et al.
Publicado: (2025)
por: Sakhovskiy, Andrey, et al.
Publicado: (2025)
Prompt-RAG: Pioneering Vector Embedding-Free Retrieval-Augmented Generation in Niche Domains, Exemplified by Korean Medicine
por: Kang, Bongsu, et al.
Publicado: (2024)
por: Kang, Bongsu, et al.
Publicado: (2024)
ParliaBench: An Evaluation and Benchmarking Framework for LLM-Generated Parliamentary Speech
por: Koniaris, Marios, et al.
Publicado: (2025)
por: Koniaris, Marios, et al.
Publicado: (2025)
PubMed Reasoner: Dynamic Reasoning-based Retrieval for Evidence-Grounded Biomedical Question Answering
por: Zhang, Yiqing, et al.
Publicado: (2026)
por: Zhang, Yiqing, et al.
Publicado: (2026)
ChronoMedKG: A Temporally-Grounded Biomedical Knowledge Graph and Benchmark for Clinical Reasoning
por: Ahmed, Md Shamim, et al.
Publicado: (2026)
por: Ahmed, Md Shamim, et al.
Publicado: (2026)
An Interpretable Ensemble of Graph and Language Models for Improving Search Relevance in E-Commerce
por: Choudhary, Nurendra, et al.
Publicado: (2024)
por: Choudhary, Nurendra, et al.
Publicado: (2024)
Computational-Assisted Systematic Review and Meta-Analysis (CASMA): Effect of a Subclass of GnRH-a on Endometriosis Recurrence
por: Tsang, Sandro
Publicado: (2025)
por: Tsang, Sandro
Publicado: (2025)
Curated AI beats frontier LLMs at pharma asset discovery
por: Kidziński, Łukasz, et al.
Publicado: (2026)
por: Kidziński, Łukasz, et al.
Publicado: (2026)
GPT has become financially literate: Insights from financial literacy tests of GPT and a preliminary test of how people use it as a source of advice
por: Niszczota, Paweł, et al.
Publicado: (2023)
por: Niszczota, Paweł, et al.
Publicado: (2023)
Agentic Framework for Political Biography Extraction
por: Zhu, Yifei, et al.
Publicado: (2026)
por: Zhu, Yifei, et al.
Publicado: (2026)
Large language models in finance : what is financial sentiment?
por: Kirtac, Kemal, et al.
Publicado: (2025)
por: Kirtac, Kemal, et al.
Publicado: (2025)
Legal RAG Bench: an end-to-end benchmark for legal RAG
por: Butler, Abdur-Rahman, et al.
Publicado: (2026)
por: Butler, Abdur-Rahman, et al.
Publicado: (2026)
Weakly Supervised Segmentation Framework for Thyroid Nodule Based on High-confidence Labels and High-rationality Losses
por: Chi, Jianning, et al.
Publicado: (2025)
por: Chi, Jianning, et al.
Publicado: (2025)
DALL-M: Context-Aware Clinical Data Augmentation with LLMs
por: Hsieh, Chihcheng, et al.
Publicado: (2024)
por: Hsieh, Chihcheng, et al.
Publicado: (2024)
Mind the Gap: Aligning Knowledge Bases with User Needs to Enhance Mental Health Retrieval
por: Chan, Amanda, et al.
Publicado: (2025)
por: Chan, Amanda, et al.
Publicado: (2025)
The Table of Media Bias Elements: A sentence-level taxonomy of media bias types and propaganda techniques
por: Menzner, Tim, et al.
Publicado: (2026)
por: Menzner, Tim, et al.
Publicado: (2026)
Current and future roles of artificial intelligence in retinopathy of prematurity
por: Jafarizadeh, Ali, et al.
Publicado: (2024)
por: Jafarizadeh, Ali, et al.
Publicado: (2024)
INESC-ID @ eRisk 2025: Exploring Fine-Tuned, Similarity-Based, and Prompt-Based Approaches to Depression Symptom Identification
por: Nunes, Diogo A. P., et al.
Publicado: (2025)
por: Nunes, Diogo A. P., et al.
Publicado: (2025)
Diagnosis extraction from unstructured Dutch echocardiogram reports using span- and document-level characteristic classification
por: Arends, Bauke, et al.
Publicado: (2024)
por: Arends, Bauke, et al.
Publicado: (2024)
TransResAI: A Compound AI System for Coastal Transportation Resilience
por: Pu, Qingwen, et al.
Publicado: (2026)
por: Pu, Qingwen, et al.
Publicado: (2026)
EasyNER: A Customizable Easy-to-Use Pipeline for Deep Learning- and Dictionary-based Named Entity Recognition from Medical and Life Science Text
por: Ahmed, Rafsan, et al.
Publicado: (2023)
por: Ahmed, Rafsan, et al.
Publicado: (2023)
When Content is Goliath and Algorithm is David: The Style and Semantic Effects of Generative Search Engine
por: Ma, Lijia, et al.
Publicado: (2025)
por: Ma, Lijia, et al.
Publicado: (2025)
Attention Asymmetry in AI Layoff Discourse on X: A Computational Analysis of Capital vs Labour Amplification
por: Bose, Joy
Publicado: (2026)
por: Bose, Joy
Publicado: (2026)
Towards Adaptive Context Management for Intelligent Conversational Question Answering
por: Perera, Manoj Madushanka, et al.
Publicado: (2025)
por: Perera, Manoj Madushanka, et al.
Publicado: (2025)
MLDocRAG: Multimodal Long-Context Document Retrieval Augmented Generation
por: Zhang, Yongyue, et al.
Publicado: (2026)
por: Zhang, Yongyue, et al.
Publicado: (2026)
Survey and Experiments on Mental Disorder Detection via Social Media: From Large Language Models and RAG to Agents
por: Ge, Zhuohan, et al.
Publicado: (2025)
por: Ge, Zhuohan, et al.
Publicado: (2025)
FinGround: Detecting and Grounding Financial Hallucinations via Atomic Claim Verification
por: Guo, Dongxin, et al.
Publicado: (2026)
por: Guo, Dongxin, et al.
Publicado: (2026)
FACTUM: Mechanistic Detection of Citation Hallucination in Long-Form RAG
por: Dassen, Maxime, et al.
Publicado: (2026)
por: Dassen, Maxime, et al.
Publicado: (2026)
A Multimodal Pipeline for Clinical Data Extraction: Applying Vision-Language Models to Scans of Transfusion Reaction Reports
por: Schäfer, Henning, et al.
Publicado: (2025)
por: Schäfer, Henning, et al.
Publicado: (2025)
On Fusing ChatGPT and Ensemble Learning in Discon-tinuous Named Entity Recognition in Health Corpora
por: Chen, Tzu-Chieh, et al.
Publicado: (2024)
por: Chen, Tzu-Chieh, et al.
Publicado: (2024)
Mix and Match: Context Pairing for Scalable Topic-Controlled Educational Summarisation
por: Yodthapa, Nathikan, et al.
Publicado: (2026)
por: Yodthapa, Nathikan, et al.
Publicado: (2026)
A Diagrammatic Calculus for a Functional Model of Natural Language Semantics
por: Boyer, Matthieu Pierre
Publicado: (2025)
por: Boyer, Matthieu Pierre
Publicado: (2025)
Artificial intelligence applications in Parkinson's disease via retinal imaging
por: Jafarizadeh, Ali, et al.
Publicado: (2026)
por: Jafarizadeh, Ali, et al.
Publicado: (2026)
The Foundational Capabilities of Large Language Models in Predicting Postoperative Risks Using Clinical Notes
por: Alba, Charles, et al.
Publicado: (2024)
por: Alba, Charles, et al.
Publicado: (2024)
Multi-Method Validation of Large Language Model Medical Translation Across High- and Low-Resource Languages
por: Anyaegbuna, Chukwuebuka, et al.
Publicado: (2026)
por: Anyaegbuna, Chukwuebuka, et al.
Publicado: (2026)
Can Large Language Models Imitate Human Speech for Clinical Assessment? LLM-Driven Data Augmentation for Cognitive Score Prediction
por: Ketir, Si-Belkacem Yamine, et al.
Publicado: (2026)
por: Ketir, Si-Belkacem Yamine, et al.
Publicado: (2026)
The Provenance Gap in Clinical AI: Evidence-Traceable Temporal Knowledge Graphs for Rare Disease Reasoning
por: Ahmed, Md Shamim, et al.
Publicado: (2026)
por: Ahmed, Md Shamim, et al.
Publicado: (2026)
Integrating clinical reasoning into large language model-based diagnosis through etiology-aware attention steering
por: Li, Peixian, et al.
Publicado: (2025)
por: Li, Peixian, et al.
Publicado: (2025)
Forgotten Words: Benchmarking NeoBERT for Dementia Detection in Low-Resource Conversational Filipino and English Speech
por: Floresca, Rez Samantha Z., et al.
Publicado: (2026)
por: Floresca, Rez Samantha Z., et al.
Publicado: (2026)
Ejemplares similares
-
Motion Compensation for Real Time Ultrasound Scanning in Robotically Assisted Prostate Biopsy Procedures
por: Markulin, Matija, et al.
Publicado: (2026) -
BALI: Enhancing Biomedical Language Representations through Knowledge Graph and Language Model Alignment
por: Sakhovskiy, Andrey, et al.
Publicado: (2025) -
Prompt-RAG: Pioneering Vector Embedding-Free Retrieval-Augmented Generation in Niche Domains, Exemplified by Korean Medicine
por: Kang, Bongsu, et al.
Publicado: (2024) -
ParliaBench: An Evaluation and Benchmarking Framework for LLM-Generated Parliamentary Speech
por: Koniaris, Marios, et al.
Publicado: (2025) -
PubMed Reasoner: Dynamic Reasoning-based Retrieval for Evidence-Grounded Biomedical Question Answering
por: Zhang, Yiqing, et al.
Publicado: (2026)