Notes on Applicability of GPT-4 to Document Understanding
Fuente:
arXiv
Saved in:
| Main Author: | Borchmann, Łukasz |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Language Models Model Language
by: Borchmann, Łukasz
Published: (2025)
by: Borchmann, Łukasz
Published: (2025)
In Case You Missed It: ARC 'Challenge' Is Not That Challenging
by: Borchmann, Łukasz
Published: (2024)
by: Borchmann, Łukasz
Published: (2024)
Query and Conquer: Execution-Guided SQL Generation
by: Borchmann, Łukasz, et al.
Published: (2025)
by: Borchmann, Łukasz, et al.
Published: (2025)
Unchecked and Overlooked: Addressing the Checkbox Blind Spot in Large Language Models with CheckboxQA
by: Turski, Michał, et al.
Published: (2025)
by: Turski, Michał, et al.
Published: (2025)
Tackling prediction tasks in relational databases with LLMs
by: Wydmuch, Marek, et al.
Published: (2024)
by: Wydmuch, Marek, et al.
Published: (2024)
Can Models Help Us Create Better Models? Evaluating LLMs as Data Scientists
by: Pietruszka, Michał, et al.
Published: (2024)
by: Pietruszka, Michał, et al.
Published: (2024)
Arctic-TILT. Business Document Understanding at Sub-Billion Scale
by: Borchmann, Łukasz, et al.
Published: (2024)
by: Borchmann, Łukasz, et al.
Published: (2024)
Dynamic Boundary Time Warping for Sub-sequence Matching with Few Examples
by: Borchmann, Łukasz, et al.
Published: (2020)
by: Borchmann, Łukasz, et al.
Published: (2020)
Text Understanding in GPT-4 vs Humans
by: Shultz, Thomas R., et al.
Published: (2024)
by: Shultz, Thomas R., et al.
Published: (2024)
LLM Probing with Contrastive Eigenproblems: Improving Understanding and Applicability of CCS
by: Schouten, Stefan F., et al.
Published: (2025)
by: Schouten, Stefan F., et al.
Published: (2025)
Understanding the Role of Temperature in Diverse Question Generation by GPT-4
by: Agarwal, Arav, et al.
Published: (2024)
by: Agarwal, Arav, et al.
Published: (2024)
Beyond Scale: Small Language Models are Comparable to GPT-4 in Mental Health Understanding
by: Jia, Hong, et al.
Published: (2025)
by: Jia, Hong, et al.
Published: (2025)
Exploiting GPT-4 Vision for Zero-shot Point Cloud Understanding
by: Sun, Qi, et al.
Published: (2024)
by: Sun, Qi, et al.
Published: (2024)
SAGA: A Participant-specific Examination of Story Alternatives and Goal Applicability for a Deeper Understanding of Complex Events
by: Vallurupalli, Sai, et al.
Published: (2024)
by: Vallurupalli, Sai, et al.
Published: (2024)
NLP-based Regulatory Compliance -- Using GPT 4.0 to Decode Regulatory Documents
by: Kumar, Bimal, et al.
Published: (2024)
by: Kumar, Bimal, et al.
Published: (2024)
From Generative Modeling to Clinical Classification: A GPT-Based Architecture for EHR Notes
by: Irany, Fariba Afrin, et al.
Published: (2026)
by: Irany, Fariba Afrin, et al.
Published: (2026)
Strategic Navigation or Stochastic Search? How Agents and Humans Reason Over Document Collections
by: Borchmann, Łukasz, et al.
Published: (2026)
by: Borchmann, Łukasz, et al.
Published: (2026)
Empathy Applicability Modeling for General Health Queries
by: Randhawa, Shan, et al.
Published: (2026)
by: Randhawa, Shan, et al.
Published: (2026)
Arctic-Text2SQL-R1: Simple Rewards, Strong Reasoning in Text-to-SQL
by: Yao, Zhewei, et al.
Published: (2025)
by: Yao, Zhewei, et al.
Published: (2025)
AgriGPT-VL: Agricultural Vision-Language Understanding Suite
by: Yang, Bo, et al.
Published: (2025)
by: Yang, Bo, et al.
Published: (2025)
100% Elimination of Hallucinations on RAGTruth for GPT-4 and GPT-3.5 Turbo
by: Wood, Michael C., et al.
Published: (2024)
by: Wood, Michael C., et al.
Published: (2024)
Enhancing Clinical Note Generation with ICD-10, Clinical Ontology Knowledge Graphs, and Chain-of-Thought Prompting Using GPT-4
by: Makohon, Ivan, et al.
Published: (2025)
by: Makohon, Ivan, et al.
Published: (2025)
Is GPT-4 Less Politically Biased than GPT-3.5? A Renewed Investigation of ChatGPT's Political Biases
by: Weber, Erik, et al.
Published: (2024)
by: Weber, Erik, et al.
Published: (2024)
DeID-GPT: Zero-shot Medical Text De-Identification by GPT-4
by: Liu, Zhengliang, et al.
Published: (2023)
by: Liu, Zhengliang, et al.
Published: (2023)
GPT-4 Technical Report
by: OpenAI, et al.
Published: (2023)
by: OpenAI, et al.
Published: (2023)
GPTEval: A Survey on Assessments of ChatGPT and GPT-4
by: Mao, Rui, et al.
Published: (2023)
by: Mao, Rui, et al.
Published: (2023)
Is GPT-4 a reliable rater? Evaluating Consistency in GPT-4 Text Ratings
by: Hackl, Veronika, et al.
Published: (2023)
by: Hackl, Veronika, et al.
Published: (2023)
SciGPT: A Large Language Model for Scientific Literature Understanding and Knowledge Discovery
by: She, Fengyu, et al.
Published: (2025)
by: She, Fengyu, et al.
Published: (2025)
Can ChatGPT Really Understand Modern Chinese Poetry?
by: Wang, Shanshan, et al.
Published: (2026)
by: Wang, Shanshan, et al.
Published: (2026)
Whose LLM is it Anyway? Linguistic Comparison and LLM Attribution for GPT-3.5, GPT-4 and Bard
by: Rosenfeld, Ariel, et al.
Published: (2024)
by: Rosenfeld, Ariel, et al.
Published: (2024)
GPT-Fathom: Benchmarking Large Language Models to Decipher the Evolutionary Path towards GPT-4 and Beyond
by: Zheng, Shen, et al.
Published: (2023)
by: Zheng, Shen, et al.
Published: (2023)
Can GPT Redefine Medical Understanding? Evaluating GPT on Biomedical Machine Reading Comprehension
by: Vatsal, Shubham, et al.
Published: (2024)
by: Vatsal, Shubham, et al.
Published: (2024)
ICLGuard: Controlling In-Context Learning Behavior for Applicability Authorization
by: Si, Wai Man, et al.
Published: (2024)
by: Si, Wai Man, et al.
Published: (2024)
Measuring and Modifying the Readability of English Texts with GPT-4
by: Trott, Sean, et al.
Published: (2024)
by: Trott, Sean, et al.
Published: (2024)
A comparison of Human, GPT-3.5, and GPT-4 Performance in a University-Level Coding Course
by: Yeadon, Will, et al.
Published: (2024)
by: Yeadon, Will, et al.
Published: (2024)
Language Writ Large: LLMs, ChatGPT, Grounding, Meaning and Understanding
by: Harnad, Stevan
Published: (2024)
by: Harnad, Stevan
Published: (2024)
The Qiyas Benchmark: Measuring ChatGPT Mathematical and Language Understanding in Arabic
by: Al-Khalifa, Shahad, et al.
Published: (2024)
by: Al-Khalifa, Shahad, et al.
Published: (2024)
ESG Classification by Implicit Rule Learning via GPT-4
by: Yun, Hyo Jeong, et al.
Published: (2024)
by: Yun, Hyo Jeong, et al.
Published: (2024)
Can GPT-4 do L2 analytic assessment?
by: Bannò, Stefano, et al.
Published: (2024)
by: Bannò, Stefano, et al.
Published: (2024)
Probability of Differentiation Reveals Brittleness of Homogeneity Bias in GPT-4
by: Lee, Messi H. J., et al.
Published: (2024)
by: Lee, Messi H. J., et al.
Published: (2024)
Similar Items
-
Language Models Model Language
by: Borchmann, Łukasz
Published: (2025) -
In Case You Missed It: ARC 'Challenge' Is Not That Challenging
by: Borchmann, Łukasz
Published: (2024) -
Query and Conquer: Execution-Guided SQL Generation
by: Borchmann, Łukasz, et al.
Published: (2025) -
Unchecked and Overlooked: Addressing the Checkbox Blind Spot in Large Language Models with CheckboxQA
by: Turski, Michał, et al.
Published: (2025) -
Tackling prediction tasks in relational databases with LLMs
by: Wydmuch, Marek, et al.
Published: (2024)