A Hybrid Architecture for Multi-Stage Claim Document Understanding: Combining Vision-Language Models and Machine Learning for Real-Time Processing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cheng, Lilu, Lu, Jingjun, Chan, Yi Xuan, Nguyen, Quoc Khai, Bi, John, Ho, Sean |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Integrating Vision-Centric Text Understanding for Conversational Recommender Systems
von: Yuan, Wei, et al.
Veröffentlicht: (2026)
von: Yuan, Wei, et al.
Veröffentlicht: (2026)
eSapiens: A Real-World NLP Framework for Multimodal Document Understanding and Enterprise Knowledge Processing
von: Shi, Isaac, et al.
Veröffentlicht: (2025)
von: Shi, Isaac, et al.
Veröffentlicht: (2025)
Markup Language Modeling for Web Document Understanding
von: Liu, Su, et al.
Veröffentlicht: (2025)
von: Liu, Su, et al.
Veröffentlicht: (2025)
Multi-Stage Field Extraction of Financial Documents with OCR and Compact Vision-Language Models
von: Jin, Yichao, et al.
Veröffentlicht: (2025)
von: Jin, Yichao, et al.
Veröffentlicht: (2025)
Tailoring Table Retrieval from a Field-aware Hybrid Matching Perspective
von: Li, Da, et al.
Veröffentlicht: (2025)
von: Li, Da, et al.
Veröffentlicht: (2025)
Real-Time Procedural Learning From Experience for AI Agents
von: Bi, Dasheng, et al.
Veröffentlicht: (2025)
von: Bi, Dasheng, et al.
Veröffentlicht: (2025)
Watermarking Large Language Model-based Time Series Forecasting
von: Yuan, Wei, et al.
Veröffentlicht: (2025)
von: Yuan, Wei, et al.
Veröffentlicht: (2025)
RecGPT: Generative Pre-training for Text-based Recommendation
von: Ngo, Hoang, et al.
Veröffentlicht: (2024)
von: Ngo, Hoang, et al.
Veröffentlicht: (2024)
Agentic Mixed-Source Multi-Modal Misinformation Detection with Adaptive Test-Time Scaling
von: Jiang, Wei, et al.
Veröffentlicht: (2026)
von: Jiang, Wei, et al.
Veröffentlicht: (2026)
On Precomputation and Caching in Information Retrieval Experiments with Pipeline Architectures
von: MacAvaney, Sean, et al.
Veröffentlicht: (2025)
von: MacAvaney, Sean, et al.
Veröffentlicht: (2025)
When Text-as-Vision Meets Semantic IDs in Generative Recommendation: An Empirical Study
von: Qiao, Shutong, et al.
Veröffentlicht: (2026)
von: Qiao, Shutong, et al.
Veröffentlicht: (2026)
SERVAL: Surprisingly Effective Zero-Shot Visual Document Retrieval Powered by Large Vision and Language Models
von: Nguyen, Thong, et al.
Veröffentlicht: (2025)
von: Nguyen, Thong, et al.
Veröffentlicht: (2025)
A Hybrid Machine Learning Approach for Graduate Admission Prediction and Combined University-Program Recommendation
von: Far, Melina Heidari, et al.
Veröffentlicht: (2026)
von: Far, Melina Heidari, et al.
Veröffentlicht: (2026)
New Method for Keyword Extraction for Patent Claims
von: Rossi, Julien
Veröffentlicht: (2024)
von: Rossi, Julien
Veröffentlicht: (2024)
Budgeted Embedding Table For Recommender Systems
von: Qu, Yunke, et al.
Veröffentlicht: (2023)
von: Qu, Yunke, et al.
Veröffentlicht: (2023)
DMAP: Human-Aligned Structural Document Map for Multimodal Document Understanding
von: Fu, ShunLiang, et al.
Veröffentlicht: (2026)
von: Fu, ShunLiang, et al.
Veröffentlicht: (2026)
An LLM-Powered Agent for Real-Time Analysis of the Vietnamese IT Job Market
von: Nguyen, Minh-Thuan, et al.
Veröffentlicht: (2025)
von: Nguyen, Minh-Thuan, et al.
Veröffentlicht: (2025)
WildClaims: Information Access Conversations in the Wild(Chat)
von: Joko, Hideaki, et al.
Veröffentlicht: (2025)
von: Joko, Hideaki, et al.
Veröffentlicht: (2025)
Counterfactual Understanding via Retrieval-aware Multimodal Modeling for Time-to-Event Survival Prediction
von: Nguyen, Ha-Anh Hoang, et al.
Veröffentlicht: (2026)
von: Nguyen, Ha-Anh Hoang, et al.
Veröffentlicht: (2026)
QUST_NLP at SemEval-2025 Task 7: A Three-Stage Retrieval Framework for Monolingual and Crosslingual Fact-Checked Claim Retrieval
von: Liu, Youzheng, et al.
Veröffentlicht: (2025)
von: Liu, Youzheng, et al.
Veröffentlicht: (2025)
Leveraging Decoder Architectures for Learned Sparse Retrieval
von: Qiao, Jingfen, et al.
Veröffentlicht: (2025)
von: Qiao, Jingfen, et al.
Veröffentlicht: (2025)
Prompt-enhanced Federated Content Representation Learning for Cross-domain Recommendation
von: Guo, Lei, et al.
Veröffentlicht: (2024)
von: Guo, Lei, et al.
Veröffentlicht: (2024)
Efficient Multimodal Streaming Recommendation via Expandable Side Mixture-of-Experts
von: Qu, Yunke, et al.
Veröffentlicht: (2025)
von: Qu, Yunke, et al.
Veröffentlicht: (2025)
Towards a Theoretical Understanding of Two-Stage Recommender Systems
von: Jaiswal, Amit Kumar
Veröffentlicht: (2024)
von: Jaiswal, Amit Kumar
Veröffentlicht: (2024)
Diversity-Augmented Negative Sampling for Implicit Collaborative Filtering
von: Xuan, Yueqing, et al.
Veröffentlicht: (2025)
von: Xuan, Yueqing, et al.
Veröffentlicht: (2025)
Evaluating and Addressing Fairness Across User Groups in Negative Sampling for Recommender Systems
von: Xuan, Yueqing, et al.
Veröffentlicht: (2023)
von: Xuan, Yueqing, et al.
Veröffentlicht: (2023)
Fast and Faithful: Real-Time Verification for Long-Document Retrieval-Augmented Generation Systems
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
Supporting Cross-language Cross-project Bug Localization Using Pre-trained Language Models
von: Chandramohan, Mahinthan, et al.
Veröffentlicht: (2024)
von: Chandramohan, Mahinthan, et al.
Veröffentlicht: (2024)
LexBoost: Improving Lexical Document Retrieval with Nearest Neighbors
von: Kulkarni, Hrishikesh, et al.
Veröffentlicht: (2024)
von: Kulkarni, Hrishikesh, et al.
Veröffentlicht: (2024)
Modality Alignment with Multi-scale Bilateral Attention for Multimodal Recommendation
von: Ren, Kelin, et al.
Veröffentlicht: (2025)
von: Ren, Kelin, et al.
Veröffentlicht: (2025)
Tackling Data Heterogeneity in Federated Time Series Forecasting
von: Yuan, Wei, et al.
Veröffentlicht: (2024)
von: Yuan, Wei, et al.
Veröffentlicht: (2024)
Combining social relations and interaction data in Recommender System with Graph Convolution Collaborative Filtering
von: Tran, Tin T., et al.
Veröffentlicht: (2025)
von: Tran, Tin T., et al.
Veröffentlicht: (2025)
HyTRec: A Hybrid Temporal-Aware Attention Architecture for Long Behavior Sequential Recommendation
von: Xin, Lei, et al.
Veröffentlicht: (2026)
von: Xin, Lei, et al.
Veröffentlicht: (2026)
Hybrid-Vector Retrieval for Visually Rich Documents: Combining Single-Vector Efficiency and Multi-Vector Accuracy
von: Kim, Juyeon, et al.
Veröffentlicht: (2025)
von: Kim, Juyeon, et al.
Veröffentlicht: (2025)
Scalable and Effective Negative Sample Generation for Hyperedge Prediction
von: Qu, Shilin, et al.
Veröffentlicht: (2024)
von: Qu, Shilin, et al.
Veröffentlicht: (2024)
Dynamic Network-Based Two-Stage Time Series Forecasting for Affiliate Marketing
von: Wang, Zhe, et al.
Veröffentlicht: (2025)
von: Wang, Zhe, et al.
Veröffentlicht: (2025)
Vision-Guided Chunking Is All You Need: Enhancing RAG with Multimodal Document Understanding
von: Tripathi, Vishesh, et al.
Veröffentlicht: (2025)
von: Tripathi, Vishesh, et al.
Veröffentlicht: (2025)
Automated Justification Production for Claim Veracity in Fact Checking: A Survey on Architectures and Approaches
von: Eldifrawi, Islam, et al.
Veröffentlicht: (2024)
von: Eldifrawi, Islam, et al.
Veröffentlicht: (2024)
ClaimCompare: A Data Pipeline for Evaluation of Novelty Destroying Patent Pairs
von: Parikh, Arav, et al.
Veröffentlicht: (2024)
von: Parikh, Arav, et al.
Veröffentlicht: (2024)
Peerispect: Claim Verification in Scientific Peer Reviews
von: Ghorbanpour, Ali, et al.
Veröffentlicht: (2026)
von: Ghorbanpour, Ali, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Integrating Vision-Centric Text Understanding for Conversational Recommender Systems
von: Yuan, Wei, et al.
Veröffentlicht: (2026) -
eSapiens: A Real-World NLP Framework for Multimodal Document Understanding and Enterprise Knowledge Processing
von: Shi, Isaac, et al.
Veröffentlicht: (2025) -
Markup Language Modeling for Web Document Understanding
von: Liu, Su, et al.
Veröffentlicht: (2025) -
Multi-Stage Field Extraction of Financial Documents with OCR and Compact Vision-Language Models
von: Jin, Yichao, et al.
Veröffentlicht: (2025) -
Tailoring Table Retrieval from a Field-aware Hybrid Matching Perspective
von: Li, Da, et al.
Veröffentlicht: (2025)