FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hildebrand, Samuel, Taylor, Curtis, Oesch, Sean, Ghawaly Jr, James M, Sadovnik, Amir, Shivers, Ryan, Schreiber, Brandon, Kurian, Kevin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Complementarity, Augmentation, or Substitutivity? The Impact of Generative Artificial Intelligence on the U.S. Federal Workforce
von: Resh, William G., et al.
Veröffentlicht: (2025)
von: Resh, William G., et al.
Veröffentlicht: (2025)
TrafficRAG: A Multimodal RAG Framework for Traffic Accident Liability Determination
von: Li, Xu, et al.
Veröffentlicht: (2026)
von: Li, Xu, et al.
Veröffentlicht: (2026)
KinyaColBERT: A Lexically Grounded Retrieval Model for Low-Resource Retrieval-Augmented Generation
von: Nzeyimana, Antoine, et al.
Veröffentlicht: (2025)
von: Nzeyimana, Antoine, et al.
Veröffentlicht: (2025)
Graphemic Normalization of the Perso-Arabic Script
von: Doctor, Raiomond, et al.
Veröffentlicht: (2022)
von: Doctor, Raiomond, et al.
Veröffentlicht: (2022)
Beyond Arabic: Software for Perso-Arabic Script Manipulation
von: Gutkin, Alexander, et al.
Veröffentlicht: (2023)
von: Gutkin, Alexander, et al.
Veröffentlicht: (2023)
MIRAGE: Scaling Test-Time Inference with Parallel Graph-Retrieval-Augmented Reasoning Chains
von: Wei, Kaiwen, et al.
Veröffentlicht: (2025)
von: Wei, Kaiwen, et al.
Veröffentlicht: (2025)
RAPTOR-AI for Disaster OODA Loop: Hierarchical Multimodal RAG with Experience-Driven Agentic Decision-Making
von: Yasuno, Takato
Veröffentlicht: (2026)
von: Yasuno, Takato
Veröffentlicht: (2026)
Grokking in the Wild: Data Augmentation for Real-World Multi-Hop Reasoning with Transformers
von: Abramov, Roman, et al.
Veröffentlicht: (2025)
von: Abramov, Roman, et al.
Veröffentlicht: (2025)
Train to Defend: First Defense Against Cryptanalytic Neural Network Parameter Extraction Attacks
von: Kurian, Ashley, et al.
Veröffentlicht: (2025)
von: Kurian, Ashley, et al.
Veröffentlicht: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
Open-TI: Open Traffic Intelligence with Augmented Language Model
von: Da, Longchao, et al.
Veröffentlicht: (2023)
von: Da, Longchao, et al.
Veröffentlicht: (2023)
Learned Relay Representations for Forward-Thinking Discrete Diffusion Models
von: Rozonoyer, Benjamin, et al.
Veröffentlicht: (2026)
von: Rozonoyer, Benjamin, et al.
Veröffentlicht: (2026)
HiPS: Hierarchical PDF Segmentation of Textbooks
von: Wehnert, Sabine, et al.
Veröffentlicht: (2025)
von: Wehnert, Sabine, et al.
Veröffentlicht: (2025)
Robustness of Large Language Models to Perturbations in Text
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
Survey Transfer Learning: Recycling Data with Silicon Responses
von: Amini, Ali
Veröffentlicht: (2025)
von: Amini, Ali
Veröffentlicht: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025)
von: Fadli, Samih
Veröffentlicht: (2025)
CBR -- Boosting Adaptive Classification By Retrieval of Encrypted Network Traffic with Out-of-distribution
von: Lukach, Amir, et al.
Veröffentlicht: (2024)
von: Lukach, Amir, et al.
Veröffentlicht: (2024)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
von: Smădu, Răzvan-Alexandru, et al.
Veröffentlicht: (2025)
von: Smădu, Răzvan-Alexandru, et al.
Veröffentlicht: (2025)
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
von: Li, Danyang, et al.
Veröffentlicht: (2025)
von: Li, Danyang, et al.
Veröffentlicht: (2025)
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
von: Lim, Shoon Kit, et al.
Veröffentlicht: (2025)
von: Lim, Shoon Kit, et al.
Veröffentlicht: (2025)
StratXplore: Strategic Novelty-seeking and Instruction-aligned Exploration for Vision and Language Navigation
von: Gopinathan, Muraleekrishna, et al.
Veröffentlicht: (2024)
von: Gopinathan, Muraleekrishna, et al.
Veröffentlicht: (2024)
HySemRAG: A Hybrid Semantic Retrieval-Augmented Generation Framework for Automated Literature Synthesis and Methodological Gap Analysis
von: Godinez, Alejandro
Veröffentlicht: (2025)
von: Godinez, Alejandro
Veröffentlicht: (2025)
Gyan: An Explainable Neuro-Symbolic Language Model
von: Srinivasan, Venkat, et al.
Veröffentlicht: (2026)
von: Srinivasan, Venkat, et al.
Veröffentlicht: (2026)
From Search to Reasoning: A Five-Level RAG Capability Framework for Enterprise Data
von: Gill, Gurbinder, et al.
Veröffentlicht: (2025)
von: Gill, Gurbinder, et al.
Veröffentlicht: (2025)
Universal Adversarial Attack on Aligned Multimodal LLMs
von: Rahmatullaev, Temurbek, et al.
Veröffentlicht: (2025)
von: Rahmatullaev, Temurbek, et al.
Veröffentlicht: (2025)
Pareto-Optimized Open-Source LLMs for Healthcare via Context Retrieval
von: Bayarri-Planas, Jordi, et al.
Veröffentlicht: (2024)
von: Bayarri-Planas, Jordi, et al.
Veröffentlicht: (2024)
A Semantic Approach to Negation Detection and Word Disambiguation with Natural Language Processing
von: Okpala, Izunna, et al.
Veröffentlicht: (2023)
von: Okpala, Izunna, et al.
Veröffentlicht: (2023)
Can AI Assist in Olympiad Coding
von: Ren, Samuel
Veröffentlicht: (2025)
von: Ren, Samuel
Veröffentlicht: (2025)
SURE-RAG: Sufficiency and Uncertainty-Aware Evidence Verification for Selective Retrieval-Augmented Generation
von: Qiu, Jingxi, et al.
Veröffentlicht: (2026)
von: Qiu, Jingxi, et al.
Veröffentlicht: (2026)
Toward Architecture-Aware Evaluation Metrics for LLM Agents
von: Souza, Débora, et al.
Veröffentlicht: (2026)
von: Souza, Débora, et al.
Veröffentlicht: (2026)
MEMTIER: Tiered Memory Architecture and Retrieval Bottleneck Analysis for Long-Running Autonomous AI Agents
von: Sidik, Bronislav, et al.
Veröffentlicht: (2026)
von: Sidik, Bronislav, et al.
Veröffentlicht: (2026)
Thinking Longer, Not Always Smarter: Evaluating LLM Capabilities in Hierarchical Legal Reasoning
von: Zhang, Li, et al.
Veröffentlicht: (2025)
von: Zhang, Li, et al.
Veröffentlicht: (2025)
A Graph-based RAG for Energy Efficiency Question Answering
von: Campi, Riccardo, et al.
Veröffentlicht: (2025)
von: Campi, Riccardo, et al.
Veröffentlicht: (2025)
Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning
von: Tong, Jingqi, et al.
Veröffentlicht: (2025)
von: Tong, Jingqi, et al.
Veröffentlicht: (2025)
Correspondence of high-dimensional emotion structures elicited by video clips between humans and Multimodal LLMs
von: Asanuma, Haruka, et al.
Veröffentlicht: (2025)
von: Asanuma, Haruka, et al.
Veröffentlicht: (2025)
PhysicsArena: The First Multimodal Physics Reasoning Benchmark Exploring Variable, Process, and Solution Dimensions
von: Dai, Song, et al.
Veröffentlicht: (2025)
von: Dai, Song, et al.
Veröffentlicht: (2025)
An Explainable Collaborative Dialogue System using a Theory of Mind
von: Cohen, Philip R., et al.
Veröffentlicht: (2023)
von: Cohen, Philip R., et al.
Veröffentlicht: (2023)
MCP: A Control-Theoretic Orchestration Framework for Synergistic Efficiency and Interpretability in Multimodal Large Language Models
von: Zhang, Luyan
Veröffentlicht: (2025)
von: Zhang, Luyan
Veröffentlicht: (2025)
GraphWalk: Enabling Reasoning in Large Language Models through Tool-Based Graph Navigation
von: Ghandi, Taraneh, et al.
Veröffentlicht: (2026)
von: Ghandi, Taraneh, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Complementarity, Augmentation, or Substitutivity? The Impact of Generative Artificial Intelligence on the U.S. Federal Workforce
von: Resh, William G., et al.
Veröffentlicht: (2025) -
TrafficRAG: A Multimodal RAG Framework for Traffic Accident Liability Determination
von: Li, Xu, et al.
Veröffentlicht: (2026) -
KinyaColBERT: A Lexically Grounded Retrieval Model for Low-Resource Retrieval-Augmented Generation
von: Nzeyimana, Antoine, et al.
Veröffentlicht: (2025) -
Graphemic Normalization of the Perso-Arabic Script
von: Doctor, Raiomond, et al.
Veröffentlicht: (2022) -
Beyond Arabic: Software for Perso-Arabic Script Manipulation
von: Gutkin, Alexander, et al.
Veröffentlicht: (2023)