V-DyKnow: A Dynamic Benchmark for Time-Sensitive Knowledge in Vision Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Mousavi, Seyed Mahed, Moiola, Christian, Rizzoli, Massimo, Alghisi, Simone, Riccardi, Giuseppe |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
DyKnow: Dynamically Verifying Time-Sensitive Factual Knowledge in LLMs
par: Mousavi, Seyed Mahed, et autres
Publié: (2024)
par: Mousavi, Seyed Mahed, et autres
Publié: (2024)
Should We Fine-Tune or RAG? Evaluating Different Techniques to Adapt LLMs for Dialogue
par: Alghisi, Simone, et autres
Publié: (2024)
par: Alghisi, Simone, et autres
Publié: (2024)
Getting to the Point: Pointing Improves LVLMs at Counting
par: Alghisi, Simone, et autres
Publié: (2026)
par: Alghisi, Simone, et autres
Publié: (2026)
What Does Loss Optimization Actually Teach, If Anything? Knowledge Dynamics in Continual Pre-training of LLMs
par: Mousavi, Seyed Mahed, et autres
Publié: (2026)
par: Mousavi, Seyed Mahed, et autres
Publié: (2026)
LLMs as Repositories of Factual Knowledge: Limitations and Solutions
par: Mousavi, Seyed Mahed, et autres
Publié: (2025)
par: Mousavi, Seyed Mahed, et autres
Publié: (2025)
[De|Re]constructing VLMs' Reasoning in Counting
par: Alghisi, Simone, et autres
Publié: (2025)
par: Alghisi, Simone, et autres
Publié: (2025)
CIVET: Systematic Evaluation of Understanding in VLMs
par: Rizzoli, Massimo, et autres
Publié: (2025)
par: Rizzoli, Massimo, et autres
Publié: (2025)
Are LLMs Robust for Spoken Dialogues?
par: Mousavi, Seyed Mahed, et autres
Publié: (2024)
par: Mousavi, Seyed Mahed, et autres
Publié: (2024)
Will LLMs Replace the Encoder-Only Models in Temporal Relation Classification?
par: Roccabruna, Gabriel, et autres
Publié: (2024)
par: Roccabruna, Gabriel, et autres
Publié: (2024)
Garbage In, Reasoning Out? Why Benchmark Scores are Unreliable and What to Do About It
par: Mousavi, Seyed Mahed, et autres
Publié: (2025)
par: Mousavi, Seyed Mahed, et autres
Publié: (2025)
When Seeing Overrides Knowing: Disentangling Knowledge Conflicts in Vision-Language Models
par: Ortu, Francesco, et autres
Publié: (2025)
par: Ortu, Francesco, et autres
Publié: (2025)
MATEO: A Multimodal Benchmark for Temporal Reasoning and Planning in LVLMs
par: Roccabruna, Gabriel, et autres
Publié: (2026)
par: Roccabruna, Gabriel, et autres
Publié: (2026)
KnowCoder-V2: Deep Knowledge Analysis
par: Li, Zixuan, et autres
Publié: (2025)
par: Li, Zixuan, et autres
Publié: (2025)
Time Sensitive Knowledge Editing through Efficient Finetuning
par: Ge, Xiou, et autres
Publié: (2024)
par: Ge, Xiou, et autres
Publié: (2024)
MVRS: The Multimodal Virtual Reality Stimuli-based Emotion Recognition Dataset
par: Mousavi, Seyed Muhammad Hossein, et autres
Publié: (2025)
par: Mousavi, Seyed Muhammad Hossein, et autres
Publié: (2025)
Enhancing Answer Reliability Through Inter-Model Consensus of Large Language Models
par: Amiri-Margavi, Alireza, et autres
Publié: (2024)
par: Amiri-Margavi, Alireza, et autres
Publié: (2024)
Know More, Know Clearer: A Meta-Cognitive Framework for Knowledge Augmentation in Large Language Models
par: Chen, Hao, et autres
Publié: (2026)
par: Chen, Hao, et autres
Publié: (2026)
KnowRL: Teaching Language Models to Know What They Know
par: Kale, Sahil, et autres
Publié: (2025)
par: Kale, Sahil, et autres
Publié: (2025)
KnowGPT: Knowledge Graph based Prompting for Large Language Models
par: Zhang, Qinggang, et autres
Publié: (2023)
par: Zhang, Qinggang, et autres
Publié: (2023)
KnowTuning: Knowledge-aware Fine-tuning for Large Language Models
par: Lyu, Yougang, et autres
Publié: (2024)
par: Lyu, Yougang, et autres
Publié: (2024)
DyABD: The Abdominal Muscle Segmentation in Dynamic MRI Benchmark
par: Belton, Niamh, et autres
Publié: (2026)
par: Belton, Niamh, et autres
Publié: (2026)
Do Large Language Models Know What They Don't Know? Kalshibench: A New Benchmark for Evaluating Epistemic Calibration via Prediction Markets
par: Nel, Lukas
Publié: (2025)
par: Nel, Lukas
Publié: (2025)
LLMs Are Not Intelligent Thinkers: Introducing Mathematical Topic Tree Benchmark for Comprehensive Evaluation of LLMs
par: Davoodi, Arash Gholami, et autres
Publié: (2024)
par: Davoodi, Arash Gholami, et autres
Publié: (2024)
DyACE: Dynamic Algorithm Co-evolution for Online Automated Heuristic Design with Large Language Model
par: Lu, Guidong, et autres
Publié: (2026)
par: Lu, Guidong, et autres
Publié: (2026)
DyVal: Dynamic Evaluation of Large Language Models for Reasoning Tasks
par: Zhu, Kaijie, et autres
Publié: (2023)
par: Zhu, Kaijie, et autres
Publié: (2023)
Vision Language Models Know Law of Conservation without Understanding More-or-Less
par: Luo, Dezhi, et autres
Publié: (2024)
par: Luo, Dezhi, et autres
Publié: (2024)
Latent Anomaly Knowledge Excavation: Unveiling Sparse Sensitive Neurons in Vision-Language Models
par: Li, Shaotian, et autres
Publié: (2026)
par: Li, Shaotian, et autres
Publié: (2026)
KnowPO: Knowledge-aware Preference Optimization for Controllable Knowledge Selection in Retrieval-Augmented Language Models
par: Zhang, Ruizhe, et autres
Publié: (2024)
par: Zhang, Ruizhe, et autres
Publié: (2024)
DyMRL: Dynamic Multispace Representation Learning for Multimodal Event Forecasting in Knowledge Graph
par: Zhao, Feng, et autres
Publié: (2026)
par: Zhao, Feng, et autres
Publié: (2026)
Do I Know This Entity? Knowledge Awareness and Hallucinations in Language Models
par: Ferrando, Javier, et autres
Publié: (2024)
par: Ferrando, Javier, et autres
Publié: (2024)
Can Large Language Models Derive New Knowledge? A Dynamic Benchmark for Biological Knowledge Discovery
par: Yang, Chaoqun, et autres
Publié: (2026)
par: Yang, Chaoqun, et autres
Publié: (2026)
What Models Know, How Well They Know It: Knowledge-Weighted Fine-Tuning for Learning When to Say "I Don't Know"
par: Lee, Joosung, et autres
Publié: (2026)
par: Lee, Joosung, et autres
Publié: (2026)
Do Retrieval Augmented Language Models Know When They Don't Know?
par: Zhou, Youchao, et autres
Publié: (2025)
par: Zhou, Youchao, et autres
Publié: (2025)
VLKEB: A Large Vision-Language Model Knowledge Editing Benchmark
par: Huang, Han, et autres
Publié: (2024)
par: Huang, Han, et autres
Publié: (2024)
Benchmarking Prompt Sensitivity in Large Language Models
par: Razavi, Amirhossein, et autres
Publié: (2025)
par: Razavi, Amirhossein, et autres
Publié: (2025)
DyGMamba: Efficiently Modeling Long-Term Temporal Dependency on Continuous-Time Dynamic Graphs with State Space Models
par: Ding, Zifeng, et autres
Publié: (2024)
par: Ding, Zifeng, et autres
Publié: (2024)
AutoBench-V: Can Large Vision-Language Models Benchmark Themselves?
par: Bao, Han, et autres
Publié: (2024)
par: Bao, Han, et autres
Publié: (2024)
KnowMol: Advancing Molecular Large Language Models with Multi-Level Chemical Knowledge
par: Yang, Zaifei, et autres
Publié: (2025)
par: Yang, Zaifei, et autres
Publié: (2025)
Benchmarking the Pedagogical Knowledge of Large Language Models
par: Lelièvre, Maxime, et autres
Publié: (2025)
par: Lelièvre, Maxime, et autres
Publié: (2025)
DySK-Attn: A Framework for Efficient, Real-Time Knowledge Updating in Large Language Models via Dynamic Sparse Knowledge Attention
par: Khan, Kabir, et autres
Publié: (2025)
par: Khan, Kabir, et autres
Publié: (2025)
Documents similaires
-
DyKnow: Dynamically Verifying Time-Sensitive Factual Knowledge in LLMs
par: Mousavi, Seyed Mahed, et autres
Publié: (2024) -
Should We Fine-Tune or RAG? Evaluating Different Techniques to Adapt LLMs for Dialogue
par: Alghisi, Simone, et autres
Publié: (2024) -
Getting to the Point: Pointing Improves LVLMs at Counting
par: Alghisi, Simone, et autres
Publié: (2026) -
What Does Loss Optimization Actually Teach, If Anything? Knowledge Dynamics in Continual Pre-training of LLMs
par: Mousavi, Seyed Mahed, et autres
Publié: (2026) -
LLMs as Repositories of Factual Knowledge: Limitations and Solutions
par: Mousavi, Seyed Mahed, et autres
Publié: (2025)