Salvato in:
| Autori principali: | Zhang, Sen, Li, Runmei, Deng, Shizhuang, Zheng, Zhichao, Zhang, Yuhe, Li, Jiani, Zhang, Kailun, Zhang, Tao, Wu, Wenjun, Wang, Qunbo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2603.27112 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Reflecting with Two Voices: A Co-Adaptive Dual-Strategy Framework for LLM-Based Agent Decision Making
di: Zhang, Wentao, et al.
Pubblicazione: (2025)
di: Zhang, Wentao, et al.
Pubblicazione: (2025)
CogRail: Benchmarking VLMs in Cognitive Intrusion Perception for Intelligent Railway Transportation Systems
di: Tian, Yonglin, et al.
Pubblicazione: (2026)
di: Tian, Yonglin, et al.
Pubblicazione: (2026)
Hot Deformation Behavior and Processing Map of Eutectoid Pearlite Rail Steel
di: Haibo Feng, et al.
Pubblicazione: (2025)
di: Haibo Feng, et al.
Pubblicazione: (2025)
Knowledge Condensation and Reasoning for Knowledge-based VQA
di: Hao, Dongze, et al.
Pubblicazione: (2024)
di: Hao, Dongze, et al.
Pubblicazione: (2024)
An LLM-based Framework for Human-Swarm Teaming Cognition in Disaster Search and Rescue
di: Ji, Kailun, et al.
Pubblicazione: (2025)
di: Ji, Kailun, et al.
Pubblicazione: (2025)
Marginal Debiased Network for Fair Visual Recognition
di: Wang, Mei, et al.
Pubblicazione: (2024)
di: Wang, Mei, et al.
Pubblicazione: (2024)
Prompt-tuning for Clickbait Detection via Text Summarization
di: Deng, Haoxiang, et al.
Pubblicazione: (2024)
di: Deng, Haoxiang, et al.
Pubblicazione: (2024)
QG-VTC: Question-Guided Visual Token Compression in MLLMs for Efficient VQA
di: Li, Shuai, et al.
Pubblicazione: (2025)
di: Li, Shuai, et al.
Pubblicazione: (2025)
DPC-VQA: Decoupling Quality Perception and Residual Calibration for Video Quality Assessment
di: Li, Xinyue, et al.
Pubblicazione: (2026)
di: Li, Xinyue, et al.
Pubblicazione: (2026)
Multi-Scale Implicit Transformer with Re-parameterize for Arbitrary-Scale Super-Resolution
di: Zhu, Jinchen, et al.
Pubblicazione: (2024)
di: Zhu, Jinchen, et al.
Pubblicazione: (2024)
RoboSurg-VQA: A Multimodal Benchmark for Surgical Segmentation-Aware Visual Question Answering
di: Zhang, Chengyi, et al.
Pubblicazione: (2026)
di: Zhang, Chengyi, et al.
Pubblicazione: (2026)
Semantic and Visual Evidence for Efficient Long-Video Reasoning: A Solution for the HD-EPIC VQA Challenge
di: Xu, Yinsong, et al.
Pubblicazione: (2026)
di: Xu, Yinsong, et al.
Pubblicazione: (2026)
MaS-VQA: A Mask-and-Select Framework for Knowledge-Based Visual Question Answering
di: Mao, Xianwei, et al.
Pubblicazione: (2026)
di: Mao, Xianwei, et al.
Pubblicazione: (2026)
Analysis of Rail Transit Operation and Maintenance Fault Recognition Considering Bayesian Knowledge Recognition Algorithm
di: Yanyan Zhang
Pubblicazione: (2025)
di: Yanyan Zhang
Pubblicazione: (2025)
Nonlinear predictive control of a quadrotor based on differential flatness
di: Runmei Zhang, et al.
Pubblicazione: (2025)
di: Runmei Zhang, et al.
Pubblicazione: (2025)
PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
di: Zhang, Xiaoman, et al.
Pubblicazione: (2023)
di: Zhang, Xiaoman, et al.
Pubblicazione: (2023)
Mechanical Mode Analysis of Centrifugal Pump Impellers Based on Numerical Simulations
di: Jiani Zhang
Pubblicazione: (2025)
di: Jiani Zhang
Pubblicazione: (2025)
Visual Robustness Benchmark for Visual Question Answering (VQA)
di: Ishmam, Md Farhan, et al.
Pubblicazione: (2024)
di: Ishmam, Md Farhan, et al.
Pubblicazione: (2024)
SpecVQA: A Benchmark for Spectral Understanding and Visual Question Answering in Scientific Images
di: Shen, Jialu, et al.
Pubblicazione: (2026)
di: Shen, Jialu, et al.
Pubblicazione: (2026)
StoryTailor:A Zero-Shot Pipeline for Action-Rich Multi-Subject Visual Narratives
di: Hu, Jinghao, et al.
Pubblicazione: (2026)
di: Hu, Jinghao, et al.
Pubblicazione: (2026)
KNVQA: A Benchmark for evaluation knowledge-based VQA
di: Cheng, Sirui, et al.
Pubblicazione: (2023)
di: Cheng, Sirui, et al.
Pubblicazione: (2023)
Study on Elevated Interval Evacuation of Metro Train considering the Influence of Fire, Railings, and Track Bed
di: Jianyao Tu, et al.
Pubblicazione: (2024)
di: Jianyao Tu, et al.
Pubblicazione: (2024)
WSI-VQA: Interpreting Whole Slide Images by Generative Visual Question Answering
di: Chen, Pingyi, et al.
Pubblicazione: (2024)
di: Chen, Pingyi, et al.
Pubblicazione: (2024)
NEUROLOGIC: From Neural Representations to Interpretable Logic Rules
di: Geng, Chuqin, et al.
Pubblicazione: (2025)
di: Geng, Chuqin, et al.
Pubblicazione: (2025)
SuperMemory-VQA: An Egocentric Visual Question-Answering Benchmark for Long-Horizon Memory
di: Alam, Samiul, et al.
Pubblicazione: (2026)
di: Alam, Samiul, et al.
Pubblicazione: (2026)
SCRA-VQA: Summarized Caption-Rerank for Augmented Large Language Models in Visual Question Answering
di: Zhang, Yan, et al.
Pubblicazione: (2025)
di: Zhang, Yan, et al.
Pubblicazione: (2025)
Benchmarking Multimodal Retrieval Augmented Generation with Dynamic VQA Dataset and Self-adaptive Planning Agent
di: Li, Yangning, et al.
Pubblicazione: (2024)
di: Li, Yangning, et al.
Pubblicazione: (2024)
PVPUFormer: Probabilistic Visual Prompt Unified Transformer for Interactive Image Segmentation
di: Zhang, Xu, et al.
Pubblicazione: (2023)
di: Zhang, Xu, et al.
Pubblicazione: (2023)
VaseVQA-3D: Benchmarking 3D VLMs on Ancient Greek Pottery
di: Zhang, Nonghai, et al.
Pubblicazione: (2025)
di: Zhang, Nonghai, et al.
Pubblicazione: (2025)
OccSTeP: Benchmarking 4D Occupancy Spatio-Temporal Persistence
di: Zheng, Yu, et al.
Pubblicazione: (2025)
di: Zheng, Yu, et al.
Pubblicazione: (2025)
DWAFM: Dynamic Weighted Graph Structure Embedding Integrated with Attention and Frequency-Domain MLPs for Traffic Forecasting
di: Shi, Sen, et al.
Pubblicazione: (2026)
di: Shi, Sen, et al.
Pubblicazione: (2026)
SpatialViz-Bench: A Cognitively-Grounded Benchmark for Diagnosing Spatial Visualization in MLLMs
di: Wang, Siting, et al.
Pubblicazione: (2025)
di: Wang, Siting, et al.
Pubblicazione: (2025)
Tetris: A Compilation Framework for VQA Applications in Quantum Computing
di: Jin, Yuwei, et al.
Pubblicazione: (2023)
di: Jin, Yuwei, et al.
Pubblicazione: (2023)
CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence
di: Ma, Dongsheng, et al.
Pubblicazione: (2026)
di: Ma, Dongsheng, et al.
Pubblicazione: (2026)
Diffusion Signals Reveal Hidden Connections: A Physics-Inspired Framework for Link Prediction via Personalized PageRank Signals
di: Deng, Huilin Wang Wenjun Zhang Weibing
Pubblicazione: (2025)
di: Deng, Huilin Wang Wenjun Zhang Weibing
Pubblicazione: (2025)
VQA$^2$: Visual Question Answering for Video Quality Assessment
di: Jia, Ziheng, et al.
Pubblicazione: (2024)
di: Jia, Ziheng, et al.
Pubblicazione: (2024)
Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages
di: Ma, Guozheng, et al.
Pubblicazione: (2023)
di: Ma, Guozheng, et al.
Pubblicazione: (2023)
Audio Outperforms Text for Visual Decoding
di: Zhang, Zhengdi, et al.
Pubblicazione: (2026)
di: Zhang, Zhengdi, et al.
Pubblicazione: (2026)
Seeing Beyond Words: MatVQA for Challenging Visual-Scientific Reasoning in Materials Science
di: Wu, Sifan, et al.
Pubblicazione: (2025)
di: Wu, Sifan, et al.
Pubblicazione: (2025)
RankDVQA: Deep VQA based on Ranking-inspired Hybrid Training
di: Feng, Chen, et al.
Pubblicazione: (2022)
di: Feng, Chen, et al.
Pubblicazione: (2022)
Documenti analoghi
-
Reflecting with Two Voices: A Co-Adaptive Dual-Strategy Framework for LLM-Based Agent Decision Making
di: Zhang, Wentao, et al.
Pubblicazione: (2025) -
CogRail: Benchmarking VLMs in Cognitive Intrusion Perception for Intelligent Railway Transportation Systems
di: Tian, Yonglin, et al.
Pubblicazione: (2026) -
Hot Deformation Behavior and Processing Map of Eutectoid Pearlite Rail Steel
di: Haibo Feng, et al.
Pubblicazione: (2025) -
Knowledge Condensation and Reasoning for Knowledge-based VQA
di: Hao, Dongze, et al.
Pubblicazione: (2024) -
An LLM-based Framework for Human-Swarm Teaming Cognition in Disaster Search and Rescue
di: Ji, Kailun, et al.
Pubblicazione: (2025)