A2Seek: Towards Reasoning-Centric Benchmark for Aerial Anomaly Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Mo, Mengjingcheng, Tong, Xinyang, Tan, Mingpi, Leng, Jiaxu, Zheng, Jiankang, Liu, Yiran, Chen, Haosheng, Gan, Ji, Li, Weisheng, Gao, Xinbo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PiercingEye: Dual-Space Video Violence Detection with Hyperbolic Vision-Language Guidance
by: Leng, Jiaxu, et al.
Published: (2025)
by: Leng, Jiaxu, et al.
Published: (2025)
Beyond Euclidean: Dual-Space Representation Learning for Weakly Supervised Video Violence Detection
by: Leng, Jiaxu, et al.
Published: (2024)
by: Leng, Jiaxu, et al.
Published: (2024)
Shape-centered Representation Learning for Visible-Infrared Person Re-identification
by: Li, Shuang, et al.
Published: (2023)
by: Li, Shuang, et al.
Published: (2023)
EHGCN: Hierarchical Euclidean-Hyperbolic Fusion via Motion-Aware GCN for Hybrid Event Stream Perception
by: Chen, Haosheng, et al.
Published: (2025)
by: Chen, Haosheng, et al.
Published: (2025)
Video-Level Language-Driven Video-Based Visible-Infrared Person Re-Identification
by: Li, Shuang, et al.
Published: (2025)
by: Li, Shuang, et al.
Published: (2025)
Causal Bootstrapped Alignment for Unsupervised Video-Based Visible-Infrared Person Re-Identification
by: Li, Shuang, et al.
Published: (2026)
by: Li, Shuang, et al.
Published: (2026)
CLIP-Driven Cloth-Agnostic Feature Learning for Cloth-Changing Person Re-Identification
by: Li, Shuang, et al.
Published: (2024)
by: Li, Shuang, et al.
Published: (2024)
Ranking-based Adaptive Query Generation for DETRs in Crowded Pedestrian Detection
by: Gao, Feng, et al.
Published: (2023)
by: Gao, Feng, et al.
Published: (2023)
SplatFont3D: Structure-Aware Text-to-3D Artistic Font Generation with Part-Level Style Control
by: Gan, Ji, et al.
Published: (2025)
by: Gan, Ji, et al.
Published: (2025)
CrashSight: A Phase-Aware, Infrastructure-Centric Video Benchmark for Traffic Crash Scene Understanding and Reasoning
by: Gan, Rui, et al.
Published: (2026)
by: Gan, Rui, et al.
Published: (2026)
Event-Centric Human Value Understanding in News-Domain Texts: An Actor-Conditioned, Multi-Granularity Benchmark
by: Wang, Yao, et al.
Published: (2026)
by: Wang, Yao, et al.
Published: (2026)
Toward Reasoning-Centric Time-Series Analysis
by: Wang, Xinlei, et al.
Published: (2025)
by: Wang, Xinlei, et al.
Published: (2025)
Spatial Reasoning in Foundation Models: Benchmarking Object-Centric Spatial Understanding
by: Mirjalili, Vahid, et al.
Published: (2025)
by: Mirjalili, Vahid, et al.
Published: (2025)
RefAerial: A Benchmark and Approach for Referring Detection in Aerial Images
by: Hu, Guyue, et al.
Published: (2026)
by: Hu, Guyue, et al.
Published: (2026)
Learning Mean-Field Games through Mean-Field Actor-Critic Flow
by: Zhou, Mo, et al.
Published: (2025)
by: Zhou, Mo, et al.
Published: (2025)
InstrAct: Towards Action-Centric Understanding in Instructional Videos
by: Yang, Zhuoyi, et al.
Published: (2026)
by: Yang, Zhuoyi, et al.
Published: (2026)
HumanVideo-MME: Benchmarking MLLMs for Human-Centric Video Understanding
by: Cai, Yuxuan, et al.
Published: (2025)
by: Cai, Yuxuan, et al.
Published: (2025)
ConceptMix++: Leveling the Playing Field in Text-to-Image Benchmarking via Iterative Prompt Optimization
by: Gan, Haosheng, et al.
Published: (2025)
by: Gan, Haosheng, et al.
Published: (2025)
From Static Benchmarks to Dynamic Protocol: Agent-Centric Text Anomaly Detection for Evaluating LLM Reasoning
by: Yoa, Seungdong, et al.
Published: (2026)
by: Yoa, Seungdong, et al.
Published: (2026)
The Quest for Efficient Reasoning: A Data-Centric Benchmark to CoT Distillation
by: Zhang, Ruichen, et al.
Published: (2025)
by: Zhang, Ruichen, et al.
Published: (2025)
Putting HUMANS first: Efficient LAM Evaluation with Human Preference Alignment
by: Gan, Woody Haosheng, et al.
Published: (2026)
by: Gan, Woody Haosheng, et al.
Published: (2026)
Semi-Supervised Learning for Anomaly Traffic Detection via Bidirectional Normalizing Flows
by: Dang, Zhangxuan, et al.
Published: (2024)
by: Dang, Zhangxuan, et al.
Published: (2024)
Towards Understanding the Safety Boundaries of DeepSeek Models: Evaluation and Findings
by: Ying, Zonghao, et al.
Published: (2025)
by: Ying, Zonghao, et al.
Published: (2025)
DPTrack:Directional Kernel-Guided Prompt Learning for Robust Nighttime Aerial Tracking
by: Zhu, Zhiqiang, et al.
Published: (2025)
by: Zhu, Zhiqiang, et al.
Published: (2025)
Non-target Divergence Hypothesis: Toward Understanding Domain Gaps in Cross-Modal Knowledge Distillation
by: Chen, Yilong, et al.
Published: (2024)
by: Chen, Yilong, et al.
Published: (2024)
Seek-and-Solve: Benchmarking MLLMs for Visual Clue-Driven Reasoning in Daily Scenarios
by: Li, Xiaomin, et al.
Published: (2026)
by: Li, Xiaomin, et al.
Published: (2026)
Towards Adaptive Open-Set Object Detection via Category-Level Collaboration Knowledge Mining
by: Ji, Yuqi, et al.
Published: (2026)
by: Ji, Yuqi, et al.
Published: (2026)
nuReasoning: A Reasoning-Centric Dataset and Benchmark for Long-Tail Autonomous Driving
by: Huang, Zhiyu, et al.
Published: (2026)
by: Huang, Zhiyu, et al.
Published: (2026)
TurnBench-MS: A Benchmark for Evaluating Multi-Turn, Multi-Step Reasoning in Large Language Models
by: Zhang, Yiran, et al.
Published: (2025)
by: Zhang, Yiran, et al.
Published: (2025)
Beyond English-Centric Training: How Reinforcement Learning Improves Cross-Lingual Reasoning in LLMs
by: Huang, Shulin, et al.
Published: (2025)
by: Huang, Shulin, et al.
Published: (2025)
RotorTM: A Flexible Simulator for Aerial Transportation and Manipulation
by: Li, Guanrui, et al.
Published: (2022)
by: Li, Guanrui, et al.
Published: (2022)
MME-Finance: A Multimodal Finance Benchmark for Expert-level Understanding and Reasoning
by: Gan, Ziliang, et al.
Published: (2024)
by: Gan, Ziliang, et al.
Published: (2024)
Towards a Dao‐Centric Orientation of Sustainability Education
by: Charlene Tan
Published: (2026)
by: Charlene Tan
Published: (2026)
Benchmarking and Understanding Compositional Relational Reasoning of LLMs
by: Ni, Ruikang, et al.
Published: (2024)
by: Ni, Ruikang, et al.
Published: (2024)
FinanceReasoning: Benchmarking Financial Numerical Reasoning More Credible, Comprehensive and Challenging
by: Tang, Zichen, et al.
Published: (2025)
by: Tang, Zichen, et al.
Published: (2025)
From Twitter to Reasoner: Understand Mobility Travel Modes and Sentiment Using Large Language Models
by: Ruan, Kangrui, et al.
Published: (2024)
by: Ruan, Kangrui, et al.
Published: (2024)
Evaluating Large Language Models for Financial Reasoning: A CFA-Based Benchmark Study
by: Yao, Xuan, et al.
Published: (2025)
by: Yao, Xuan, et al.
Published: (2025)
DocVideoQA: Towards Comprehensive Understanding of Document-Centric Videos through Question Answering
by: Wang, Haochen, et al.
Published: (2025)
by: Wang, Haochen, et al.
Published: (2025)
Meso‐Cenozoic Multiple Deformations of the Southern Margin of Sichuan Basin and Their Response to Tectonic Events in East Asia
by: Haosheng Zhang, et al.
Published: (2025)
by: Haosheng Zhang, et al.
Published: (2025)
Prompt Disentanglement via Language Guidance and Representation Alignment for Domain Generalization
by: Cheng, De, et al.
Published: (2025)
by: Cheng, De, et al.
Published: (2025)
Similar Items
-
PiercingEye: Dual-Space Video Violence Detection with Hyperbolic Vision-Language Guidance
by: Leng, Jiaxu, et al.
Published: (2025) -
Beyond Euclidean: Dual-Space Representation Learning for Weakly Supervised Video Violence Detection
by: Leng, Jiaxu, et al.
Published: (2024) -
Shape-centered Representation Learning for Visible-Infrared Person Re-identification
by: Li, Shuang, et al.
Published: (2023) -
EHGCN: Hierarchical Euclidean-Hyperbolic Fusion via Motion-Aware GCN for Hybrid Event Stream Perception
by: Chen, Haosheng, et al.
Published: (2025) -
Video-Level Language-Driven Video-Based Visible-Infrared Person Re-Identification
by: Li, Shuang, et al.
Published: (2025)