See What You Need: Query-Aware Visual Intelligence through Reasoning-Perception Loops
Fuente:
arXiv
Salvato in:
| Autori principali: | Dong, Zixuan, Peng, Baoyun, Wang, Yufei, Liu, Lin, Dong, Xinxin, Cao, Yunlong, Wang, Xiaodong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
TerraSeg: Self-Supervised Ground Segmentation for Any LiDAR
di: Lentsch, Ted, et al.
Pubblicazione: (2026)
di: Lentsch, Ted, et al.
Pubblicazione: (2026)
UNION: Unsupervised 3D Object Detection using Object Appearance-based Pseudo-Classes
di: Lentsch, Ted, et al.
Pubblicazione: (2024)
di: Lentsch, Ted, et al.
Pubblicazione: (2024)
Fairness Without Labels: Pseudo-Balancing for Bias Mitigation in Face Gender Classification
di: Dong, Haohua, et al.
Pubblicazione: (2025)
di: Dong, Haohua, et al.
Pubblicazione: (2025)
Salient Concept-Aware Generative Data Augmentation
di: Zhao, Tianchen, et al.
Pubblicazione: (2025)
di: Zhao, Tianchen, et al.
Pubblicazione: (2025)
Smelly, dense, and spreaded: The Object Detection for Olfactory References (ODOR) dataset
di: Zinnen, Mathias, et al.
Pubblicazione: (2025)
di: Zinnen, Mathias, et al.
Pubblicazione: (2025)
Learning Association via Track-Detection Matching for Multi-Object Tracking
di: Adžemović, Momir
Pubblicazione: (2025)
di: Adžemović, Momir
Pubblicazione: (2025)
Wafer Map Defect Classification Using Autoencoder-Based Data Augmentation and Convolutional Neural Network
di: Bao, Yin-Yin, et al.
Pubblicazione: (2024)
di: Bao, Yin-Yin, et al.
Pubblicazione: (2024)
Polarization-Based Eye Tracking with Personalized Siamese Architectures
di: Kalkanli, Beyza, et al.
Pubblicazione: (2026)
di: Kalkanli, Beyza, et al.
Pubblicazione: (2026)
Deep Learning-Based Multi-Object Tracking: A Comprehensive Survey from Foundations to State-of-the-Art
di: Adžemović, Momir
Pubblicazione: (2025)
di: Adžemović, Momir
Pubblicazione: (2025)
CG-HOI: Contact-Guided 3D Human-Object Interaction Generation
di: Diller, Christian, et al.
Pubblicazione: (2023)
di: Diller, Christian, et al.
Pubblicazione: (2023)
MVTamperBench: Evaluating Robustness of Vision-Language Models
di: Agarwal, Amit, et al.
Pubblicazione: (2024)
di: Agarwal, Amit, et al.
Pubblicazione: (2024)
Skullptor: High Fidelity 3D Head Reconstruction in Seconds with Multi-View Normal Prediction
di: Artru, Noé, et al.
Pubblicazione: (2026)
di: Artru, Noé, et al.
Pubblicazione: (2026)
Capacity Constraint Analysis Using Object Detection for Smart Manufacturing
di: Ahmad, Hafiz Mughees, et al.
Pubblicazione: (2024)
di: Ahmad, Hafiz Mughees, et al.
Pubblicazione: (2024)
SH17: A Dataset for Human Safety and Personal Protective Equipment Detection in Manufacturing Industry
di: Ahmad, Hafiz Mughees, et al.
Pubblicazione: (2024)
di: Ahmad, Hafiz Mughees, et al.
Pubblicazione: (2024)
A deep learning approach to track eye movements based on events
di: Seth, Chirag, et al.
Pubblicazione: (2025)
di: Seth, Chirag, et al.
Pubblicazione: (2025)
A Single Image Is All You Need: Zero-Shot Anomaly Localization Without Training Data
di: Moradi, Mehrdad, et al.
Pubblicazione: (2025)
di: Moradi, Mehrdad, et al.
Pubblicazione: (2025)
Hierarchical Point-Patch Fusion with Adaptive Patch Codebook for 3D Shape Anomaly Detection
di: Kang, Xueyang, et al.
Pubblicazione: (2026)
di: Kang, Xueyang, et al.
Pubblicazione: (2026)
FutureHuman3D: Forecasting Complex Long-Term 3D Human Behavior from Video Observations
di: Diller, Christian, et al.
Pubblicazione: (2022)
di: Diller, Christian, et al.
Pubblicazione: (2022)
Detecting AI-Generated Videos with Spiking Neural Networks
di: Jang, Minsuk, et al.
Pubblicazione: (2026)
di: Jang, Minsuk, et al.
Pubblicazione: (2026)
LRCP: Low-Rank Compressibility Guided Visual Token Pruning for Efficient LVLMs
di: Lu, Hongyu, et al.
Pubblicazione: (2026)
di: Lu, Hongyu, et al.
Pubblicazione: (2026)
Corn Ear Detection and Orientation Estimation Using Deep Learning
di: Sprague, Nathan, et al.
Pubblicazione: (2024)
di: Sprague, Nathan, et al.
Pubblicazione: (2024)
Gr-IoU: Ground-Intersection over Union for Robust Multi-Object Tracking with 3D Geometric Constraints
di: Toida, Keisuke, et al.
Pubblicazione: (2024)
di: Toida, Keisuke, et al.
Pubblicazione: (2024)
Predictive Modeling of Maritime Radar Data Using Transformer Architecture
di: Qesaraku, Bjorna, et al.
Pubblicazione: (2025)
di: Qesaraku, Bjorna, et al.
Pubblicazione: (2025)
Few-Class Arena: A Benchmark for Efficient Selection of Vision Models and Dataset Difficulty Measurement
di: Cao, Bryan Bo, et al.
Pubblicazione: (2024)
di: Cao, Bryan Bo, et al.
Pubblicazione: (2024)
Rethinking Visual Intelligence: Insights from Video Pretraining
di: Acuaviva, Pablo, et al.
Pubblicazione: (2025)
di: Acuaviva, Pablo, et al.
Pubblicazione: (2025)
Scene Detection Policies and Keyframe Extraction Strategies for Large-Scale Video Analysis
di: Korolkov, Vasilii
Pubblicazione: (2025)
di: Korolkov, Vasilii
Pubblicazione: (2025)
Heart Failure Prediction using Modal Decomposition and Masked Autoencoders for Scarce Echocardiography Databases
di: Bell-Navas, Andrés, et al.
Pubblicazione: (2025)
di: Bell-Navas, Andrés, et al.
Pubblicazione: (2025)
Transforming faces into video stories -- VideoFace2.0
di: Brkljač, Branko, et al.
Pubblicazione: (2025)
di: Brkljač, Branko, et al.
Pubblicazione: (2025)
Deep Spectral Meshes: Multi-Frequency Facial Mesh Processing with Graph Neural Networks
di: Kosk, Robert, et al.
Pubblicazione: (2024)
di: Kosk, Robert, et al.
Pubblicazione: (2024)
Logits-Constrained Framework with RoBERTa for Ancient Chinese NER
di: Hua, Wenjie, et al.
Pubblicazione: (2025)
di: Hua, Wenjie, et al.
Pubblicazione: (2025)
GIQ: Benchmarking 3D Geometric Reasoning of Vision Foundation Models with Simulated and Real Polyhedra
di: Michalkiewicz, Mateusz, et al.
Pubblicazione: (2025)
di: Michalkiewicz, Mateusz, et al.
Pubblicazione: (2025)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
di: Semenov, Andrei, et al.
Pubblicazione: (2024)
di: Semenov, Andrei, et al.
Pubblicazione: (2024)
SeMi: When Imbalanced Semi-Supervised Learning Meets Mining Hard Examples
di: Wang, Yin, et al.
Pubblicazione: (2025)
di: Wang, Yin, et al.
Pubblicazione: (2025)
AQFusionNet: Multimodal Deep Learning for Air Quality Index Prediction with Imagery and Sensor Data
di: Kushal, Koushik Ahmed, et al.
Pubblicazione: (2025)
di: Kushal, Koushik Ahmed, et al.
Pubblicazione: (2025)
ARTPS: Depth-Enhanced Hybrid Anomaly Detection and Learnable Curiosity Score for Autonomous Rover Target Prioritization
di: Baydemir, Poyraz
Pubblicazione: (2025)
di: Baydemir, Poyraz
Pubblicazione: (2025)
TRACES: Temporal Recall with Contextual Embeddings for Real-Time Video Anomaly Detection
di: Siddiqui, Yousuf Ahmed, et al.
Pubblicazione: (2025)
di: Siddiqui, Yousuf Ahmed, et al.
Pubblicazione: (2025)
SpectralCA: Bi-Directional Cross-Attention for Next-Generation UAV Hyperspectral Vision
di: Brovko, D. V.
Pubblicazione: (2025)
di: Brovko, D. V.
Pubblicazione: (2025)
Single-Step Reconstruction-Free Anomaly Detection and Segmentation via Diffusion Models
di: Moradi, Mehrdad, et al.
Pubblicazione: (2025)
di: Moradi, Mehrdad, et al.
Pubblicazione: (2025)
License Plate Detection and Character Recognition Using Deep Learning and Font Evaluation
di: Vargoorani, Zahra Ebrahimi, et al.
Pubblicazione: (2024)
di: Vargoorani, Zahra Ebrahimi, et al.
Pubblicazione: (2024)
StatsMerging: Statistics-Guided Model Merging via Task-Specific Teacher Distillation
di: Merugu, Ranjith, et al.
Pubblicazione: (2025)
di: Merugu, Ranjith, et al.
Pubblicazione: (2025)
Documenti analoghi
-
TerraSeg: Self-Supervised Ground Segmentation for Any LiDAR
di: Lentsch, Ted, et al.
Pubblicazione: (2026) -
UNION: Unsupervised 3D Object Detection using Object Appearance-based Pseudo-Classes
di: Lentsch, Ted, et al.
Pubblicazione: (2024) -
Fairness Without Labels: Pseudo-Balancing for Bias Mitigation in Face Gender Classification
di: Dong, Haohua, et al.
Pubblicazione: (2025) -
Salient Concept-Aware Generative Data Augmentation
di: Zhao, Tianchen, et al.
Pubblicazione: (2025) -
Smelly, dense, and spreaded: The Object Detection for Olfactory References (ODOR) dataset
di: Zinnen, Mathias, et al.
Pubblicazione: (2025)