Persona-aware and Explainable Bikeability Assessment: A Vision-Language Model Approach
Fuente:
arXiv
Salvato in:
| Autori principali: | Dai, Yilong, Wang, Ziyi, Wang, Chenguang, Zhou, Kexin, Qian, Yiheng, Xu, Susu, Yan, Xiang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Beyond RGB: Leveraging Vision Transformers for Thermal Weapon Segmentation
di: Kambhatla, Akhila, et al.
Pubblicazione: (2025)
di: Kambhatla, Akhila, et al.
Pubblicazione: (2025)
The Ethics Engine: A Modular Pipeline for Accessible Psychometric Assessment of Large Language Models
di: Van Clief, Jake, et al.
Pubblicazione: (2025)
di: Van Clief, Jake, et al.
Pubblicazione: (2025)
The Impact of Image Resolution on Face Detection: A Comparative Analysis of MTCNN, YOLOv XI and YOLOv XII models
di: Ömercikoğlu, Ahmet Can, et al.
Pubblicazione: (2025)
di: Ömercikoğlu, Ahmet Can, et al.
Pubblicazione: (2025)
CADE 2.5 - ZeResFDG: Frequency-Decoupled, Rescaled and Zero-Projected Guidance for SD/SDXL Latent Diffusion Models
di: Rychkovskiy, Denis
Pubblicazione: (2025)
di: Rychkovskiy, Denis
Pubblicazione: (2025)
Explaining What Machines See: XAI Strategies in Deep Object Detection Models
di: Seyedmomeni, FatemehSadat, et al.
Pubblicazione: (2025)
di: Seyedmomeni, FatemehSadat, et al.
Pubblicazione: (2025)
Feature Based Methods in Domain Adaptation for Object Detection: A Review Paper
di: Mohamadi, Helia, et al.
Pubblicazione: (2024)
di: Mohamadi, Helia, et al.
Pubblicazione: (2024)
A Novel Compression Framework for YOLOv8: Achieving Real-Time Aerial Object Detection on Edge Devices via Structured Pruning and Channel-Wise Distillation
di: Sabaghian, Melika, et al.
Pubblicazione: (2025)
di: Sabaghian, Melika, et al.
Pubblicazione: (2025)
Architectural Insights into Knowledge Distillation for Object Detection: A Comprehensive Review
di: Golizadeh, Mahdi, et al.
Pubblicazione: (2025)
di: Golizadeh, Mahdi, et al.
Pubblicazione: (2025)
Enhancing Small Object Detection with YOLO: A Novel Framework for Improved Accuracy and Efficiency
di: Moghadami, Mahila, et al.
Pubblicazione: (2025)
di: Moghadami, Mahila, et al.
Pubblicazione: (2025)
Unlocking UML Class Diagram Understanding in Vision Language Models
di: Naboichenko, Artem, et al.
Pubblicazione: (2026)
di: Naboichenko, Artem, et al.
Pubblicazione: (2026)
Fairness Without Labels: Pseudo-Balancing for Bias Mitigation in Face Gender Classification
di: Dong, Haohua, et al.
Pubblicazione: (2025)
di: Dong, Haohua, et al.
Pubblicazione: (2025)
A deep learning approach to track eye movements based on events
di: Seth, Chirag, et al.
Pubblicazione: (2025)
di: Seth, Chirag, et al.
Pubblicazione: (2025)
Point, Detect, Count: Multi-Task Medical Image Understanding with Instruction-Tuned Vision-Language Models
di: Gautam, Sushant, et al.
Pubblicazione: (2025)
di: Gautam, Sushant, et al.
Pubblicazione: (2025)
RailSafeNet: Visual Scene Understanding for Tram Safety
di: Valach, Ondřej, et al.
Pubblicazione: (2025)
di: Valach, Ondřej, et al.
Pubblicazione: (2025)
BG-YOLO: A Bidirectional-Guided Method for Underwater Object Detection
di: Zhang, Jian, et al.
Pubblicazione: (2024)
di: Zhang, Jian, et al.
Pubblicazione: (2024)
AVATAAR: Agentic Video Answering via Temporal Adaptive Alignment and Reasoning
di: Patel, Urjitkumar, et al.
Pubblicazione: (2025)
di: Patel, Urjitkumar, et al.
Pubblicazione: (2025)
QSilk: Micrograin Stabilization and Adaptive Quantile Clipping for Detail-Friendly Latent Diffusion
di: Rychkovskiy, Denis
Pubblicazione: (2025)
di: Rychkovskiy, Denis
Pubblicazione: (2025)
Polarization-Based Eye Tracking with Personalized Siamese Architectures
di: Kalkanli, Beyza, et al.
Pubblicazione: (2026)
di: Kalkanli, Beyza, et al.
Pubblicazione: (2026)
Beyond RNNs: Benchmarking Attention-Based Image Captioning Models
di: Yanambakkam, Hemanth Teja, et al.
Pubblicazione: (2025)
di: Yanambakkam, Hemanth Teja, et al.
Pubblicazione: (2025)
HelloMeme: Integrating Spatial Knitting Attentions to Embed High-Level and Fidelity-Rich Conditions in Diffusion Models
di: Zhang, Shengkai, et al.
Pubblicazione: (2024)
di: Zhang, Shengkai, et al.
Pubblicazione: (2024)
μ-Net: A Deep Learning-Based Architecture for μ-CT Segmentation
di: Bruno, Pierangela, et al.
Pubblicazione: (2024)
di: Bruno, Pierangela, et al.
Pubblicazione: (2024)
Adversarial Patch Attacks on Vision-Based Cargo Occupancy Estimation via Differentiable 3D Simulation
di: Hedna, Mohamed Rissal, et al.
Pubblicazione: (2025)
di: Hedna, Mohamed Rissal, et al.
Pubblicazione: (2025)
Archival Faces: Detection of Faces in Digitized Historical Documents
di: Vaško, Marek, et al.
Pubblicazione: (2025)
di: Vaško, Marek, et al.
Pubblicazione: (2025)
Sat-JEPA-Diff: Bridging Self-Supervised Learning and Generative Diffusion for Remote Sensing
di: Komurcu, Kursat, et al.
Pubblicazione: (2026)
di: Komurcu, Kursat, et al.
Pubblicazione: (2026)
Underwater SONAR Image Classification and Analysis using LIME-based Explainable Artificial Intelligence
di: Natarajan, Purushothaman, et al.
Pubblicazione: (2024)
di: Natarajan, Purushothaman, et al.
Pubblicazione: (2024)
WildfireVLM: AI-powered Analysis for Early Wildfire Detection and Risk Assessment Using Satellite Imagery
di: Ayanzadeh, Aydin, et al.
Pubblicazione: (2026)
di: Ayanzadeh, Aydin, et al.
Pubblicazione: (2026)
Few-Class Arena: A Benchmark for Efficient Selection of Vision Models and Dataset Difficulty Measurement
di: Cao, Bryan Bo, et al.
Pubblicazione: (2024)
di: Cao, Bryan Bo, et al.
Pubblicazione: (2024)
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
di: Patel, Hitesh Laxmichand, et al.
Pubblicazione: (2025)
di: Patel, Hitesh Laxmichand, et al.
Pubblicazione: (2025)
VideoHEDGE: Entropy-Based Hallucination Detection for Video-VLMs via Semantic Clustering and Spatiotemporal Perturbations
di: Gautam, Sushant, et al.
Pubblicazione: (2026)
di: Gautam, Sushant, et al.
Pubblicazione: (2026)
EventFlow: Real-Time Neuromorphic Event-Driven Classification of Two-Phase Boiling Flow Regimes
di: Chang, Sanghyeon, et al.
Pubblicazione: (2025)
di: Chang, Sanghyeon, et al.
Pubblicazione: (2025)
Customizing Graph Neural Networks using Path Reweighting
di: Chen, Jianpeng, et al.
Pubblicazione: (2021)
di: Chen, Jianpeng, et al.
Pubblicazione: (2021)
AQFusionNet: Multimodal Deep Learning for Air Quality Index Prediction with Imagery and Sensor Data
di: Kushal, Koushik Ahmed, et al.
Pubblicazione: (2025)
di: Kushal, Koushik Ahmed, et al.
Pubblicazione: (2025)
ARTPS: Depth-Enhanced Hybrid Anomaly Detection and Learnable Curiosity Score for Autonomous Rover Target Prioritization
di: Baydemir, Poyraz
Pubblicazione: (2025)
di: Baydemir, Poyraz
Pubblicazione: (2025)
TRACES: Temporal Recall with Contextual Embeddings for Real-Time Video Anomaly Detection
di: Siddiqui, Yousuf Ahmed, et al.
Pubblicazione: (2025)
di: Siddiqui, Yousuf Ahmed, et al.
Pubblicazione: (2025)
YOLO Ensemble for UAV-based Multispectral Defect Detection in Wind Turbine Components
di: Svystun, Serhii, et al.
Pubblicazione: (2025)
di: Svystun, Serhii, et al.
Pubblicazione: (2025)
FastGS: Training 3D Gaussian Splatting in 100 Seconds
di: Ren, Shiwei, et al.
Pubblicazione: (2025)
di: Ren, Shiwei, et al.
Pubblicazione: (2025)
SEGS-SLAM: Structure-enhanced 3D Gaussian Splatting SLAM with Appearance Embedding
di: Wen, Tianci, et al.
Pubblicazione: (2025)
di: Wen, Tianci, et al.
Pubblicazione: (2025)
RCI: A Score for Evaluating Global and Local Reasoning in Multimodal Benchmarks
di: Agarwal, Amit, et al.
Pubblicazione: (2025)
di: Agarwal, Amit, et al.
Pubblicazione: (2025)
Predictive Modeling of Maritime Radar Data Using Transformer Architecture
di: Qesaraku, Bjorna, et al.
Pubblicazione: (2025)
di: Qesaraku, Bjorna, et al.
Pubblicazione: (2025)
Moral Semantics Survive Machine Translation: Cross-Lingual Evidence from Moral Foundations Corpora
di: Skorski, Maciej
Pubblicazione: (2026)
di: Skorski, Maciej
Pubblicazione: (2026)
Documenti analoghi
-
Beyond RGB: Leveraging Vision Transformers for Thermal Weapon Segmentation
di: Kambhatla, Akhila, et al.
Pubblicazione: (2025) -
The Ethics Engine: A Modular Pipeline for Accessible Psychometric Assessment of Large Language Models
di: Van Clief, Jake, et al.
Pubblicazione: (2025) -
The Impact of Image Resolution on Face Detection: A Comparative Analysis of MTCNN, YOLOv XI and YOLOv XII models
di: Ömercikoğlu, Ahmet Can, et al.
Pubblicazione: (2025) -
CADE 2.5 - ZeResFDG: Frequency-Decoupled, Rescaled and Zero-Projected Guidance for SD/SDXL Latent Diffusion Models
di: Rychkovskiy, Denis
Pubblicazione: (2025) -
Explaining What Machines See: XAI Strategies in Deep Object Detection Models
di: Seyedmomeni, FatemehSadat, et al.
Pubblicazione: (2025)