SideSeeing: A multimodal dataset and collection of tools for sidewalk assessment
Fuente:
arXiv
Guardado en:
| Autores principales: | Damaceno, R. J. P., Ferreira, L., Miranda, F., Hosseini, M., Cesar Jr, R. M. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Exploring the dynamic interplay of cognitive load and emotional arousal by using multimodal measurements: Correlation of pupil diameter and emotional arousal in emotionally engaging tasks
por: Kosel, C., et al.
Publicado: (2024)
por: Kosel, C., et al.
Publicado: (2024)
Deep Umbra: A Generative Approach for Sunlight Access Computation in Urban Spaces
por: Omar, Kazi Shahrukh, et al.
Publicado: (2024)
por: Omar, Kazi Shahrukh, et al.
Publicado: (2024)
Seeing Candidates at Scale: Multimodal LLMs for Visual Political Communication on Instagram
por: Achmann-Denkler, Michael, et al.
Publicado: (2026)
por: Achmann-Denkler, Michael, et al.
Publicado: (2026)
Can Multimodal LLMs See Science Instruction? Benchmarking Pedagogical Reasoning in K-12 Classroom Videos
por: Shen, Yixuan, et al.
Publicado: (2026)
por: Shen, Yixuan, et al.
Publicado: (2026)
Consensus and Subjectivity of Skin Tone Annotation for ML Fairness
por: Schumann, Candice, et al.
Publicado: (2023)
por: Schumann, Candice, et al.
Publicado: (2023)
Seeing the Intangible: Survey of Image Classification into High-Level and Abstract Categories
por: Pandiani, Delfina Sol Martinez, et al.
Publicado: (2023)
por: Pandiani, Delfina Sol Martinez, et al.
Publicado: (2023)
Real-Time Automated donning and doffing detection of PPE based on Yolov4-tiny
por: Verma, Anusha, et al.
Publicado: (2024)
por: Verma, Anusha, et al.
Publicado: (2024)
Who's in and who's out? A case study of multimodal CLIP-filtering in DataComp
por: Hong, Rachel, et al.
Publicado: (2024)
por: Hong, Rachel, et al.
Publicado: (2024)
Responsible AI for Earth Observation
por: Ghamisi, Pedram, et al.
Publicado: (2024)
por: Ghamisi, Pedram, et al.
Publicado: (2024)
Headset: Human emotion awareness under partial occlusions multimodal dataset
por: Lohesara, Fatemeh Ghorbani, et al.
Publicado: (2024)
por: Lohesara, Fatemeh Ghorbani, et al.
Publicado: (2024)
A multimodal gesture recognition dataset for desktop human-computer interaction
por: Wang, Qi, et al.
Publicado: (2024)
por: Wang, Qi, et al.
Publicado: (2024)
AI-based System for Transforming text and sound to Educational Videos
por: ElAlami, M. E., et al.
Publicado: (2026)
por: ElAlami, M. E., et al.
Publicado: (2026)
MObyGaze: a film dataset of multimodal objectification densely annotated by experts
por: Tores, Julie, et al.
Publicado: (2025)
por: Tores, Julie, et al.
Publicado: (2025)
BLEnD-Vis: Benchmarking Multimodal Cultural Understanding in Vision Language Models
por: Tan, Bryan Chen Zhengyu, et al.
Publicado: (2025)
por: Tan, Bryan Chen Zhengyu, et al.
Publicado: (2025)
Video-based Pedestrian and Vehicle Traffic Analysis During Football Games
por: Fleischer, Jacques P., et al.
Publicado: (2024)
por: Fleischer, Jacques P., et al.
Publicado: (2024)
AquaMonitor: A multimodal multi-view image sequence dataset for real-life aquatic invertebrate biodiversity monitoring
por: Impiö, Mikko, et al.
Publicado: (2025)
por: Impiö, Mikko, et al.
Publicado: (2025)
MULTIAQUA: A multimodal maritime dataset and robust training strategies for multimodal semantic segmentation
por: Muhovič, Jon, et al.
Publicado: (2025)
por: Muhovič, Jon, et al.
Publicado: (2025)
FigSIM: A Dataset for Fine-grained Suicide Severity and Figurative Language in Suicide Memes
por: Chen, Liuliu, et al.
Publicado: (2026)
por: Chen, Liuliu, et al.
Publicado: (2026)
Climatic & Anthropogenic Hazards to the Nasca World Heritage: Application of Remote Sensing, AI, and Flood Modelling
por: Sakai, Masato, et al.
Publicado: (2024)
por: Sakai, Masato, et al.
Publicado: (2024)
Copycats: the many lives of a publicly available medical imaging dataset
por: Jiménez-Sánchez, Amelia, et al.
Publicado: (2024)
por: Jiménez-Sánchez, Amelia, et al.
Publicado: (2024)
GeoLocator: a location-integrated large multimodal model for inferring geo-privacy
por: Yang, Yifan, et al.
Publicado: (2023)
por: Yang, Yifan, et al.
Publicado: (2023)
EHWGesture -- A dataset for multimodal understanding of clinical gestures
por: Amprimo, Gianluca, et al.
Publicado: (2025)
por: Amprimo, Gianluca, et al.
Publicado: (2025)
Deploying Rapid Damage Assessments from sUAS Imagery for Disaster Response
por: Manzini, Thomas, et al.
Publicado: (2025)
por: Manzini, Thomas, et al.
Publicado: (2025)
URBAN-SPIN: A street-level bikeability index to inform design implementations in historical city centres
por: Ding, Haining, et al.
Publicado: (2026)
por: Ding, Haining, et al.
Publicado: (2026)
AnxietyFaceTrack: A Smartphone-Based Non-Intrusive Approach for Detecting Social Anxiety Using Facial Features
por: Sahu, Nilesh Kumar, et al.
Publicado: (2025)
por: Sahu, Nilesh Kumar, et al.
Publicado: (2025)
BenDFM: A taxonomy and synthetic CAD dataset for manufacturability assessment in sheet metal bending
por: Ballegeer, Matteo, et al.
Publicado: (2026)
por: Ballegeer, Matteo, et al.
Publicado: (2026)
LLM-Driven Completeness and Consistency Evaluation for Cultural Heritage Data Augmentation in Cross-Modal Retrieval
por: Zhang, Jian, et al.
Publicado: (2025)
por: Zhang, Jian, et al.
Publicado: (2025)
Exploring the Adversarial Robustness of Face Forgery Detection with Decision-based Black-box Attacks
por: Chen, Zhaoyu, et al.
Publicado: (2023)
por: Chen, Zhaoyu, et al.
Publicado: (2023)
EgoPrivacy: What Your First-Person Camera Says About You?
por: Li, Yijiang, et al.
Publicado: (2025)
por: Li, Yijiang, et al.
Publicado: (2025)
Surgeons Awareness, Expectations, and Involvement with Artificial Intelligence: a Survey Pre and Post the GPT Era
por: Arboit, Lorenzo, et al.
Publicado: (2025)
por: Arboit, Lorenzo, et al.
Publicado: (2025)
How Many Visual Levers Drive Urban Perception? Interventional Counterfactuals via Multiple Localised Edits
por: Tang, Jason, et al.
Publicado: (2026)
por: Tang, Jason, et al.
Publicado: (2026)
Generalized People Diversity: Learning a Human Perception-Aligned Diversity Representation for People Images
por: Srinivasan, Hansa, et al.
Publicado: (2024)
por: Srinivasan, Hansa, et al.
Publicado: (2024)
A Weak Supervision Learning Approach Towards an Equitable Mobility Estimation
por: Aidoo, Theophilus, et al.
Publicado: (2025)
por: Aidoo, Theophilus, et al.
Publicado: (2025)
Generalizable Slum Detection from Satellite Imagery with Mixture-of-Experts
por: Lee, Sumin, et al.
Publicado: (2025)
por: Lee, Sumin, et al.
Publicado: (2025)
ChaosBench: A Multi-Channel, Physics-Based Benchmark for Subseasonal-to-Seasonal Climate Prediction
por: Nathaniel, Juan, et al.
Publicado: (2024)
por: Nathaniel, Juan, et al.
Publicado: (2024)
The Iconicity of the Generated Image
por: van Noord, Nanne, et al.
Publicado: (2025)
por: van Noord, Nanne, et al.
Publicado: (2025)
From Drone Imagery to Livability Mapping: AI-powered Environment Perception in Rural China
por: Deng, Weihuan, et al.
Publicado: (2025)
por: Deng, Weihuan, et al.
Publicado: (2025)
Architecture inside the mirage: evaluating generative image models on architectural style, elements, and typologies
por: Magrill, Jamie, et al.
Publicado: (2026)
por: Magrill, Jamie, et al.
Publicado: (2026)
EmoAssist: Emotional Assistant for Visual Impairment Community
por: Qi, Xingyu, et al.
Publicado: (2025)
por: Qi, Xingyu, et al.
Publicado: (2025)
PLAS-Net: Pixel-Level Area Segmentation for UAV-Based Beach Litter Monitoring
por: Liu, Yongying, et al.
Publicado: (2026)
por: Liu, Yongying, et al.
Publicado: (2026)
Ejemplares similares
-
Exploring the dynamic interplay of cognitive load and emotional arousal by using multimodal measurements: Correlation of pupil diameter and emotional arousal in emotionally engaging tasks
por: Kosel, C., et al.
Publicado: (2024) -
Deep Umbra: A Generative Approach for Sunlight Access Computation in Urban Spaces
por: Omar, Kazi Shahrukh, et al.
Publicado: (2024) -
Seeing Candidates at Scale: Multimodal LLMs for Visual Political Communication on Instagram
por: Achmann-Denkler, Michael, et al.
Publicado: (2026) -
Can Multimodal LLMs See Science Instruction? Benchmarking Pedagogical Reasoning in K-12 Classroom Videos
por: Shen, Yixuan, et al.
Publicado: (2026) -
Consensus and Subjectivity of Skin Tone Annotation for ML Fairness
por: Schumann, Candice, et al.
Publicado: (2023)