Demonstration of MaskSearch: Efficiently Querying Image Masks for Machine Learning Workflows
Fuente:
arXiv
Salvato in:
| Autori principali: | Wei, Lindsey Linxi, Yeung, Chung Yik Edward, Yu, Hongjian, Zhou, Jingchuan, He, Dong, Balazinska, Magdalena |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MaskSearch: Querying Image Masks at Scale
di: He, Dong, et al.
Pubblicazione: (2023)
di: He, Dong, et al.
Pubblicazione: (2023)
RACOON: An LLM-based Framework for Retrieval-Augmented Column Type Annotation with a Knowledge Graph
di: Wei, Lindsey Linxi, et al.
Pubblicazione: (2024)
di: Wei, Lindsey Linxi, et al.
Pubblicazione: (2024)
NeedleDB: A Generative-AI Based System for Accurate and Efficient Image Retrieval using Complex Natural Language Queries
di: Erfanian, Mahdi, et al.
Pubblicazione: (2026)
di: Erfanian, Mahdi, et al.
Pubblicazione: (2026)
Vidformer: Drop-in Declarative Optimization for Rendering Video-Native Query Results
di: Winecki, Dominik, et al.
Pubblicazione: (2026)
di: Winecki, Dominik, et al.
Pubblicazione: (2026)
A Review of Media Copyright Management using Blockchain Technologies from the Academic and Business Perspectives
di: García, Roberto, et al.
Pubblicazione: (2023)
di: García, Roberto, et al.
Pubblicazione: (2023)
Controllable Text-to-Speech Synthesis with Masked-Autoencoded Style-Rich Representation
di: Wang, Yongqi, et al.
Pubblicazione: (2025)
di: Wang, Yongqi, et al.
Pubblicazione: (2025)
QoS-QoE Translation with Large Language Model
di: Yu, Yingjie, et al.
Pubblicazione: (2026)
di: Yu, Yingjie, et al.
Pubblicazione: (2026)
PETLP: A Privacy-by-Design Pipeline for Social Media Data in AI Research
di: Oh, Nick, et al.
Pubblicazione: (2025)
di: Oh, Nick, et al.
Pubblicazione: (2025)
A Conceptual Model of Intelligent Multimedia Data Rendered using Flying Light Specks
di: Yazdani, Nima, et al.
Pubblicazione: (2024)
di: Yazdani, Nima, et al.
Pubblicazione: (2024)
Interdependency Matters: Graph Alignment for Multivariate Time Series Anomaly Detection
di: Wang, Yuanyi, et al.
Pubblicazione: (2024)
di: Wang, Yuanyi, et al.
Pubblicazione: (2024)
Multimodal LLM-based Query Paraphrasing for Video Search
di: Wu, Jiaxin, et al.
Pubblicazione: (2024)
di: Wu, Jiaxin, et al.
Pubblicazione: (2024)
QPT V2: Masked Image Modeling Advances Visual Scoring
di: Xie, Qizhi, et al.
Pubblicazione: (2024)
di: Xie, Qizhi, et al.
Pubblicazione: (2024)
HeGTa: Leveraging Heterogeneous Graph-enhanced Large Language Models for Few-shot Complex Table Understanding
di: Jin, Rihui, et al.
Pubblicazione: (2024)
di: Jin, Rihui, et al.
Pubblicazione: (2024)
Galley: Modern Query Optimization for Sparse Tensor Programs
di: Deeds, Kyle, et al.
Pubblicazione: (2024)
di: Deeds, Kyle, et al.
Pubblicazione: (2024)
FreeMask: Rethinking the Importance of Attention Masks for Zero-Shot Video Editing
di: Cai, Lingling, et al.
Pubblicazione: (2024)
di: Cai, Lingling, et al.
Pubblicazione: (2024)
MaskCD: Mitigating LVLM Hallucinations by Image Head Masked Contrastive Decoding
di: Deng, Jingyuan, et al.
Pubblicazione: (2025)
di: Deng, Jingyuan, et al.
Pubblicazione: (2025)
LazyVLM: Neuro-Symbolic Approach to Video Analytics
di: Jian, Xiangru, et al.
Pubblicazione: (2025)
di: Jian, Xiangru, et al.
Pubblicazione: (2025)
EditMGT: Unleashing Potentials of Masked Generative Transformers in Image Editing
di: Chow, Wei, et al.
Pubblicazione: (2025)
di: Chow, Wei, et al.
Pubblicazione: (2025)
Self-Enhancing Video Data Management System for Compositional Events with Large Language Models [Technical Report]
di: Zhang, Enhao, et al.
Pubblicazione: (2024)
di: Zhang, Enhao, et al.
Pubblicazione: (2024)
Scaling and Masking: A New Paradigm of Data Sampling for Image and Video Quality Assessment
di: Liu, Yongxu, et al.
Pubblicazione: (2024)
di: Liu, Yongxu, et al.
Pubblicazione: (2024)
Harnessing Multimodal Large Language Models for Personalized Product Search with Query-aware Refinement
di: Zhang, Beibei, et al.
Pubblicazione: (2025)
di: Zhang, Beibei, et al.
Pubblicazione: (2025)
Continual Multimodal Knowledge Graph Construction
di: Chen, Xiang, et al.
Pubblicazione: (2023)
di: Chen, Xiang, et al.
Pubblicazione: (2023)
SmartFreeEdit: Mask-Free Spatial-Aware Image Editing with Complex Instruction Understanding
di: Sun, Qianqian, et al.
Pubblicazione: (2025)
di: Sun, Qianqian, et al.
Pubblicazione: (2025)
KathDB: Explainable Multimodal Database Management System with Human-AI Collaboration
di: Xiao, Guorui, et al.
Pubblicazione: (2025)
di: Xiao, Guorui, et al.
Pubblicazione: (2025)
CAMSIC: Content-aware Masked Image Modeling Transformer for Stereo Image Compression
di: Zhang, Xinjie, et al.
Pubblicazione: (2024)
di: Zhang, Xinjie, et al.
Pubblicazione: (2024)
SMC++: Masked Learning of Unsupervised Video Semantic Compression
di: Tian, Yuan, et al.
Pubblicazione: (2024)
di: Tian, Yuan, et al.
Pubblicazione: (2024)
SyncLipMAE: Contrastive Masked Pretraining for Audio-Visual Talking-Face Representation
di: Ling, Zeyu, et al.
Pubblicazione: (2025)
di: Ling, Zeyu, et al.
Pubblicazione: (2025)
XY-Cut++: Advanced Layout Ordering via Hierarchical Mask Mechanism on a Novel Benchmark
di: Liu, Shuai, et al.
Pubblicazione: (2025)
di: Liu, Shuai, et al.
Pubblicazione: (2025)
Generalizable Deepfake Detection Based on Forgery-aware Layer Masking and Multi-artifact Subspace Decomposition
di: Zhang, Xiang, et al.
Pubblicazione: (2026)
di: Zhang, Xiang, et al.
Pubblicazione: (2026)
Multi-Objective Agentic Rewrites for Unstructured Data Processing
di: Wei, Lindsey Linxi, et al.
Pubblicazione: (2025)
di: Wei, Lindsey Linxi, et al.
Pubblicazione: (2025)
Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots
di: Zheng, Guangting, et al.
Pubblicazione: (2025)
di: Zheng, Guangting, et al.
Pubblicazione: (2025)
StegoHound: A Novel Multi-Approaches Method for Efficient and Effective Identification and Extraction of Digital Evidence Masked by Steganographic Techniques in WAV and MP3 Files
di: Ghanem, Mohamed C., et al.
Pubblicazione: (2023)
di: Ghanem, Mohamed C., et al.
Pubblicazione: (2023)
COutfitGAN: Learning to Synthesize Compatible Outfits Supervised by Silhouette Masks and Fashion Styles
di: Zhou, Dongliang, et al.
Pubblicazione: (2025)
di: Zhou, Dongliang, et al.
Pubblicazione: (2025)
Progressive Confident Masking Attention Network for Audio-Visual Segmentation
di: Wang, Yuxuan, et al.
Pubblicazione: (2024)
di: Wang, Yuxuan, et al.
Pubblicazione: (2024)
Probabilistic Temporal Masked Attention for Cross-view Online Action Detection
di: Xie, Liping, et al.
Pubblicazione: (2025)
di: Xie, Liping, et al.
Pubblicazione: (2025)
MU-MAE: Multimodal Masked Autoencoders-Based One-Shot Learning
di: Liu, Rex, et al.
Pubblicazione: (2024)
di: Liu, Rex, et al.
Pubblicazione: (2024)
MaskAnyone Toolkit: Offering Strategies for Minimizing Privacy Risks and Maximizing Utility in Audio-Visual Data Archiving
di: Owoyele, Babajide Alamu, et al.
Pubblicazione: (2024)
di: Owoyele, Babajide Alamu, et al.
Pubblicazione: (2024)
Self-supervised Spatio-Temporal Graph Mask-Passing Attention Network for Perceptual Importance Prediction of Multi-point Tactility
di: He, Dazhong, et al.
Pubblicazione: (2024)
di: He, Dazhong, et al.
Pubblicazione: (2024)
Rethinking Masking Strategies for Masked Prediction-based Audio Self-supervised Learning
di: Niizumi, Daisuke, et al.
Pubblicazione: (2026)
di: Niizumi, Daisuke, et al.
Pubblicazione: (2026)
PAME: Self-Supervised Masked Autoencoder for No-Reference Point Cloud Quality Assessment
di: Shan, Ziyu, et al.
Pubblicazione: (2024)
di: Shan, Ziyu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
MaskSearch: Querying Image Masks at Scale
di: He, Dong, et al.
Pubblicazione: (2023) -
RACOON: An LLM-based Framework for Retrieval-Augmented Column Type Annotation with a Knowledge Graph
di: Wei, Lindsey Linxi, et al.
Pubblicazione: (2024) -
NeedleDB: A Generative-AI Based System for Accurate and Efficient Image Retrieval using Complex Natural Language Queries
di: Erfanian, Mahdi, et al.
Pubblicazione: (2026) -
Vidformer: Drop-in Declarative Optimization for Rendering Video-Native Query Results
di: Winecki, Dominik, et al.
Pubblicazione: (2026) -
A Review of Media Copyright Management using Blockchain Technologies from the Academic and Business Perspectives
di: García, Roberto, et al.
Pubblicazione: (2023)