Sketch Input Method Editor: A Comprehensive Dataset and Methodology for Systematic Input Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Guangming, Wang, Siyuan, Cheng, Qing, Wu, Kelong, Li, Hao, Zhang, Liang |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Content-Conditioned Generation of Stylized Free hand Sketches
by: Liu, Jiajun, et al.
Published: (2024)
by: Liu, Jiajun, et al.
Published: (2024)
Sketch and Text Synergy: Fusing Structural Contours and Descriptive Attributes for Fine-Grained Image Retrieval
by: Wang, Siyuan, et al.
Published: (2026)
by: Wang, Siyuan, et al.
Published: (2026)
Survey on Hand Gesture Recognition from Visual Input
by: Linardakis, Manousos, et al.
Published: (2025)
by: Linardakis, Manousos, et al.
Published: (2025)
Holistic Surgical Phase Recognition with Hierarchical Input Dependent State Space Models
by: Wu, Haoyang, et al.
Published: (2025)
by: Wu, Haoyang, et al.
Published: (2025)
ALERT Open Dataset and Input-Size-Agnostic Vision Transformer for Driver Activity Recognition using IR-UWB
by: Park, Jeongjun, et al.
Published: (2025)
by: Park, Jeongjun, et al.
Published: (2025)
Deep Learning in Dental Image Analysis: A Systematic Review of Datasets, Methodologies, and Emerging Challenges
by: Zhou, Zhenhuan, et al.
Published: (2025)
by: Zhou, Zhenhuan, et al.
Published: (2025)
Representation Learning of Point Cloud Upsampling in Global and Local Inputs
by: Zhang, Tongxu, et al.
Published: (2025)
by: Zhang, Tongxu, et al.
Published: (2025)
ViSketch-GPT: Collaborative Multi-Scale Feature Extraction for Sketch Recognition and Generation
by: Federico, Giulio, et al.
Published: (2025)
by: Federico, Giulio, et al.
Published: (2025)
SpatialDreamer: Self-supervised Stereo Video Synthesis from Monocular Input
by: Lv, Zhen, et al.
Published: (2024)
by: Lv, Zhen, et al.
Published: (2024)
FreeCond: Free Lunch in the Input Conditions of Text-Guided Inpainting
by: Hsiao, Teng-Fang, et al.
Published: (2024)
by: Hsiao, Teng-Fang, et al.
Published: (2024)
Improving the Transferability of Adversarial Attacks by an Input Transpose
by: Wan, Qing, et al.
Published: (2025)
by: Wan, Qing, et al.
Published: (2025)
Discovering Hidden Visual Concepts Beyond Linguistic Input in Infant Learning
by: Ke, Xueyi, et al.
Published: (2025)
by: Ke, Xueyi, et al.
Published: (2025)
Improving Diagnostic Performance on Small and Imbalanced Datasets Using Class-Based Input Image Composition
by: Azzeddine, Hlali, et al.
Published: (2025)
by: Azzeddine, Hlali, et al.
Published: (2025)
SketchGraphNet: A Memory-Efficient Hybrid Graph Transformer for Large-Scale Sketch Corpora Recognition
by: Chen, Shilong, et al.
Published: (2026)
by: Chen, Shilong, et al.
Published: (2026)
VidSketch: Hand-drawn Sketch-Driven Video Generation with Diffusion Control
by: Jiang, Lifan, et al.
Published: (2025)
by: Jiang, Lifan, et al.
Published: (2025)
Towards Attributions of Input Variables in a Coalition
by: Zheng, Xinhao, et al.
Published: (2023)
by: Zheng, Xinhao, et al.
Published: (2023)
FlexGen: Flexible Multi-View Generation from Text and Image Inputs
by: Xu, Xinli, et al.
Published: (2024)
by: Xu, Xinli, et al.
Published: (2024)
SketchRef: a Multi-Task Evaluation Benchmark for Sketch Synthesis
by: Lin, Xingyue, et al.
Published: (2024)
by: Lin, Xingyue, et al.
Published: (2024)
Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input
by: Liu, Jiajun, et al.
Published: (2024)
by: Liu, Jiajun, et al.
Published: (2024)
Learning Content-Aware Multi-Modal Joint Input Pruning via Bird's-Eye-View Representation
by: Li, Yuxin, et al.
Published: (2024)
by: Li, Yuxin, et al.
Published: (2024)
Crafter: A Multi-Agent Harness for Editable Scientific Figure Generation from Diverse Inputs
by: Zhao, Haozhe, et al.
Published: (2026)
by: Zhao, Haozhe, et al.
Published: (2026)
OmniEval: A Benchmark for Evaluating Omni-modal Models with Visual, Auditory, and Textual Inputs
by: Zhang, Yiman, et al.
Published: (2025)
by: Zhang, Yiman, et al.
Published: (2025)
VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models
by: Balakrishnan, Ravikumar, et al.
Published: (2025)
by: Balakrishnan, Ravikumar, et al.
Published: (2025)
Development of a Dual-Input Neural Model for Detecting AI-Generated Imagery
by: Gallagher, Jonathan, et al.
Published: (2024)
by: Gallagher, Jonathan, et al.
Published: (2024)
VISOR: Visual Input-based Steering for Output Redirection in Vision-Language Models
by: Phute, Mansi, et al.
Published: (2025)
by: Phute, Mansi, et al.
Published: (2025)
Robustness of Vision Language Models Against Split-Image Harmful Input Attacks
by: Rashid, Md Rafi Ur, et al.
Published: (2026)
by: Rashid, Md Rafi Ur, et al.
Published: (2026)
Unbiased Visual Reasoning with Controlled Visual Inputs
by: Li, Zhaonan, et al.
Published: (2025)
by: Li, Zhaonan, et al.
Published: (2025)
Input-Adaptive Generative Dynamics in Diffusion Models
by: Xing, Yucheng, et al.
Published: (2024)
by: Xing, Yucheng, et al.
Published: (2024)
InfiniteVL: Synergizing Linear and Sparse Attention for Highly-Efficient, Unlimited-Input Vision-Language Models
by: Tao, Hongyuan, et al.
Published: (2025)
by: Tao, Hongyuan, et al.
Published: (2025)
Mitigating the Reasoning Tax in Vision-Language Fine-Tuning with Input-Adaptive Depth Aggregation
by: Ren, Yiming, et al.
Published: (2026)
by: Ren, Yiming, et al.
Published: (2026)
CoProSketch: Controllable and Progressive Sketch Generation with Diffusion Model
by: Zhan, Ruohao, et al.
Published: (2025)
by: Zhan, Ruohao, et al.
Published: (2025)
Towards Robust Object Detection: Identifying and Removing Backdoors via Module Inconsistency Analysis
by: Zhang, Xianda, et al.
Published: (2024)
by: Zhang, Xianda, et al.
Published: (2024)
Sketch-to-Layout: Sketch-Guided Multimodal Layout Generation
by: Brioschi, Riccardo, et al.
Published: (2025)
by: Brioschi, Riccardo, et al.
Published: (2025)
Empowering Backbone Models for Visual Text Generation with Input Granularity Control and Glyph-Aware Training
by: Li, Wenbo, et al.
Published: (2024)
by: Li, Wenbo, et al.
Published: (2024)
Robust Dynamic Facial Expression Recognition
by: Liu, Feng, et al.
Published: (2025)
by: Liu, Feng, et al.
Published: (2025)
InterSketch: An Interleaved Reasoning Model with Self-correcting Visual Sketch and Stepwise Reward
by: Ning, Zhiwei, et al.
Published: (2026)
by: Ning, Zhiwei, et al.
Published: (2026)
CMOSE: Comprehensive Multi-Modality Online Student Engagement Dataset with High-Quality Labels
by: Wu, Chi-hsuan, et al.
Published: (2023)
by: Wu, Chi-hsuan, et al.
Published: (2023)
MROSS: Multi-Round Region-based Optimization for Scene Sketching
by: Liang, Yiqi, et al.
Published: (2024)
by: Liang, Yiqi, et al.
Published: (2024)
LostPaw: Finding Lost Pets using a Contrastive Learning-based Transformer with Visual Input
by: Voinea, Andrei, et al.
Published: (2023)
by: Voinea, Andrei, et al.
Published: (2023)
Single-Input Multi-Output Model Merging: Leveraging Foundation Models for Dense Multi-Task Learning
by: Giraldo, Juan Garcia, et al.
Published: (2025)
by: Giraldo, Juan Garcia, et al.
Published: (2025)
Similar Items
-
Content-Conditioned Generation of Stylized Free hand Sketches
by: Liu, Jiajun, et al.
Published: (2024) -
Sketch and Text Synergy: Fusing Structural Contours and Descriptive Attributes for Fine-Grained Image Retrieval
by: Wang, Siyuan, et al.
Published: (2026) -
Survey on Hand Gesture Recognition from Visual Input
by: Linardakis, Manousos, et al.
Published: (2025) -
Holistic Surgical Phase Recognition with Hierarchical Input Dependent State Space Models
by: Wu, Haoyang, et al.
Published: (2025) -
ALERT Open Dataset and Input-Size-Agnostic Vision Transformer for Driver Activity Recognition using IR-UWB
by: Park, Jeongjun, et al.
Published: (2025)