SketchThinker-R1: Towards Efficient Sketch-Style Reasoning in Large Multimodal Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Ruiyang, Zhou, Dongzhan, Zheng, Zhedong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation
by: Zhang, Ruiyang, et al.
Published: (2024)
by: Zhang, Ruiyang, et al.
Published: (2024)
Sketch3D: Style-Consistent Guidance for Sketch-to-3D Generation
by: Zheng, Wangguandong, et al.
Published: (2024)
by: Zheng, Wangguandong, et al.
Published: (2024)
Uncertainty-o: One Model-agnostic Framework for Unveiling Uncertainty in Large Multimodal Models
by: Zhang, Ruiyang, et al.
Published: (2025)
by: Zhang, Ruiyang, et al.
Published: (2025)
Sketch Down the FLOPs: Towards Efficient Networks for Human Sketch
by: Sain, Aneeshan, et al.
Published: (2025)
by: Sain, Aneeshan, et al.
Published: (2025)
AutoSketch: VLM-assisted Style-Aware Vector Sketch Completion
by: Chin, Hsiao-Yuan, et al.
Published: (2025)
by: Chin, Hsiao-Yuan, et al.
Published: (2025)
Text to Sketch Generation with Multi-Styles
by: Li, Tengjie, et al.
Published: (2025)
by: Li, Tengjie, et al.
Published: (2025)
Sketch-to-Layout: Sketch-Guided Multimodal Layout Generation
by: Brioschi, Riccardo, et al.
Published: (2025)
by: Brioschi, Riccardo, et al.
Published: (2025)
DiffThinker: Towards Generative Multimodal Reasoning with Diffusion Models
by: He, Zefeng, et al.
Published: (2025)
by: He, Zefeng, et al.
Published: (2025)
Multi-Style Facial Sketch Synthesis through Masked Generative Modeling
by: Sun, Bowen, et al.
Published: (2024)
by: Sun, Bowen, et al.
Published: (2024)
SketchAgent: Language-Driven Sequential Sketch Generation
by: Vinker, Yael, et al.
Published: (2024)
by: Vinker, Yael, et al.
Published: (2024)
SketchGPT: Autoregressive Modeling for Sketch Generation and Recognition
by: Tiwari, Adarsh, et al.
Published: (2024)
by: Tiwari, Adarsh, et al.
Published: (2024)
InterSketch: An Interleaved Reasoning Model with Self-correcting Visual Sketch and Stepwise Reward
by: Ning, Zhiwei, et al.
Published: (2026)
by: Ning, Zhiwei, et al.
Published: (2026)
Latent Sketchpad: Sketching Visual Thoughts to Elicit Multimodal Reasoning in MLLMs
by: Zhang, Huanyu, et al.
Published: (2025)
by: Zhang, Huanyu, et al.
Published: (2025)
VSearcher: Long-Horizon Multimodal Search Agent via Reinforcement Learning
by: Zhang, Ruiyang, et al.
Published: (2026)
by: Zhang, Ruiyang, et al.
Published: (2026)
It's All About Your Sketch: Democratising Sketch Control in Diffusion Models
by: Koley, Subhadeep, et al.
Published: (2024)
by: Koley, Subhadeep, et al.
Published: (2024)
SwiftSketch: A Diffusion Model for Image-to-Vector Sketch Generation
by: Arar, Ellie, et al.
Published: (2025)
by: Arar, Ellie, et al.
Published: (2025)
SketchAnimator: Animate Sketch via Motion Customization of Text-to-Video Diffusion Models
by: Yang, Ruolin, et al.
Published: (2025)
by: Yang, Ruolin, et al.
Published: (2025)
Sketch2Anim: Towards Transferring Sketch Storyboards into 3D Animation
by: Zhong, Lei, et al.
Published: (2025)
by: Zhong, Lei, et al.
Published: (2025)
PS-StyleGAN: Illustrative Portrait Sketching using Attention-Based Style Adaptation
by: Jain, Kushal Kumar, et al.
Published: (2024)
by: Jain, Kushal Kumar, et al.
Published: (2024)
Approaching Outside: Scaling Unsupervised 3D Object Detection from 2D Scene
by: Zhang, Ruiyang, et al.
Published: (2024)
by: Zhang, Ruiyang, et al.
Published: (2024)
Harnessing Uncertainty-aware Bounding Boxes for Unsupervised 3D Object Detection
by: Zhang, Ruiyang, et al.
Published: (2024)
by: Zhang, Ruiyang, et al.
Published: (2024)
U-Sketch: An Efficient Approach for Sketch to Image Diffusion Models
by: Mitsouras, Ilias, et al.
Published: (2024)
by: Mitsouras, Ilias, et al.
Published: (2024)
Sketch-1-to-3: One Single Sketch to 3D Detailed Face Reconstruction
by: Wen, Liting, et al.
Published: (2025)
by: Wen, Liting, et al.
Published: (2025)
Sketch and Refine: Towards Fast and Accurate Lane Detection
by: Chen, Chao, et al.
Published: (2024)
by: Chen, Chao, et al.
Published: (2024)
CustomSketching: Sketch Concept Extraction for Sketch-based Image Synthesis and Editing
by: Xiao, Chufeng, et al.
Published: (2024)
by: Xiao, Chufeng, et al.
Published: (2024)
Sketch2PoseNet: Efficient and Generalized Sketch to 3D Human Pose Prediction
by: Wang, Li, et al.
Published: (2025)
by: Wang, Li, et al.
Published: (2025)
SketchFusion: Learning Universal Sketch Features through Fusing Foundation Models
by: Koley, Subhadeep, et al.
Published: (2025)
by: Koley, Subhadeep, et al.
Published: (2025)
From Sketch to Fresco: Efficient Diffusion Transformer with Progressive Resolution
by: Zheng, Shikang, et al.
Published: (2026)
by: Zheng, Shikang, et al.
Published: (2026)
SketchVideo: Sketch-based Video Generation and Editing
by: Liu, Feng-Lin, et al.
Published: (2025)
by: Liu, Feng-Lin, et al.
Published: (2025)
Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs
by: Tong, Jintao, et al.
Published: (2025)
by: Tong, Jintao, et al.
Published: (2025)
SketchingReality: From Freehand Scene Sketches To Photorealistic Images
by: Bourouis, Ahmed, et al.
Published: (2026)
by: Bourouis, Ahmed, et al.
Published: (2026)
How to Handle Sketch-Abstraction in Sketch-Based Image Retrieval?
by: Koley, Subhadeep, et al.
Published: (2024)
by: Koley, Subhadeep, et al.
Published: (2024)
Sketch2Arti: Sketch-based Articulation Modeling of CAD Objects
by: Yang, Yi, et al.
Published: (2026)
by: Yang, Yi, et al.
Published: (2026)
Towards Interactive Image Inpainting via Sketch Refinement
by: Liu, Chang, et al.
Published: (2023)
by: Liu, Chang, et al.
Published: (2023)
CoProSketch: Controllable and Progressive Sketch Generation with Diffusion Model
by: Zhan, Ruohao, et al.
Published: (2025)
by: Zhan, Ruohao, et al.
Published: (2025)
SketchGraphNet: A Memory-Efficient Hybrid Graph Transformer for Large-Scale Sketch Corpora Recognition
by: Chen, Shilong, et al.
Published: (2026)
by: Chen, Shilong, et al.
Published: (2026)
SketchDeco: Training-Free Latent Composition for Precise Sketch Colourisation
by: Utintu, Chaitat, et al.
Published: (2024)
by: Utintu, Chaitat, et al.
Published: (2024)
StrandDesigner: Towards Practical Strand Generation with Sketch Guidance
by: Zhang, Na, et al.
Published: (2025)
by: Zhang, Na, et al.
Published: (2025)
Let Human Sketches Help: Empowering Challenging Image Segmentation Task with Freehand Sketches
by: Zang, Ying, et al.
Published: (2025)
by: Zang, Ying, et al.
Published: (2025)
AirSketch: Generative Motion to Sketch
by: Lim, Hui Xian Grace, et al.
Published: (2024)
by: Lim, Hui Xian Grace, et al.
Published: (2024)
Similar Items
-
VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation
by: Zhang, Ruiyang, et al.
Published: (2024) -
Sketch3D: Style-Consistent Guidance for Sketch-to-3D Generation
by: Zheng, Wangguandong, et al.
Published: (2024) -
Uncertainty-o: One Model-agnostic Framework for Unveiling Uncertainty in Large Multimodal Models
by: Zhang, Ruiyang, et al.
Published: (2025) -
Sketch Down the FLOPs: Towards Efficient Networks for Human Sketch
by: Sain, Aneeshan, et al.
Published: (2025) -
AutoSketch: VLM-assisted Style-Aware Vector Sketch Completion
by: Chin, Hsiao-Yuan, et al.
Published: (2025)