ClickDiff: Click to Induce Semantic Contact Map for Controllable Grasp Generation with Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Peiming, Wang, Ziyi, Liu, Mengyuan, Liu, Hong, Chen, Chen |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ClickVOS: Click Video Object Segmentation
by: Guo, Pinxue, et al.
Published: (2024)
by: Guo, Pinxue, et al.
Published: (2024)
ClickAttention: Click Region Similarity Guided Interactive Segmentation
by: Xu, Long, et al.
Published: (2024)
by: Xu, Long, et al.
Published: (2024)
Lens Privacy Sealing: A New Benchmark and Method for Physical Privacy-Preserving Action Recognition
by: Liu, Mengyuan, et al.
Published: (2026)
by: Liu, Mengyuan, et al.
Published: (2026)
ClickSeg3D: Few-Click Interactive Segmentation via Semantic Embeddings
by: Kang, Xueyang, et al.
Published: (2026)
by: Kang, Xueyang, et al.
Published: (2026)
UST-SSM: Unified Spatio-Temporal State Space Models for Point Cloud Video Modeling
by: Li, Peiming, et al.
Published: (2025)
by: Li, Peiming, et al.
Published: (2025)
Click to Grasp: Zero-Shot Precise Manipulation via Visual Diffusion Descriptors
by: Tsagkas, Nikolaos, et al.
Published: (2024)
by: Tsagkas, Nikolaos, et al.
Published: (2024)
Superman: Unifying Skeleton and Vision for Human Motion Perception and Generation
by: Wang, Xinshun, et al.
Published: (2026)
by: Wang, Xinshun, et al.
Published: (2026)
Universal Skeleton Understanding via Differentiable Rendering and MLLMs
by: Wang, Ziyi, et al.
Published: (2026)
by: Wang, Ziyi, et al.
Published: (2026)
GoClick: Lightweight Element Grounding Model for Autonomous GUI Interaction
by: Li, Hongxin, et al.
Published: (2026)
by: Li, Hongxin, et al.
Published: (2026)
RegionGrasp: A Novel Task for Contact Region Controllable Hand Grasp Generation
by: Wang, Yilin, et al.
Published: (2024)
by: Wang, Yilin, et al.
Published: (2024)
Structured Click Control in Transformer-based Interactive Segmentation
by: Xu, Long, et al.
Published: (2024)
by: Xu, Long, et al.
Published: (2024)
ClickRemoval: An Interactive Open-Source Tool for Object Removal in Diffusion Models
by: Zhang, Ledun, et al.
Published: (2026)
by: Zhang, Ledun, et al.
Published: (2026)
Recognizing Actions from Robotic View for Natural Human-Robot Interaction
by: Wang, Ziyi, et al.
Published: (2025)
by: Wang, Ziyi, et al.
Published: (2025)
Click2Graph: Interactive Panoptic Video Scene Graphs from a Single Click
by: Ruschel, Raphael, et al.
Published: (2025)
by: Ruschel, Raphael, et al.
Published: (2025)
Follow-Your-Click: Open-domain Regional Image Animation via Short Prompts
by: Ma, Yue, et al.
Published: (2024)
by: Ma, Yue, et al.
Published: (2024)
AdaptiveClick: Clicks-aware Transformer with Adaptive Focal Loss for Interactive Image Segmentation
by: Lin, Jiacheng, et al.
Published: (2023)
by: Lin, Jiacheng, et al.
Published: (2023)
Click-to-Ask: An AI Live Streaming Assistant with Offline Copywriting and Online Interactive QA
by: Yu, Ruizhi, et al.
Published: (2026)
by: Yu, Ruizhi, et al.
Published: (2026)
FocalClick-XL: Towards Unified and High-quality Interactive Segmentation
by: Chen, Xi, et al.
Published: (2025)
by: Chen, Xi, et al.
Published: (2025)
Learning Trimaps via Clicks for Image Matting
by: Zhang, Chenyi, et al.
Published: (2024)
by: Zhang, Chenyi, et al.
Published: (2024)
CRS-Diff: Controllable Remote Sensing Image Generation with Diffusion Model
by: Tang, Datao, et al.
Published: (2024)
by: Tang, Datao, et al.
Published: (2024)
DiffMap: Enhancing Map Segmentation with Map Prior Using Diffusion Model
by: Jia, Peijin, et al.
Published: (2024)
by: Jia, Peijin, et al.
Published: (2024)
ClickDiffusion: Harnessing LLMs for Interactive Precise Image Editing
by: Helbling, Alec, et al.
Published: (2024)
by: Helbling, Alec, et al.
Published: (2024)
TextDiff: Mask-Guided Residual Diffusion Models for Scene Text Image Super-Resolution
by: Liu, Baolin, et al.
Published: (2023)
by: Liu, Baolin, et al.
Published: (2023)
ClickTrack: Towards Real-time Interactive Single Object Tracking
by: Wang, Kuiran, et al.
Published: (2024)
by: Wang, Kuiran, et al.
Published: (2024)
Clore: Interactive Pathology Image Segmentation with Click-based Local Refinement
by: Wang, Tiantong, et al.
Published: (2026)
by: Wang, Tiantong, et al.
Published: (2026)
DiffPlace: Street View Generation via Place-Controllable Diffusion Model Enhancing Place Recognition
by: Li, Ji, et al.
Published: (2026)
by: Li, Ji, et al.
Published: (2026)
DualDiff: Dual-branch Diffusion Model for Autonomous Driving with Semantic Fusion
by: Li, Haoteng, et al.
Published: (2025)
by: Li, Haoteng, et al.
Published: (2025)
Click-Calib: A Robust Extrinsic Calibration Method for Surround-View Systems
by: Wang, Lihao
Published: (2025)
by: Wang, Lihao
Published: (2025)
StereoDiff: Stereo-Diffusion Synergy for Video Depth Estimation
by: Li, Haodong, et al.
Published: (2025)
by: Li, Haodong, et al.
Published: (2025)
Diff-BGM: A Diffusion Model for Video Background Music Generation
by: Li, Sizhe, et al.
Published: (2024)
by: Li, Sizhe, et al.
Published: (2024)
PiClick: Picking the desired mask from multiple candidates in click-based interactive segmentation
by: Yan, Cilin, et al.
Published: (2023)
by: Yan, Cilin, et al.
Published: (2023)
Clicks2Line: Using Lines for Interactive Image Segmentation
by: Lee, Chaewon, et al.
Published: (2024)
by: Lee, Chaewon, et al.
Published: (2024)
DiffBench Meets DiffAgent: End-to-End LLM-Driven Diffusion Acceleration Code Generation
by: jiao, Jiajun, et al.
Published: (2026)
by: jiao, Jiajun, et al.
Published: (2026)
UniMotion: A Unified Framework for Motion-Text-Vision Understanding and Generation
by: Wang, Ziyi, et al.
Published: (2026)
by: Wang, Ziyi, et al.
Published: (2026)
DiffDoctor: Diagnosing Image Diffusion Models Before Treating
by: Wang, Yiyang, et al.
Published: (2025)
by: Wang, Yiyang, et al.
Published: (2025)
Zoom in, Click out: Unlocking and Evaluating the Potential of Zooming for GUI Grounding
by: Jiang, Zhiyuan, et al.
Published: (2025)
by: Jiang, Zhiyuan, et al.
Published: (2025)
DiffCalib: Reformulating Monocular Camera Calibration as Diffusion-Based Dense Incident Map Generation
by: He, Xiankang, et al.
Published: (2024)
by: He, Xiankang, et al.
Published: (2024)
DiffFAS: Face Anti-Spoofing via Generative Diffusion Models
by: Ge, Xinxu, et al.
Published: (2024)
by: Ge, Xinxu, et al.
Published: (2024)
DiffVL: Diffusion-Based Visual Localization on 2D Maps via BEV-Conditioned GPS Denoising
by: Gao, Li, et al.
Published: (2025)
by: Gao, Li, et al.
Published: (2025)
DiffSemanticFusion: Semantic Raster BEV Fusion for Autonomous Driving via Online HD Map Diffusion
by: Sun, Zhigang, et al.
Published: (2025)
by: Sun, Zhigang, et al.
Published: (2025)
Similar Items
-
ClickVOS: Click Video Object Segmentation
by: Guo, Pinxue, et al.
Published: (2024) -
ClickAttention: Click Region Similarity Guided Interactive Segmentation
by: Xu, Long, et al.
Published: (2024) -
Lens Privacy Sealing: A New Benchmark and Method for Physical Privacy-Preserving Action Recognition
by: Liu, Mengyuan, et al.
Published: (2026) -
ClickSeg3D: Few-Click Interactive Segmentation via Semantic Embeddings
by: Kang, Xueyang, et al.
Published: (2026) -
UST-SSM: Unified Spatio-Temporal State Space Models for Point Cloud Video Modeling
by: Li, Peiming, et al.
Published: (2025)