VisualChef: Generating Visual Aids in Cooking via Mask Inpainting
Fuente:
arXiv
Saved in:
| Main Authors: | Kuzyk, Oleh, Li, Zuoyue, Pollefeys, Marc, Wang, Xi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sat2Scene: 3D Urban Scene Generation from Satellite Images with Diffusion
by: Li, Zuoyue, et al.
Published: (2024)
by: Li, Zuoyue, et al.
Published: (2024)
From Inpainting to Editing: Unlocking Robust Mask-Free Visual Dubbing via Generative Bootstrapping
by: He, Xu, et al.
Published: (2025)
by: He, Xu, et al.
Published: (2025)
OpenFrontier: General Navigation with Visual-Language Grounded Frontiers
by: Padilla-Cerdio, Esteban, et al.
Published: (2026)
by: Padilla-Cerdio, Esteban, et al.
Published: (2026)
LEAP-VO: Long-term Effective Any Point Tracking for Visual Odometry
by: Chen, Weirong, et al.
Published: (2024)
by: Chen, Weirong, et al.
Published: (2024)
CLII: Visual-Text Inpainting via Cross-Modal Predictive Interaction
by: Zhao, Liang, et al.
Published: (2024)
by: Zhao, Liang, et al.
Published: (2024)
Chain-of-Cooking:Cooking Process Visualization via Bidirectional Chain-of-Thought Guidance
by: Xu, Mengling, et al.
Published: (2025)
by: Xu, Mengling, et al.
Published: (2025)
ImLoc: Revisiting Visual Localization with Image-based Representation
by: Jiang, Xudong, et al.
Published: (2026)
by: Jiang, Xudong, et al.
Published: (2026)
SemanticMIM: Marring Masked Image Modeling with Semantics Compression for General Visual Representation
by: Yuan, Yike, et al.
Published: (2024)
by: Yuan, Yike, et al.
Published: (2024)
Visual Jenga: Discovering Object Dependencies via Counterfactual Inpainting
by: Bhattad, Anand, et al.
Published: (2025)
by: Bhattad, Anand, et al.
Published: (2025)
FreeInpaint: Tuning-free Prompt Alignment and Visual Rationality Enhancement in Image Inpainting
by: Gong, Chao, et al.
Published: (2025)
by: Gong, Chao, et al.
Published: (2025)
R-SCoRe: Revisiting Scene Coordinate Regression for Robust Large-Scale Visual Localization
by: Jiang, Xudong, et al.
Published: (2025)
by: Jiang, Xudong, et al.
Published: (2025)
Generalize Polyp Segmentation via Inpainting across Diverse Backgrounds and Pseudo-Mask Refinement
by: Ma, Jiajian, et al.
Published: (2024)
by: Ma, Jiajian, et al.
Published: (2024)
Active Visual Localization for Multi-Agent Collaboration: A Data-Driven Approach
by: Hanlon, Matthew, et al.
Published: (2023)
by: Hanlon, Matthew, et al.
Published: (2023)
VIGS-SLAM: Visual Inertial Gaussian Splatting SLAM
by: Zhu, Zihan, et al.
Published: (2025)
by: Zhu, Zihan, et al.
Published: (2025)
Controllable Egocentric Video Generation via Occlusion-Aware Sparse 3D Hand Joints
by: Zhang, Chenyangguang, et al.
Published: (2026)
by: Zhang, Chenyangguang, et al.
Published: (2026)
FrontierNet: Learning Visual Cues to Explore
by: Sun, Boyang, et al.
Published: (2025)
by: Sun, Boyang, et al.
Published: (2025)
Geometry of the Visual Cortex with Applications to Image Inpainting and Enhancement
by: Ballerin, Francesco, et al.
Published: (2023)
by: Ballerin, Francesco, et al.
Published: (2023)
E-Commerce Inpainting with Mask Guidance in Controlnet for Reducing Overcompletion
by: Li, Guandong
Published: (2024)
by: Li, Guandong
Published: (2024)
InstaInpaint: Instant 3D-Scene Inpainting with Masked Large Reconstruction Model
by: You, Junqi, et al.
Published: (2025)
by: You, Junqi, et al.
Published: (2025)
Inpainting-Driven Mask Optimization for Object Removal
by: Shimosato, Kodai, et al.
Published: (2024)
by: Shimosato, Kodai, et al.
Published: (2024)
Improving Adversarial Robustness via Decoupled Visual Representation Masking
by: Liu, Decheng, et al.
Published: (2024)
by: Liu, Decheng, et al.
Published: (2024)
SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs
by: Miao, Yang, et al.
Published: (2024)
by: Miao, Yang, et al.
Published: (2024)
ChefFusion: Multimodal Foundation Model Integrating Recipe and Food Image Generation
by: Li, Peiyu, et al.
Published: (2024)
by: Li, Peiyu, et al.
Published: (2024)
Face Mask Removal with Region-attentive Face Inpainting
by: Yang, Minmin
Published: (2024)
by: Yang, Minmin
Published: (2024)
DCText: Scheduled Attention Masking for Visual Text Generation via Divide-and-Conquer Strategy
by: Song, Jaewoo, et al.
Published: (2025)
by: Song, Jaewoo, et al.
Published: (2025)
MR.NAVI: Mixed-Reality Navigation Assistant for the Visually Impaired
by: Pfitzer, Nicolas, et al.
Published: (2025)
by: Pfitzer, Nicolas, et al.
Published: (2025)
Token Painter: Training-Free Text-Guided Image Inpainting via Mask Autoregressive Models
by: Jiang, Longtao, et al.
Published: (2025)
by: Jiang, Longtao, et al.
Published: (2025)
Prompt-Guided Image Editing with Masked Logit Nudging in Visual Autoregressive Models
by: El-Ghoussani, Amir, et al.
Published: (2026)
by: El-Ghoussani, Amir, et al.
Published: (2026)
MTADiffusion: Mask Text Alignment Diffusion Model for Object Inpainting
by: Huang, Jun, et al.
Published: (2025)
by: Huang, Jun, et al.
Published: (2025)
Marten: Visual Question Answering with Mask Generation for Multi-modal Document Understanding
by: Wang, Zining, et al.
Published: (2025)
by: Wang, Zining, et al.
Published: (2025)
DWIM: Towards Tool-aware Visual Reasoning via Discrepancy-aware Workflow Generation & Instruct-Masking Tuning
by: Ke, Fucai, et al.
Published: (2025)
by: Ke, Fucai, et al.
Published: (2025)
Multi Activity Sequence Alignment via Implicit Clustering
by: Kwon, Taein, et al.
Published: (2025)
by: Kwon, Taein, et al.
Published: (2025)
I Dream My Painting: Connecting MLLMs and Diffusion Models via Prompt Generation for Text-Guided Multi-Mask Inpainting
by: Fanelli, Nicola, et al.
Published: (2024)
by: Fanelli, Nicola, et al.
Published: (2024)
MICDrop: Masking Image and Depth Features via Complementary Dropout for Domain-Adaptive Semantic Segmentation
by: Yang, Linyan, et al.
Published: (2024)
by: Yang, Linyan, et al.
Published: (2024)
Real-Time Cooked Food Image Synthesis and Visual Cooking Progress Monitoring on Edge Devices
by: Gupta, Jigyasa, et al.
Published: (2025)
by: Gupta, Jigyasa, et al.
Published: (2025)
CAR: Controllable Autoregressive Modeling for Visual Generation
by: Yao, Ziyu, et al.
Published: (2024)
by: Yao, Ziyu, et al.
Published: (2024)
CSF-Net: Context-Semantic Fusion Network for Large Mask Inpainting
by: Heo, Chae-Yeon, et al.
Published: (2025)
by: Heo, Chae-Yeon, et al.
Published: (2025)
Robust 3D Brain MRI Inpainting with Random Masking Augmentation
by: Zhang, Juexin, et al.
Published: (2025)
by: Zhang, Juexin, et al.
Published: (2025)
Masked Diffusion Captioning for Visual Feature Learning
by: Feng, Chao, et al.
Published: (2025)
by: Feng, Chao, et al.
Published: (2025)
Benchmarking Egocentric Visual-Inertial SLAM at City Scale
by: Krishnan, Anusha, et al.
Published: (2025)
by: Krishnan, Anusha, et al.
Published: (2025)
Similar Items
-
Sat2Scene: 3D Urban Scene Generation from Satellite Images with Diffusion
by: Li, Zuoyue, et al.
Published: (2024) -
From Inpainting to Editing: Unlocking Robust Mask-Free Visual Dubbing via Generative Bootstrapping
by: He, Xu, et al.
Published: (2025) -
OpenFrontier: General Navigation with Visual-Language Grounded Frontiers
by: Padilla-Cerdio, Esteban, et al.
Published: (2026) -
LEAP-VO: Long-term Effective Any Point Tracking for Visual Odometry
by: Chen, Weirong, et al.
Published: (2024) -
CLII: Visual-Text Inpainting via Cross-Modal Predictive Interaction
by: Zhao, Liang, et al.
Published: (2024)