CookAnything: A Framework for Flexible and Consistent Multi-Step Recipe Image Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Ruoxuan, Wen, Bin, Xie, Hongxia, Yao, Yi, Zuo, Songhan, Jiang-Lin, Jian-Yu, Shuai, Hong-Han, Cheng, Wen-Huang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
RecipeGen: A Step-Aligned Multimodal Benchmark for Real-World Recipe Generation
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025)
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025)
RecipeGen: A Benchmark for Real-World Recipe Image Generation
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025)
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025)
EmoArt: A Multidimensional Dataset for Emotion-Aware Artistic Generation
di: Zhang, Cheng, et al.
Pubblicazione: (2025)
di: Zhang, Cheng, et al.
Pubblicazione: (2025)
Aligning Progress and Feasibility: A Neuro-Symbolic Dual Memory Framework for Long-Horizon LLM Agents
di: Wen, Bin, et al.
Pubblicazione: (2026)
di: Wen, Bin, et al.
Pubblicazione: (2026)
Perspective-Aware Teaching: Adapting Knowledge for Heterogeneous Distillation
di: Lin, Jhe-Hao, et al.
Pubblicazione: (2025)
di: Lin, Jhe-Hao, et al.
Pubblicazione: (2025)
MindPower: Enabling Theory-of-Mind Reasoning in VLM-based Embodied Agents
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025)
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025)
Future Sight and Tough Fights: Revolutionizing Sequential Recommendation with FENRec
di: Huang, Yu-Hsuan, et al.
Pubblicazione: (2024)
di: Huang, Yu-Hsuan, et al.
Pubblicazione: (2024)
Single Document Image Highlight Removal via A Large-Scale Real-World Dataset and A Location-Aware Network
di: Pan, Lu, et al.
Pubblicazione: (2025)
di: Pan, Lu, et al.
Pubblicazione: (2025)
PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos
di: Zhou, Zhiyu, et al.
Pubblicazione: (2026)
di: Zhou, Zhiyu, et al.
Pubblicazione: (2026)
Quick No-Cook Recipes / Melissa Clark
di: Clark, Melissa
Pubblicazione: (2003)
di: Clark, Melissa
Pubblicazione: (2003)
Generalized Step-Chirp Sequences With Flexible Bandwidth
di: Du, Cheng, et al.
Pubblicazione: (2024)
di: Du, Cheng, et al.
Pubblicazione: (2024)
Beyond Success: Refining Elegant Robot Manipulation from Mixed-Quality Data via Just-in-Time Intervention
di: Mao, Yanbo, et al.
Pubblicazione: (2025)
di: Mao, Yanbo, et al.
Pubblicazione: (2025)
PizzaCommonSense: Learning to Model Commonsense Reasoning about Intermediate Steps in Cooking Recipes
di: Diallo, Aissatou, et al.
Pubblicazione: (2024)
di: Diallo, Aissatou, et al.
Pubblicazione: (2024)
MindClaw: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2026)
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2026)
OSCAR: Object Status and Contextual Awareness for Recipes to Support Non-Visual Cooking
di: Li, Franklin Mingzhe, et al.
Pubblicazione: (2025)
di: Li, Franklin Mingzhe, et al.
Pubblicazione: (2025)
Exploring Object Status Recognition for Recipe Progress Tracking in Non-Visual Cooking
di: Li, Franklin Mingzhe, et al.
Pubblicazione: (2025)
di: Li, Franklin Mingzhe, et al.
Pubblicazione: (2025)
Recipe Generation from Unsegmented Cooking Videos
di: Nishimura, Taichi, et al.
Pubblicazione: (2022)
di: Nishimura, Taichi, et al.
Pubblicazione: (2022)
The Fabrication of Reality and Fantasy: Scene Generation with LLM-Assisted Prompt Interpretation
di: Yao, Yi, et al.
Pubblicazione: (2024)
di: Yao, Yi, et al.
Pubblicazione: (2024)
Cooking Up World History: Multicultural Recipes and Resources.
di: Marden, Patricia C., et al.
Pubblicazione: (1994)
di: Marden, Patricia C., et al.
Pubblicazione: (1994)
Foreground Focus: Enhancing Coherence and Fidelity in Camouflaged Image Generation
di: Chen, Pei-Chi, et al.
Pubblicazione: (2025)
di: Chen, Pei-Chi, et al.
Pubblicazione: (2025)
CookingDiffusion: Cooking Procedural Image Generation with Stable Diffusion
di: Wang, Yuan, et al.
Pubblicazione: (2025)
di: Wang, Yuan, et al.
Pubblicazione: (2025)
Cook2LTL: Translating Cooking Recipes to LTL Formulae using Large Language Models
di: Mavrogiannis, Angelos, et al.
Pubblicazione: (2023)
di: Mavrogiannis, Angelos, et al.
Pubblicazione: (2023)
EmoVIT: Revolutionizing Emotion Insights with Visual Instruction Tuning
di: Xie, Hongxia, et al.
Pubblicazione: (2024)
di: Xie, Hongxia, et al.
Pubblicazione: (2024)
Using LLMs to Extract Food Entities from Cooking Recipes
di: Pitsilou, Vasiliki, et al.
Pubblicazione: (2024)
di: Pitsilou, Vasiliki, et al.
Pubblicazione: (2024)
Losses that Cook: Topological Optimal Transport for Structured Recipe Generation
di: Ottoborgo, Mattia, et al.
Pubblicazione: (2026)
di: Ottoborgo, Mattia, et al.
Pubblicazione: (2026)
Cooking Up U. S. History: Recipes and Research to Share with Children.
di: Barchers, Suzanne I., et al.
Pubblicazione: (1991)
di: Barchers, Suzanne I., et al.
Pubblicazione: (1991)
DataChef: Cooking Up Optimal Data Recipes for LLM Adaptation via Reinforcement Learning
di: Chen, Yicheng, et al.
Pubblicazione: (2026)
di: Chen, Yicheng, et al.
Pubblicazione: (2026)
Multi-rater Prompting for Ambiguous Medical Image Segmentation
di: Wang, Jinhong, et al.
Pubblicazione: (2024)
di: Wang, Jinhong, et al.
Pubblicazione: (2024)
Insert Anything: Image Insertion via In-Context Editing in DiT
di: Song, Wensong, et al.
Pubblicazione: (2025)
di: Song, Wensong, et al.
Pubblicazione: (2025)
Tiny-YOLOSAM: Fast Hybrid Image Segmentation
di: Xu, Kenneth, et al.
Pubblicazione: (2025)
di: Xu, Kenneth, et al.
Pubblicazione: (2025)
FastDrag: Manipulate Anything in One Step
di: Zhao, Xuanjia, et al.
Pubblicazione: (2024)
di: Zhao, Xuanjia, et al.
Pubblicazione: (2024)
Say Anything with Any Style
di: Tan, Shuai, et al.
Pubblicazione: (2024)
di: Tan, Shuai, et al.
Pubblicazione: (2024)
MambaReg: Mamba-Based Disentangled Convolutional Sparse Coding for Unsupervised Deformable Multi-Modal Image Registration
di: Wen, Kaiang, et al.
Pubblicazione: (2024)
di: Wen, Kaiang, et al.
Pubblicazione: (2024)
Lightweight Deep Learning for Resource-Constrained Environments: A Survey
di: Liu, Hou-I, et al.
Pubblicazione: (2024)
di: Liu, Hou-I, et al.
Pubblicazione: (2024)
Crossing Boundaries: Leveraging Semantic Divergences to Explore Cultural Novelty in Cooking Recipes
di: Carichon, Florian, et al.
Pubblicazione: (2025)
di: Carichon, Florian, et al.
Pubblicazione: (2025)
Cooking Up U.S. History: Recipes and Research To Share with Children. Second Edition.
di: Barchers, Suzanne I., et al.
Pubblicazione: (1999)
di: Barchers, Suzanne I., et al.
Pubblicazione: (1999)
Traffic Scene Small Target Detection Method Based on YOLOv8n-SPTS Model for Autonomous Driving
di: Wu, Songhan
Pubblicazione: (2025)
di: Wu, Songhan
Pubblicazione: (2025)
One Pool Is Not Enough: Multi-Cluster Memory for Practical Test-Time Adaptation
di: Tseng, Yu-Wen, et al.
Pubblicazione: (2026)
di: Tseng, Yu-Wen, et al.
Pubblicazione: (2026)
A Recipe for Success: Co‐Design and Operationalisation of a Regional Research Capacity Building Programme Using Cooke's Framework
di: Tracy Flenady, et al.
Pubblicazione: (2026)
di: Tracy Flenady, et al.
Pubblicazione: (2026)
Towards Multi-View Consistent Style Transfer with One-Step Diffusion via Vision Conditioning
di: Zuo, Yushen, et al.
Pubblicazione: (2024)
di: Zuo, Yushen, et al.
Pubblicazione: (2024)
Documenti analoghi
-
RecipeGen: A Step-Aligned Multimodal Benchmark for Real-World Recipe Generation
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025) -
RecipeGen: A Benchmark for Real-World Recipe Image Generation
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025) -
EmoArt: A Multidimensional Dataset for Emotion-Aware Artistic Generation
di: Zhang, Cheng, et al.
Pubblicazione: (2025) -
Aligning Progress and Feasibility: A Neuro-Symbolic Dual Memory Framework for Long-Horizon LLM Agents
di: Wen, Bin, et al.
Pubblicazione: (2026) -
Perspective-Aware Teaching: Adapting Knowledge for Heterogeneous Distillation
di: Lin, Jhe-Hao, et al.
Pubblicazione: (2025)