How Does India Cook Biryani?
Fuente:
arXiv
Salvato in:
| Autori principali: | Goel, Shubham, S, Farzana, Rishi, C V, Arun, Aditya, Jawahar, C V |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PhyEduVideo: A Benchmark for Evaluating Text-to-Video Models for Physics Education
di: M, Megha Mariam K., et al.
Pubblicazione: (2026)
di: M, Megha Mariam K., et al.
Pubblicazione: (2026)
Attend to what I say: Highlighting relevant content on slides
di: M, Megha Mariam K, et al.
Pubblicazione: (2026)
di: M, Megha Mariam K, et al.
Pubblicazione: (2026)
IndicSTR12: A Dataset for Indic Scene Text Recognition
di: Lunia, Harsh, et al.
Pubblicazione: (2024)
di: Lunia, Harsh, et al.
Pubblicazione: (2024)
Advancing Question Answering on Handwritten Documents: A State-of-the-Art Recognition-Based Model for HW-SQuAD
di: Pal, Aniket, et al.
Pubblicazione: (2024)
di: Pal, Aniket, et al.
Pubblicazione: (2024)
Source-free Video Domain Adaptation by Learning from Noisy Labels
di: Dasgupta, Avijit, et al.
Pubblicazione: (2023)
di: Dasgupta, Avijit, et al.
Pubblicazione: (2023)
Prompt2LVideos: Exploring Prompts for Understanding Long-Form Multimodal Videos
di: Jahagirdar, Soumya Shamarao, et al.
Pubblicazione: (2025)
di: Jahagirdar, Soumya Shamarao, et al.
Pubblicazione: (2025)
Real-Time Human Reconstruction and Animation using Feed-Forward Gaussian Splatting
di: Chatterjee, Devdoot, et al.
Pubblicazione: (2026)
di: Chatterjee, Devdoot, et al.
Pubblicazione: (2026)
EW-DETR: Evolving World Object Detection via Incremental Low-Rank DEtection TRansformer
di: Monga, Munish, et al.
Pubblicazione: (2026)
di: Monga, Munish, et al.
Pubblicazione: (2026)
HW-MLVQA: Elucidating Multilingual Handwritten Document Understanding with a Comprehensive VQA Benchmark
di: Pal, Aniket, et al.
Pubblicazione: (2025)
di: Pal, Aniket, et al.
Pubblicazione: (2025)
PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction
di: Mishra, Naman, et al.
Pubblicazione: (2026)
di: Mishra, Naman, et al.
Pubblicazione: (2026)
Unifying Scientific Communication: Fine-Grained Correspondence Across Scientific Media
di: M, Megha Mariam K., et al.
Pubblicazione: (2026)
di: M, Megha Mariam K., et al.
Pubblicazione: (2026)
Face Time Traveller : Travel Through Ages Without Losing Identity
di: Kar, Purbayan, et al.
Pubblicazione: (2026)
di: Kar, Purbayan, et al.
Pubblicazione: (2026)
Reading Between the Lanes: Text VideoQA on the Road
di: Tom, George, et al.
Pubblicazione: (2023)
di: Tom, George, et al.
Pubblicazione: (2023)
Can Reasons Help Improve Pedestrian Intent Estimation? A Cross-Modal Approach
di: Khindkar, Vaishnavi, et al.
Pubblicazione: (2024)
di: Khindkar, Vaishnavi, et al.
Pubblicazione: (2024)
How Does Bilateral Ear Symmetry Affect Deep Ear Features?
di: Ozturk, Kagan, et al.
Pubblicazione: (2025)
di: Ozturk, Kagan, et al.
Pubblicazione: (2025)
DriveSafe: A Framework for Risk Detection and Safety Suggestions in Driving Scenarios
di: Artham, Sainithin, et al.
Pubblicazione: (2026)
di: Artham, Sainithin, et al.
Pubblicazione: (2026)
Pedestrian Intention and Trajectory Prediction in Unstructured Traffic Using IDD-PeD
di: Bokkasam, Ruthvik, et al.
Pubblicazione: (2025)
di: Bokkasam, Ruthvik, et al.
Pubblicazione: (2025)
AI-Generated Lecture Slides for Improving Slide Element Detection and Retrieval
di: Maniyar, Suyash, et al.
Pubblicazione: (2025)
di: Maniyar, Suyash, et al.
Pubblicazione: (2025)
IDD-X: A Multi-View Dataset for Ego-relative Important Object Localization and Explanation in Dense and Unstructured Traffic
di: Parikh, Chirag, et al.
Pubblicazione: (2024)
di: Parikh, Chirag, et al.
Pubblicazione: (2024)
Multiple Instance Learning for Glioma Diagnosis using Hematoxylin and Eosin Whole Slide Images: An Indian Cohort Study
di: Chauhan, Ekansh, et al.
Pubblicazione: (2024)
di: Chauhan, Ekansh, et al.
Pubblicazione: (2024)
The More You See in 2D, the More You Perceive in 3D
di: Han, Xinyang, et al.
Pubblicazione: (2024)
di: Han, Xinyang, et al.
Pubblicazione: (2024)
Towards Accurate Lip-to-Speech Synthesis in-the-Wild
di: Hegde, Sindhu, et al.
Pubblicazione: (2024)
di: Hegde, Sindhu, et al.
Pubblicazione: (2024)
Chain-of-Cooking:Cooking Process Visualization via Bidirectional Chain-of-Thought Guidance
di: Xu, Mengling, et al.
Pubblicazione: (2025)
di: Xu, Mengling, et al.
Pubblicazione: (2025)
Towards Deployable OCR models for Indic languages
di: Mathew, Minesh, et al.
Pubblicazione: (2022)
di: Mathew, Minesh, et al.
Pubblicazione: (2022)
A Dataset for Semantic Segmentation in the Presence of Unknowns
di: Laskar, Zakaria, et al.
Pubblicazione: (2025)
di: Laskar, Zakaria, et al.
Pubblicazione: (2025)
Towards Safer and Understandable Driver Intention Prediction
di: Karuppasamy, Mukilan, et al.
Pubblicazione: (2025)
di: Karuppasamy, Mukilan, et al.
Pubblicazione: (2025)
Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind
di: Plizzari, Chiara, et al.
Pubblicazione: (2024)
di: Plizzari, Chiara, et al.
Pubblicazione: (2024)
Cross-Domain Evaluation of Few-Shot Classification Models: Natural Images vs. Histopathological Images
di: Sekhar, Ardhendu, et al.
Pubblicazione: (2024)
di: Sekhar, Ardhendu, et al.
Pubblicazione: (2024)
Hide and Seek: How Does Watermarking Impact Face Recognition?
di: Yao, Yuguang, et al.
Pubblicazione: (2024)
di: Yao, Yuguang, et al.
Pubblicazione: (2024)
Utilizing Multi-Agent Reinforcement Learning with Encoder-Decoder Architecture Agents to Identify Optimal Resection Location in Glioblastoma Multiforme Patients
di: Arun, Krishna, et al.
Pubblicazione: (2025)
di: Arun, Krishna, et al.
Pubblicazione: (2025)
Designing Production-Scale OCR for India: Multilingual and Domain-Specific Systems
di: Faraz, Ali, et al.
Pubblicazione: (2026)
di: Faraz, Ali, et al.
Pubblicazione: (2026)
Curvature Informed Furthest Point Sampling
di: Bhardwaj, Shubham, et al.
Pubblicazione: (2024)
di: Bhardwaj, Shubham, et al.
Pubblicazione: (2024)
CookingDiffusion: Cooking Procedural Image Generation with Stable Diffusion
di: Wang, Yuan, et al.
Pubblicazione: (2025)
di: Wang, Yuan, et al.
Pubblicazione: (2025)
PRECISe : Prototype-Reservation for Explainable Classification under Imbalanced and Scarce-Data Settings
di: Ganatra, Vaibhav, et al.
Pubblicazione: (2024)
di: Ganatra, Vaibhav, et al.
Pubblicazione: (2024)
Cook and Clean Together: Teaching Embodied Agents for Parallel Task Execution
di: Liang, Dingkang, et al.
Pubblicazione: (2025)
di: Liang, Dingkang, et al.
Pubblicazione: (2025)
VisualChef: Generating Visual Aids in Cooking via Mask Inpainting
di: Kuzyk, Oleh, et al.
Pubblicazione: (2025)
di: Kuzyk, Oleh, et al.
Pubblicazione: (2025)
How Does Audio Influence Visual Attention in Omnidirectional Videos? Database and Model
di: Zhu, Yuxin, et al.
Pubblicazione: (2024)
di: Zhu, Yuxin, et al.
Pubblicazione: (2024)
Real-Time Cooked Food Image Synthesis and Visual Cooking Progress Monitoring on Edge Devices
di: Gupta, Jigyasa, et al.
Pubblicazione: (2025)
di: Gupta, Jigyasa, et al.
Pubblicazione: (2025)
Novel View Synthesis using DDIM Inversion
di: Singh, Sehajdeep, et al.
Pubblicazione: (2025)
di: Singh, Sehajdeep, et al.
Pubblicazione: (2025)
V-Nutri: Dish-Level Nutrition Estimation from Egocentric Cooking Videos
di: Yue, Chengkun, et al.
Pubblicazione: (2026)
di: Yue, Chengkun, et al.
Pubblicazione: (2026)
Documenti analoghi
-
PhyEduVideo: A Benchmark for Evaluating Text-to-Video Models for Physics Education
di: M, Megha Mariam K., et al.
Pubblicazione: (2026) -
Attend to what I say: Highlighting relevant content on slides
di: M, Megha Mariam K, et al.
Pubblicazione: (2026) -
IndicSTR12: A Dataset for Indic Scene Text Recognition
di: Lunia, Harsh, et al.
Pubblicazione: (2024) -
Advancing Question Answering on Handwritten Documents: A State-of-the-Art Recognition-Based Model for HW-SQuAD
di: Pal, Aniket, et al.
Pubblicazione: (2024) -
Source-free Video Domain Adaptation by Learning from Noisy Labels
di: Dasgupta, Avijit, et al.
Pubblicazione: (2023)