Papers › LLaVA-Chef: A Multi-modal Generative Model for Food Recipes
LLaVA-Chef: A Multi-modal Generative Model for Food Recipes
Fnu Mohbat, Mohammed J. Zaki
In the rapidly evolving landscape of online recipe sharing within a globalized context, there has been a notable surge in research towards comprehending and generating food recipes. Recent advancements in large language models (LLMs) like GPT-2 and LLaVA have paved the way for Natural Language Processing (NLP) approaches to delve deeper into various facets of food-related tasks, encompassing ingredient recognition and comprehensive recipe generation. Despite impressive performance and multi-modal adaptability of LLMs, domain-specific training remains paramount for their effective application. This work evaluates existing LLMs for recipe generation and proposes LLaVA-Chef, a novel model trained on a curated dataset of diverse recipe prompts in a multi-stage approach. First, we refine the mapping of visual food image embeddings to the language space. Second, we adapt LLaVA to the food domain by fine-tuning it on relevant recipe data. Third, we utilize diverse prompts to enhance the model's recipe comprehension. Finally, we improve the linguistic quality of generated recipes by penalizing the model with a custom loss function. LLaVA-Chef demonstrates impressive improvements over pretrained LLMs and prior works. A detailed qualitative analysis reveals that LLaVA-Chef generates more detailed recipes with precise ingredient mentions, compared to existing approaches.
In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.
Code
Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.
Code Syntology ran Syntology
Not run by Syntology. Nothing on this page verifies that the listed code works.
Tasks
Results from the paper archive 2025-07-28
| Task | Dataset | Model | Metric | Value | Rank at snapshot | Leaderboard | Report |
|---|---|---|---|---|---|---|---|
| Recipe Generation | Food.com | LLaVA-Chef | BLEU-1 | 29 | #1 of 2 | Archive leaderboard | report |
| Recipe Generation | Food.com | LLaVA-Chef | BLEU-4 | 6 | #1 of 2 | Archive leaderboard | report |
| Recipe Generation | Food.com | LLaVA-Chef | BPE Perplexity | 2.6 | #1 of 2 | Archive leaderboard | report |
| Recipe Generation | Food.com | LLaVA-Chef | D-1 | 0 | #1 of 2 | Archive leaderboard | report |
| Recipe Generation | Food.com | LLaVA-Chef | D-2 | 0 | #1 of 2 | Archive leaderboard | report |
| Recipe Generation | Food.com | LLaVA-Chef | Rouge-L | 18.4 | #1 of 2 | Archive leaderboard | report |
| Recipe Generation | Now You're Cooking! | LLaVA-Chef | Perplexity | 2.6 | #1 of 2 | Archive leaderboard | report |
| Recipe Generation | allrecipes.com | LLaVA-Chef | BLEU | 6.0 | #2 of 2 | Archive leaderboard | report |
| Recipe Generation | allrecipes.com | LLaVA-Chef | Perplexity | 2.6 | #2 of 2 | Archive leaderboard | report |
Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.
Methods
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections