DeepSData
Dataset guide · Machine learning & corpora

20846_Groups_Image_Caption_Data_of_Cookbook

This dataset contains 20,846 groups of cookbook image-caption data, each group includes 4-18 images and corresponding text descriptions in Chinese and English, covering Chinese, Western, Korean, Japanese cuisines, etc., released by DatatangBeijing.

← Back to dataset library · 中文版

Machine learning & corporaAccess: check the source page

Key facts

InstitutionDatatangBeijing
CoverageChinese, Western, Korean, Japanese cuisines, etc.
Time spanSee official page
Scale20,846 groups, 4-18 images per group
LicenseTerms: check the source page
AccessModelScope platform

Contents & fields

This dataset contains 20,846 groups of cookbook data, each group includes 4-18 images and corresponding text descriptions. Descriptions are in Chinese and English, with Chinese descriptions no less than 15 characters and English no less than 30 words. Cuisines include Chinese, Western, Korean, Japanese, etc. No field-level details are disclosed on the source page.

Research uses

Can be used for recipe recommendation, culinary education, multimodal image-text understanding, image caption generation, etc.

Information comes from the source page. Please check that page for current details and terms.

Keywords

cookbookimage-captionbilingualmultimodalcuisinedataset

Access & license

License: Terms: check the source page | Access: check the source page

Why this is hard to get on your own

When obtaining large-scale, multi-cuisine, bilingual image-caption data for multimodal model training, data is often scattered and annotations are inconsistent.

Related datasets

Same domain

Need this data retrieved and prepared?

Tell us your hard requirements. We first assess availability, then retrieve for real — and if it truly cannot be obtained, we say so plainly.