The Massively Multilingual Image Dataset (MMID)
收藏aws亚马逊开源数据集2024-03-07 收录
数据链接:
官方服务:
资源简介:
MMID is a large-scale, massively multilingual dataset of images paired with the words they represent collected at the University of Pennsylvania. The dataset is doubly parallel: for each language, words are stored parallel to images that represent the word, and parallel to the word's translation into English (and corresponding images.)
提供机构:
Penn NLP


