five

Free Sample Dataset - 1000 High Resolution Images & Metadata|图像数据数据集|机器学习数据集

收藏
Databricks2024-05-09 收录
图像数据
机器学习
下载链接:
https://marketplace.databricks.com/details/3da47c97-478c-4935-a0e7-c61502c7c7b7/Shutterstock_Free-Sample-Dataset---1000-High-Resolution-Images-&-Metadata
下载链接
链接失效反馈
资源简介:
**Overview** This is a free sample dataset consisting of 1000 images and accompanying metadata sourced from our +550 million image library. Image types for this sample include photos, vectors, and illustrations across a vast range of content categories and settings. This sample includes a wide range of metadata fields including content descriptions and keywords that make it ideal for powering a wide variety of machine learning use cases. If you’d like to start licensing data from the full range of imagery and metadata available at Shutterstock please reach out directly to our team at sales.databricks@shutterstock.com to start using our tailored services to help ideate, curate and customize datasets for your unique business needs. **Use cases** This type of data can be licensed from Shutterstock for a wide variety of use cases including powering machine learning models that have generative capabilities. **Metadata** Sample metadata fields included in this dataset are listed below, for a full list of all metadata available from Shutterstock please contact our team at sales.databricks@shutterstock.com. **asset metadata:** id keywords image_type is_creative mature_flag date_submitted date_captured asset_location popularity_score german_description spanish_description french_description korean_description japanese_description labels moderation_labels has_model_release has_people primary_category
 **file metadata:** asset_id asset_file_size asset_file_extension asset_file_size_in_bytes width height orientation
 **model metadata:** asset_id model_release_id age_range age_in_years gender ethnicity **Our data** Shutterstock offers the largest, highest quality and most diverse collection of creative content with best-in-class metadata, giving technology businesses the scale and accuracy they need to build and sustain a wide variety of machine learning models. Our growing library of +550M images, +40M videos, +4M music and audio tracks, and +1.2M 3D models and data is human-reviewed for accuracy and IP infringement, allowing you to use our data worry-free and avoid unwanted or unlawful content. We ethically source all content from over 2 million creators in +150 countries and with over 60 million new assets added annually, our ever-growing library gives you access to fresh and diverse datasets that can be refreshed regularly to meet all your data needs.
提供机构:
Shutterstock
用户留言
有没有相关的论文或文献参考?
这个数据集是基于什么背景创建的?
数据集的作者是谁?
能帮我联系到这个数据集的作者吗?
这个数据集如何下载?
点击留言
数据主题
具身智能
数据集  4099个
机构  8个
大模型
数据集  439个
机构  10个
无人机
数据集  37个
机构  6个
指令微调
数据集  36个
机构  6个
蛋白质结构
数据集  50个
机构  8个
空间智能
数据集  21个
机构  5个
5,000+
优质数据集
54 个
任务类型
进入经典数据集
热门数据集

中国区域交通网络数据集

该数据集包含中国各区域的交通网络信息,包括道路、铁路、航空和水路等多种交通方式的网络结构和连接关系。数据集详细记录了各交通节点的位置、交通线路的类型、长度、容量以及相关的交通流量信息。

data.stats.gov.cn 收录

Wafer Defect

该数据集包含了七个主要类别的晶圆缺陷,分别是:BLOCK ETCH、COATING BAD、PARTICLE、PIQ PARTICLE、PO CONTAMINATION、SCRATCH和SEZ BURNT。这些类别涵盖了晶圆在生产过程中可能出现的多种缺陷类型,每一种缺陷都有其独特的成因和表现形式。数据集不仅在类别数量上具有多样性,而且在样本的多样性和复杂性上也展现了其广泛的应用潜力。每个类别的样本均经过精心标注,确保了数据的准确性和可靠性。

github 收录

YouTube-English

该数据集包含从各种YouTube频道提取的英语音频片段以及相应的转录元数据。数据用于训练自动语音识别(ASR)模型。数据来源于YouTube频道,处理过程包括下载、分割和保存音频及元数据。数据集总结部分详细列出了每个频道的视频数量、持续时间和占总数据集的百分比。

huggingface 收录

Stanford Cars

Cars数据集包含196类汽车的16,185图像。数据被分成8,144训练图像和8,041测试图像,其中每个类被大致分成50-50。类别通常在品牌,型号,年份,例如2012特斯拉Model S或2012 BMW M3 coupe的级别。

OpenDataLab 收录

波士顿房价数据集

波士顿房价数据集是一个经典的机器学习数据集,通常用于回归任务,尤其是房价预测。下方文档中有所有字段顺序的描述。

阿里云天池 收录