five

Free Sample Dataset - 1000 High Resolution Images & Metadata|图像数据数据集|机器学习数据集

收藏
Databricks2024-05-09 收录
图像数据
机器学习
下载链接:
https://marketplace.databricks.com/details/3da47c97-478c-4935-a0e7-c61502c7c7b7/Shutterstock_Free-Sample-Dataset---1000-High-Resolution-Images-&-Metadata
下载链接
链接失效反馈
资源简介:
**Overview** This is a free sample dataset consisting of 1000 images and accompanying metadata sourced from our +550 million image library. Image types for this sample include photos, vectors, and illustrations across a vast range of content categories and settings. This sample includes a wide range of metadata fields including content descriptions and keywords that make it ideal for powering a wide variety of machine learning use cases. If you’d like to start licensing data from the full range of imagery and metadata available at Shutterstock please reach out directly to our team at sales.databricks@shutterstock.com to start using our tailored services to help ideate, curate and customize datasets for your unique business needs. **Use cases** This type of data can be licensed from Shutterstock for a wide variety of use cases including powering machine learning models that have generative capabilities. **Metadata** Sample metadata fields included in this dataset are listed below, for a full list of all metadata available from Shutterstock please contact our team at sales.databricks@shutterstock.com. **asset metadata:** id keywords image_type is_creative mature_flag date_submitted date_captured asset_location popularity_score german_description spanish_description french_description korean_description japanese_description labels moderation_labels has_model_release has_people primary_category
 **file metadata:** asset_id asset_file_size asset_file_extension asset_file_size_in_bytes width height orientation
 **model metadata:** asset_id model_release_id age_range age_in_years gender ethnicity **Our data** Shutterstock offers the largest, highest quality and most diverse collection of creative content with best-in-class metadata, giving technology businesses the scale and accuracy they need to build and sustain a wide variety of machine learning models. Our growing library of +550M images, +40M videos, +4M music and audio tracks, and +1.2M 3D models and data is human-reviewed for accuracy and IP infringement, allowing you to use our data worry-free and avoid unwanted or unlawful content. We ethically source all content from over 2 million creators in +150 countries and with over 60 million new assets added annually, our ever-growing library gives you access to fresh and diverse datasets that can be refreshed regularly to meet all your data needs.
提供机构:
Shutterstock
用户留言
有没有相关的论文或文献参考?
这个数据集是基于什么背景创建的?
数据集的作者是谁?
能帮我联系到这个数据集的作者吗?
这个数据集如何下载?
点击留言
数据主题
具身智能
数据集  4099个
机构  8个
大模型
数据集  439个
机构  10个
无人机
数据集  37个
机构  6个
指令微调
数据集  36个
机构  6个
蛋白质结构
数据集  50个
机构  8个
空间智能
数据集  21个
机构  5个
5,000+
优质数据集
54 个
任务类型
进入经典数据集
热门数据集

The MaizeGDB

The MaizeGDB(Maize Genetics and Genomics Database)是一个专门为玉米(Zea mays)基因组学研究提供数据和工具的在线资源。该数据库包含了玉米的基因组序列、基因注释、遗传图谱、突变体信息、表达数据、以及与玉米相关的文献和研究工具。MaizeGDB旨在支持玉米遗传学和基因组学的研究,为科学家提供了一个集成的平台来访问和分析玉米的遗传和基因组数据。

www.maizegdb.org 收录

OpenSonarDatasets

OpenSonarDatasets是一个致力于整合开放源代码声纳数据集的仓库,旨在为水下研究和开发提供便利。该仓库鼓励研究人员扩展当前的数据集集合,以增加开放源代码声纳数据集的可见性,并提供一个更容易查找和比较数据集的方式。

github 收录

Paper III (Walker et al. 2024)

Data products used in 3-D CMZ Paper III, Walker et al. (2024). The full cloud catalogue is provided in tabular format, along with a full CMZ map showing the clouds and their assigned IDs. For each cloud ID in the published catalogue there are: - Individual cube cutouts from the MOPRA 3mm CMZ survey (HC3N, HCN, and HNCO). - Individual cube cutouts from the APEX 1mm CMZ survey (13CO, C18O, and H2CO). - Cloud-averaged spectra of the ATCA H2CO 4.83 GHz line. - PV slices of the ATCA H2CO 4.83 GHz line, taken across the major axis of the source. - Where applicable, there are mask files which correspond to the different velocity components of the cloud. In these cases, there are two mask files per velocity component, corresponding to the different masking approaches described in the paper.

DataCite Commons 收录

Stanford Cars

Cars数据集包含196类汽车的16,185图像。数据被分成8,144训练图像和8,041测试图像,其中每个类被大致分成50-50。类别通常在品牌,型号,年份,例如2012特斯拉Model S或2012 BMW M3 coupe的级别。

OpenDataLab 收录

EcoInvent

EcoInvent是一个生命周期评估(LCA)数据库,包含了大量产品的环境影响数据。它提供了详细的产品生命周期数据,包括原材料提取、生产、使用和废弃处理等各个阶段的环境影响信息。

www.ecoinvent.org 收录