遇见数据集

lzzmm/BurstGPT

收藏
Hugging Face2025-07-28 更新2025-10-25 收录
官方服务:

资源简介:

这是一个包含ChatGPT(GPT-3.5)和GPT-4模型工作负载跟踪的实际世界数据集,用于优化大规模语言模型服务系统。数据集记录了连续四个月的121天,总共有约5.29M行数据,大小约为188MB。数据集提供了请求和响应令牌长度、模型类型、日志类型等信息,并分为两个时间段,每个时间段都有成功和失败的数据。

This is a real-world workload trace dataset for ChatGPT(GPT-3.5) and GPT-4 models, used to optimize LLM serving systems. The dataset covers 121 consecutive days over 4 consecutive months, with a total of approximately 5.29M lines and a size of about 188MB. The dataset provides information such as request and response token length, model type, log type, etc., and is divided into two periods, each with successful and failed data.

提供机构:
lzzmm
搜集汇总
数据集介绍
lzzmm/BurstGPT 数据集图片
背景与挑战
背景概述
该数据集是实际世界中的ChatGPT和GPT-4模型工作负载跟踪数据,用于优化大语言模型服务系统。它包含连续四个月(121天)约5.29M行记录,提供请求/响应令牌长度、模型类型、日志类型等信息,并区分了成功与失败的数据。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务