遇见数据集

VEUSES

收藏
DataCite Commons2020-09-02 更新2024-07-25 收录
官方服务:

资源简介:

EUSES is the most frequently used spreadsheet corpus, and contains 4,037 spreadsheets. These spreadsheets were extracted from World Wide Web. We applied SpreadCluster to the EUSES and manually validated all groups. Based on the validated result, we built the VEUSES corpus, containing 177 evolution groups and 363 spreadsheets.VEUSES is published associated with our MSR 2017 paper in May 2017. <br>Liang Xu, Wensheng Dou, Chushu Gao, Jie Wang, Jun Wei, Hua Zhong, Tao Huang. SpreadCluster: Recovering Versioned Spreadsheets through Similarity-Based Clustering. In <i>Proceedings of the 14th International Conference on Mining Software Repositories</i> (<b><i>MSR 2017</i></b>), May 2017.<br>

提供机构:
figshare
创建时间:
2017-03-29
二维码
社区交流群
二维码
科研交流群
商业服务