EDA Corpus
收藏资源简介:
EDA Corpus是由亚利桑那州立大学和纽约大学合作创建的一个面向OpenROAD的开源数据集,包含超过1000个数据点,分为问题与答案对和代码与脚本对两种格式。该数据集旨在通过提供丰富的训练数据,促进大型语言模型在电子设计自动化(EDA)领域的应用,特别是在物理设计任务中。数据集的创建过程涉及从OpenROAD的GitHub问题、讨论和文档中收集和验证数据,确保每条数据的高质量和相关性。EDA Corpus的应用领域主要集中在提高芯片设计的自动化水平,帮助新老设计师更有效地理解和使用OpenROAD工具。
EDA Corpus is an open-source dataset dedicated to OpenROAD, jointly developed by Arizona State University and New York University. It comprises over 1,000 data points in two formats: question-answer pairs and code-and-script pairs. The core objective of this dataset is to facilitate the application of large language models (LLMs) in the field of electronic design automation (EDA), particularly for physical design tasks, by providing abundant training data. The dataset was constructed by collecting and validating data sourced from OpenROAD's GitHub issues, community discussions and official documentation, ensuring the high quality and contextual relevance of each data entry. The main application scenarios of EDA Corpus focus on improving the automation level of chip design, helping both novice and experienced designers more effectively understand and utilize the OpenROAD tools.

- 1EDA Corpus: A Large Language Model Dataset for Enhanced Interaction with OpenROAD亚利桑那州立大学 · 2024年



