A Dataset for Suggesting Variable Orderings for Cylindrical Algebraic Decompositions
摘要
Data have been playing a vital role in many successful applications of artificial intelligence. To embrace the power of modern artificial intelligence (AI) technology to accelerate symbolic algorithms, it is indispensable to have a large amount of data of high quality. Unfortunately, such datasets are rare in the area of symbolic computation. Indeed, generation of a large dataset from scratch often costs a lot of time and effort. In this work, we make public a random dataset containing more than 20K labelled examples on suggesting the variable ordering of cylindrical algebraic decomposition (CAD). We report in detail the design, generation as well as statistical information of the dataset. The value of this dataset is demonstrated by evaluating several heuristic methods with it and training and testing on it with several machine learning (ML) models.