Reproducibility dataset for a large experimental survey on word embeddings and ontology-based methods for word similarity.
Ontology highlight
ABSTRACT: This data article introduces a reproducibility dataset with the aim of allowing the exact replication of all experiments, results and data tables introduced in our companion paper (Lastra-Díaz et al., 2019), which introduces the largest experimental survey on ontology-based semantic similarity methods and Word Embeddings (WE) for word similarity reported in the literature. The implementation of all our experiments, as well as the gathering of all raw data derived from them, was based on the software implementation and evaluation of all methods in HESML library (Lastra-Díaz et al., 2017), and their subsequent recording with Reprozip (Chirigati et al., 2016). Raw data is made up by a collection of data files gathering the raw word-similarity values returned by each method for each word pair
SUBMITTER: Lastra-Diaz JJ
PROVIDER: S-EPMC6736772 | biostudies-literature | 2019 Oct
REPOSITORIES: biostudies-literature
ACCESS DATA