This dataset is a PostgreSQL dump of a database generated using autocode_big from the data in the LexiRumah lexical database of eastern Indonesia and Timor-Leste. The dataset contains the forms from the parent dataset, together with automatically generated sound correspondence scorers, pairwise similarity scores, cognate classes, and alignments, all created using Lingpy's LexStat algorithm.