AppsApps

Building a Gold Standard for a Russian Collocations Database

Date
Speaker
  1. Maria Khokhlova
Abstract

The talk focuses on the process of building a gold standard that will include data from Russian dictionaries and corpora. The standard is being prepared for a Russian Collocations Database that already includes information on words’ collocability and was extracted from text corpora by statistical measures and linguistic filters. The gold standard will be also used for the evaluation of the extracted collocations and for marking them as “true” collocations with references to the dictionaries.