Dmitry Menshikov

3papers

34citations

Novelty15%

AI Score22

Ranked #186,768 of 201,326 authors (top 93%)#12,756 in AI (top 90%)

3 Papers

IRAug 25, 2022Code

Lib-SibGMU -- A University Library Circulation Dataset for Recommender Systems Developmen

Eduard Zubchuk, Mikhail Arhipkin, Dmitry Menshikov et al.

We opensource under CC BY 4.0 license Lib-SibGMU - a university library circulation dataset - for a wide research community, and benchmark major algorithms for recommender systems on this dataset. For a recommender architecture that consists of a vectorizer that turns the history of the books borrowed into a vector, and a neighborhood-based recommender, trained separately, we show that using the fastText model as a vectorizer delivers competitive results.

ASMar 30, 2021Code

MediaSpeech: Multilanguage ASR Benchmark and Dataset

Rostislav Kolobov, Olga Okhapkina, Olga Omelchishina et al.

The performance of automated speech recognition (ASR) systems is well known to differ for varied application domains. At the same time, vendors and research groups typically report ASR quality results either for limited use simplistic domains (audiobooks, TED talks), or proprietary datasets. To fill this gap, we provide an open-source 10-hour ASR system evaluation dataset NTR MediaSpeech for 4 languages: Spanish, French, Turkish and Arabic. The dataset was collected from the official youtube channels of media in the respective languages, and manually transcribed. We estimate that the WER of the dataset is under 5%. We have benchmarked many ASR systems available both commercially and freely, and provide the benchmark results. We also open-source baseline QuartzNet models for each language.

AIFeb 8, 2022

Using a Language Model in a Kiosk Recommender System at Fast-Food Restaurants

Eduard Zubchuk, Dmitry Menshikov, Nikolay Mikhaylovskiy

Kiosks are a popular self-service option in many fast-food restaurants, they save time for the visitors and save labor for the fast-food chains. In this paper, we propose an effective design of a kiosk shopping cart recommender system that combines a language model as a vectorizer and a neural network-based classifier. The model performs better than other models in offline tests and exhibits performance comparable to the best models in A/B/C tests.