• A
  • A
  • A
  • ABC
  • ABC
  • ABC
  • А
  • А
  • А
  • А
  • А
Regular version of the site

Researchers Rank Recommendation Algorithms Using Sports Tournament Model

Researchers Rank Recommendation Algorithms Using Sports Tournament Model

© iStock

Researchers from the AI and Digital Science Institute at the HSE Faculty of Computer Science have developed an approach for selecting recommendation algorithms more effectively. Their approach uses pairwise comparisons of algorithms to create a tournament table, with the overall ranking based on their performance across all datasets in the tournament. This can reduce the number of algorithms that need to be tested when developing new services, saving both time and money. The study was presented at the 32nd ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD 2026).

Recommendation systems determine which products, films, songs, or publications to show users. To do this, they analyse users’ past behaviour and predict what they might be interested in next. For example, recommendation systems identify users with similar interests and take into account the sequence of their views or purchases.

However, there is no universal recommendation algorithm. A method that works well for an online store may not be suitable for an online cinema. Therefore, algorithms are usually pre-tested on existing datasets, and their performance is averaged across them. This approach, however, has its limitations: the resulting ranking does not account for the specifics of individual datasets and may be unstable, while testing all algorithms online with real users is costly and risky.

Researchers at HSE University have developed a methodology for comparing recommendation algorithms based on the Bradley–Terry model, which is used to rank competitors based on the outcomes of pairwise ‘matches.’ In this study, the recommendation algorithms were the ‘players,’ while the ‘matches’ were tests of the algorithms on different datasets. Two algorithms were compared using a selected metric, such as recommendation accuracy, and the one with the higher score was declared the winner. Based on the results of all pairwise comparisons, the model estimated the relative strength of each algorithm. The researchers trained and tested 14 algorithms on 89 datasets from different fields. 

The results showed that the same algorithms could occupy different positions in the ranking depending on the type of data used in the ‘matches.’ For example, on sequential datasets, where the order of user actions is important, SASRec and GASATF ranked as the top performers. However, when a dataset did not contain an explicit sequence, these algorithms dropped to tenth and eleventh place, respectively, while LightGCN and ALS topped the ranking.

Additionally, the researchers tested the robustness of the rankings to incomplete data. The ranking produced by the Bradley–Terry model remained stable even when some comparisons were missing, meaning that not every algorithm was compared with every other algorithm. The authors also tested an extended version of the model on additional datasets, taking the context into account. In this case, the model had to predict the winners without directly comparing the recommendation algorithms. 

'The idea of using a sports model came to us thanks to the work of our senior colleague Vladimir Spokoiny. To draw an analogy with a sports tournament, the outcome of a match is influenced not only by the players themselves but also by conditions such as the city or the weather. In our study, the characteristics of the dataset served as this context, including the number of users and items, the average length of users’ histories, and other parameters. If we train the model to take this context into account alongside the results of previous comparisons, it can use the characteristics of a new dataset to predict in advance which algorithms are likely to perform best,' says co-author Anton Lysenko, Expert at the International Laboratory of Stochastic Algorithms and High-Dimensional Inference.

In 78% of cases, the algorithm ranked first by the model was indeed among the top three. With the conventional approach of averaging performance metrics, this was the case in only 16% of cases. The authors attribute this difference to the fact that, unlike heuristic comparison methods, their model has a sound theoretical foundation, making the resulting rankings more reliable.

'Our approach helps identify which algorithms are best suited to a particular dataset before they are tested online. For example, if a bank needs to recommend loyalty programmes, it can feed the characteristics of a new dataset into the model to identify the most promising algorithms. Only these algorithms would then need to be tested on real users, rather than all available options. This saves both time and resources,' says co-author Sergey Samsonov, Head of the International Laboratory of Stochastic Algorithms and High-Dimensional Inference.

The study was carried out as part of a programme implemented by the HSE AI Research Centre and supported by a grant from the Ministry of Economic Development of the Russian Federation.

See also:

How to Assess Students’ Knowledge in the Age of AI

A researcher at HSE University has proposed a flowchart to help lecturers decide how to assess students who use artificial intelligence. It shows where the use of AI should be restricted and where it can be incorporated into the learning process. The article has been published in IT Professional.

HSE University Expands Cooperation with Malaysia in Technology Foresight

HSE University researchers will take part in a study of the future of engineering education in Malaysia, while the Malaysian Industry-Government Group for High Technology (MIGHT) will use the iFORA big-data analysis system to validate the findings of its foresight research. These are the outcomes of a visit by HSE representatives to Kuala Lumpur.

Scientists Train Neural Network to Generate Process Plans from 3D Models

Researchers at the HSE FCS AI and Digital Science Institute have developed CAD2TechSpec, a framework that converts 3D models of mechanical parts into machining process plans—step-by-step instructions for machine tools. The solution aims to reduce the time required for the design and preparation of technical process documentation in mechanical engineering, aircraft manufacturing, and other high-tech industries. The study findings have been published in PeerJ Computer Science.

Biologists Discover 'Molecular Fingerprint' of Preeclampsia

Researchers at HSE University employed a new method to model hypoxia in placental cells during pregnancies complicated by preeclampsia and identified molecular markers of tissue hypoxia. Since hypoxia is one of the key mechanisms underlying preeclampsia, these findings are important for a more accurate and timely diagnosis of the disease and for the development of effective treatment methods. The paper has been published in Placenta.

Laboratory of Future Networks: HSE Telecommunications Research Institute Develops 5G/6G Research Testbed

The 5G/6G testbed at the HSE Telecommunications Research Institute is becoming a research platform, an educational laboratory, and a foundation for developing new software components for future networks. It makes it possible not only to observe how a mobile network operates, but also to change its operating conditions and measure the results: data-transmission speed, latency, errors, radio-resource utilisation, and other parameters. Based on the testbed, researchers plan to develop MIMO, O-RAN, xApp, and IAB technologies, as well as experiment with artificial intelligence.

‘Hedgehog’ Versus ‘Relatives’: Researchers Measure How the Brain Responds to Unexpected Words During Natural Speech

Russian neurophysiologists, including researchers from HSE University, have demonstrated the feasibility of using event-related fields (ERFs) to study brain activity during natural speech perception. The researchers showed that this approach can be applied not only to individual words but also to continuous speech. Their findings indicate that words whose meanings differ significantly from the preceding context require longer processing times. The study also reveals that the brain processes function words in two stages: first, it identifies their grammatical role and then uses this information to predict the next word. The study has been published in Frontiers in Human Neuroscience.

HSE Researchers Create New Corpus of Early Child Speech in Russian

Researchers at the HSE Centre for Language and Brain have presented RusLan-M, an open multimedia corpus that makes it possible to trace the development of early child speech in Russian from first words to the emergence of complex grammatical constructions. The database contains around 41 hours of video recordings and more than 35,000 child utterances. The new resource will help researchers study more precisely how children acquire Russian and, in the longer term, develop more reliable tools for assessing speech development. The study has been published in Language Resources and Evaluation.

Hybrid Intelligence: Competencies in the Age of AI Discussed at Technoprom-2026

Artificial intelligence is not creating new professions, but rather transforming the nature of existing ones. This was the conclusion reached by participants in the panel session ‘Hybrid Intelligence: Digital and Human Drivers of Development,’ organised by the Institute for Statistical Studies and Economics of Knowledge (ISSEK) at HSE University as part of the 13th International Forum of Technological Development (Technoprom-2026). The experts discussed how the nature of work is changing, which skills are becoming increasingly sought after, and what prevents companies from fully capitalising on new technologies.

Scientists Develop New Solution for 6G Communication Systems

A terahertz neuromorphic circuit developed by scientists at HSE University could make 6G communication systems both more accurate and energy-efficient. The circuit enables indoor tracking of mobile devices with an accuracy of up to 99%. The results were presented at PIERS 2026, an international symposium on Photonics and Electromagnetism held in China.

Scientists Develop Algorithm for More Reliable Processors in Data Centres

Researchers from HSE MIEM and Samara University have developed the LRF-3D algorithm to automatically bypass idle nodes in three-dimensional networks-on-chip. Thanks to its hierarchical architecture, the algorithm outperforms existing solutions in both speed and path accuracy, improving processor reliability for use in data centres, supercomputers, and AI computing. The source code and test results are publicly available.