Discussion about this post

User's avatar
Wigan's avatar

I'd love to see some bots as well, but I'm not optimistic we'll see any for a very simple reason. The number of people who will participate in this contest just isn't big enough for their to be overlap with HoS subscribers. There might be a few thousand optimistically in both categories, and that's just not enough for a chance of a meaningful overlap.

I won't be participating myself, but just in case anyone's curious I have built systems like this in the past. I don't think the new class of LLMs will change things much, other than they could make it easier to build "prediction slop". The state-of-the-art is probably still machine learning prediction models like XGBoost, Random Forests, Cat Boost, etc... These are models that are built off of historical data from prior statistics about the players and outcomes from previous matches and / or aggregated predictions from other sources like gambling websites. I'd expect that "and / or" to be the major differentiator between winning bots and also-rans, but I couldn't tell you what the right answer would be without actually doing the hard work.

The other wrinkle on what I mentioned above is that there are a new class of machine learning models called "pre-trained models" that would probably outperform the above mentioned "XGBoost, Random Forests, Cat Boost". These are kind of interesting because they've been fed information about tens of thousands of other predictions tasks, like "predict how much people will spend on music next year" "predict which hospital patients die" or "predict what the weather will be next week" and somehow they can use this apparently unrelated information to somehow get smarter on a new task, like "predict who will the US Open". If I had to guess somebody using one of those pre-trained models and absorbing player stats and / or betting odds from gambling sites will win.

No posts

Ready for more?