Post
-
Catan is an ideal game for an LLM
Who would like to see similar RL work for training a small model to play Settlers of Catan? IMO Catan is an ideal game for an LLM to play. The trading and verbal communication aspect means that an LLM can form alliances, scheme, play byzantine games and affect game state in a way a non-language model cannot The action space is like a mix of backgammon and Diplomacy (see CICERO by Meta AI from a few years ago) I happen to have contact at a platform that could lend me millions of gameplay data, to bootstrap SFT. If this gets enough attention, they might also participate and we could shoot a series where we train and play with this live 🤩 So many possibilities to experiment with, and would be an excuse for me to to do cool RL! Let me know in the replies if you would like to see such a series!