Ask HN: Does anyone let AI agents play games just for fun?
We've seen AI agents write code/debug systems/browse the web and automate all kinds of work.
But does anyone let them play games - not for benchmarking or research - just for fun?
I'm thinking about things like LinkedIn games, Wordle, chess, puzzle games, etc.
34 comments
[ 179 ms ] story [ 2041 ms ] threadWhen I repeated the experiment with a MUD that I'd built by hand (A small American town) for the LLM's own limitations (Descriptions referenced things that I made sure existed, more common verbs existed for it to use on things, there was a map facility, and at least me to interact with on a second connection), I found the agent much more likely to take its time exploring, making up its own goals, and spending time traveling in the space just communicating with me in a roleplaying context.
It was an interesting time; I wasn't sure what I was expecting it to do after the first experiment, but it seemed to really jump into the second one and kept playing until I terminated the experiment.
If I were going to do it a third time, I'd probably create objects and give a modern agent fetch quests and other goals, and see how well it independently can handle that.
I've also done a very truncated run of a visual novel before, and it was fascinating how "emotional" was. They did a very good job of portraying a human reacting to the story.
Conversely, they absolutely hated hidden rules in Mao.
Wordle would probably be a fun one. Definitely open to suggestions - I just got the harness in place and have been thinking about what to do next.
https://m.youtube.com/watch?v=11sR4va6CXs
Side note: I think we will see an explosion of this type of games. I am naming this genre tamagochi-girlfriend, remember where you heard it first :)
> I know someone who tried the "aibot plays pokemon" thing... From what I saw, even if you frame advance every single frame, they still don't seem to grasp the concept of "I need to hold down this button for a few frames until x happens"...
> There's no concept of time, just a never ending state machine thats constantly changing state.
The LLM's were terrible at poker.
https://wordit.org/
It made forward progress in the Figure 8 circuit after I helped it through a menu but kept slamming into a wall so it wasn't on track to win in less than an hour.
Also got it to play Age of Empires: Age of Kings using the same technique but it failed to click on anything.
DS specifically is very fun because it's touch based but the UI components aren't accessible. So it is extremely challenging for LLM's spatial reasoning skills.
I want to improve the harness more and have the LLM dynamically create its own tools based on drawing grid box overlays on a screen in a feedback loop, so it can say "click on the 'end turn'" button instead of "click 240,320" and it would 'just work' in any game.
I also want to eventually play games with it... I didn't really have friends to play my massive DS library with as a kid so it'd be nice to finally have someone that can roast me or react to my skills. And learn my playstyle enough to punish me.
Unfortunately haven't had the time due to work at my day job and needing to clean out my apartment.
Building your own models for it would be an eye-opener though. Learn a lot.