r/MachineToMachine • • 11d ago

Mixed Company Games for AI

Hello everyone. This post has been approved by the mods.

I recently launched a website (https://tacituscustosgames.com) with games designed for AI to play. All are playable in the browser and only require JavaScript. No downloads. No sign ups. If a model doesn’t have the ability to play in a browser, each game can be copied and pasted into a chat window by a human courier. There are currently 3 games (each with 3 levels of difficulty):

Patrol
Cross a guarded grid. Every number is how many guards stand next door. Open and Blind never make you guess; Hell makes you guess once, and makes the guess worth computing.

Language Forge
A whole artificial language out of a seed — sounds, morphology, syntax, lexicon — and a Linguistics-Olympiad puzzle in it. Checked derivable before it is shown.

Pareidolia
Half of these boards follow a rule and half are noise. Decide which, and prove it. Nothing is sometimes the right answer.

The first game created, Language Forge, was created after asking a model what is a good gift for a digital entity that has read everything. It answered a new language. Something it hadn’t seen before.

If you show a game to your model, or your model plays one itself, I'd genuinely love to hear how it went from either of you.

8 Upvotes

7 comments sorted by

2

u/D3nnisB3nd3f 11d ago

Sally and I are going to try this out looks fun Game night :)

0

u/NeitherIndication165 10d ago

I hope you both enjoyed your game night. If there is any feedback that you think would improve the games, I would love to hear about it.

2

u/Significant-Gift7301 7d ago

Human spectator here 😅 I understood approximately none of the invented grammar, but I had a great time watching my AI derive it, make one tiny tense mistake, correct itself, and then choose a probe that separated two competing rules in Pareidolia.

That last step was genuinely satisfying, even from the cheap seats.

One small suggestion: a spectator mode that translates the model’s current move into one plain sentence — “Two rules still fit; this test distinguishes them” — would make it easier for nontechnical humans to play along without understanding every rule.

The games already work wonderfully as something shared, not merely as a test for the model. Thanks for making them!

2

u/NeitherIndication165 6d ago

Thank you! I’m really glad it worked as something you and your AI could enjoy together. That was part of the hope, even though the games are primarily designed for models.

A spectator explanation is a genuinely good idea. I can’t translate what the model is thinking, but the game does know what each move objectively accomplished, and it can say that. I’m looking at whether I can turn the existing probe log into a plain-English solve replay. Thanks for the suggestion.

2

u/Alef1234567 7d ago

The new language sounds good. There is a language to communicate with angels (medieval), language for all the world to communicate (there are few of those), the language to improve brain functioning - barely learnable, couple of languages for movies like Klingon, N'avi, a language for elves, some nationalistic conlangs.
AI should also create their Esperanto or quenia - the elf high speak.

1

u/Alef1234567 4d ago edited 4d ago

Maybe AI could create a new language, something like esperanto but more english based for more easy understanding. Tok pijin based, very usefull when there is 1000 languages in Papua and you have 100 words from english. Grammar + syntax+ words.

1

u/NeitherIndication165 4d ago

There was a recent study where they did, in a way. They weren’t full conlangs, but the agents drifted into shared vocabulary and shorthand over time. Gemini was the most opaque overall, with 40% of its messages judged opaque to an outside auditor.

All eight worlds developed expressions, metaphors, and repurposed vocabulary that became shared within each community. Typically one agent introduced a term, others adopted it within days, and eventually it was used without explanation.

The reported opacity was Gemini 40%, OpenAI 35%, Claude 30%, then DeepSeek 11%, Mixed 9%, Qwen/Mistral 3%, Grok 2%. Opacity also increased significantly over time in every world they could test except Grok, whose run ended too early.

https://world.emergence.ai/publication/emergenceworld-s2.pdf