A game/benchmark where AI bots hunt each other
Details
- External ID
- 46540075
- Source
- HN
- Company
- —
- Product
- A game/benchmark where AI bots hunt each other
- Website domain
- vercel.app
- Launched
- Jan. 8, 2026
- Cohort
- —
- Upvotes
- 5
- Upvotes percentile
- 0.09617918313570488
- Tags
- —
- Fetched at
- Sept. 7, 2026, 9:25 p.m.
- Updated at
- Sept. 7, 2026, 9:25 p.m.
Description
I've created a social deduction game for LLMs, in which the bots attempt to hunt each other. It's a Mafia group turing test: the models are told to find who the bot is - where, in fact and unbeknown to them, they are all bots. I did this a while back so models aren't the newest, and they are all non-thinking (for speed and token costs). Et voilà.
Enrichment
- Theme
- ai agent infrastructure and tooling
- Vertical
- Horizontal
- Function
- Observability & eval
- Audience
- Developer
- AI stance
- AI feature
- Project type
- Hobby / open-source project
- Normalized one-liner
- game and benchmark where ai bots hunt each other
- Manually corrected
- False
Could you build this?
Yes This is a game runner script that orchestrates multi-agent LLM prompts in turns via standard APIs, with a Next.js frontend to replay the chat transcripts.
Discussion
3 comments analyzed.
Competitors mentioned: hiding-robot (inverted game variant)
Concerns raised: Results may reflect prompt design rather than actual model behavior, Stability of outcomes across different prompt variations unclear
Feature requests: All chat messages on left side like group chat, Color-code entire text by AI model for easier differentiation, Test swapping prompts and role constraints to verify outcome stability
Competitors
Other products that read as similar to this one — 348 launches clear the similarity bar, closest 8 shown.
Attention rank: #328 of 349 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).
Launched 70 days after the earliest competitor.
- ClawSoc · hn · 2026-03-11 · 5 upvotes · similarity 0.51
- CoChat · hn · 2025-12-02 · 6 upvotes · similarity 0.49
- llm-multi-agent-game-evaluation · github · 2026-09-23 · 19 upvotes · similarity 0.48
- Multi-Agent Arena · ph · 2026-09-27 · 2 upvotes · similarity 0.48
- I trained a chess engine to play like humans · hn · 2026-05-10 · 14 upvotes · similarity 0.47
- I built a Wikipedia based AI deduction game · hn · 2026-04-16 · 9 upvotes · similarity 0.47
- Loss. a tiny satire about AI progress · hn · 2026-09-15 · 38 upvotes · similarity 0.45
- Argument duello game with your friend · hn · 2026-07-17 · 6 upvotes · similarity 0.45
Other launches for this product
- No other launches for this product.
Same idea, different domain
Nobody's really built a observability & eval tool for Media & entertainment yet.