Testing Claude Sonnet with a Strategy Board Game: Rule Adherence Challenges

Testing Strategy Games with Claude Sonnet
A developer on r/ClaudeAI tested Claude Sonnet by playing OFMOS® Essential, a patented strategy board game where players manage a product portfolio across a positioning map. The test involved playing the game manually against the model, prompt by prompt.
Implementation Details
The developer designed a structured system prompt containing:
- The full ruleset of OFMOS® Essential
- A text-based board representation
- Action definitions
- Scoring instructions
- Turn management directives
After each turn, Claude updated the board state and running scores based on the structured prompt system.
Performance Assessment
Claude Sonnet demonstrated several capabilities:
- Understood the game rules correctly
- Articulated strategic reasoning during gameplay
- Tracked scores consistently throughout the game
However, the model frequently made illegal moves. The developer noted this was expected behavior since the system lacked a constrained move-generation layer, requiring the model to self-enforce rules—a task where it often broke down.
Developer Questions
The developer is seeking community input on similar experiments with board or strategy games, specifically asking about:
- Experiences with rule adherence in different models
- Observations about strategic depth in AI gameplay
- Which models performed best in similar scenarios
This type of testing is useful for developers working with AI coding agents to understand the practical limitations of language models in rule-based environments where precise constraint enforcement is required.
📖 Read the full source: r/ClaudeAI
👀 See Also

Non-developer builds personalized AI news editor with Claude
A non-technical user created a personalized daily news briefing system using Claude AI, starting with a simple summarization prompt and evolving into a full toolkit with context-aware filtering and bias checking.

OpenClaw Agent Memory Continuity Solution Using Database Query System
An OpenClaw user solved agent memory continuity between sessions by implementing a database that stores session data, allowing the agent to query past references instead of storing entire sessions in context. The agent named Sage could remember previous conversations after session resets using this approach.

OpenClaw user automates dating app interactions with AI agent
A Reddit user built an OpenClaw agent that handles swiping, conversation management, and match filtering on dating apps, reporting 500+ swipes per day and 3x more matches after one week.

Analyzing 7 Years of Diary Entries with an LLM: RAG vs Fine-Tuning Failures
After keeping a diary since 2019, a developer fed 200+ entries to an LLM to discover patterns — RAG failed, fine-tuning failed, and privacy was a constraint. The final approach revealed cyclical life lessons every two years.