Game or device state
Simulation today; visual device path as the deployment direction.
mixed maturityNPLUS / Kodo
Kodo rebuilds Brawl Stars as a controllable game environment first, then uses that system as the foundation for reinforcement learning.
OVERVIEW / CURRENT DIRECTION
The current architecture direction keeps a stable actor-facing contract while the observation source changes between simulation and a device perception path. Exact model and library choices remain under review.
Learning may use information unavailable to the deployed actor. Any such branch must stay visibly separate from actor input.
STACK / WORKING ARCHITECTURE
This is the public contract, not a frozen implementation spec. Components show their role and authority state; exact algorithm, model-family, dimension, and library choices are deliberately withheld until locked.
Simulation today; visual device path as the deployment direction.
mixed maturityA representation the policy can consume across execution regimes.
working designTemporal and entity representation are still under architectural review.
under reviewThe policy family is a current direction rather than a locked public claim.
under reviewActions return to the simulated game or, later, the device-control path.
working designPLAN / WORKING ROADMAP
The latest simplified planning revision is used here as a roadmap, not as proof that every stage or stack choice is already implemented.
Training results do not count until the current environment and observation blockers are closed.
details held until confirmedCollect representative gameplay from the real game.
working routeInfer observation and action labels from recorded play.
working routeUse human data to initialise a policy before reinforcement learning.
plannedContinue learning against a changing population of agents.
plannedEvaluate perception, delay, aim control, and the sim-to-device gap.
test directionPUBLIC SIMPLIFICATION Detailed blocker IDs, exact model names, action dimensions, observation dimensions, and hyperparameters are intentionally not reproduced here.
RETRAINING RISK Device-side changes may invalidate earlier checkpoints; that dependency stays visible in the roadmap rather than being hidden as implementation detail.
STATUS/ TRUTH LAYER
Kodo now separates what has been reported implemented from architecture that is likely but not yet locked.
CONFIRMED BUILD
WORKING DIRECTION
PUBLICATION POLICY
AUTHORITY BASIS / NPLUS teammate reply + revised planning diagram µ 2026.08.22 µ deleted chat content is not used as evidence.