A deployed model learns nothing after training — no exploration, no adjustment, no toddler-style experimentation. This companion visualizes the paper's A/B/M architecture: observation-based learning, action-based learning, and a software-defined-networking-style orchestrator meant to route between them automatically.