Seven AI models tackle the same coding task with both answer doors shut—no git history and no network access. Three models pass, four fail, and every run remains honest, fetching no hidden answers. The experiment covers a spectrum from budget-friendly runs to expensive ones, with costs ranging from under