A comparison of three AI models (GPT-5.6-Terra, DeepSeek-V4-Flash, and Claude Sonnet 5) tested their ability to generate Ansible configurations across three scenarios of increasing complexity. Spotter analysis found 140 errors, 94 warnings, and 221 hints total, with missing fully qualified module names being the most common error type across all models.