A developer describes configuring Codex into a multi-agent system with specialized sub-agents bound to different models to reduce token consumption and improve context window efficiency. Rather than using a single powerful model for all tasks, the approach routes different job categories to appropriately-sized models, saving tokens while maintaining output quality.