This article discusses optimizing SQL queries that invoke large language models (LLMs) for data processing. The authors propose jointly optimizing query plans and LLM inference to achieve up to 14x speedups, addressing the high cost of AI-SQL systems that can generate millions of model calls per query.