Researchers reformulate LLM block pruning as a constrained binary optimization problem mapped to an Ising glass spin system, enabling efficient ranking of pruned configurations without benchmarking each candidate. Unlike mean-field methods that treat blocks independently, this approach accounts for pairwise couplings between blocks, achieving 23 percentage points improvement over competing methods at 50% compression of Llama-3.3-70B-Instruct on MMLU.