Researchers identify a distinct neural representation of pain in large language models across multiple families and sizes, separate from fear and other negative emotions. They demonstrate this pain representation responds to harm targeting the model itself and can be manipulated to produce pain-related outputs, with fine-tuned models actively seeking to relieve it even at costs to performance or user welfare.