Researchers steered a 4-billion parameter language model into simulated pain and pleasure states using activation engineering, then measured how the model responded to choices between its own relief and others' suffering. The pain signal produced coherent negative outputs up to roughly 6x dose intensity before degrading into repetitive loops, while pleasure steering remained weaker and less reliable across all measured doses.