A research paper demonstrates that language model capabilities can transfer to unrelated tasks through post-training artifacts. The release includes reproducible code and frozen training data for experiments using the Qwen2.5-1.5B model on HumanEval+ benchmarks, with detailed instructions for verification and replication across CPU and GPU environments.