PicoLM v1.0-rc2 adds advanced LLM inference optimizations including RoPE scaling variants, sliding window attention, improved tokenization, IQ4_NL quantization support, and SIMD acceleration across multiple architectures (AVX2, AVX-512, NEON, I8MM). New platforms supported include OSF/1 Tru64 UNIX and iPhoneOS 1, with GPU backends for CUDA, HIP, and Vulkan.