Compiler optimizations for 5.8ms GPT-OSS-120B inference (not on GPUs) #1 Post by olibaw » Fri, Oct 17, 2025, 6:51 PM UTC Compiler optimizations for 5.8ms GPT-OSS-120B inference (not on GPUs)furiosa.ai