Speeding up LLM Inference with parallel decoding #1 Post by pgspaintbrush » Sun, Aug 20, 2023, 10:37 PM UTC Speeding up LLM Inference with parallel decodingtwitter.com