Cutting LLM Batch Inference Time by Half with Dynamic Prefix Bucketing #1 Post by DISCURSIVE » Thu, Nov 20, 2025, 10:25 PM UTC Cutting LLM Batch Inference Time by Half with Dynamic Prefix Bucketingdaft.ai