Fastllm: A LLM inference library that runs DeepSeek-V4 with 10GB VRAM #1 Post by nogajun » Tue, Jun 30, 2026, 3:47 AM UTC Fastllm: A LLM inference library that runs DeepSeek-V4 with 10GB VRAMgithub.com