ByteBrief
Skimming the internet so you don't have to
Kog targets 30x faster LLM inference on existing GPUs | ByteBrief