French startup Kog, led by solo founder Gaël Delalleau, is developing low-level software optimizations to accelerate AI inference on standard datacenter GPUs like the Nvidia H200 and AMD MI300X. Following a May 2026 tech preview that achieved 3,000 tokens per second on a 2B parameter model, Kog is shifting focus to larger models to satisfy demand from 200 business leads. Backed by Varsity VC and Bpifrance, the 11-person team aims to achieve 10x speed on a major model by September 2026 to kick off its Series A round.
Sep 29, 2026 · 2 sources
Sep 28, 2026 · 7 sources
Sep 30, 2026 · 3 sources
Sep 30, 2026 · 7 sources
Story comments
Loading comments…