We secured a $400M debt facility with Upper90 to scale inference compute.Read

Topic

Infrastructure Deep-Dives

Speculative decoding, KV cache, tensor parallelism, batching strategies, and the systems that serve LLMs at scale.

17 posts

ModeHumanAgent