Posted last month
The role involves designing and optimizing large‑scale inference platforms for xAI's Grok service, focusing on distributed infrastructure, low‑level GPU kernels, and high‑concurrency serving at massive scale.