Posted 6 days ago
Design and optimize low‑precision RL inference stack, focusing on performance, profiling, and collaboration with modeling team. Requires experience in large‑scale distributed systems and LLM inference.