Posted last month
Design, build, and optimize massive GPU clusters for AI training and inference; develop low-level CUDA kernels, Linux kernel internals, and custom orchestration; collaborate with research teams to accelerate supercomputer performance.