Advice on how to land GPU/ML Systems interviews [D]
Mirrored from r/MachineLearning for archival readability. Support the source by reading on the original site.
| Hi, I’m a recent MS in Data Science grad seeking entry-level GPU/ML Systems or Inference Engineering roles. (Currently on F1 OPT). Background: CUDA, Metal, and Triton work - built a FlashAttention kernel in Metal for Apple Silicon (tiled, online softmax, fp16 with simdgroup matrix intrinsics), a distributed Mixture-of-Experts layer using NCCL alltoall for expert-parallel token dispatch with a custom Triton kernel, and various CUDA kernels using shared memory tiling and warp-level primitives. Experience: Worked at a FinTech for a year then came to the US for Master’s. No job experience as a HPC engineer. Few things I'd appreciate help with:
Happy to share more details if useful. Thanks for reading. If anyone is looking for similar roles or is in a similar situation, I’d like to connect with you, feel free to DM me! [link] [comments] |
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.