Advice on how to land GPU/ML Systems interviews [D]
Mirrored from r/MachineLearning for archival readability. Support the source by reading on the original site.
| Hi, I’m a recent MS in Data Science grad seeking entry-level GPU/ML Systems or Inference Engineering roles. (Currently on F1 OPT). Background: CUDA, Metal, and Triton work - built a FlashAttention kernel in Metal for Apple Silicon (tiled, online softmax, fp16 with simdgroup matrix intrinsics), a distributed Mixture-of-Experts layer using NCCL alltoall for expert-parallel token dispatch with a custom Triton kernel, and various CUDA kernels using shared memory tiling and warp-level primitives. Experience: Worked at a FinTech for a year then came to the US for Master’s. No job experience as a HPC engineer. Few things I'd appreciate help with:
Happy to share more details if useful. Thanks for reading. If anyone is looking for similar roles or is in a similar situation, I’d like to connect with you, feel free to DM me! [link] [comments] |
More from r/MachineLearning
-
A collision-entropy floor for watermark/retrieval AI-text detection. Looking for a sanity check before I take this further [D]
Aug 14
-
Are supervised and unsupervised learning still relevant today? [D]
Aug 14
-
TMLR Relevance and Prestige [D]
Aug 13
-
Reproducible canvas-aligned low-level patterns in somerandomllm-generated images and their possible relation to iterative editing artifacts [D]
Aug 13
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.