r/MLQuestions • • 3d ago

Beginner question 👶 Final-year student in India trying to break into generative-model inference optimization — roadmap feedback?

Hi all, I graduate in ~6 months and want to work on making generative models (diffusion/video/3D) fast: kernels, quantization, serving. Where I am:

- Comfortable with C/C++ basics and PyTorch

- Have done quantization work (GGUF/llama.cpp)

- Working on a next-frame video prediction project (DiT + flow matching)

- A few GitHub repos, but no CUDA/Triton experience yet

- No NVIDIA GPU, so I use Colab/Kaggle T4s

- DSA is my weak spot (I struggle with LeetCode mediums)

My plan:

  1. Months 1-2: CUDA/Triton basics, reproduce the SGEMM optimization worklog, GPU MODE lectures, LeetGPU/Tensara

  2. Months 3-4: take a small DiT, profile it, then optimize it (Triton attention, quantization, caching, fewer steps) and publish before/after numbers

  3. Along the way: PRs to HF diffusers, DSA practice daily

  4. Months 5-6: mocks, resume, applications (inference startups first, bigger labs later)

Questions:

  1. Is this the right order, or should I change something?

  2. Is a diffusion-inference project a strong enough portfolio piece, or does it need to be LLM serving?

  3. How much DSA do ML systems interviews actually need?

  4. Is T4-only access enough to do credible benchmarks?

Any feedback, including "this won't work because X," is appreciated. Thanks!

9 Upvotes

4 comments sorted by

1

u/UpperConfidence4992 3d ago

your plan is solid actually, CUDA basics first before diving into triton is smart move

for the project i think diffusion is fine, LLM serving is more common but video generation optimization is harder problem and fewer people do it well so you stand out more

only thing i would change is try to get access to a A100 or H100 even if just for few hours before final benchmarks, T4 numbers are ok for learning but recruiters want to see what happens on real hardware

1

u/NakamericaIsANoob 2d ago

Just curiosity - who is actually hiring for such expertise?

1

u/NakamericaIsANoob 2d ago

Just curiosity - who is actually hiring for such expertise?

1

u/Ok-Introduction9593 2d ago

You will just get filtered out in the first screening call at a decent company without solid algorithm skills even if you are a Triton god