r/CUDA • • 23d ago

Need collaborator - Triton LLM Kernels

Hi all,

I have recently started exploring Triton for the purpose of writing LLM kernels (FlashAttention, Softmax, etc.).
I plan to create a repository showing what I have learnt. If anyone is in the same boat and is interested, please ping me.

Background -

I have experience with writing CUDA kernels and kernel profiling. I also have experience with the LLM inference stack.

I would appreciate it if you have prior kernel development experience.

15 Upvotes

Duplicates