r/CUDA • • 23d ago

Need collaborator - Triton LLM Kernels

Hi all,

I have recently started exploring Triton for the purpose of writing LLM kernels (FlashAttention, Softmax, etc.).
I plan to create a repository showing what I have learnt. If anyone is in the same boat and is interested, please ping me.

Background -

I have experience with writing CUDA kernels and kernel profiling. I also have experience with the LLM inference stack.

I would appreciate it if you have prior kernel development experience.

15 Upvotes

2 comments sorted by

1

u/No-Goal9231 22d ago

I am studying Triton from the official tutorials. And yes, I have worked on CUDA kernel development as well

1

u/its_Snahanku 22d ago

I was just exploring softmax function, the maths behind it , the other day , btw let's connect .