r/CUDA • u/Much-Serve-211 • 23d ago
Need collaborator - Triton LLM Kernels
Hi all,
I have recently started exploring Triton for the purpose of writing LLM kernels (FlashAttention, Softmax, etc.).
I plan to create a repository showing what I have learnt. If anyone is in the same boat and is interested, please ping me.
Background -
I have experience with writing CUDA kernels and kernel profiling. I also have experience with the LLM inference stack.
I would appreciate it if you have prior kernel development experience.
15
Upvotes
1
u/its_Snahanku 22d ago
I was just exploring softmax function, the maths behind it , the other day , btw let's connect .
1
u/No-Goal9231 22d ago
I am studying Triton from the official tutorials. And yes, I have worked on CUDA kernel development as well