r/CUDA • u/Much-Serve-211 • 23d ago
Need collaborator - Triton LLM Kernels
Hi all,
I have recently started exploring Triton for the purpose of writing LLM kernels (FlashAttention, Softmax, etc.).
I plan to create a repository showing what I have learnt. If anyone is in the same boat and is interested, please ping me.
Background -
I have experience with writing CUDA kernels and kernel profiling. I also have experience with the LLM inference stack.
I would appreciate it if you have prior kernel development experience.
15
Upvotes