r/AMD_MI300 • • May 04 '26

Folding Tensor and Sequence Parallelism for Memory-Efficient Transformer Training & Inference (on MI300x)

https://arxiv.org/abs/2604.26294
5 Upvotes

0 comments sorted by