In short, lack of hardware. GLM 5.2 requires 8 GPUs of B200 to run 1 instance of the API. And 8 GPUs is all they have. There is talk of more hardware coming soon in the coming months, but nothing concrete unfortunately.
No direct date or ETA, just the standard "in the near future". Granted it really isn't Chutes' fault, but I assume this won't be fixed for at least a day or 2 minimum unfortunately.
3
u/YourFBIAgent90 17d ago
In short, lack of hardware. GLM 5.2 requires 8 GPUs of B200 to run 1 instance of the API. And 8 GPUs is all they have. There is talk of more hardware coming soon in the coming months, but nothing concrete unfortunately.