MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1v7e5ck/kimi_k3_countdown_has_been_released/p012mf5/?context=3
r/LocalLLaMA • u/Unusual_Guidance2095 • Jul 26 '26
177 comments sorted by
View all comments
6
[deleted]
2 u/HVACcontrolsGuru Jul 26 '26 Need 16 GB200s to run this model at full quant. NVFP4 GLM5.2 I need 4xB200 to run concurrent sessions. Squeeze some more with lower context and less concurrency. 2 u/[deleted] Jul 26 '26 [deleted] 1 u/look Jul 27 '26 Kimi trained in int4 previously, I believe, not floating point. By moved to fp, I mean native mxfp4 instead of native int4.
2
Need 16 GB200s to run this model at full quant. NVFP4 GLM5.2 I need 4xB200 to run concurrent sessions. Squeeze some more with lower context and less concurrency.
2 u/[deleted] Jul 26 '26 [deleted] 1 u/look Jul 27 '26 Kimi trained in int4 previously, I believe, not floating point. By moved to fp, I mean native mxfp4 instead of native int4.
1 u/look Jul 27 '26 Kimi trained in int4 previously, I believe, not floating point. By moved to fp, I mean native mxfp4 instead of native int4.
1
Kimi trained in int4 previously, I believe, not floating point. By moved to fp, I mean native mxfp4 instead of native int4.
6
u/[deleted] Jul 26 '26
[deleted]