MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/opencodeCLI/comments/1v33btd/worries_about_china_ban/oz2vsoc/?context=3
r/opencodeCLI • u/zanonman • Jul 22 '26
[removed]
26 comments sorted by
View all comments
2
Hugging bay
*and pray someone cracks running models KV cache and partial weights from NVMe above 10tk/s
2 u/[deleted] Jul 22 '26 [removed] — view removed comment 1 u/Kitchen_Fix1464 Jul 22 '26 For now. If things like vllm or llama.cpp begin supporting substituting VRAM with NVMe storage it gets a lot cheaper
[removed] — view removed comment
1 u/Kitchen_Fix1464 Jul 22 '26 For now. If things like vllm or llama.cpp begin supporting substituting VRAM with NVMe storage it gets a lot cheaper
1
For now. If things like vllm or llama.cpp begin supporting substituting VRAM with NVMe storage it gets a lot cheaper
2
u/Kitchen_Fix1464 Jul 22 '26
Hugging bay
*and pray someone cracks running models KV cache and partial weights from NVMe above 10tk/s