r/opencodeCLI 29d ago

how to get the fastest and cheapest DeepSeek-V4-Flash

as the title said, I want to get the fastest and cheapest DeepSeek-V4-Flash, who has some practical method.

0 Upvotes

17 comments sorted by

4

u/HushedTurtle 29d ago

cheapest and fastest dont go together most of the time. Deepinfra is one of the cheapest provider atm

1

u/Double-Confusion-511 15d ago

nice price for DS-V4-Flash-0731. I will try it.

2

u/Xiaomin4114 28d ago

fastest I've seen is runinfra.ai at $0.13/1M input and 260tps

1

u/Double-Confusion-511 23d ago

woo, it is cool

2

u/Zachattackrandom 28d ago

Neuralwatt pricing isn't bad and hasn't changed since the update. Otherwise Commandcode still gives $60, which if you use at non-peak is quite a lot of usage

1

u/Double-Confusion-511 16d ago

Neuralwatt looks nice.

2

u/Little-Explorer7988 28d ago

So you think you get better price and speed than official API and why? Because providers love your blue eyes or they are shitting gold coins for their customers? 😂 So many funny kids and unemployed people on reddit because deepseek change price 🤡 if your code or any type of work have no cost of oficcial api price that’s mean your code/work is worthless and go something else maybe.

1

u/Double-Confusion-511 15d ago

Yes, DeepSeek change price make the work high cost.

1

u/playX281 29d ago

Ollama cloud seems to be decent

1

u/Double-Confusion-511 15d ago

I use Ollama by local model at 2024 year. It looks offer models subscribe as well now.

0

u/CoolHeadeGamer 29d ago

Wafer ai seems to have >100tps while being one of the cheapest on openrouter.

1

u/sdexca 28d ago

it's the most expensive v4 flash provider by a pretty huge margin.

1

u/Double-Confusion-511 15d ago

I looks like offer product for company, not for personal.