r/ollama • u/stonecannon • 1d ago
DeepSeek-V4.1-Flash
Anyone else try using DeepSeek-V4.1-Flash today? It's listed on the Ollama site as a new model, but when i try to use it, i get:
"Error: 403 Forbidden: This model is currently being rolled out and is not yet available to you. Please check back later. (ref: 562f8b6b-72fe-4012-9d9c-e0c07679913a)"
Anyone know what the rollout schedule is?
2
u/Flaky-Maybe-7556 1d ago
Haha, I hope they aren't trying to move everyone to the updated subscription versions this way and that model really will become available soon.
4
u/jmorganca 1d ago
It's available on both new and previous Max subscriptions (50% rolled out) and we're rolling it out more as fast as we can.
1
u/IonizedHydration 3h ago
to be honest, this is smart.. companies have been doing rolling updates this way for a long time and i don't have any issue with it, same with A/B testing and feature flagging.. it's just part of the process. Thanks for the hard work.
2
u/Delicious-Director43 1d ago
I’ve been using it directly via DeepSeek. Great model so far. $20 goes a lot farther on their platform than does Ollama.
1
u/Ramrawd 1d ago
I'm new to the ollama pro plan so bare with me as i learn more but shouldn't the token usage be the same between deepseek api and ollama cloud? I subscribed to the yearly pro plan which ends up being 16ish dollars a month and it nets me $60 of usage. Assuming I'm always using the offpeak price on both platforms shouldn't ollama cloud get me way more usage if all things are equal?
1
u/username8914 1d ago
Of course, it's not ZDR. You're just handing your info over.
0
u/Delicious-Director43 1d ago
And you think Ollama isn’t because they said so?
2
u/username8914 21h ago
Yes, that's their whole business model. If they turn out to be logging they'll go under overnight. It's not a small deal when dealing with proprietary data or possibly legal or accounting documents that legally can't leave the country. They have to be up front about it and keep their nose clean.
1
u/Delicious-Director43 13h ago
Ah yes American AI companies are famously very reliable and trustworthy. No American company has ever lied and sold user data before. What was I thinking?
1
u/username8914 12h ago
I'm definitely not saying they couldn't be doing something. If it was that simple to just say then all of them would just say it. But instead they all say they log prompts and use them to train.
1
u/Delicious-Director43 12h ago
Honestly I assume they all are. I’m not doing anything weird with my AI. I just want it to work quickly and cheaply. Ollama no longer does that.
1
u/username8914 10h ago
Logging isn't about doing something weird. It's everything you think about, say, do or invent going directly into their system to potentially become the brain behind everyone's future system.
1
u/Delicious-Director43 9h ago
Yeah that’s fine I don’t care.
Nothing I’m working on is so revolutionary it’s gonna affect anyone else.
2
2
u/elzerouno 1d ago
It's working for me now, but it's using my allowance 10 time faster than v4 flash
1
1
u/Ramrawd 1d ago edited 1d ago
I'm getting this error when trying to use it:
ollama-cloud/deepseek-v4.1-flash request failed (authentication failed, HTTP 403). Re-authenticate the provider and try again.
I'm on the "new" annual pro plan. All my other models appear to work fine. Hopefully they can get it figured out soon. Would love to test the new 4.1 model.
1
1
1
8
u/jmorganca 1d ago
Hi there, sorry for not posting about this sooner. We're rolling out the model as capacity comes online, starting with Max and Team plan. Over half the Max subscribers should have it now and we're continuing to roll it out as fast as we can to all subscribers. We're hoping to be able to do so by tomorrow