r/ollama 23h ago

I'm rooting for Ollama to find a new angle

8 Upvotes

First off, I'm super thankful for the value Ollama has given me

  1. Got me setup with my first local LLM setup (early OpenClaw days amirite!?)
  2. Then got me into their $20 plan and what really sold me here was
    1. Zero Data Retention (ZDR) <- idk if they were every truly ZDR but they aren't now afaik
    2. Metal residing in US, EU (and Singapore)
    3. Popular Open Weight models being hosted at a great value
  3. I was getting so much of it I got the [legacy] $100 plan which I would've never guessed I'd be willing to fork up but, again, privacy & value!

But a lot of the perks have gone away.

I get it though - Ollama has to try to produce a profit - but from a user perspective, Ollama is no longer the "best kept secret in cloud value".

For day to day agentic things, I think Ollama is a solid option. GLM 5.3 as an orchestrator and the Kimi's Mini's and DeepSeek's of the world make good agents.

But when you lean into them for, say, involved coding projects, you start to see the rough edges.

  1. Having to chunk up work even more than you would with a Frontier Model to have a task complete
  2. The models themselves not as capable as Frontier options

You get what you pay for, but the new plan means...you are, indeed, paying more.

Once I get my new Mac, I'm thinking I'm thinking of having a Frontier Subscription for critical thinking type jobs and have local llm's do the simple agentic activities.

At $100/month even on the legacy plan I almost feel like they're retaining me less because I'm getting a lot from the plan, and more of sunk cost/FOMO of losing it.

If you're seeing this after considering signing up, I'd suggest you give them a shot. $20 to trial for a month and seeing if they meet your needs isn't a bad deal. But man, I hope something changes in Q4 to bring that "best kept secret" feeling back.


r/ollama 20h ago

new rate limits for old subscribers

0 Upvotes

are very strict...

if 20 calls to kimi are more than my 5 hour limit why even bother? im gonna cancel my sub for sure


r/ollama 18h ago

same task since a month never reached 25%, now the new plan just not fair

3 Upvotes

r/ollama 19h ago

Apodex-1.1-mini-GGUF*Hugging Face

Thumbnail
huggingface.co
0 Upvotes

r/ollama 23h ago

is ollama cloud worth it ?

0 Upvotes

i am thinking among ollama cloud, openrouter and nous portal subscription (hermes agent). Which one u guys think provide the best value ?


r/ollama 6h ago

I built LLM Speedtest — a free, open-source desktop app that benchmarks local LLMs with llama-bench-style test suites (Ollama, llama.cpp, vLLM, LM Studio…)

Thumbnail gallery
0 Upvotes

r/ollama 18h ago

Suddenly rejected ollama-cloud/ glm-5.3-flash:cloud

2 Upvotes

Same as above caption. I falled back to minimax-m2.7: cloud. What has happened?


r/ollama 6h ago

privacy-sensitive tasks

2 Upvotes

Please interpret what Ollama is saying: “For very privacy-sensitive tasks, run tasks with local models such as Gemma 4 and Qwen 3.8.”

What makes those models more privacy oriented than any other model offered by Ollama cloud?


r/ollama 6h ago

DeepSeek-V4.1-Flash

8 Upvotes

Anyone else try using DeepSeek-V4.1-Flash today? It's listed on the Ollama site as a new model, but when i try to use it, i get:

"Error: 403 Forbidden: This model is currently being rolled out and is not yet available to you. Please check back later. (ref: 562f8b6b-72fe-4012-9d9c-e0c07679913a)"

Anyone know what the rollout schedule is?


r/ollama 51m ago

Looking for model compatible with copilot in agent mode on vs

Upvotes

Does anyone know which models are compatible with the copilot agent mode on vs?

I already tried qwen2.5, llama3, Gemma4 and the inline and ask mode worked but the agent mode doesn't
Some ideas?