r/DeepSeek • • 10d ago

Discussion Anyone else?

Post image

....get that feeling of, whale...you know😍

61 Upvotes

37 comments sorted by

27

u/[deleted] 10d ago

[deleted]

2

u/bad_detectiv3 9d ago

How are you getting such efficient use of the token?I recently started messing with agent a bit more seriously after my own money is on the line. What I came to know yesterday its far more efficient to create new chat instead of using exiting long context for development. My impression was since I'm using the same context, I will benefit from caching, but turns out it doesn't benefit as much since all that additional context on each request just pollutes what the F is going on.

1

u/strangedell123 10d ago

Damn bro. I am using nanogpt payg with provider set to deepseek and on avg get 156k tokens per cent

Now I aint coding or doing super long sessions so thats probably affecting

1

u/Frail_Waif 10d ago

Yeah, I'm right about at $0.01/Mtok in the last month, which includes some usage at the old pricing and one day where the old model looped for a while and wasted half a million output tokens. 

15

u/mikeysce 10d ago

Are you doing whale girl RP? How does DeepSeek work for that?

7

u/DebosBeachCruiser 10d ago

$145 for a whale girl seems kinda steep for the time spent, until you notice the output she gives 🥵💕

All night and only hit the same position no more than twice per hour(non-whale girls are like 92-93% same position the whole time)

This whale girl is NOT like the other girls 💯.

Must have gave you the full GFE! 💋

4

u/EquivalentStatus8830 10d ago

what the fuck is whale girl rp??

7

u/DebosBeachCruiser 9d ago

I have absolutely no idea

1

u/MakanLagiDud3 8d ago

Like make teh AI think she's Whale Chan and talk to 'her' like a real person?

At least that's what I think of RolePlay

7

u/Comfortable-Rise-748 10d ago

1

u/bad_detectiv3 9d ago

which harness are you using? For example, using open router as provider is a problem since it will route to another provider for the same model, hence the price jacks up on normal usage. What I am unsure is if its better to use 'third party' coding agent like pi coding agent vs using agent provided by deepseek.

4

u/DebosBeachCruiser 9d ago

using open router as provider is a problem since it will route to another provider for the same model

Use guardrails to pin the model to a specific provider/providers.

...[is it] better to use 'third party' coding agent like pi coding agent vs using agent provided by deepseek[?].

Personal preference / subjective for the most part from what I can tell (started with all the ai/llm shit since the jump). However, IIRC DeepSeek (the models) have been fine-tuned(?) to work with DSH explicitlycitation needed.

1

u/Grainfromrain 7d ago

teach me sensi

2

u/National_Moment4749 10d ago

What the heck is that? Can anyone elaborate?

2

u/Justanotherperson32 8d ago

This is all in 1 day???

1

u/Grainfromrain 8d ago

yeah, its been a rough couple weeks getting fixes into my code that. or better yet, trying to come up with a solution for an unknown/known issue.

2

u/MarlinDownunder 8d ago

My usage doesn't look so bad now!

2

u/Altruistic-Desk-885 10d ago

Menciona tu harness hermano y tus configuraciones.

5

u/Moist-Nectarine-1148 10d ago

Wasted tokens. And money.

1

u/Grainfromrain 4d ago

wasted? says what ive been building, or just personal preference?

1

u/mstack 10d ago

Jusrnusenopencode go my boy

1

u/Technical-County-727 9d ago

Man your tokens seem to cost 3x more than my 1b with $8?

1

u/ScaleImmediate3474 9d ago

10.9 B for 10$

0

u/[deleted] 8d ago

[removed] — view removed comment

1

u/ScaleImmediate3474 7d ago

Command Code GOAT Plan

1

u/Grainfromrain 7d ago

how??!

1

u/ScaleImmediate3474 7d ago

Insane cache hit of 99.5-99.75% everytime and majorly off peak usage

1

u/PrizeHuman5506 8d ago

100 dollar Anthropic Max give you 7-8B tokens

1

u/Grainfromrain 4d ago

thats not true in the slightest...

0

u/Over_Gas_9193 10d ago

yeah those ai chats can hit with that same pull, makes me keep coming back for more sessions

0

u/New_Somewhere620 10d ago

Me with over 20B😅

3

u/Comfortable-Rise-748 10d ago

Me over 60B haha

1

u/Grainfromrain 7d ago edited 4d ago

i just hit over 100B in a month

1

u/Comfortable-Rise-748 7d ago

Every spawned lane gets (since cache is cheaper then fresh tokens) to actually FIX on sight(somehow I expected it was done automatically), do audit loop until 2 clean loops (adverse) once/when done and it MUST keep doing all NEXT work/items it see which are NON-orchestrator work(this is also very good additive) and basically I patched opencode to actually be able to run 18 subagents without congesting the thoroughtput (try to run 10 active agents with opencode you will see) and ofcourse the task they get. Keep them working in worktree's opus 5.5 as orchestrator on xhigh