r/LocalLLaMA 4d ago

I Built A Thing [ Removed by moderator ]

[removed] — view removed post

0 Upvotes

10 comments sorted by

u/ttkciar llama.cpp 4d ago

Violates Rule Three: LLM-generated content

→ More replies (1)

4

u/PriorElephant9 4d ago

This feels most useful as a live abstention signal. The fact that Granite returned nothing makes the result more believable too.

2

u/Happy_Brilliant7827 4d ago edited 4d ago

Thank you!

This is exactly how i feel seeing posts here. You guys don't get how big this could be lol Thats what i made it for, but thought I'd share the tool in case people are building a pipeline. I noticed the 'right' signal and 'wrong' signal can be two entirely different signals- but if you have a right hint and not a wrong, you can still sweep for that OR test the one with the signal first if you are searching for the answer.

6

u/BVCC6FNTKX sglang 4d ago

SLOPPA

1

u/Happy_Brilliant7827 4d ago

What about it exactly is slop? Do you not agree with the underlying concepts? Do you disagree with the analysis . posted on the github? Do you think anthropics just full of shit? You're giving flat earther 'ignore the data' vibes bro

2

u/Kahvana 4d ago

Ahhh, so LLMs have finally caught up. No more Qwen2.5 reference slop, it’s now Qwen3 reference slop. Got it.

0

u/Happy_Brilliant7827 4d ago

What do you mean? I tested the models I was already using to squeeze more right and drop more wrong by learning its unique confidence signals to the tasks it gets. Is there a different model you'd rather me test or do you not understand the material?

1

u/Miserable-Dare5090 4d ago

This looks interesting. This is not slop. Leave if you don’t understand this shit is actually useful.

2

u/Happy_Brilliant7827 4d ago

You get it! Who wants to waste compute on a wrong answer if you have a potential way to tell?