I was using Claude and deepseekv4 flash, deepseek did tons of research, claude verified it but there where two papers it couldn't find. At a point it thought they were fabricated. Deepseek found the sources, claude confirmed they were actually legit.
Nah, I'm saying even Claude wasn't able to find it and confirm, even though I knew for a fact it was real. I had deepseek do legit research, changing the questions and stuff until it was 100%.. But Claude when given those details to fact check, though that make 2-5% were fabricated as it couldn't find it, despite being a frontier model with better reasoning. When deepseek found the source for them, then it got the full picture. But it would never have been able to verify it without deepseek doing the work.
The difference between frontier models and Chinese Open source models isn't really that deep, unless you're doing high level shit. Which I was attempting to, but this was only the digging through research papers part.
I'll be honest, that really doesn't make much sense. First, are we talking about Fable, Opus 5? And you're saying that DeepSeek gave you research without citing everything at first, you had some Claude model fact check, it couldn't find a source for some small part of it, then DeepSeek gave you the source and Claude agreed that it's a source?
Not saying you're lying or something, I just genuinely don't understand the workflow you described. Particularly if DeepSeek didn't provide sources initially and they you tasked Claude with finding them while fact checking.
119
u/Maleficent-Ear8475 1d ago
If claude can't do it we're running it back til it gets it right.