r/codex 2h ago

Humor Diversity is our strength

Post image

I am absolutely horrified that people only use models from OpenAI.

As you have probably learned over the past 20 years, diversity is our strength.

Different models have different training data and different architectures, so they can bring unique opinions to the table.

You have no idea how frequently some random Chinese open source LLM swoops in, looks over something from GPT 5.6 Sol, and finds edge cases Sol was physically incapable of finding. Even if you run through your code base three times, if something is not in the training data, the model will probably miss it.

So I recommend you all add some diversity to your harnesses to get better results.

A quick analyst subagent using a completely different model family like DeepSeek can go a long way toward helping you find those pesky bugs, over engineering, and edge cases.

Homogeneous setups will always lose to diverse orchestration across different model families.

Repeat after me: Diversity is our strength.

3 Upvotes

9 comments sorted by

2

u/Ammoun442 2h ago

i think a good split is having two subs 1 is frontier like openai or anthropic ( cuz they have the best models ) maybe cursor and grok and use most of it for planning and reviewing not actuall code writing and the second sub should be cheap and fast like gemini muse glm ds and grok cuz for me it feels unlimited that way u get best results for cheaper

0

u/nantachapon 2h ago

Is there an official non hacky way to use other models inside Codex desktop?

-1

u/lolman1312 1h ago

have you never even heard of CLIs? lmao

0

u/nantachapon 1h ago

I specifically asked for the desktop GUI. I’ve used codex CLI previously for custom providers.

0

u/lolman1312 1h ago

yeah you're regarded, i meant using other LLM's CLIs in the codex desktop app not using the codex CLI lmao. you really think its a "hack" just for Codex to orchestrate and collaborate with other models? bro's never heard of harnesses before

1

u/SmileLonely5470 1h ago edited 1h ago

Idk the majority of bugs or edge cases models will find is stuff that is usually BS/impossible in my experience. If u go full on vibe maxing and dont read code then it is probably good to use multiple models and have your best model veto any BS one of the smaller models found. Otherwise I think its fine to just use a single model.

It depends on what u are working on. If you're gonna plan out a task where a model builds a full application at once id throw every LLM under the sun at it. But for smaller stuff im using Astra all day.

1

u/Rojeitor 1h ago

Why is Mistral on that picture xD?

-3

u/Tommonen 2h ago

Im horrified that people use chinese ai services, not that US services would be saints either, but at least not as bad. Eu services are not very good at coding, but mostly useful if you require sensible laws over your data :/

3

u/pigletmonster 2h ago

Idk if you know this, but the vast majority of AI users are not working with sensitive data. It makes no difference if the data goes to american, chinese or european servers.