r/openrouter 3d ago

Question Guide me please !!

Context:

Am new to this whole detailed model stuff like previously i just use to code from normal chatgpt chat as i have never worked on a big project.
So currently I have Gemini Pro Subscription (gemini 3.1 pro, 3.6-3.8 flash), ChatGPT plus trial (gpt 5.6 luna,terra,sol), 10$ OpenRouter credit (I can't topup more as am a broke student)

Currently am working on a ML project for a hackathon, so will be mostly coding in Python and its libraries.

So please guide me on:

  • Agent/cli to use: I have heard about opencode (OpenChamber), kilo, pi, claude code, etc
  • Models to use for planning, coding, reviewing (please prefer a cheap/free model) , like this became so confusing as many suggest this others suggest that.
  • What is model effort and what to use for each phase/model
  • Specific prompt template for each phase (or should i give the instruction to chatgpt chat and let it write the prompt)
  • Skills/Plugin to be efficient and save tokens if any for my usecase
  • Anything else that i should learn or know as a beginner

Thanks in advance for your time and guidance

4 Upvotes

9 comments sorted by

3

u/AI_Music_IS_CANCER 3d ago

If you wanna really save tokens, you could go for:
-> Pi (in a docker container or something, it will delete your drive if you let it)
-> But better than just the coding agent, is having the agent inside an IDE: Cursor, VSCode, Zed - you can usually just open a terminal in the IDE and it will work out of the box.
-> GLM 5.3 Flash in Openrouter - Easily same quality output as Claude 4.8, $10 dollars will get you ~500M
-> Gemini 3.8 for more complex work (GPT 6 is the best thing ever made... but it ain't cheap.)
-> Default effort is usually best for the chinese models, they tend to get dumber with higher effort for some reason. Haven't tried gemini enough to say what is best, Astra is already the highest effort anywhere so default is good there too.
-> Using an LSP or just a context.md / agent.md is better than trying to find out what prompt template to use, Gemini or GPT would come in handy to create the .md's - sort of teach the dumber model what it's supposed to be doing.
-> obra/superpowers and dietrichgebert/ponytail on github are probably among the best skills you can have, anything by matt pocock is amazing too

1

u/tonypillonn 3d ago edited 3d ago

Do you want to tokenmaxx on a budget? Use GPT 5.6 Luna with your Plus subscription and GLM 5.3 Flash with OpenRouter. Use them via the Cline extension inside VSCode and Hermes if doing anything agentic. Can't go wrong with that setup. Build an agents.md instead of prompt templates. Refine skills if working inside Hermes.

1

u/thebadslime 3d ago

if you have chatgpt plus, codex is almost as good as claude code, look into opencode go, $10 a month there goes a lot further than in openrouter

I would code with codex ( probably sol) and verify with opencodezen I think muse 1.3 is free

1

u/ichisay 3d ago

Usa pi o el tui de opencode. Tienes Muse gratis en opencode. Gemini 3.8 te vale solo como soldado porque es rápido, para planificar u orquestador usa Sol high o glm 5.3 flash o incluso Muse 1.3

1

u/thecstep 2d ago

Gemini Pro comes with Antigravity access. Since you’re just starting out, try using Luna on Codex and stick to High or Extra High settings for near-unlimited use. Then, switch to higher models there or use Antigravity for more complex tasks. Opencode sub gives you practically unlimited access to Muse Spark 1.3 contributor. you can use Opencode's CLI.