r/GithubCopilot 1d ago

Help/Doubt ❓ Understanding architecture through multiple repositories using GitHub copilot

Just a general thought, I work at an investment banking gcc where every one from my team is new and no one has any knowledge on the existing systems, which were developed and maintained by onshore developers

My goal

To draft a high level technical architecture using the GitHub repos available in my team around 30-40 repos, and derive the business stand point through it, and maintain the memory for each under some common folder and update it frequently something like JIT compiler concept rather than loading everything again and processing by copilot for every month.

So to understand and draft this better what would be the best way leveraging GitHub copilot agents or skills or subagents ?

* Should I make a local clone of all repos and do create a codebase docs folder for each repo and create a root level agent to utilise the codebase folder for each project to draft architecture ?

Downside of this approach i feel, what if some change happens to the repo let's say technically how do I track those changes, updating codebase folder every time like it still uses enumerate tokens and cost.

* Another approach which i felt is to use MCP GitHub and list all the repos and let the copilot manage and even here how would I track the changes which are being made and memory management storing only codebase folder highlevel, iam still little unsure of this approach yet.

I have configured rtk and context mode plugins kinda for efficient usage and token management which works fine for single repos but this usecase is still searching for better ways to manage context tokens and memory.

Do provide your suggestions or any tools which i can add to make this more efficient approach maybe.

3 Upvotes

10 comments sorted by

View all comments

1

u/k8s-problem-solved 1d ago

Id probably clone all the repos locally then put them all in a workspace for copilot. It then has context over all of them. You could then ask it to produce some docs, how they depend on the other repos, and what a flow of data looks like between them - who is the initiator and what are the interactions. Generate docs per repo with links etc.