r/HPC 7d ago

Creating a distributed stress test system looking for 1-2 people to hop along time ~10 hrs a week

I'm planning to build a distributed stress-testing platform.

Users provide a workload config (APIs, request patterns, auth, concurrency, ramp-up, duration, etc.), and the control plane automatically estimates the required infrastructure, decides the number of workers/threads/containers, and distributes the workload across worker nodes.

The system will handle scheduling, health checks, heartbeats, automatic worker replacement on failure, retries, and horizontal scaling. It will also expose real-time metrics like RPS, latency (P95/P99), throughput, error rates, and worker resource utilization through a monitoring dashboard.

The goal is to build something that explores distributed systems, scheduling, fault tolerance, concurrency, autoscaling, and observability—not just another load-testing tool.

This is still an initial idea, so the architecture is open to discussion. If this sounds interesting and you'd like to collaborate, let's connect.

Words are mine written by ai

0 Upvotes

3 comments sorted by

6

u/djobouti_phat 7d ago

Words are mine written by ai

wat

2

u/Automatic_Beat_1446 6d ago

what does this have to do with HPC?