Fault-tolerant distributed job scheduler simulator
A visual simulation of job scheduling, worker failure, retries and recovery across a small distributed computing cluster.

Project brief
Project scope
Project components, tools, requirements, testing, and limitations will be added when this project is scoped.
Current scopeAdvanced
The project tools will be confirmed during scoping.
Included
- 01Go simulator for dependent jobs, heterogeneous workers, resources, labels, and deterministic events
- 02First-come, shortest-job, priority, and resource-aware scheduling policies
- 03Worker loss, worker recovery, task failure, retry, backoff, reassignment, and checkpoint recovery
- 0462 reproducible policy, recovery, fairness, and scaling experiment runs
- 05Local trace viewer with metrics, worker timelines, task outcomes, resource use, and ordered events
- 06Automated tests with clean race, vulnerability, and static security checks
- 07Complete source code in a private GitHub repository
- 0881-page project documentation in PDF and editable Word formats
- 0910-page setup and usage guide in PDF and editable Word formats
Project record
No information is collected on this page.
- Permanent project ID
- GP-CS-03NOUGU
- Catalogued
- 21 Aug 2026
- Completed
- 23 Aug 2026
- Verified
- 23 Aug 2026
- Demonstration
- Included in repository
Handover
After purchase
- 01Payment is confirmed
The project is marked unavailable and cannot be purchased again.
- 02Repository access is granted
The buyer's submitted GitHub account receives access to the private repository.
- 03The purchase record is delivered
The certification sheet is prepared from the reviewed buyer details and sent privately by email.