This simulation illustrates how container orchestration platforms like Kubernetes manage and scale applications within a cloud environment. A stream of pods, each requesting a slice of CPU and memory, arrives at the cluster; a real bin-packing scheduler — first-fit, best-fit or worst-fit — places every pod onto the node that best satisfies its policy, stacking pods as 3D blocks whose height tracks CPU request and whose colour tracks memory share. Watch resource allocation happen live: nodes fill up, the cluster autoscaler adds capacity once average load crosses 75% and quietly reclaims idle nodes once they sit empty, and killing a node demonstrates service discovery and self-healing as its pods are evicted into the pending queue and immediately rescheduled elsewhere.