Current limits on balances are documented here.
Let’s take a shared-2x machine as our example. For each vCPU, the machine get’s 5ms of baseline quota per 80ms period. It can accumulate 2x 500s of balance.
Each vCPU running at 100% will consume 75ms (period - baseline) of balance. So a single core could run at 100% for 13,333 periods (1000s/75ms) which is 18 minutes. Running both vCPUs at 100% would consume 150ms of balance per period and could be sustained for 6,667 periods (1000s/150ms) which is 9 minutes.
There are several charts, some of which show per-CPU load. The ones showing combined CPU load show the average between vCPUs.
That’s TBD. Something that’s getting lost in this conversation is that there’s nothing wrong with getting throttled. Many apps will chose to run as hard as they’re allowed to and that’s a totally fine way to use the platform. We’ll be looking at ways to notify users whose machines are working at their limit without unnecessarily bothering users who are doing so intentionally.