LLM eval jobs spent 82% of their time waiting for a container. Mercor’s job-level compute scheduler got 6x to 20x more throughput from the same quota.