Performance-modeled launches
Don't guess your parallelism config.
- The performance model searches the whole space — sharding, replicas, batch size, placement — before a single GPU-hour burns.
- Set a cost target or a deadline; get the plan that hits it, with receipts.