Kick off the deploy model pipeline.
Which GPU tier a deployment runs on: H100 (normal) or B200 (turbo).
Traffic level a deployment is sized for; picks its replica count.
Successful Response
True when an existing merged model was reused and the merge stage was skipped.