GPU pilot brief and scorecard: a practical selection worksheet
Download an editable worksheet to agree scope, compare configurations and decide whether to stop, adjust or scale. An evaluation method, not a customer case study.
· Arvica Cloud · Analysis & buying guide
Download the blank worksheets
CSV files for Excel or Google Sheets, with no email required. Result fields are blank until you run and document a test.
Write the decision first: can this model meet the response target, can this scene run correctly, or does an eight-GPU job justify the node cost? Choose one primary acceptance measure and a small set of quality, reliability and budget constraints. Do not change the pass threshold after seeing the results.
Prepare the reproducible brief
Fill in model or scene revision, framework, container or environment, precision, input shape, dataset permission, region and intended workload. Name the customer reviewer and a technical contact. Record what is unknown so that assumptions can be resolved before activation, not hidden in a configuration label.
Agree the commercial envelope
Record provider, infrastructure location, tenancy, GPU and storage configuration, duration, cost cap and what happens when compute stops. A pilot must have an accepted quote; no capacity is reserved merely by downloading this worksheet or submitting an enquiry. Obtain permission before using any real customer dataset.
Record every comparison consistently
Use one worksheet row per run. Keep configuration, input revision, warm-up and cache conditions, sample count, output quality, peak memory, time or latency, errors and total billable cost. Save a reference to the raw evidence. Leave measured fields blank until a test is actually run; the downloadable template contains no benchmark results.
Make a stop, adjust or scale decision
Stop when an essential compatibility or quality condition fails. Adjust when an identified bottleneck can be tested within the agreed scope and budget. Consider scaling when the accepted workload meets its criteria and the larger configuration has a justified test plan. A passed single-node pilot does not establish multi-node scaling.
Turn evidence into a publishable case only with permission
A real case should identify test conditions, what was measured, limitations and the permitted customer attribution. Do not replace blank fields with invented results. Until a test is completed and publication is approved, share this method and blank worksheet as planning resources.
Original evaluation guidance; not a customer case study or measured benchmark.
Compare scope, allocation, storage, networking and pilot acceptance before committing to a B300 node. Includes a worked cost formula, not a supplier price.
Build a representative inference test with model, precision, context, concurrency and quality targets. Connect the measured result to a transparent GPU budget.