Separate training, fine-tuning, and inference needs
Use the workload phase to guide the planning conversation. H100 80GB each can frame training discussions, H200 141GB each can frame high-memory AI, and L40S 48GB can frame inference, vision, or rendering. These are examples only, not confirmed availability.
- State whether the work is pre-training, supervised fine-tuning, evaluation, or serving.
- Capture model size, sequence length, precision, batch targets, and memory headroom.
- Identify the datasets, checkpoints, artifact storage, and data-movement constraints involved.