You are building an image-recognition model in PyTorch using the ResNet50 architecture. Your code runs successfully on a small subsample on your local laptop. The complete dataset contains 200k labeled images. You need to scale the training workload quickly while keeping costs as low as possible, and you plan to use 4 V100 GPUs. What should you do?
Community Discussion
No comments yet. Be the first to start the discussion!
Community Discussion