Big vs. small GPU clouds for fine-tuning LLMs
https://news.ycombinator.com/item?id=37101579Hi everyone,
I am looking to fine-tune a Llama 2 (the 7B and 70B to see if there is a big difference), and I am looking at the different Cloud options for GPUs.
There are of course the big cloud providers like AWS, and the smaller ones like Paperspace and co.
I am trying to benchmark each in terms of price, ease of use, quick availability of GPUs, and feature-richness.
Could you share the insights on big vs small cloud providers when training a LLM? If you have other criteria to make a decision I would be interested too!