QLoRA: Efficient Finetuning of Quantized LLMs 3 years ago
The /g/ board on 4chan has a /lmg/ general that focuses on running models locally. They regularly discuss fine tuning models, quantization tech, and building apps on text-generation-webui/kobold.
You might get some interest but it's also 4chan...