275
I Put a Datacenter GPU in My Gaming PC for £200
(blog.tymscar.com)
This is a most excellent place for technology news and articles.
Yes, that's basically what the article is about. They run the LLM across both GPUs.
But that's a feature of llama.cpp. SLI doesn't really exist any more, and NVlink requires a specific setup, which the 4080 is not part of (the 3090 was the last consumer one, apparently). So you couldn't pool the VRAM.