Hi r/LocalLLaMA, Affiliation Disclosure: I am the creator of this open-source project. Like many independent researchers and homelab builders here, I heavily rely on the modded RTX 2080 Ti 22GB cards due to their high VRAM-to-cost ratio. However, running modern models like Lance on older Turing arch