So I have a RTX 3080 10GB VRAM which I've been using with Qwen2.5 Coder and Gemm...

		Johnny_Bonk 14 days ago \| parent \| context \| favorite \| on: Running local LLMs offline on a ten-hour flight So I have a RTX 3080 10GB VRAM which I've been using with Qwen2.5 Coder and Gemma 4 E2B. Im wondering what models you have tried with which quants.