Will this LLM fit my GPU?

Estimate: weights at the quant's average bits + FP16 KV cache + ~1 GB runtime. Model shapes from each model's published config.json.by CyberMax