Can you run mradermacher/Llama-GitVac-Turbo-3B-GGUF on RTX 3060 12GB?
mradermacher/Llama-GitVac-Turbo-3B-GGUF needs ~3.98 GB, which is within the 12 GB on RTX 3060 12GB by memory capacity (~8.02 GB headroom). Computed from stated size and quant, not a runtime guarantee.
mradermacher/Llama-GitVac-Turbo-3B-GGUF needs ~3.98 GB, which is within the 12 GB on RTX 3060 12GB by memory capacity (~8.02 GB headroom). Computed from stated size and quant, not a runtime guarantee.
Memory-capacity estimate across common GPUs
| GPU | Verdict | VRAM | Needs |
|---|---|---|---|
| 8GB laptop (no dedicated GPU) | Fits (by memory) | 8 GB | ~3.98 GB |
| 16GB laptop (no dedicated GPU) | Fits (by memory) | 16 GB | ~3.98 GB |
| NVIDIA T4 16GB (free Colab) | Fits (by memory) | 16 GB | ~3.98 GB |
| NVIDIA L4 24GB | Fits (by memory) | 24 GB | ~3.98 GB |
| RTX 3060 12GB | Fits (by memory) | 12 GB | ~3.98 GB |
| RTX 4080 16GB | Fits (by memory) | 16 GB | ~3.98 GB |
| RTX 3090 24GB | Fits (by memory) | 24 GB | ~3.98 GB |
| RTX 4090 24GB | Fits (by memory) | 24 GB | ~3.98 GB |
| A100 40GB | Fits (by memory) | 40 GB | ~3.98 GB |
| A100 80GB | Fits (by memory) | 80 GB | ~3.98 GB |
| H100 80GB | Fits (by memory) | 80 GB | ~3.98 GB |
| Apple M-series 16GB (unified) | Fits (by memory) | 16 GB | ~3.98 GB |
| Apple M-series 32GB (unified) | Fits (by memory) | 32 GB | ~3.98 GB |
| Apple M-series 64GB (unified) | Fits (by memory) | 64 GB | ~3.98 GB |
| Apple M-series 128GB (unified) | Fits (by memory) | 128 GB | ~3.98 GB |
| CPU / 32GB system RAM | Fits (by memory) | 32 GB | ~3.98 GB |
| CPU / 64GB system RAM | Fits (by memory) | 64 GB | ~3.98 GB |
How this was computed
- Weights + KV cache at the stated context + a runtime/overhead allowance; memory capacity only, not a benchmark.
Next steps
- Open mradermacher/Llama-GitVac-Turbo-3B-GGUF on Hugging Bay — files, license, hosting, reviews.
- Raw fit report (JSON)
Methodology: weights + KV-cache at 8192 tokens + a runtime overhead allowance, from the recommended download size, parameter count, and quant. Computed from stated size and quant, not a runtime guarantee. Not a benchmark; runtimes differ.
