▦ gpuroom
Will that LLM fit on your GPU?
GitHub ↗
Model
Quantization
Context —
KV cache
Copy share link
Fit matrix
Compare GPUs
Fine-tune
GPU A
GPU B
Method