ymrtech
login
about
services
services
security
vpn
infra
nix-config
audit
git
status
writing
contact
login
DARK
Writing · llm
[ <- all writing ]
/
[ all categories & tags ]
1 post tagged
llm
, newest first.
Running a 27B model on a 3090: a budget, not a flag
[2026-10-16]
/ infrastructure
Fitting a 27B model on one 24GB card is two arithmetic unknowns - weight memory at your quantisation and KV cache at your context - and a curve, not a number.
[ read ]
llm
cuda
quantisation
benchmarking
ymrtech@ymrtech
DARK
|
NixOS
|
UTF-8