Skip to content

Tag

#local-llms

Every story tagged local-llms, newest first.

A GPU graphics card circuit board lying on a wooden workbench with tools and electronics visible in the blurred background
Guide · aiDeep read

Which local LLMs fit in 8, 12, 16, 24 or 32GB of VRAM

At Q4_K_M and 8K context, 8GB comfortably holds Gemma 2 9B (5.8 GB) and 24GB holds Gemma 3 27B (17.5 GB). Full tables for every common VRAM tier, computed the same way our GPU checker computes them.

Ahmad J · Sep 6, 2026 · 8 min read