Public preview · figures are illustrative fixture data, attributed to their source · data to May 2025 · read-only
System · desktop

CPU-only Ryzen 9 7950X (128 GB DDR5)

Runs models up to ~170B parameters at 4-bit in system RAM (CPU only). No GPU: system RAM only.

GPU memory
System RAM
128 GB
Runs
18 of 19model variants, 8K
Estimated cost
~$1,600

What it runs

Largest models that fit without spilling into system memory, at an 8K contextHow much text the model can consider at once, counted in tokens — roughly ¾ of a word each..

Components

Measured performance

Model · quantizationRuntimeContextMeasurementsSource
Llama 3.1 8B Instruct Q4_K_M (LM Studio Community)llama.cppcpu · b46004K
92 tok/s · Prompt processing (512)
12.4 tok/s · Generation (128)
Mutinai illustrative fixtures

Community results

Contributions are not open yet, so there is nothing here from members.

Reviews

Contributions are not open yet, so there is nothing here from members.