30-SECOND SUMMARY
What to take away
- CPU-only tests can begin with a quantized 1B–3B model.
- Check instruction-set support, storage, cooling, and RAM together.
- Judge reuse by task completion, not headline speed.
Test the old laptop first
Measure before replacing it.
- 01INSPECT
CPU, RAM, storage
- 02SMALL
Choose 1B–3B Q4
- 03RUN
Build a CPU baseline
- 04DECIDE
Check task success
Inventory the machine
Record the OS, exact CPU, RAM, and free storage, then compare them with the runtime requirements. An old CPU can have enough RAM yet lack a required instruction set.
Connect power and clear the vents before sustained inference.
For verification, save the model and runtime versions, source input, relevant settings, and observed output together. Repeat the step while changing only one factor, and record unexpected results and untested limits as carefully as successes before applying the guidance to private or production data.
Create a small baseline
Close memory-heavy apps, use a 1B–3B Q4 model and short context, then repeat the same prompt.
The working rule for “Create a small baseline” is: Check instruction-set support, storage, cooling, and RAM together. Complete one small test with public data first, then record each changed condition and its effect before increasing scope.
Record the current version and settings before the example, then verify the expected response, file, or process afterward. Preserve the error and return to the smallest working command before adding options; this separates installation failures from input and integration failures.
ollama run gemma3:1b
ollama psChoose realistic work
Test short rewriting, classification, and summaries that a person can verify. Long documents and concurrent users may require newer hardware.
The working rule for “Choose realistic work” is: Judge reuse by task completion, not headline speed. Complete one small test with public data first, then record each changed condition and its effect before increasing scope.
Define completion with an observable result instead of a general impression. Repeat the same input, and if the output changes, isolate whether the model, runtime settings, or source data changed before moving to the next stage.
Decide whether to upgrade
Extra RAM or an SSD can help when the CPU remains supported. Replacement is more rational when compatibility, battery, or cooling costs dominate.
The working rule for “Decide whether to upgrade” is: CPU-only tests can begin with a quantized 1B–3B model. Complete one small test with public data first, then record each changed condition and its effect before increasing scope.
For verification, save the model and runtime versions, source input, relevant settings, and observed output together. Repeat the step while changing only one factor, and record unexpected results and untested limits as carefully as successes before applying the guidance to private or production data.
Frequently asked questions
Is a GPU required?
No. Small models can run on CPU, although slowly. For a practical check, follow the “Inventory the machine” section, change one condition at a time, and record the result.
Can 8GB work?
It may support small, short tests if other applications are closed. Check instruction-set support, storage, cooling, and RAM together. For a practical check, follow the “Create a small baseline” section, change one condition at a time, and record the result.
Should I use private files first?
No. Validate stability and the data path with public material first. For a practical check, follow the “Choose realistic work” section, change one condition at a time, and record the result.
Primary sources
Check the original documentation for version-specific details.
llama.cpp repository Ollama hardware support LM Studio requirements