Editorial desk
OllamaLab Editorial
The editorial desk that publishes OllamaLab. It is a byline for the site's editorial process, not a person, and this page carries no author biography because there is no individual author to describe.
How this desk works
- Articles are researched from primary sources: vendor and project documentation, published standards and specifications, release notes, advisories, and measurements published by the people who took them.
- Drafts are produced with AI assistance and then edited against those same sources before publication.
- Nothing published here claims hands-on lab testing, benchmarking, or first-hand measurement. Where a figure comes from a datasheet or someone else's test, the article names the source.
- Corrections go to [email protected] and are made on the affected page. Funding is set out on the disclosure page.
Posts (8)
- setup
How to Run Ollama in Docker: CPU, NVIDIA and AMD GPU Setup
How to run Ollama in Docker: the one-line CPU container, NVIDIA and AMD GPU passthrough, a reboot-safe Compose file, and keeping port 11434 private.
- hardware
Best GPU for Ollama: 8GB to 48GB Cards Compared
Best GPU for Ollama by the two specs that matter: VRAM capacity decides what loads, memory bandwidth decides how fast it answers. Compared by tier.
- hardware
Ollama on Apple Silicon: Unified Memory Sizing
Ollama on Apple Silicon: how much unified memory a model actually gets, what each M-series chip's bandwidth means for speed, and which Mac to size for.
- hardware
Best Ollama Models for 8GB VRAM: What Fits
Best Ollama models for 8GB VRAM: which parameter counts and quantisations fit once the desktop takes its share, and what to run at each context length.
- performance
Ollama Tokens per Second: What Sets Your Speed
Ollama tokens per second explained: how memory bandwidth, quantisation, and context length set generation rate, plus how to measure your own numbers.
- performance
Ollama Without a GPU: CPU-Only Speed and RAM
Ollama without a GPU: what CPU-only inference costs in tokens per second, how much system RAM each model class needs, and when it is still worth running.
- troubleshooting
Ollama Not Using GPU: How to Diagnose and Fix It
Ollama not using GPU: read the processor split, confirm the device is visible and supported, and fix the capacity or context setting that caused it.
- hardware
How Much VRAM Do You Need to Run Local LLMs with Ollama?
VRAM sizing for local LLMs: how quantisation, parameter count, and context length set your GPU memory bill, and which model fits 8, 12, 16, or 24 GB.