LLM Inference Lab

Running and optimizing open-source LLMs on a single Apple Silicon laptop — and measuring what actually matters: speed, memory, and real task quality (coding & math). Everything here is reproducible; no cloud GPU required. New to the terms? Every i explains itself, and there's a full Glossary.

Headline findings

Explore