Independent systems + ML research · Apple Silicon
Running LLMs larger than memory
on a consumer Mac
Can an existing pretrained model be transformed post-training into an execution
representation whose instantaneous working set is dramatically smaller than the full checkpoint —
while preserving most of its capabilities? This platform exposes the complete paper trail:
hypotheses, notes, code, benchmarks, and raw results.
total model size ≠ resident size ≠ bytes read per token ≠ parameters required for this token