Benedict Patrick

I build my own models.

Core-1 is a language model designed to run on ordinary CPUs instead of rented GPUs. Two ideas, stacked.

32bits per weight

A network is millions of decimals.

Every weight is a 32 bit number, and every step of inference multiplies them. That is why models need big GPUs.

16×smaller than fp32

Core-1 allows only three values.

Every weight is forced to −1, 0 or +1 during training. Packed at two bits each, the model gets sixteen times smaller.

0multiplications per weight

So multiplying becomes adding.

Zero means skip. Plus one means add. Minus one means subtract. Plain integer math that cheap processors are good at.

O(1)memory per new token

And memory stays flat.

Attention keeps a cache that grows with every token. Core-1 uses a state space backbone with a fixed size state instead.

Status: proof of concept. A small ternary state space language model that learns and packs sixteen times smaller, validated at toy scale.

Cipheron

Released

A small coding model trained for secure code review. Point it at source code and it flags vulnerabilities and suggests safer implementations, fully offline.

Strongest on SQL and command injection. The model card lists where it still falls short.

parameters
0.5B
safetensors
BF16
license
Apache 2.0
Hugging Face

Say hello

Selective about projects. Ambitious about the ones I take: models, engines, and the tools around them.

Start a project →benedictpatrickjohn@gmail.com