Lamb Labs · one-bit research

shaun

A full-model binary Qwen3-1.7B, trained to survive at roughly 1.13 bits per weight—and small enough to fit in a 237 MiB GGUF.

waking Shaun…

Shaun runs on CPU from the same packed Q1_0 artifact you can download and run locally. The public demo is deliberately capped at 256 output tokens and one generation at a time.

experimental model · short answers work best · prompts are not logged
Try one of these
Enter to send