SYSF.IO · CHAT

A 1-bit model. Free conversation. In your browser.

Bonsai 1.7B is an open model by Prism ML: built from Qwen3-1.7B and trained end-to-end for 1-bit weights — 0.24 GB instead of 3.4 GB in FP16, running entirely in your browser on WebGPU. This is the free-chat variant: talk to the model directly and see what a 1-bit 1.7B can (and can’t) do. Nothing you type leaves your device. A sysf.io demonstration of sovereign, on-device inference.

1.7B PARAMS · 1-BIT Q1_0 (≈1.125 BPW) · 0.24 GB · 32K CONTEXT · APACHE-2.0

MODEL © PRISM ML (APACHE-2.0) · ENGINE: WEBML-COMMUNITY (MIT), PORTED TO QWEN3 BY SYSF.IO · TRIAGE VARIANT

0% REQUESTING WEBGPU DEVICE 0.00 / 0.24 GB
SYSF.IO · CHAT
READY
TRIAGE VARIANT ↗

What is on your mind?

You’re chatting with a 1-bit, 1.7-billion-parameter model running on your GPU. Expect quick, concise answers — and the occasional small-model stumble. Nothing leaves your device.