Skip to main content
Bonsai 1.7B is the smallest model in the family. At 0.25 GB (1-bit), it targets smart glasses, wearables, and always-on background tasks where other models are too heavy.

Specifications

Artifacts

The FP16 reference weights (3.45 GB) are in the ternary GGUF repo.

Run it

Through the demo repo:
Or directly with llama.cpp / MLX: