Concentrating intelligence

Large models can't fit on smartphones. Datacenters can't sustain them. PrismML is building ultra dense intelligence to solve both.
14× less memory
8× faster
5× less energy
Supported by:
Benchmark palette

Intelligence density

Negative log of the model's error rate divided by the model size

Model benchmark comparison

Average score across all benchmarks

Throughput

Tokens per second across hardware platforms (higher is better)

Energy consumption

Milliwatt-hours per token across hardware (lower is better)

Explore other Bonsai models:

Bonsai Image

16.0 GB 16-bit (standard)

1-bit Bonsai 8B

Centering AI research on efficiency

Successful artificial intelligence isn’t just about making models larger, but also smarter. Utilizing breakthrough research at Caltech, PrismML is pushing the frontier of intelligence density by reshaping how models are designed, prioritizing intelligence per bit over sheer parameter count.