Sangama → Results

Results

Test write-ups, newest first. Raw data and full reports are in the repository.

DateTitle
30 Sep 2026Qwen3.5-397B on 20 GPUs, through a relay
A 397-billion-parameter model on 20 GPUs with 16 GB each; from 0.6 to 4.8 tokens per second.
29 Sep 2026A small model across CPUs, GPUs and the Internet
Identical output on every engine and device; 3.1 tokens per second between a home Mac and a US GPU.

Sangama is open source under the MIT License. Contact: hello@sangama.co · Discord