I haven't compared it for perf, but at first glance this appears to be an LLM (text-only) while both binary and ternary bonsai 27B models are VLMs (where the vision tower and the adapter MLP weights are unquantized and kept in float16).