Pretty impressive so far, but needs more testing.

It is more useful than Qwen3.8:27b (which is already quite good) and runs faster on my 7900 XTX / 64 GB DDR4 system.

Local LLM is getting more exciting every day!

Interested to know throughput on 7900 XTX and what setup you're using?