It wouldn't be the first time ollama's llama.cpp fork reintroduced bugs and was missing important optimizations.