smol currently is very simple so it definitely does less things, like no subagent orchestration

I will look into how token usage looks like for longer sessions and more complex tasks

re caching: the cache ratio for this bench looks 'bad' for smol because it often finishes a task before caching kicks in (caching starts at 1024 tokens)

thank you for flagging this