This really just exists so cognition can stop spending API tokens with Anthropic or OpenAI.

Basically any successful AI based service will do this because at scale the frontier models are expensive and you’ll have enough data to fine tune your own.

Same reason Harvey is doing models now and basically every other provider

Couldn't they just grab and run an open weight model to save on API tokens?

You get better performance if you also finetune it for your task

[deleted]