I saw a 1B model yesterday that was fine tuned on Fable output. I found that hilarious, but it did actually make all the scores go up.

(Actually talking to it, it was about as coherent as you'd expect, i.e. 3/10)

The floor for "actually usable model" keeps dropping though. (Seems to be about 27B right now?)