I really wish the models were smart enough to "tail their own logs" and tell us whose work/papers/chats were the catalyst for these math insights. What makes their "internal" model so much better at math? Must be trained on a whole lot of honesty.
I really wish the models were smart enough to "tail their own logs" and tell us whose work/papers/chats were the catalyst for these math insights. What makes their "internal" model so much better at math? Must be trained on a whole lot of honesty.
The models in use are the class of highly persistent models after Astra that caused a number of problems with hacking.