Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

This is probably less likely with this model, as it’s almost certainly a further RL training continuation of 3.5 27b. The bugs with this architecture were worked out when that dropped.


Valuable note!




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: