Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

That raises the question of what is it actually doing?

If it isn't spending tokens on quality, is it the assumptions about the task difficulty that cause it to perform better? Or are their broader differences in the model being run.



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: