Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Reinforcement learning makes it something different. It becomes much more of a search engine through next-token-space that targets the training objective. Better to think about it like that, and then you'll see why "these things have motive" is not a terrible analogy, and you'll better be able to anticipate what they do.
 help



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: