Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

It's interesting that it makes that mistake, but then catches it a few lines later.

A common complaint about LLMs is that once they make a mistake, they will keep making it and write the rest of their completion under the assumption that everything before was correct. Even if they've been RLHF to take human feedback into account and the human points out the mistake, their answer is "Certainly! Here's the corrected version" and then they write something that makes the same mistake.

So it's interesting that this model does something that appears to be self-correction.



Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: