logoalt Hacker News

famouswafflesyesterday at 6:43 PM1 replyview on HN

It can lead to hidden reasoning, if the looping allows it to stuff enough information outside visible CoT. Open AI demostrates such an ability by asking it to solve problems while thinking about something else entirely. All the other models are unable to do this except Astra. It doesn't have to be a substitute for CoT to cause monitorability issues.


Replies

cmatoday at 1:09 AM

If you ask it not to think about something that doesn't cause the pink elephant issue?

show 1 reply