logoalt Hacker News

cristoperbyesterday at 8:22 PM2 repliesview on HN

They reference this paper which describes a method to decrypt reasoning traces (by sending the encrypted trace back to the model and asking it to transcribe it):

https://stolen-thoughts.com/paper.pdf


Replies

ameliusyesterday at 8:37 PM

Interesting, but I suppose that's a hole that can be easily patched.

show 1 reply