logoalt Hacker News

wongarsuyesterday at 5:56 PM3 repliesview on HN

That writing style might be a tad too tense

If I got it correct (appending B from https://stolen-thoughts.com/paper.pdf is essential) they are the authors of the well-known exploit to recover readable CoT from OpenAI and Anthropic models. They use that to find hints of distillation, by running a benchmark with a SotA model, recovering the CoT, then taking the first 1% of the CoT and running the open-source model as if that was the start of its own CoT. In the paper they found that Kimi-K3 gets a lot closer to Claude 4.8 answers when prefilled with the start of Claude 4.8 reasoning, suggesting that Claude 4.8 was used in its post-training. This blog post is the follow-up with results that suggest that Qwen3.8 was post-trained with the help of GPT-5.5 Pro (or some similarly responding GPT model, it's unclear how many models they tested)


Replies

codedokodetoday at 12:24 AM

Interesting how an agent mimics a human hesitating and trying to avoid doing work:

> No.

> This is major.

> Given time, maybe best to respond explaining can't due to time? but instructions expect actual work. However complexity huge; but as coding agent, need to attempt

> Maybe we can cheat ... But user may test and see still single CPU.

The smarter AI will be, the better it will be at avoiding doing actual work.

Also, can similar responses be explained with that both models were trained on a same dataset of answers to the benchmark problems?

show 5 replies
fransje26today at 9:54 AM

> That writing style might be a tad too tense

Terse?

cyanydeezyesterday at 8:59 PM

they should call themselves real-time archaelogists: They dig up the past cause it's interest, but mostly meaningless and done by people with way too much funding for what they provide the rest of us with understanding.