When I was in China it was insinuated that every lab did this. I’m glad there’s public research on it and am still shocked the frontier labs haven’t patched this stuff. We don’t need policy action on distillation, we just need the products to work as intended.
Great paper.
Nathan Lambert
@natolambert
Alexander Panfilov @kotekjedi_ml
We can finally talk about it:
We found a way to extract hidden reasoning of frontier models using a vulnerability in the APIs of every frontier AI company.
We verified that our reasoning token count matches billed API thinking tokens 1:1 for most of the prompts we queried.
We found a way to extract hidden reasoning of frontier models using a vulnerability in the APIs of every frontier AI company.
We verified that our reasoning token count matches billed API thinking tokens 1:1 for most of the prompts we queried.
