跳到正文
原文
Rohan Paul· @rohanpaul_ai · X·· 1 小时前AI 评分71
AI 导读

OpenAI 官方首次指控一个与中国有关的 AI 实验室试图提取其模型的隐藏推理内容进行对抗性蒸馏。操作者未破解加密,而是通过操纵模型交互复制加密推理并让其他模型解密转写。

正文

OpenAI officially accuses a Chinese AI lab of distillation for the first time, says the lab tried to extract its models' hidden reasoning.

The operators (China linked developers) couldn't crack any encryption; instead they copied encrypted reasoning from one conversation and asked a model elsewhere to decrypt and transcribe it.

"The activity began on July 1, initially at a low volume until we observed high-volume spikes on July 24 and 25 consisting of 16,000 requests1 using a relevant extraction pattern from over 4,000 users. Further investigation identified related prompt-pattern activity across a cluster of more than 15,000 users, which we fully disrupted by July 28."

In response, OpenAI closed the replay pathway, added checks that hold streamed output which might expose reasoning, banned accounts and worked with third-party services the operators used.

来源:Rohan Paul · x.com