logoalt Hacker News

milkshakeslast Saturday at 9:15 PM2 repliesview on HN

https://www.anthropic.com/news/detecting-and-preventing-dist...

Moonshot AI Scale: Over 3.4 million exchanges

The operation targeted:

Agentic reasoning and tool use Coding and data analysis Computer-use agent development Computer vision Moonshot (Kimi models) employed hundreds of fraudulent accounts spanning multiple access pathways. Varied account types made the campaign harder to detect as a coordinated operation. We attributed the campaign through request metadata, which matched the public profiles of senior Moonshot staff. In a later phase, Moonshot used a more targeted approach, attempting to extract and reconstruct Claude’s reasoning traces.


Replies

InsideOutSantalast Saturday at 10:19 PM

I'm assuming you posted that as evidence for the claim that "empirically, it appears that distillation of a more advanced model is a required first step", but I don't think it is. It's just evidence that Moonshot distills Anthropic's models, which, yes, they do.

show 1 reply
idiotsecantlast Saturday at 10:10 PM

>request metadata, which matched the public profiles of senior Moonshot staff

Translation: we have the machinery in place to identify our users, and actively do so.