Key takeaways

  • When it comes to Chinese AI labs using distillation techniques to extract knowledge from frontier model makers, Y Combinator CEO Garry Tan…
  • Distillation is when a model maker extensively prompts another model in order to learn how it works and reasons.
  • He feels it’s an overreach for AI labs to dictate what their customers can do with the information their models share with them.

What happened

When it comes to Chinese AI labs using distillation techniques to extract knowledge from frontier model makers, Y Combinator CEO Garry Tan is hoping regulators stay out of it. S. AI labs should perhaps play the same game. “I would do nothing,” he told CNBC in an interview earlier this week. S. a more robust set of open-weight options that aren’t Chinese.

Tan, who is himself such an avid AI user that once described himself as having cyber psychosis, wants to see a balance between open-weight AI labs and frontier labs. “They are at the frontier and driving it forward. We want that to be fundable, and be a great business model ongoing,” he told CNBC.

Why it matters

Distillation is when a model maker extensively prompts another model in order to learn how it works and reasons. It is commonly, and legitimately, used by AI labs to help train new models. Anthropic this week released its second report alleging that Chinese labs are engaged in “illicit distillation attacks,” hiding their identities to distill without permission and relying on fraud and stolen credentials to do so. S.

regulators to crack down on distillation. It’s notable that the commander of Silicon Valley’s prestigious and prolific startup accelerator doesn’t agree. To be clear, Tan isn’t advocating for American AI labs to use stolen credentials to distill. He wants them to be free to come in the front door. In fact, his argument is twofold.

He feels it’s an overreach for AI labs to dictate what their customers can do with the information their models share with them. He also notes that the proprietary AI labs didn’t ask permission when they vacuumed up as much human knowledge as they could to train their models. They famously ingested plenty of copyrighted material without the permission of those intellectual property holders.

“Controlling what users and customers do with API calls to closed weight models feels constraining, and there’s a role government can play here to normalize the fact that access to intelligence that was trained on broad public access data should itself also be more a form of a public good than something locked away behind restrictive terms of service,” he told TechCrunch when asked why American labs should be free to distill, too.

What to watch

” To him, the true AI doomer scenario is for all the immense power of frontier AI to wind up in the hands of a single powerful, proprietary provider. “The nightmare scenario, the doomer scenario for AI is that there’s just one company,” he said. “It has the best access to capital. It has the best AI researchers. It runs away with it and suddenly there’s one company that’s monolithic.

” When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.