My favourite kind of research is the kind that breaks a default everyone’s relying on.
This one does that for synthetic training data. The team’s latest research drop is live! 👇
The biggest hurdle in frontier AI is curating high-quality, diverse data.
What happens if you need to post-train a model for a new capability, but have zero data?
Today, we release the technical report for Invent a Dataset.