Vibe code giveth, vibe code taketh.
- https://www.theverge.com/cs/features/831818/ai-mercor-handshake-scale-surge-staffing-companies
- https://outlier.ai/math/en-us
- https://www.opentrain.ai/
- https://www.pin.com/blog/ai-labs-hiring-train-models/
Much of this is data annotation, reasoning trace evaluation, and problem set curation. But there is no way they haven't atleast paid some mathematicians to work on research grade problems in tandem with their models, and then used that for training data. - AI companies do not share their custom internal harnesses.
- AI companies do not share their custom internal training data.
- AI companies do not share how much compute they allocate to trying to solve problems of this nature.
- AI companies are primarily marketing their models to investors as human-replacing rather than human-augmenting.
- AI companies are under enormous financial pressure to make their business work.
The last two points incentivize them to find these types of "first proof" successes as aggressively as they can, and I'm sure they've thrown the whole book at it.