i’d like to revise my earlier comment: 2022 may have been the last time we had pure humans win a Fields Medal.
I’m fairly certain this batch's winners used LLMs for research, lit-revews, reviewing work, and calculations... perhaps not enough to count as a co-author, but still enough to handle a lot of the grunt work.
Who would have imagined the pace of progress in LLM-powered math..
“The Anthropic team for their incredible coding models (Fable-5 wrote every line of code in this project), and the Claude Code harness.”
Source: the repo
Shameless plug: We're building this. Our goal is to provide AI pentesting agents that run continuously, because the reality is that companies (eg: those doing SOC 2) typically get a point-in-time pentest once a year while furiously shipping code via Cursor/Claude Code and changing infrastructure daily.
I like how Terence Tao framed this [0]: blue teams (builders aka 'vibe-coders') and red teams (attackers) are dual to each other. AI is often better suited for the red team role, critiquing, probing, and surfacing weaknesses, rather than just generating code (In this case, I feel hallucinations are more of a feature than a bug).
We have an early version and are looking for companies to try it out. If you'd like to chat, I'm at [email protected].
Tangent: Did Windsurf actually get acquired by OpenAI? I would have imagined some sort of announcement from OpenAI at the very least? Bloomberg was the one to break that news too, but haven't seen any follow up.
Does anyone have insight into Neon's financials - specifically their revenue, COGS, and gross margins? I'm trying to understand what made Databricks value them at $1B. Was it strong unit economics, rapid growth, or mostly strategic/tech value?
Wow that had been quite a journey - thanks for the detailed response. We’ve been using cel internally in a golang codebase and been pretty happy with it. I’ve only know about starlark in the Bazel context - I’ve learned a couple of things from your post. Thanks :)
Now: On sabbatical