ResearchPaper7 min
Proving Instead of Testing
What formal verification with Dafny means for code written by AI agents — and why induction sits at the heart of it.
Alignment as an operational question rather than a philosophical one: what a model does when the training signal and the intent come apart, and how anyone would notice.
ResearchPaper7 min
What formal verification with Dafny means for code written by AI agents — and why induction sits at the heart of it.
ResearchPaper4 min
A new protocol proposes a way to make AI-assisted claims independently challengeable by ensuring transparent and falsifiable publication records.
ResearchDeep dive9 min
Nineteen teams spent a month building AI lie detectors. The winning trick turned out to be less clever than it looked — and that's the interesting part.
ResearchPaper7 min
A new benchmark shows language models know what's wrong but not how people actually react to it — and the bias may come from alignment itself.
ResearchNews5 min
OpenAI's achievement in solving a Millennium Prize Problem raises questions about the role of AI in mathematics and the ethics of collaboration.
ResearchNews5 min
The emergence of reasoning language models presents both unprecedented opportunities and unique challenges. Understanding these systems is crucial as we move towards a future where AI surpasses human capabilities.