Research

Published, with the code and the open questions.

What exists, what it measured, and what it did not settle. The last of those is the honest measure of where the work is.

Research areas

Four areas, one at a time.

Narrow on purpose. A question taken to a measurable answer is worth more than four taken to an opinion.

  • Model and data evaluationCurrent

    Whether a number can be trusted, and what it costs when it cannot. The published paper and its code sit here.

  • Physical AI

    Models that act on the world rather than describe it, where a wrong answer moves something.

  • Robotics

    Hardware in the loop, so the same control code runs against a simulator and against the real thing without being rewritten.

  • Aviation

    Where the tolerance for an unexplained failure is lowest, which makes it the hardest honest test of everything above.

A team of two, one question at a time.

We finish before we start something else. That is slower than it sounds, and it is the reason anything here carries a measurement rather than a claim.