Independent research · One person · Established 2024

Dingal AI Research Independent research on how AI systems behave
Menu

All notes

  1. Report

    Failure compounds faster than per-step accuracy suggests

    Measuring how language-model agents fail over long tasks. Per-step accuracy barely separates the systems; end-to-end success separates them a lot.

  2. Release

    The evaluation harness is now public

    The code I use to run and record my own evaluations, released so the results here can be reproduced.

  3. Note

    A year of interpretability work, and what held up

    Which methods for looking inside a model produced findings I would still defend, and which did not survive being checked.

  4. Note

    Why this exists

    What this practice is for, what it will publish, and the things it will not do.

Everything here is also on the RSS feed.