About Me

I’m an AI researcher and technical leader interested in building intelligent systems that work reliably in the real world—and in figuring out how we know when they do.

My work spans AI research, cognitive science, and engineering, with a particular interest in reasoning, agent behavior, and evaluation. I like problems that start out underspecified: turning a research question or emerging capability into something concrete enough to build, measure, test, and improve. Over the course of my career, that has meant everything from studying how AI systems reason about people’s beliefs and goals to designing evaluation frameworks, building AI prototypes and research infrastructure, and leading interdisciplinary technical teams.

I’m especially drawn to work at the boundary between research and application: understanding what new AI capabilities make possible, identifying where they break down, and turning that understanding into better systems, tools, and products. Lately, I’ve been particularly interested in robust evaluation, AI reliability and safety, and the ways representation and context shape model behavior.

Research

Understanding intelligent behavior

  • Reasoning
  • Evaluation
  • Representation

Building

Turning ideas into working systems

  • AI evaluations
  • Agents
  • Research tooling
  • Knowledge systems

Leading

Turning ambiguous problems into research programs

  • Technical strategy
  • Interdisciplinary teams
  • Government R&D
  • Mentorship

Looking for my old site? It’s archived here.

Latest from the blog

Hello, World

I’ve been meaning to start writing more about AI outside of papers and work projects, so… here goes.

This blog is a place for things I’m thinking about, working on, or find interesting—AI research, things I learn while building with AI, questions about how we evaluate these systems, and probably the occasional opinion about where the field is going.

Some posts will be technical. Some won’t be. Some may grow out of projects I’m working on; others may just be things I found interesting enough to write down.