Publications

Papers and preprints

Current entries are in preparation and should not be read as accepted publications.

CrashDiag: Mechanically Verified Reinforcement Learning for Infrastructure Repair

In preparation

Research writeup planned around bounded JSON actions, state-based verifiers, answer-free GRPO rewards, held-out mechanical evaluation, and fail-closed promotion gates.