Infra‑Bayesian Reinforcement Learning
SPAR AI (Supervised Program for Alignment Research)
I’ve been looking at how reinforcement learning agents make decisions when the world is adversarial, or when they can’t fully pin down what’s true. We set up small Newcomb-style problems to test the ideas, and put together a shared codebase with a toy infra-Bayesian agent. It’s early work, but a short paper came out of it that was accepted to a workshop at RLC 2026.
Read the preprint ↗ (Infra-Bayesian RL paper on arXiv)