—Ideas—Projects—Common Room—About

systems look neutral
until you trace who benefits
and who disappears

— M.

Hi, I’m Melanie

An emerging researcher across AI safety and AI governance, asking what better governance could look like for countries that didn’t set the rules of the game.

—Ideas—Projects—Common Room—About

Latest Notes

01
Aug 2026Political Science

AI Grand Strategy for Middle Powers

Sovereign AI, how?

→
02
Aug 2026AI Safety

Interpretability and Evals

Attribution graphs, linear probes, natural language autoencoders, capability and propensity evaluations, and alignment auditing as tools for making model behavior legible.

→
03
Jul 2026Game Theory

Cooperative Game

A note on how the stag hunt, the prisoner's dilemma, and Pareto efficiency frame cooperation among learning agents.

→
04
Jun 2026Benchmarking

The Science of Benchmarking

A technical note on benchmark design and evaluation quality

→
05
Jun 2026AI Safety

Inner Alignment

Deception, reward tampering, mesa-optimization, goal misgeneralization, and why learned objectives may diverge from training objectives.

→