Current project: “Mechanism”

Game theoretic online intervention defense for language model agents that studies how to coordinated critics and observers stop harmful agent trajectories before irreversible damage.

The public development record preserves intentionally incomplete snapshots from three stages of the project, including open questions, TBD experiments, and changes to the final empirical claim.

Read “Mechanism” development snapshots: v0.1 → v0.2 → v0.3

Bayesian Temporal Equillibrium Identifiabiltiy of Episodic Coordination in Point Processes
Post-Covid19 Exploits Post-Covid19 means of exploits
Dispersion plot for emotional triggers Tracking in-session anxiety
Bots Classifier Mapping bot likelihood with k-means
Agency Network Graph Literature Review Network Graph
modeling conflict Pairwise feature classification accuracy