AI Futures Simulator
A Monte Carlo simulator of AI-related futures (2026–2050) — capability growth, plot events, and society feedback compounding into emergent outcome regions.
Read more → ActiveClaims-Bench
Characterizing language-model agents' implicit moral and stakeholder commitments — what value profile does a model reveal when stakes are unclear and reasonable people disagree?
Read more → ActiveCompute-Pause Verification
Can cross-node communication size and shape separate benign LLM pretraining from inference? A hardware-based mechanism for verifying AI compute-governance agreements under low trust.
Read more → ProposalShortcut Forensics
Does a single interpretable direction in a coding agent's residual stream drive shortcut-taking under pressure, or is it several independent mechanisms wearing the same disguise?
Read more → ProposalTheory of Mind / Empathy Dissociation
Does an open-weight LLM represent cognitive theory-of-mind and affective empathy as separate, dissociable directions — the way TPJ/mPFC and insula/ACC are separate in the human brain?
Read more → EssayGoverning the Machine Economy
AI agents are becoming autonomous economic actors. The infrastructure for that economy is being built now — the governance layer that ensures it benefits human society is not.
Read more →