Published Articles
Explore peer-reviewed open access articles published in Ai Safety Governance. All articles are permanently available to readers worldwide under CC BY 4.0.
Research Article
2026-03-10
Dr. Evelyn Vance, Dr. Charles Zhang
This paper introduces a probing framework to extract factual beliefs and latent reasoning paths from LLMs, addressing hallucination and alignment.
Research Article
2026-05-02
Prof. Julian Alistair, Dr. Hana Tanaka
We present a sandbox environment to evaluate coordination and negotiation safety in multi-agent networks, outlining systemic risks.