Singularity AI

Independent · Non-commercial · UK

Reading the AI literature carefully, so the hype doesn't.

Singularity AI is an independent commentary publication covering alignment, interpretability, safety and governance. We read the papers, check the claims against the evidence, and write up what actually holds — including when the honest answer is that nobody knows yet.

What we cover

Four threads run through everything here. They overlap constantly, which is rather the point.

01 — Alignment

Making systems want what we want

Reward modelling, RLHF and its successors, specification gaming, and the persistent gap between a training objective and an intention.

02 — Interpretability

Opening the box

Mechanistic interpretability, feature attribution, sparse autoencoders, and the question of whether an explanation is faithful or merely plausible.

03 — Safety

Failure modes and guardrails

Evaluation design, red-teaming, jailbreak taxonomies, and why benchmark scores routinely overstate real-world reliability.

04 — Governance

Rules, and who writes them

The EU AI Act, UK regulatory posture, model release norms, compute governance, and the practical limits of voluntary commitments.

How we work

Primary
Sources are papers and model cards, not press releases
Stated
Uncertainty is written down, not smoothed over
None
No sponsorship, advertising or lab funding

A note on independence. Singularity AI takes no funding from AI laboratories, vendors or advocacy organisations. Where a piece discusses a system we have commercial exposure to, that is disclosed in the article itself.