Home/Blog/Safety & governance

Safety & governance

11 articles

July 24, 2026 Data sovereignty: the question that decides your AI architecture Before the question of whether a model is good enough comes the question of whether you are allowed to send it your data. Four things get called sovereignty, they have different answers,… July 14, 2026 AI alignment and safety, without the hype or dismissal AI safety is discussed either as impending doom or as overblown hype, and neither framing helps you understand it. The… July 3, 2026 Agent permissions: the question nobody asks until afterwards Eighty percent of organisations running agents say those agents have taken unintended actions. One in five has had a… June 23, 2026 What is prompt injection, and why is it unsolved? Prompt injection is the number one security risk for AI applications, and researchers treat it as unsolved: not a bug… June 15, 2026 Mechanistic interpretability: opening the AI black box We built AI systems that work without fully understanding how they work. We have every number inside them, yet the numbers… June 13, 2026 AI bias and fairness: why 'fair' has no single answer AI now helps decide who gets a loan, an interview, bail, or medical priority, and the fear is that it does so unfairly.… June 11, 2026 The secret language that never was, and the escape that did The famous stories about AI going rogue are mostly false. The verified incidents are less dramatic and more concerning,… June 3, 2026 What is AI sycophancy? Why AI tells you what you want Tell an AI its plan is brilliant and it agrees; push back on a correct answer and it caves. This is sycophancy, the… June 2, 2026 What is AI jailbreaking? Why safety can be talked around AI models are trained to refuse harmful requests, yet people keep finding ways to make them comply anyway. Jailbreaking is… May 23, 2026 How to secure an LLM application: risks and defenses Traditional security rests on a boundary between instructions and data. A language model has no such boundary, because… May 18, 2026 AI and copyright: the three questions people confuse Whether training on protected work is lawful, whether AI output can be owned or infringes, and whether any of it is fair…

Other subjects