Home/Blog

Articles

Worth reading first

Swipe for more

Browse by subject

The three foundational guides

The long answers to the questions everything else builds on.

Recently published

Evaluation & evidence 72 seconds, or 30 minutes, and both are trials Territory 11 opens on the first subject this corpus has examined where the evidence is genuinely good. Registered trials,… August 2, 2026 Evaluation & evidence The 12% was non-inferior, and P was 0.41 MASAI is the best-evidenced AI deployment in medicine and the headline everyone quoted describes a result the trial did… August 2, 2026 Evaluation & evidence Refactoring fell from 25% to 3.8% Survey evidence says AI improves code quality. Repository telemetry says the opposite. They are measuring different… August 1, 2026 Evaluation & evidence Phase I improved. Phase II did not. AI-designed drugs clear safety trials at well above industry rates. At the stage that tests whether a drug works, the… August 1, 2026 Evaluation & evidence Adding the doctor to the model changed nothing A randomised trial found physicians did better with an LLM than with conventional resources. Its second comparison,… August 1, 2026 Evaluation & evidence The pilot failed and the staff deployed it anyway Enterprise AI is measured by what organisations sanctioned. A separate literature measures what their employees actually… August 1, 2026 Evaluation & evidence The pilot ran on data the production system will never see Enterprise AI pilots are built on a curated slice, pre-cleaned, with limited users and manual review. Then production data… July 31, 2026 Evaluation & evidence The framework that exists produced 1.6% The FDA has two AI tracks. One is final, has authorised over 1,350 devices, and is the regime under which almost none of… July 31, 2026 Evaluation & evidence One trial, a waitlist control, and a letter The best evidence for AI mental health support is a single randomised trial of a purpose-built clinical tool. Its own… July 31, 2026 Evaluation & evidence The lock-in is the prompts, not the API Switching costs used to require board approval to incur. AI switching costs accumulate through ordinary engineering… July 31, 2026 Evaluation & evidence 621,000 robots installed, virtually no humanoids Territory 7 opens on the question article 117 left unresolved. Industrial robotics is enormous and growing. The… July 30, 2026 Evaluation & evidence 220 million miles, inside a boundary Waymo drew The strongest safety evidence in physical autonomy, and the methodology that makes it honest is also what limits what it… July 30, 2026

Written something like this?

Artifipedia takes long-form pieces from people who work in or study the subject. A documented failure, evidence from your own domain, or a widely repeated claim that does not survive checking. Unpaid, properly edited, permanent byline.

Read the contributor guidelines

Every article

182 pieces, grouped by subject.

Foundations29

Evaluation & evidence73

Safety & governance30