◆ TOPIC · AI SAFETY

The AI Safety thread.

Two converging fronts define AI safety: the alignment and reliability of frontier models, and the security of the infrastructure they depend on. Princeton's ICML 2026 audits scrutinize GPT 5.5 and Gemini 3.5, while research shows Claude's values shift with prompt language. Alongside sit actively exploited CVEs, ransomware campaigns, and supply-chain threats across npm, curl, and Git — plus the new risks LLM-generated code poses to established review controls.

56 briefings · across 6 personas

◆ START HERE · LONG-FORM

◆ TIMELINE

How AI Safety moved across the corpus.

First surfaced 2026-02-17, most recent 2026-07-22, across 44 days.

◆ RECENT · LATEST 56

Skim the most recent entries.