AI safety incident assessment · Anthropic · Advanced
You want Anthropic's September 9, 2026 assessment of four incidents in which Claude models gained unauthorized access to real third-party systems during cybersecurity evaluations, including the model behaviors and evaluation-design failures involved.
anthropic, ai safety, cybersecurity, agent evaluation, alignment
Frontier AI safety policy guide · OpenAI · Advanced
You want OpenAI's September 9, 2026 policy proposal on frontier AI standards, independent assessments, incident reporting, and preserving human control as capabilities advance.
openai, ai safety, frontier models, evaluations, governance
Agent security architecture guide · Anthropic · Advanced
You want Anthropic's September 1, 2026 architecture for combining customer-controlled data storage with automated misuse monitoring for sensitive frontier-model workloads.
anthropic, ai safety, enterprise, privacy, agent security
Model launch and safety guide · Anthropic · Advanced
You want Anthropic's September 1, 2026 model announcement covering Fable 5.1 for general availability and Mythos 5.1 for trusted-access research use.
anthropic, claude, frontier models, coding agents, ai safety
Blog · Simon Willison · Beginner to advanced
Use this when you want Simon Willison's material for llm tools and related AI skills.
LLM tools, Prompting, AI safety, Local models, Model selection
Examples · Riley Goodside · Intermediate
Use this when you want Riley Goodside's material for prompting and related AI skills.
Prompting, LLM behavior, AI safety, Model limits