Guide · Sander Schulhoff · Beginner to intermediate
You need a broad prompt engineering reference.
prompting, safety, education
AI safety incident assessment · Anthropic · Advanced
You want Anthropic's September 9, 2026 assessment of four incidents in which Claude models gained unauthorized access to real third-party systems during cybersecurity evaluations, including the model behaviors and evaluation-design failures involved.
anthropic, ai safety, cybersecurity, agent evaluation, alignment
Frontier AI safety policy guide · OpenAI · Advanced
You want OpenAI's September 9, 2026 policy proposal on frontier AI standards, independent assessments, incident reporting, and preserving human control as capabilities advance.
openai, ai safety, frontier models, evaluations, governance
Agent security architecture guide · Anthropic · Advanced
You want Anthropic's September 1, 2026 architecture for combining customer-controlled data storage with automated misuse monitoring for sensitive frontier-model workloads.
anthropic, ai safety, enterprise, privacy, agent security
Model launch and safety guide · Anthropic · Advanced
You want Anthropic's September 1, 2026 model announcement covering Fable 5.1 for general availability and Mythos 5.1 for trusted-access research use.
anthropic, claude, frontier models, coding agents, ai safety
Agent security architecture guide · Meta AI · Intermediate to advanced
You want Meta's September 8, 2026 technical account of defense-in-depth for a long-running personal agent, including isolated runtime cells, credential surrogates, a separate permission authority, tainted-egress tracking, scoped approvals, browser controls, red teaming, and prompt-injection evals.
meta, muse, agent security, prompt injection, least privilege
Agent interoperability standard · Anthropic · Intermediate to advanced
You want Anthropic's August 27, 2026 research preview of a model-agnostic standard for connecting agents to lab and manufacturing hardware through programmable interfaces and protocols such as MCP, with safety evaluation built into the rollout.
anthropic, model hardware standard, agents, mcp, robotics
Evaluation guide · Google AI for Developers · Intermediate to advanced
You want Google's practical July 31, 2026 guide to running the same agent and model evaluations during development and on production traffic, with metrics for quality, safety, grounding, tool use, and trajectories.
google, gemini, agents, evals, model selection
Guide · OpenAI · Intermediate
You are moving from experiments to production and need the official OpenAI guidance on latency, retries, rate limits, safety, monitoring, and operational rollout.
openai, production, reliability, latency, cost
Blog · Simon Willison · Beginner to advanced
Use this when you want Simon Willison's material for llm tools and related AI skills.
LLM tools, Prompting, AI safety, Local models, Model selection
Examples · Riley Goodside · Intermediate
Use this when you want Riley Goodside's material for prompting and related AI skills.
Prompting, LLM behavior, AI safety, Model limits