AI directory search

Search across educators, skills, and resources.

Use this when you know the topic you need: Claude Code, MCP, evals, RAG, agents, product, coding, prompting, foundations, or model internals.

12 matches for "security"

Video matches

Watch first when you want a fast feel for the topic before opening courses, docs, or profiles.

Promptfoo red teaming video thumbnail ►

Promptfoo red teaming

Promptfoo · evals, prompt testing, red teaming, security

Educators

Sander Schulhoff profile photo

Sander Schulhoff

Learn Prompting · Beginner to intermediate

Dense prompt engineering reference with beginner-friendly structure.

Skills

Prompting, AI literacy, Security

Agent tools and skill directories

AI worker platform

Mission Control AI

Mission Control AI · Preconfigured AI workers

Use this when you want role-specific AI workers with SOPs, integrations, and governance policies already built in.

Start

Map one operational process, identify data and approval boundaries, and evaluate whether a prebuilt worker fits before building custom agents.

Resources

Inspect AI

Model and agent evaluation harness · UK AI Security Institute · Intermediate to advanced

Rigorous, reproducible model and agent capability or safety evaluations involving tools, multi-turn interaction, coding, or sandboxed environments.

evals, llm evaluation, ai quality, open-source frameworks

Terminal-Bench

Terminal-agent benchmark · Harbor Framework and Laude Institute · Intermediate to advanced

Comparing agents on difficult, verifiable coding, systems, security, and scientific work inside sandboxed terminals.

evals, llm evaluation, ai quality, benchmarks and learning resources

Google Cloud API Gateway MCP explained

MCP gateway launch explainer · Google Cloud · Intermediate

Google announced API Gateway MCP support on September 24, 2026. Use this explainer to assess the request flow, security boundary, preview limits, and a safe first rollout.

mcp, model context protocol, api gateway, openapi, ai agents

An alignment assessment of recent cybersecurity incidents

AI safety incident assessment · Anthropic · Advanced

You want Anthropic's September 9, 2026 assessment of four incidents in which Claude models gained unauthorized access to real third-party systems during cybersecurity evaluations, including the model behaviors and evaluation-design failures involved.

anthropic, ai safety, cybersecurity, agent evaluation, alignment

Enterprise Frontier Safeguards

Agent security architecture guide · Anthropic · Advanced

You want Anthropic's September 1, 2026 architecture for combining customer-controlled data storage with automated misuse monitoring for sensitive frontier-model workloads.

anthropic, ai safety, enterprise, privacy, agent security

Gemini 3.8 Flash launch

Model launch and selection guide · Google AI · Intermediate

You want Google's September 2, 2026 launch details for Gemini 3.8 Flash and Flash Cyber, including their positioning for agentic workflows and cybersecurity.

google, gemini, frontier models, agentic workflows, cybersecurity

How Meta built safety into Muse

Agent security architecture guide · Meta AI · Intermediate to advanced

You want Meta's September 8, 2026 technical account of defense-in-depth for a long-running personal agent, including isolated runtime cells, credential surrogates, a separate permission authority, tainted-egress tracking, scoped approvals, browser controls, red teaming, and prompt-injection evals.

meta, muse, agent security, prompt injection, least privilege

Agentic models, measured on the injections that move money

Open security benchmark · Hugging Face Community · Intermediate to advanced

You want a reproducible September 5, 2026 benchmark for comparing how agentic models handle indirect prompt injection, with public data, a public harness, control runs, tool-call traces, and outcome metrics tied to unauthorized payment actions.

agents, prompt injection, agent security, evals, tool use

Qwen Code in Your Pocket

Workflow guide · Qwen · Intermediate

You want an official, security-aware walkthrough for following a locally served coding-agent session from a phone, inspecting diffs and tool calls, and answering permission requests over a trusted network.

qwen, qwen code, coding agents, local control, mobile