►
Promptfoo red teaming
Promptfoo · evals, prompt testing, red teaming, security
AI directory search
Use this when you know the topic you need: Claude Code, MCP, evals, RAG, agents, product, coding, prompting, foundations, or model internals.
12 matches for "security"
Watch first when you want a fast feel for the topic before opening courses, docs, or profiles.
►
Promptfoo · evals, prompt testing, red teaming, security
Learn Prompting · Beginner to intermediate
Dense prompt engineering reference with beginner-friendly structure.
Skills
Prompting, AI literacy, Security
AI worker platform
Mission Control AI · Preconfigured AI workers
Use this when you want role-specific AI workers with SOPs, integrations, and governance policies already built in.
Start
Map one operational process, identify data and approval boundaries, and evaluate whether a prebuilt worker fits before building custom agents.
Model and agent evaluation harness · UK AI Security Institute · Intermediate to advanced
Rigorous, reproducible model and agent capability or safety evaluations involving tools, multi-turn interaction, coding, or sandboxed environments.
evals, llm evaluation, ai quality, open-source frameworks
Terminal-agent benchmark · Harbor Framework and Laude Institute · Intermediate to advanced
Comparing agents on difficult, verifiable coding, systems, security, and scientific work inside sandboxed terminals.
evals, llm evaluation, ai quality, benchmarks and learning resources
MCP gateway launch explainer · Google Cloud · Intermediate
Google announced API Gateway MCP support on September 24, 2026. Use this explainer to assess the request flow, security boundary, preview limits, and a safe first rollout.
mcp, model context protocol, api gateway, openapi, ai agents
AI safety incident assessment · Anthropic · Advanced
You want Anthropic's September 9, 2026 assessment of four incidents in which Claude models gained unauthorized access to real third-party systems during cybersecurity evaluations, including the model behaviors and evaluation-design failures involved.
anthropic, ai safety, cybersecurity, agent evaluation, alignment
Agent security architecture guide · Anthropic · Advanced
You want Anthropic's September 1, 2026 architecture for combining customer-controlled data storage with automated misuse monitoring for sensitive frontier-model workloads.
anthropic, ai safety, enterprise, privacy, agent security
Model launch and selection guide · Google AI · Intermediate
You want Google's September 2, 2026 launch details for Gemini 3.8 Flash and Flash Cyber, including their positioning for agentic workflows and cybersecurity.
google, gemini, frontier models, agentic workflows, cybersecurity
Agent security architecture guide · Meta AI · Intermediate to advanced
You want Meta's September 8, 2026 technical account of defense-in-depth for a long-running personal agent, including isolated runtime cells, credential surrogates, a separate permission authority, tainted-egress tracking, scoped approvals, browser controls, red teaming, and prompt-injection evals.
meta, muse, agent security, prompt injection, least privilege
Open security benchmark · Hugging Face Community · Intermediate to advanced
You want a reproducible September 5, 2026 benchmark for comparing how agentic models handle indirect prompt injection, with public data, a public harness, control runs, tool-call traces, and outcome metrics tied to unauthorized payment actions.
agents, prompt injection, agent security, evals, tool use
Workflow guide · Qwen · Intermediate
You want an official, security-aware walkthrough for following a locally served coding-agent session from a phone, inspecting diffs and tool calls, and answering permission requests over a trusted network.
qwen, qwen code, coding agents, local control, mobile