►
How To Test AI Agents With Simulations
Hamel Husain · 2026, agent simulations, evals, production traces
AI directory search
Use this when you know the topic you need: Claude Code, MCP, evals, RAG, agents, product, coding, prompting, foundations, or model internals.
92 matches for "agents"
Watch first when you want a fast feel for the topic before opening courses, docs, or profiles.
►
Hamel Husain · 2026, agent simulations, evals, production traces
►
Alexey Grigorev · 2026, AI engineering careers, coding agents, evals
►
Bitwise AI · 2026, Prime Agent, Rust, multi-agent coding
►
AI Engineer · 2026, GitHub Copilot, custom agents, GitHub Actions
►
Graham Neubig · 2026, deep research, search agents, retrieval
►
AI Engineer · 2026, reinforcement learning, small models, financial agents
►
AI Engineer · 2026, CLI design, MCP, agent tools
►
AI Engineer · 2026, agent evaluation, customer feedback, browser agents
►
MadeForCloud · 2026, agent security, OWASP, goal hijacking
AI Engineer · Intermediate to advanced
A useful hub for the emerging AI engineering stack and practitioner talks.
Skills
AI engineering, Agents, Developer tools
DeepLearning.AI Short Courses · Beginner to advanced
Structured, practical courses from prompt engineering through agentic workflows.
Skills
Prompting, Agents, RAG, ML foundations
DAIR.AI Prompt Engineering Guide · Beginner to advanced
A broad prompt engineering guide with many concrete examples and patterns.
Skills
Prompting, RAG, Reasoning, Agents
Lilian Weng's AI posts · Advanced
Deep, well-structured technical summaries of major AI topics.
Skills
Agents, RAG, ML research
AI Hero · Beginner to advanced
Practical developer-focused AI education across LLM fundamentals, AI SDK app development, MCP, Claude Code workflows, agent-ready codebases, evals, TDD, handoffs, and reusable skills such as /teach, /grill-me, /to-prd, /to-issues, /tdd, /triage, and /handoff.
Skills
AI coding, Claude Skills, Agentic workflows, AI SDK, MCP, LLM fundamentals, Personalized learning
School of AI Automation · Beginner to intermediate
Skool community aimed at people who want reliable AI agent systems and their first clients without starting as engineers.
Skills
AI agents, Client acquisition, Templates, Automation systems
Club Jam AI for Entrepreneurs · Beginner to intermediate
Skool profile describes AI training for entrepreneurs and small business owners with no tech and no hype.
Skills
ChatGPT, Claude, AI agents, Small business AI
AI Founders Labs · Beginner to intermediate
Skool hub collecting tools, files, walkthroughs, prompts, and lightweight interfaces for AI founders.
Skills
AI agents, Ready-made projects, Dashboards, Prompts
Linda Mutricy on Maven · Beginner to intermediate
Maven bootcamp for operations leaders who want to build agents and automated workflows without prior coding knowledge.
Skills
AI automation, Operations workflows, AI agents, No-code automation
Matt Burton on Maven · Beginner to intermediate
Role-friendly overview of the AI landscape for leaders who need to understand assistants, avatars, automations, and agents.
Skills
AI leadership, Assistants, Avatars, Automations, Agents
Vidhi Chugh on Maven · Beginner to intermediate
Business-leader course on what AI agents are, where to use them, and how to prioritize initiatives and ROI.
Skills
Agentic AI, AI strategy, ROI, Opportunity prioritization
Stephanie Nyarko on Maven · Beginner to intermediate
Practical no-code AI-agent lesson from a product leader focused on helping founders and coaches automate and grow.
Skills
n8n, AI agents, No-code platforms, Business automation
Cole Medin AI agents and local AI · Beginner to intermediate
Consistently practical agent, RAG, local AI, and automation teaching with a strong bias toward workflows you can actually reproduce.
Skills
AI agents, RAG, Local AI, n8n, Claude Code
Forward Future · Beginner to intermediate
Useful when you want creator-led walkthroughs, current open-source experiments, and practical explanations without reading raw docs all day.
Skills
AI agents, Open-source AI, Vibe coding, MCP, AI workflows
PraisonAI · Intermediate
A practical open-source builder path for learners who want agent patterns, examples, and tool integrations they can inspect directly.
Skills
AI agents, Multi-agent workflows, MCP, RAG, Open source
LlamaIndex · Intermediate
Strong practical material for connecting LLMs to private data, documents, retrieval, and workflows.
Skills
RAG, Agents, Document workflows, Context augmentation
The Data Exchange · Intermediate
Good practitioner interviews across data, ML, and AI engineering.
Skills
Data systems, ML engineering, AI trends
LangChain and LangGraph tutorials · Intermediate
Hands-on notebooks for LangGraph, agent patterns, retrieval, and stateful LLM apps.
Skills
LangGraph, Agents, RAG, LLM orchestration
James Briggs AI tutorials · Beginner to intermediate
Useful practical explanations of embeddings, retrieval, LangChain, Pinecone, and agent workflows.
Skills
Vector search, RAG, Agents, Embeddings
Weaviate education videos · Beginner to intermediate
Good interviews and tutorials around vector databases, retrieval systems, and AI app patterns.
Skills
Vector search, RAG, Hybrid search, Agents
Dave Ebbelaar AI automation · Beginner to intermediate
Practical AI automation videos for workflows, agents, and business process improvements.
Skills
Automation, Agents, AI workflows, No-code tools
OpenAI and AI engineering talks · Intermediate
Useful for thinking about AI-native software, structured workflows, and developer experience.
Skills
AI engineering, Developer tools, Agents, Structured data
AI Tinkerers One-Shot · Intermediate
High-signal practitioner talks and demos that show what builders are actually trying, where agent systems break, and how workflows get stitched together in the real world.
Topics
Agents, Browser automation, AI engineering, Applied demos, Community learning
Hugging Face Learn · Beginner to advanced
One of the best free ecosystems for learning open-source AI by building with models, datasets, spaces, agents, context engineering, and MCP workflows.
Topics
Agents, Context engineering, MCP, Transformers, Post-training, Open models
Anthropic Academy and Claude docs · Beginner to advanced
Official material for learning Anthropic's current Claude family, model tradeoffs, Claude Code, MCP, computer use, practical prompt workflows, and team model-governance workflows without relying on third-party summaries.
Topics
Claude models, Claude Code, MCP, Computer use, AI fluency, Frontier model selection, Model governance, Release notes
LangChain Academy · Intermediate
Useful when you are ready to build multi-step LLM applications, agents, and graph-based workflows.
Topics
LangGraph, Agents, LLM orchestration, RAG
OpenAI docs, Academy, and Cookbook · Beginner to advanced
Official model and implementation material for learning GPT-6 Astra and cost-sensitive GPT-5.6 choices, Codex workflows, subagents, memories, agent evals, MCP and connector patterns, retrieval, background jobs, prompt engineering, production best practices, model optimization, structured outputs, and OpenAI's Academy learning path.
Topics
GPT-6 Astra, GPT models, Reasoning models, Model selection, Agents, Subagents, RAG, Structured outputs, MCP, Evals, Memories
AI Agents for Beginners · Beginner to intermediate
A structured lesson path for understanding when to use agents and how to build simple agentic systems.
Topics
Agents, Multi-agent workflows, Tool use
AssemblyAI YouTube · Beginner to intermediate
Practical developer tutorials and clear overviews of current AI engineering topics.
Topics
Speech AI, LLMs, Agents, ML concepts
DeepLearning.AI · Beginner to advanced
A broad catalog of structured AI courses, including many short practical courses from tool creators.
Topics
Generative AI, Deep learning, Prompting, Agents
AWS AI and ML training · Beginner to advanced
Best for people learning AI services, generative AI apps, and ML workflows on AWS.
Topics
Cloud AI, Agents, MLOps, Generative AI
Microsoft Learn AI · Beginner to advanced
Useful for Azure AI services, Copilot extensibility, and Microsoft-focused AI app development.
Topics
Azure AI, Agents, Copilot, Cloud AI
Codecademy AI courses · Beginner
Good for beginners who want guided coding exercises while learning AI concepts.
Topics
AI foundations, Python, Prompting, LLM apps
n8n AI workflow tutorials · Intermediate
Good for building agentic workflows that connect LLMs with APIs and business systems.
Topics
Automation, Agents, Tool use, Workflow design
Relevance AI Academy · Beginner to intermediate
Useful for teams learning agent workflows without building every piece from scratch.
Topics
Agents, Automation, AI workflows, No-code tools
Gumloop tutorials · Beginner to intermediate
Template-led way to learn AI automations for research, scraping, enrichment, and repetitive operations.
Topics
Automation, No-code AI, AI workflows, Agents
CrewAI docs · Intermediate
Useful for learning role-based multi-agent patterns and orchestration concepts.
Topics
Agents, Multi-agent workflows, Tool use, Automation
Microsoft AutoGen docs · Intermediate
Good for learning multi-agent design patterns and conversation-based orchestration.
Topics
Agents, Multi-agent workflows, Tool use, AI engineering
Gemini API model docs · Beginner to advanced
Official Gemini material for learning Gemini 3.8 Flash and the current stable, preview, latest, and experimental model lineup, plus the now-GA Interactions API, Lyria music generation, prompt design, function calling, background execution, Deep Research, Computer Use, Hooks, Live API, File Search, coding-agent setup, multimodal tradeoffs, and AI Studio workflows.
Topics
Gemini 3.8 Flash, Gemini models, Multimodal AI, Long context, Model selection, AI Studio, Interactions API, Background execution, Deep Research, Computer Use, Hooks, Live API, Coding agents, Music generation, File Search, API examples
Mistral models docs · Beginner to advanced
Official material for comparing current Mistral families such as Devstral, Magistral, Voxtral, OCR, and newer Small and Medium models across coding, reasoning, and multimodal use cases.
Topics
Mistral models, Open models, Model selection, Agents, Coding, Reasoning, Multimodal AI
Official Qwen material for learning the Qwen3 family, multilingual and multimodal capabilities, local deployment paths, and function-calling behavior.
Topics
Qwen, Open models, Multilingual AI, Coding agents, Multimodal AI, Model selection
Skill system
OpenClaw · Official docs
Use this to understand the OpenClaw skill format: markdown instruction files in directories with SKILL.md frontmatter and tool-use guidance.
Start
Read the official skills docs before installing community skills, then test one low-risk local skill in a disposable workspace.
Agent harness
OpenClaw · Open-source agent harness
Use this when you want an agent runtime you can shape with local skills, plugins, and workspace-level behavior.
Start
Start with the docs, install the minimum useful setup, then add one skill only after you understand its file access and commands.
AI worker platform
Mission Control AI · Preconfigured AI workers
Use this when you want role-specific AI workers with SOPs, integrations, and governance policies already built in.
Start
Map one operational process, identify data and approval boundaries, and evaluate whether a prebuilt worker fits before building custom agents.
Engineering automation platform
VIKTOR.AI · AI-powered engineering apps and agents
Use this when the agent problem is domain engineering workflow automation, not general business admin.
Start
Find a repeatable engineering calculation, review, or reporting workflow and evaluate whether a domain app is safer than a general agent.
Founders, operations teams, developers
Learn first
Good matches
Open next
Developers using coding agents and tool-connected workflows
Learn first
Good matches
Open next
GitHub repo · OpenAI · Beginner to advanced
You need implementation examples rather than theory.
api, examples, rag, agents
GitHub repo · Microsoft · Beginner to intermediate
You want a structured agent learning path with code.
agents, workshops, beginner
Guide · DAIR.AI · Beginner to advanced
You want examples of prompting techniques and patterns.
prompting, rag, reasoning, agents
Workshop · Matt Pocock · Intermediate
You want a structured AI SDK v6 course that covers model choice, text and object generation, UI streams, agents, persistence, context engineering, evals, and advanced app patterns.
ai sdk, llm apps, agents, streaming, evals
Free tutorial · Matt Pocock · Beginner
You need clear mental models for system prompts, tokens, context windows, tools, and agents before building or using AI systems seriously.
llm fundamentals, tokens, context windows, tools, agents
Free tutorial · Matt Pocock · Beginner to intermediate
You want to build TypeScript LLM apps with Vercel's AI SDK, including streaming, structured outputs, model switching, embeddings, tool calls, and agents.
ai sdk, typescript, streaming, structured outputs, tool calling
Free tutorial · Matt Pocock · Intermediate
You want to understand MCP and build TypeScript MCP servers over stdio or HTTP, connect Claude Code to tools, use MCP prompts, and package servers for distribution.
mcp, typescript, claude code, tool calling, agents
Dictionary · Matt Pocock · Beginner to intermediate
You want plain-English definitions for agentic coding concepts such as context windows, tools, MCP, handoffs, skills, subagents, feedback loops, and agent-ready work.
ai coding, agent vocabulary, context engineering, claude skills
Guide · Matt Pocock · Intermediate
You want to write project instructions that help coding agents understand commands, conventions, architecture, and working boundaries.
agents.md, context engineering, ai coding, agentic workflows
Guide · Matt Pocock · Intermediate
You want to improve a codebase so AI agents can navigate it, run checks, make smaller changes, and recover from mistakes more reliably.
ai coding, agentic workflows, codebase architecture, developer tools
Guide · Matt Pocock · Intermediate
You want TypeScript checks, tests, linters, and review loops that help agents produce better code and catch regressions quickly.
ai coding, typescript, testing, feedback loops
Short course · DeepLearning.AI · Intermediate
You want a focused course on building stateful AI agents and agent workflows with LangGraph.
agents, langgraph, tool use, agentic workflows, ai engineering
Short course · DeepLearning.AI · Beginner to intermediate
You want a practical introduction to role-based multi-agent systems and task orchestration.
agents, multi-agent workflows, crewai, automation
Short course · DeepLearning.AI · Intermediate
You want a hands-on MCP course for connecting tools, context, and Claude-powered apps.
mcp, anthropic, tool use, context engineering, agents
Free course · Hugging Face · Beginner to intermediate
You want a free structured MCP path with concepts, assignments, SDKs, and a certificate route.
mcp, model context protocol, agents, tools, integrations
Short course · DeepLearning.AI · Intermediate
You need to test, trace, and improve agent workflows instead of judging only single LLM responses.
agent evals, evals, agents, reliability, tracing
Short course · DeepLearning.AI · Intermediate
You want to understand how coding agents use tools, inspect code, run commands, and iterate on software tasks.
coding agents, tool execution, developer tools, agents, ai coding
Short course · DeepLearning.AI · Beginner to intermediate
You want a structured Claude Code course before using it on a serious codebase.
claude code, coding agents, developer tools, ai coding, anthropic
Short course · DeepLearning.AI · Beginner to intermediate
You want a fast introduction to building LLM applications with chains, retrieval, and tools.
llm apps, langchain, agents, rag, ai engineering
Agent testing and red-team framework · Giskard AI · Intermediate to advanced
Behavioral tests and adversarial scans of multi-turn agents, chatbots, and RAG systems from a pytest-compatible workflow.
evals, llm evaluation, ai quality, open-source frameworks
Cloud evaluation portal and SDK · Microsoft · Intermediate to advanced
Microsoft Foundry teams evaluating models, RAG systems, and single- or multi-turn agents for quality, safety, task completion, and tool use.
evals, llm evaluation, ai quality, cloud and provider evals
Agent evaluation and optimization harness · Harbor Framework Team · Intermediate to advanced
Running coding and computer-use agents against reproducible, sandboxed task suites at local or cloud scale.
evals, llm evaluation, ai quality, benchmarks and learning resources
Software-engineering agent benchmark · SWE-bench team · Intermediate to advanced
Measuring whether coding agents can resolve real GitHub issues by producing repository patches that pass executable tests.
evals, llm evaluation, ai quality, benchmarks and learning resources
Terminal-agent benchmark · Harbor Framework and Laude Institute · Intermediate to advanced
Comparing agents on difficult, verifiable coding, systems, security, and scientific work inside sandboxed terminals.
evals, llm evaluation, ai quality, benchmarks and learning resources
Agent-evaluation learning resource · Anthropic · Intermediate to advanced
Product and engineering teams designing practical evaluations for multi-turn, tool-using agents.
evals, llm evaluation, ai quality, benchmarks and learning resources
Certification learning path · Google Cloud · Beginner to intermediate
Use this five-course learning path to prepare leaders to explain generative AI, apply Google's AI tools to business innovation, and prepare for the Generative AI Leader certification.
generative ai leader, business transformation, gemini, ai agents, certification
Desktop agent safety guide · GitHub · Intermediate
GitHub released Copilot computer use in public preview on October 1, 2026. Use this guide to decide when visual desktop control is appropriate, enable it deliberately, and test its permission and stop controls before real work.
github copilot, computer use, desktop agents, copilot cli, gui automation
Model migration explainer · Anthropic · Intermediate to advanced
Anthropic released Claude Sonnet 5.5 on September 28, 2026. Use this guide to understand the efficiency claims, breaking API changes, and a safe migration test from Sonnet 5.
claude sonnet 5.5, claude code, coding agents, model migration, adaptive thinking
Reasoning-efficient coding model explainer · Fireworks AI · Intermediate to advanced
Fireworks released Ember-1 on September 23, 2026 as a Kimi K3-based model trained to use fewer reasoning tokens. Use this explainer to evaluate the quality, cost, latency, and context-growth claims on your own agent workload.
ember-1, kimi k3, reasoning models, reasoning tokens, cost optimization
MCP gateway launch explainer · Google Cloud · Intermediate
Google announced API Gateway MCP support on September 24, 2026. Use this explainer to assess the request flow, security boundary, preview limits, and a safe first rollout.
mcp, model context protocol, api gateway, openapi, ai agents
AI-assisted science explainer · Anthropic · Intermediate
Anthropic reported ART on September 23, 2026. Use this explainer to separate what its agents found, what the lab confirmed, and what remains a hypothesis.
claude, ai agents, ai for science, biology, genome mining
Frontier coding model launch explainer · SpaceXAI · Intermediate to advanced
SpaceXAI released Grok 4.7 on September 21, 2026 for coding, agentic tasks, and knowledge work. Use this guide to compare its task reliability and total cost with your current model.
grok 4.7, coding agents, long-running agents, model selection, evals
Multi-model routing launch explainer · Unbiased · Intermediate to advanced
Union Alpha was revealed as Unbiased's Pareto 26.9 on September 17, 2026: a hosted system that routes work across several models and returns one checked answer.
pareto 26.9, union alpha, model routing, ensembles, coding agents
Omnimodal model launch explainer · Qwen · Intermediate
Qwen3.8-Omni-Flash launched September 18, 2026 for long audio and video analysis, text answers, web search, and tool-calling workflows.
qwen3.8, omnimodal models, audio, video, agents
Coding agent evaluation guide · Google Developers Blog · Intermediate to advanced
You want Google's September 9, 2026 guide to evaluating coding agents with small behavioral checks, outcome-based assertions, and batch runs that catch regressions without treating a single benchmark score as the whole story.
coding agents, evals, behavioral evaluations, regression testing, harness engineering
Managed agent API launch and guide · OpenAI · Intermediate to advanced
You want OpenAI's September 10, 2026 launch guide for building long-running cloud agents with the managed Codex harness, context compaction, tool search, multi-agent delegation, and your choice of hosted or self-managed sandbox.
openai, agents, codex, context compaction, multi-agent