AI directory search

Search across educators, skills, and resources.

Use this when you know the topic you need: Claude Code, MCP, evals, RAG, agents, product, coding, prompting, foundations, or model internals.

6 matches for "traces"

Providers and platforms

Langfuse profile photo

Langfuse

Langfuse Docs · Intermediate

Good operational material for tracing, scoring, and improving production LLM apps.

Topics

Observability, Prompt management, Evals, Tracing

Resources

Agentic models, measured on the injections that move money

Open security benchmark · Hugging Face Community · Intermediate to advanced

You want a reproducible September 5, 2026 benchmark for comparing how agentic models handle indirect prompt injection, with public data, a public harness, control runs, tool-call traces, and outcome metrics tied to unauthorized payment actions.

agents, prompt injection, agent security, evals, tool use

Give Your Coding Agents a Memory You Own

Open-source guide and repo · Hugging Face · Intermediate

You want a September 3, 2026 walkthrough of funes, an open-source local memory layer that indexes agent traces, preserves provenance, and lets Claude Code, Codex, pi, and Hermes recall decisions across sessions and machines.

hugging face, coding agents, agent memory, codex, claude code

How to Evaluate AI Agents

Evaluation guide · Hugging Face Community · Intermediate

You need a practical evaluation plan built around representative tasks, controlled environments, observable traces, outcome and constraint metrics, repeated trials, and production failures turned into regression tests.

hugging face, agents, evals, evaluation harnesses, trajectory metrics

OpenAI Evaluate agent workflows

Guide · OpenAI · Intermediate

You need the current OpenAI path for tracing, grading, and regression-testing agent workflows instead of only single-prompt evals.

openai, agents, evals, traces, graders

AI Engineering: Code Is the New Agent Harness

Beehiiv post · Sumanth P · Intermediate

You want a concise technical briefing on why code, traces, tests, and harnesses matter for real agent systems.

beehiiv, agents, ai engineering, evals, tracing