AI directory search

Search across educators, skills, and resources.

Use this when you know the topic you need: Claude Code, MCP, evals, RAG, agents, product, coding, prompting, foundations, or model internals.

1 matches for "llm judges"

Resources

OpenRouter Ori Eval

Evaluation guide · OpenRouter · Intermediate to advanced

You want to turn real project prompts and data into repeatable cross-provider agent evals that check answers, tool calls, completion, latency, and cost, then rerun them in CI when models change.

openrouter, evals, model selection, coding agents, ci