Stanford CS229
Stanford CS229 Machine Learning · Intermediate
A strong foundation for people who need the math and modeling basics under applied AI.
Topics
ML foundations, Supervised learning, Unsupervised learning, Model evaluation
AI directory search
Use this when you know the topic you need: Claude Code, MCP, evals, RAG, agents, product, coding, prompting, foundations, or model internals.
5 matches for "model evaluation"
Stanford CS229 Machine Learning · Intermediate
A strong foundation for people who need the math and modeling basics under applied AI.
Topics
ML foundations, Supervised learning, Unsupervised learning, Model evaluation
Fireworks AI model catalog · Intermediate to advanced
Official material for comparing serverless and dedicated model serving, training specialized models, and measuring quality, token use, cost, and task duration on production-shaped evaluations.
Topics
Ember-1, Kimi K3, Serverless inference, Dedicated deployment, Model training, Model evaluation, Reasoning efficiency
Language-model benchmark runner · EleutherAI · Intermediate to advanced
Reproducible few-shot and zero-shot evaluation of base or instruction-tuned language models on established academic benchmarks.
evals, llm evaluation, ai quality, benchmarks and learning resources
Tabular foundation model explainer · NVIDIA · Intermediate to advanced
NVIDIA released Kumo Tabular on September 29, 2026. Use this guide to understand its in-context prediction workflow, reproduce a baseline, and test its vendor-reported results on your own tables.
kumo tabular, tabular foundation models, classification, regression, in-context learning
Evaluation guide · Google AI for Developers · Intermediate to advanced
You want Google's practical July 31, 2026 guide to running the same agent and model evaluations during development and on production traffic, with metrics for quality, safety, grounding, tool use, and trajectories.
google, gemini, agents, evals, model selection