Skip to main content

4 docs tagged with "evaluation"

View all tags

Evaluating Local Models

A reproducible protocol for measuring whether one exact local model configuration clears a real workflow threshold.

Evidence and Bias in AI Notes

A practical way to label specifications, vendor claims, benchmark results, observations, and recommendations without pretending they prove the same thing.

Model & API Radar

Choose the model tier by task, then combine a small set of still-usable free APIs when a prototype must cost nothing.