LLM Model Notes

Independent reference

LLM Model Notes

Clear, practical guides to evaluation design, inference constraints, and the trade-offs behind model comparisons.

Updated · See the project log

Current notes

Evaluation · 9 min

Reading benchmark results carefully

Define the evaluation contract, inspect uncertainty, and check prompt sensitivity before treating a score as evidence.

Inference · 9 min

Latency, throughput, and context

Measure first-token delay, generation cadence, and capacity under a workload another person can reproduce.

Methods · 10 min

Repeatable model comparisons

Turn a product question into a pinned manifest, paired observations, and a compact result bundle.