Signal & Noise
Menu

Inference

Running a trained model to get outputs (as opposed to training it). Where all production LLM cost and latency lives.

Related terms

← Back to the full glossary