Cheapest LLM API Pricing, Answered

84 models · 79 with published list rates · verified 2 October 2026

Everyone else publishes a pricing table. This is a sorted answer, cut by the constraint that actually decides which model you should pick. Every row is computed from the published dataset, not typed in.

Pick your constraint

The number that matters is usually output. On most models the output rate is 3–8× the input rate. A workload with short prompts and long answers is an output workload. Rank by output unless your prompts are the long part.