Learn
3 pages on how language models behave at inference time, and what that costs.
LLM fundamentals & architecture
What a language model actually does when it answers you, and which parts of the architecture you can feel in latency, quality and cost.