LLM
LLM Inference Infrastructure Provisioning Course, Part 1: Basic Concepts and Inference Speed
Hello! Today we begin a multi-part series on inference infrastructure provisioning — GPU provisioning in particular — for building your own LLM service in-house. Introduction As LLMs have evolved, the use of public LLMs such as ChatGPT and Claude has spread rapidly through the enterprise. Forward-looking companies, however, are moving beyond these