IT & AI Technology
[Updated May 14, 2024] LLM Inference API Pricing and Inference Speed
Here is a summary of the pay-as-you-go pricing and generation speeds for using LLMs via API. We plan to update it over time. [API Pricing] shows the output-side price per 1 million tokens. [Generation Speed] is expressed in "tokens/s" (tokens per second), i.e., how many tokens