One line. Many voicesSeek and you shall find

injuly.in faviconInference cost at scale with napkin math

kept by

Napkin math shows serving a 32B LLM on an NVIDIA B200 GPU costs ~$9.36 per user per month when 300 users share the GPU with typical idle duty cycles.

read later

For all the tabs you promised to read.
Save to read. Read to clear.

Close tabs. Keep links.