Self-hosted LLM vs API break-even math Fanout ProContinue readingThis source-rich field note is available with Fanout Pro.Get Fanout ProRelated articlesFP8 attention vs FP8 KV cacheGoodput vs throughput in LLM inferenceThe scaled dot-product attention equationContinue learningGradient DescentInference engineering