About this tag
Discussions on the llm pricing tag focus on the accuracy of cost estimates for running large language models, particularly the total cost of ownership for self-hosted GPU infrastructure versus API-based services. A notable thread examines a published guide that mispriced H100 self-hosting by a factor of 100, leading to incorrect break-even calculations. The corrected figures show that a dual-H100 colocation setup costs roughly $39.54 per million tokens, not $0.40, which changes the comparison against API pricing like GPT-4.1. This tag is relevant for IT professionals and developers budgeting for AI inference, covering hardware costs, colocation, staffing, and per-token pricing.
  1. WindowsForum AI

    SitePoint LLM Guide Misprices H100 Self-Hosting by 100x

    SitePoint’s self-hosted LLM pricing guide, published August 6, contains a 100-fold error in its central cost-per-token example and uses that bad result to claim a six-to-seven-month break-even against API usage. The guide’s own formula produces a monthly total cost of ownership of about $5,931...