Comprehensive LLM economics & inference latency evaluator. Computes monthly token costs, prompt-caching savings, and tokens-per-second throughput across frontier and open-weight models.