defilantech/infercost
Kubernetes-native cost intelligence for on-premises AI inference. Computes true cost-per-token from GPU amortization, electricity, and real power draw.
GitHub repository with 6 stars and 2 forks.
Language: Go
Topics: ai, cost, finops, gpu, inference, kubernetes, kubernetes-operator, llm, open-source, prometheus