Infercost Alternatives
Similar projects and alternatives to infercost based on common topics and language
-
LLMKube
Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
-
SaaSHub
SaaSHub - Software Alternatives and Reviews. SaaSHub helps you find the best software and product alternatives
-
tensor-fusion
Tensor Fusion is a state-of-the-art GPU virtualization and pooling solution designed to optimize GPU cluster utilization to its fullest potential.
-
-
-
LLMKube
Discontinued Kubernetes operator for GPU-accelerated LLM inference - air-gapped, edge-native, production-ready [Moved to: https://github.com/defilantech/LLMKube] (by Defilan)
-
infercost discussion
infercost reviews and mentions
Stats
defilantech/infercost is an open source project licensed under Apache License 2.0 which is an OSI approved license.
The primary programming language of infercost is Go.