Your GPU is not 90% busy. truthscale measures what each node can actually deliver, reports the gap, and makes that the signal you scale on.

0 stars 0 forks 0 watchers Go Apache License 2.0
autoscaling cloud-native dcgm finops go gpu gpu-monitoring gpu-utilization inference-serving kubernetes kv-cache llm-inference mlops nvidia observability platform-engineering prometheus rke2 sre vllm
2 Open Issues Need Help Last updated: Aug 20, 2026

Open Issues Need Help

View All on GitHub

Your GPU is not 90% busy. truthscale measures what each node can actually deliver, reports the gap, and makes that the signal you scale on.

Go
#autoscaling#cloud-native#dcgm#finops#go#gpu#gpu-monitoring#gpu-utilization#inference-serving#kubernetes#kv-cache#llm-inference#mlops#nvidia#observability#platform-engineering#prometheus#rke2#sre#vllm

Your GPU is not 90% busy. truthscale measures what each node can actually deliver, reports the gap, and makes that the signal you scale on.

Go
#autoscaling#cloud-native#dcgm#finops#go#gpu#gpu-monitoring#gpu-utilization#inference-serving#kubernetes#kv-cache#llm-inference#mlops#nvidia#observability#platform-engineering#prometheus#rke2#sre#vllm