Why is local inferencing still metered?
Because local inferencing consumes capacity even when it doesn't consume cloud tokens. Metering local consumption surfaces which skills are driving GPU load, when capacity is approaching saturation, and what the imputed cost of local versus cloud routing actually is. The platf