How fast is local inferencing in Centralpoint?
Latency depends on the model and hardware, but typical local inference on properly-sized hardware completes in hundreds of milliseconds to a few seconds for most enterprise tasks. Centralpoint can also use speculative decoding, KV caching, and continuous batching to further compress latency. Th