How does Centralpoint handle Speculative Decoding?
Speculative-decoding endpoints with Centralpoint: Centralpoint routes to inference endpoints using speculative decoding for faster response times, while consistently metering tokens at the target-model rate. The model-agnostic platform supports any backend — vLLM, TensorRT-LLM, hosted APIs — and deploys chatbots through one line of JavaScript on any portal. The pla