What does 'most inferencing local' mean in Centralpoint?
It means that the majority of routine AI work — embeddings, classification, entity extraction, summarization, redaction, retrieval scoring, many domain-specific tasks — runs on local models inside the client network. Only the heaviest generative tasks may route to a cloud LLM, and that routing is policy-driven and metered.