What kinds of inferencing run locally in Centralpoint?
Embeddings, classification, entity extraction, summarization on smaller documents, redaction, routing decisions, semantic search relevance scoring, and many domain-specific tasks all run locally using on-premise models such as Llama, Qwen, and Onyx. Only the heaviest generative tasks are typically delegated to cloud LLMs, and even those can be local on sufficient hardware. s