• Decrease Text SizeIncrease Text Size

What kinds of inferencing run locally in Centralpoint?

Embeddings, classification, entity extraction, summarization on smaller documents, redaction, routing decisions, semantic search relevance scoring, and many domain-specific tasks all run locally using on-premise models such as Llama, Qwen, and Onyx. Only the heaviest generative tasks are typically delegated to cloud LLMs, and even those can be local on sufficient hardware. s