What this proves for your project
That a fine-tuned small model plus retrieval delivers a domain assistant without per-token bills or data leaving your infrastructure, running on one rented server rather than a GPU cluster — and that the person offering to build yours has already built one in production, in public, with the trade-offs published instead of hidden. Click the brain on the homepage and interrogate it yourself. If your documents need the same treatment, start with the RAG pilot — a fixed three-week engagement; if compliance means the model has to live on hardware you own outright, that is self-hosted AI . Both sit on the full price ladder .