White-label private AI · for GPU integrators
Enclave is the white-label application layer that turns your hardware quote into a working private AI deployment — a fixed-price line item on your paper, delivered under your brand. You margin every deal. We never touch your customer relationship.
The stall
That's not a flaw — it's the right call for a hardware business. But it leaves your customer with a $30–100k system and nobody to make it do the thing they bought it for.
"Private ChatGPT on our documents. Our hardware, our building. Nothing leaves."
The NVIDIA RAG Blueprint gets them a strong pilot — by design. Its interface is a sample for evaluation, and identity, document permissions, and connectors are left to the application layer. Someone still has to build that layer.
The box underdelivers, a freelancer half-finishes it, and the customer remembers who sold them the hardware.
Enclave is the layer that monetizes what your warranty disclaims.
Built, not promised
This isn't a services pitch with a deck. Enclave is a built platform — four hardened services, 1,900+ automated tests, a scored evaluation harness — and the standing demo runs the full story on a single machine in about two minutes: sign-in, document sync, permission-scoped cited answers, sandboxed data analysis, audit export. With the Wi-Fi visibly off.
Branded chat application, single sign-on against the customer's identity provider, role-based access, cited answers with a visible reasoning trail, guardrails, admin training.
SharePoint and file-share connectors with permission sync, ingestion tuned on the real corpus, an evaluation gold set, monitoring, backup/restore runbooks, admin console with audit export.
The platform installs at your integration bench, so the box ships working. Data onboarding happens after delivery, remotely, inside the customer's network.
And it is more than a chat window on a vector search. The application brain is a LangGraph supervisor loop that routes every turn — retrieval, sandboxed computation, or live systems — and calls the stock NVIDIA RAG Blueprint as its retrieval engine. This is the layer between the cluster you deliver and the outcome your customer bought:
The SKU
Separable line item — the customer can strike it and the box still ships. Or attach it as a scripted follow-up two weeks after ship notification. Your reps choose per deal.
| Package | What the customer gets | Price |
|---|---|---|
| Assessment mandatory first step | Corpus survey, permission-model review, sizing, written findings + firm quote | $2,500 (credited) |
| Pilot | Single box (2× NVIDIA RTX™ PRO 6000-class), one corpus, branded chat UI, SSO, ≤25 users, guardrails, training, scored acceptance demo. Pre-priced Production option: 60 days, 50% pilot credit | $25,000 fixed |
| Production | + SharePoint / file-share connectors with document-permission sync, ingestion tuning on the real corpus, evaluation gold set, monitoring, runbook, admin console | $55–95k fixed-scope |
| Managed AI | Patching, model refreshes, index maintenance, quarterly quality report, next-business-day support | $3.5–4.5k / mo |
Built on the NVIDIA RAG Blueprint: NVIDIA NIM™ microservices, NVIDIA NeMo™ Retriever embedding and reranking, guardrails, GPU-accelerated hybrid search — under an NVIDIA AI Enterprise subscription. For compliance-driven buyers who need the supported-stack story — and you resell the per-GPU NVIDIA AI Enterprise licenses for additional margin on every deployment.
Fully open-source substrate (MIT/Apache), no per-GPU licensing, comparable capability — for cost-driven buyers. Same services pricing, same white-label rights.
Partner economics
"For buyers whose data can't touch the cloud — CUI and DFARS 7012 environments, no-cloud-AI policies, air-gapped networks — cloud assistants aren't on the menu. And a cloud assistant never sold anyone a box."
WHY THIS SKU EXISTS
Why us
Enclave's founder built a U.S. state treasury's private AI platform from scratch: multi-agent RAG over sensitive financial data, local models, enterprise SSO and role-based access control, zero-trust architecture, Kubernetes — the exact architecture this SKU deploys.
Next step
20 minutes with the founder, including the live demo. If it's not a fit for your quote flow, you'll know by minute ten.
Book a partner call