Private Token Factories: How Rafay and Protopia AI Let Sensitive Workloads Run on Shared GPU Capacity

The direct win is being able to serve workloads that used to require dedicated, single-tenant endpoints at multi-tenant utilization instead. In Protopia’s modeled deployment economics, that shift meaningfully lowers annual infrastructure cost for those workloads, since they’re no longer sitting in underutilized carve-outs (actual savings depend on workload mix and scale).
Sensitive Data, Open Models, and the AI Factory’s Inference Privacy Layer

Protopia AI Stained Glass Transform (SGT) is now available for NVIDIA Nemotron 3 Super and Nemotron 3 Nano Omni. As an integrated data-privacy layer, SGT is designed to minimize plaintext exposure from the inference path, so an enterprise’s most sensitive workloads can run on the multi-tenant AI factory infrastructure where supported.
Zero-Trust AI Factories Unlock Sensitive AI Anywhere with Protopia AI Stained Glass Transform and NVIDIA Confidential Containers

NVIDIA Confidential Computing, combined with Protopia AI’s Stained Glass Transform, helps organizations protect sensitive inference data and model IP, enabling more secure AI deployment across multi-tenant infrastructure.
Half Your AI Factory Is Sitting Idle, here is the Blueprint That Fixes It

AI Factory ROI depends on using high-value data at full fidelity and running compute at full capacity. Privacy constraints force operators to choose one or the other. HPE and Protopia AI have validated a blueprint that eliminates that tradeoff through private multi-tenant inference.
MGT, Protopia AI, and HPE Launch AI Experience Center to Accelerate Secure AI across State, Local, and Education Government (SLED)

MGT, Protopia AI, and HPE introduce a national AI Experience Center for SLED agencies to scale governed AI while protecting sensitive data.
Secure, Scalable AI Factories

Protopia Stained Glass Transform (SGT) is now compatible with NVIDIA NIM microservices. Together, SGT and NIM microservices enable organizations to run sensitive, high-value workloads on efficient shared infrastructure based on the NVIDIA AI Factory for Government reference design.