
Sensitive Data, Open Models, and the AI Factory’s Inference Privacy Layer
Protopia AI Stained Glass Transform (SGT) is now available for NVIDIA Nemotron 3 Super and Nemotron 3 Nano Omni. As an integrated data-privacy layer, SGT is designed to minimize plaintext exposure from the inference path, so an enterprise’s most sensitive workloads can run on the multi-tenant AI factory infrastructure where supported.
Nemotron 3 models provide leading accuracy with fastest throughput, open weights, published recipes, enterprise-ready agents, and model customization for control, trust, and AI specialization.
Protopia and Rafay deliver multi-tenancy for shared GPU AI factories

Enterprise AI infrastructure providers are turning to multi-tenancy paired with upstream data protection to convert idle GPU capacity into secure, token-metered services that enterprises will actually adopt at scale.
Private Token Factories: How Rafay and Protopia AI Let Sensitive Workloads Run on Shared GPU Capacity
Rafay’s Token Factory makes it brutally simple for neoclouds to publish AI models as OpenAI-compatible, token-metered endpoints on shared GPU capacity. Developers love it. Finance loves the chargeback.
But, regulated enterprises with clinical records, financial transactions, or proprietary source code may struggle to leverage this opportunity. Typical inference endpoints receive inputs in plaintext. For a healthcare system, a bank, or a government agency, plaintext on shared infrastructure can be a non-starter.
