As enterprises accelerate AI adoption, managing inference workloads, token costs, and infrastructure complexity has become a critical priority. Spectro Cloud has introduced PaletteAI Inference Launchpad, a turnkey solution designed to help organizations run AI inference closer to their data while reducing token costs by up to 70%.
Alongside the launch, the company announced expanded support for AMD-powered AI infrastructure, enabling enterprises, neoclouds, and sovereign cloud providers to deploy and manage production AI environments across both AMD and NVIDIA platforms from a unified infrastructure management platform.
As AI moves into large-scale production, organizations are facing increasing operational expenses driven by token consumption. According to Goldman Sachs Research, AI token usage is expected to grow 24-fold by 2030, reaching approximately 120 quadrillion tokens per month as enterprise and consumer AI adoption accelerates.
PaletteAI Inference Launchpad is designed to address this challenge by providing a turnkey, locally managed AI inference solution that eliminates the need to build and maintain a custom inference stack. Organizations can intelligently route workloads between local AI models and external frontier models while optimizing performance and controlling operational costs.
The new platform enables enterprises to deploy AI inference closer to applications, users, and data while maintaining governance and operational control. PaletteAI Inference Launchpad includes capabilities for token usage metering, governance policy enforcement, quota management, and intelligent workload routing across heterogeneous AI environments.
By reducing dependence on external AI inference services where local models are more efficient, organizations can significantly lower operational expenses while improving performance and maintaining compliance requirements.
Spectro Cloud has also expanded PaletteAI support for AMD-powered AI infrastructure, including AMD GPUs, AMD GPU Operator, ROCm runtime, the AMD enterprise AI reference stack, and AMD-optimized models available through the AMD Inference Microservices (AIMs) catalog.
These enhancements provide enterprises, neocloud providers, and sovereign cloud operators with a validated path for deploying and managing AMD-based AI infrastructure while maintaining flexibility across diverse hardware environments.
PaletteAI delivers full-stack lifecycle management spanning bare metal infrastructure, Kubernetes, GPUs, AI runtimes, models, inference services, and token-level governance. The platform enables organizations to deploy, operate, govern, and scale AI environments consistently across both AMD and NVIDIA infrastructure.
"Enterprises and cloud providers are looking for open, scalable AI infrastructure that gives them more control over cost, performance and deployment choice," said Kumaran Siva, Corporate Vice President, Enterprise AI at AMD. "Spectro Cloud’s support for AMD-powered infrastructure in PaletteAI and PaletteAI Inference Launchpad helps customers accelerate production AI deployments across flexible, open AI stacks."
To support organizations operating in regulated industries and sovereign cloud environments, Spectro Cloud is partnering with infrastructure providers specializing in secure AI deployments. NexusIgnite is among the first partners helping customers deploy AI inference infrastructure where data residency, governance, and operational control are essential.
"NexusIgnite is focused on delivering sovereign AI infrastructure for organizations with strict data residency, governance and compliance requirements," said Greg Forrest, CEO, NexusIgnite. "As enterprise AI moves into production, customers need inference services that are easier to deploy, operate and govern across their GPU infrastructure. Spectro Cloud's PaletteAI Inference Launchpad fits our managed AI infrastructure strategy and gives those customers a faster path to production within trusted sovereign environments."
PaletteAI Inference Launchpad and expanded AMD support for PaletteAI are available now. Customers can contact Spectro Cloud to schedule a live demonstration. The company will also showcase live demonstrations during AMD Advancing AI on July 22–23, 2026.
Spectro Cloud helps platform teams and cloud providers modernize and manage infrastructure for the AI era without adding more tools or operational complexity.
With PaletteAI, enterprises, public sector organizations, neoclouds and sovereign clouds can build, deploy, manage, govern and scale full-stack environments across VMs, Kubernetes, edge, regulated and air-gapped locations, and AI infrastructure. PaletteAI helps teams start quickly with Launchpads — turnkey, locally managed solutions for urgent outcomes such as VM modernization, token cost control and edge operations — then scale into centralized lifecycle management, governance and fleet operations on the same platform.