Red Hat and Cisco have jointly engineered this solution to deliver a preintegrated and ready-to-use platform that combines high-performance hardware with a flexible, cloud-native application and AI platform. It provides organizations with everything they need to streamline time to value as they move from pilot to production, including unified infrastructure, streamlined deployment, AI-optimized security features, and more.
Unified, scalable architecture for any AI workload
This comprehensive AI platform from Red Hat and Cisco brings together high-performance compute, networking, and management capabilities and a consistent, cloud-native foundation for building, training, deploying, monitoring, and scaling AI models across complex environments.
It provides lossless, low-latency networking and sustained throughput for AI model training, fine-tuning, and inferencing, and supports the evolving needs of modern AI workloads through a modular scale-unit design and the proven scalability of Kubernetes-powered Red Hat OpenShift.
Streamlined deployment through validated integration
As part of Cisco’s collection of CVDs, this reference architecture offers fully integrated, tested, and validated AI infrastructure, with 100+ pages of documentation, including Ansible playbooks for automating deployment and Day 2 operations.
In tandem with its plug-and-play deployment, this helps organizations reduce setup time and complexity, including minimizing misconfigurations and mitigating related risk.
Flexible and efficient operations
Cisco AI PODs with Red Hat offer the flexibility to run AI-powered applications in containers or virtual machines (VMs), with centralized and unified management for traditional and emerging workloads. It reinforces this flexibility with Red Hat OpenShift delivering consistency across hybrid or multicloud environments, enterprise-wide automation capabilities offered by Ansible Automation Platform, and streamlined lifecycle management through Cisco Intersight.
Cisco Intersight provides a unified, cloud-native platform for managing the entire compute infrastructure, including Cisco UCS servers and GPU-optimized systems, with policy-based provisioning, deployment, and monitoring. This allows IT teams to centrally manage fleets at scale, automate complex tasks, and reduce setup time for AI workloads across datacenters, colocation facilities, and distributed edge environments.
Reliable, security-focused infrastructure
This solution is jointly engineered and validated for interoperability and performance, and designed to provide consistent policy enforcement, governance, and observability across complex IT environments.
It provides an embedded and validated focus on security and governance at every layer, and offers the option to further reinforce this with proactive, AI-optimized security capabilities from Red Hat AI, Cisco AI Defense, Cisco Hypershield, Red Hat OpenShift and a range of tested and integrated third-party security solutions.
Simplified choice and flexibility
Offered as a fully configurable solution within a range of Cisco AI POD bundle options, this flexible offering gives customers the ability to tailor their AI infrastructure to their specific use cases and infrastructure strategy.
It offers the flexibility to choose between a number of operational models. This includes a choice of operational models for hardware (e.g., Cisco UCS, AI accelerators) and software (e.g., Red Hat OpenShift, Red Hat AI), backed up by a comprehensive ecosystem of integrated partner hardware and software solutions, and 2 key operational models for network management, depending on an organization’s needs.
- Air-gapped, on-premise networking management with Cisco Nexus Switches and Cisco Nexus Dashboard, and a broad range of configurability options. This operational model, which is backed by Red Hat AI and the ability to support air-gapped, on-premise deployments, meets the needs of organizations with high compliance requirements, such as those in the financial, government, or healthcare sectors.
- A cloud-managed operational model that offers Cisco Nexus Hyperfabric as the network management platform, while keeping all data on premise. This provides a cloud-like experience for customers repatriating workloads, with a centralized dashboard for building design templates, autogenerating blueprints and bills of materials (BOMs), and accessing mobile-friendly deployment guides, as well as real-time connection validation and lifetime monitoring.