A bank is planning to deploy an AI inference solution at its branch locations. The solution must support workloads that require a balance of compute, memory, and potential GPU acceleration, and it must be suitable for installation by nontechnical onsite resources. The requirements are:
high availability of the chassis management plane.
support for up to 768 GB of memory for in-memory model storage and processing
remote server launch requiring no on-site IT staff
use of Cisco Intersight for SaaS-based infrastructure lifecycle management
Which solution meets the requirements?
What describes inference traffic patterns in AI deployments?
What is a purpose of Cisco AI PODs?
A global enterprise is deploying a new AI-driven analytics platform that requires high-performance GPU acceleration, large memory capacity, and robust virtualization support. The current compute environment must co-exist with the newly purposed GPU-enabled workload. This environment will continue to grow, so the customer wants to scale out the resources as needed.
Which Cisco product meets the requirements?
An engineer must optimize the performance of an AI inference workload running on a Cisco UCS C-Series rack server. The workload experiences intermittent latency spikes, and Intersight system logs show frequent thermal event entries.
Which troubleshooting action must be taken first to address the performance issues on the Cisco UCS server?
An engineer configures quality of service in the Cisco ACI fabric to connect VAST storage servers.
Which combination of attributes must be selected?
An organization deploys a new AI training fabric that uses RoCEv2 for GPU communication. The network architect designs the QoS configuration to ensure reliable RDMA transport and must meet these requirements:
Support 256 GPU servers with RDMA connectivity.
Prevent any packet loss that causes RDMA connection failures.
Maintain consistent low-latency communication with a target of less than 10 microseconds.
Use industry-standard protocols and configurations.
Which configuration ensures that RoCEv2 operates as a lossless transport?
An Intersight administrator plans to deploy new server solutions to several small branch offices across the country. Each location needs at least one chassis. The servers must have redundant CPUs, memory, and 200 Gbps of unified fabric connectivity per compute node.
Which hybrid AI compute solution meets the requirements?
Which set of statements describes Quantized Congestion Notification?
Which type of Cisco Intersight profile, when deployed to fabric interconnects, includes the configuration settings for ports, VLANs, and VSANs?