A bank is planning to deploy an AI inference solution at its branch locations. The solution must support workloads that require a balance of compute, memory, and potential GPU acceleration, and it must be suitable for installation by nontechnical onsite resources. The requirements are:
high availability of the chassis management plane.
support for up to 768 GB of memory for in-memory model storage and processing
remote server launch requiring no on-site IT staff
use of Cisco Intersight for SaaS-based infrastructure lifecycle management
Which solution meets the requirements?
What describes inference traffic patterns in AI deployments?
What is a purpose of Cisco AI PODs?
A global enterprise is deploying a new AI-driven analytics platform that requires high-performance GPU acceleration, large memory capacity, and robust virtualization support. The current compute environment must co-exist with the newly purposed GPU-enabled workload. This environment will continue to grow, so the customer wants to scale out the resources as needed.
Which Cisco product meets the requirements?
An engineer must optimize the performance of an AI inference workload running on a Cisco UCS C-Series rack server. The workload experiences intermittent latency spikes, and Intersight system logs show frequent thermal event entries.
Which troubleshooting action must be taken first to address the performance issues on the Cisco UCS server?
An engineer configures quality of service in the Cisco ACI fabric to connect VAST storage servers.
Which combination of attributes must be selected?
An organization deploys a new AI training fabric that uses RoCEv2 for GPU communication. The network architect designs the QoS configuration to ensure reliable RDMA transport and must meet these requirements:
Support 256 GPU servers with RDMA connectivity.
Prevent any packet loss that causes RDMA connection failures.
Maintain consistent low-latency communication with a target of less than 10 microseconds.
Use industry-standard protocols and configurations.
Which configuration ensures that RoCEv2 operates as a lossless transport?
An Intersight administrator plans to deploy new server solutions to several small branch offices across the country. Each location needs at least one chassis. The servers must have redundant CPUs, memory, and 200 Gbps of unified fabric connectivity per compute node.
Which hybrid AI compute solution meets the requirements?
Which set of statements describes Quantized Congestion Notification?
Which type of Cisco Intersight profile, when deployed to fabric interconnects, includes the configuration settings for ports, VLANs, and VSANs?
What does workload distribution offer in an AI infrastructure with local and external resources?
A Cisco UCS C885A M8 server contains a GPU sled configured with six power supply units (PSUs) in an N+2 redundancy configuration. The workload on the GPU sled is currently drawing a load of 4500W.
Using Power Save Mode, how many PSUs can be placed into standby mode to improve power efficiency without compromising redundancy?
Which result is provided through image recognition using AI?
A server administrator must monitor the performance and utilization of GPUs in Cisco UCS servers running AI workloads.
Which two metrics must be used in Cisco Intersight to accomplish these goals? (Choose two.)
A network engineer logs into Nexus Dashboard and sees an anomaly showing that a BGP peer connection has gone down. Next to the anomaly, a yellow bubble displays “Correlated.”
What does Nexus Dashboard indicate by “Correlated” in this case?
How does RAG enhance the capabilities of LLMs?
A customer is deploying a new AI fabric with all new NVIDIA GPUs and must verify the performance between the nodes in the network.
Which tool must be used to validate the customer benchmarks?
An engineer deploys an AI fabric on Cisco Nexus 9000 Series Switches connected to Cisco UCS nodes with NVIDIA GPUs. The requirements call for optimal network performance for training workloads. Adaptive routing is enabled on the NICs, and per-packet load balancing is enabled on the switches.
Which other configuration must be implemented for this integration to be fully operational?