Archivo de la categoría: inference

IBM Cloud and Together AI expand AI infrastructure with NVIDIA

IBM is preparing to deploy a large NVIDIA-based AI computing cluster on IBM Cloud under a multiyear agreement with Together AI worth $240 million. The cluster is expected to become available in the first quarter of 2027 and will use NVIDIA HGX B300 systems connected through NVIDIA Spectrum-X Ethernet networking. Together AI plans to use […]

The post IBM Cloud and Together AI expand AI infrastructure with NVIDIA appeared first on Cloud Computing News.

Google Cloud unveils AI-optimised infrastructure enhancements

Google Cloud has announced significant advancements in its AI-optimised infrastructure, including fifth-generation TPUs and A3 VMs based on NVIDIA H100 GPUs. Traditional approaches to designing and constructing computing systems are proving inadequate for the surging demands of workloads like generative AI and large language models (LLMs). Over the last five years, the parameters in LLMs… Read more »

The post Google Cloud unveils AI-optimised infrastructure enhancements appeared first on Cloud Computing News.