Colossus MK2 GC200

The Colossus MK2 GC200 is Graphcore’s second-generation Intelligence Processing Unit (IPU), a processor architected from the ground up for machine intelligence workloads rather than repurposed graphics rendering. It targets organizations training large-scale AI models—such as natural language proces

TPU 8i

The TPU 8i is Google’s eighth-generation Tensor Processing Unit, purpose-built for low-latency inference workloads. It is designed for organizations deploying large language models and AI agents that require fast, responsive APIs and efficient handling of long-context sequences.

Mx Series FracTLcore Server Systems

Cornami’s Mx Series FracTLcore Server Systems are a next-generation computing platform designed for real-time, scalable performance across AI, data analytics, and secure computing workloads. They target enterprises and developers needing high-throughput, low-latency processing for encrypted AI, zero

NVIDIA Jetson

NVIDIA Jetson is a family of compact edge AI computers designed for robotics and edge AI applications, supported by the NVIDIA JetPack SDK for accelerated software development. It targets developers, […]

Cerebras Inference Service

Cerebras Inference Service is a cloud-based AI inference platform built on the company’s wafer-scale chip technology, designed for developers and enterprises that need the fastest possible model response times. It targets users running large open-weight models for coding, reasoning, voice, and agent

CS-3 Wafer-Scale Engine

The Cerebras CS-3 is a wafer-scale AI system built around the WSE-3, the largest processor ever manufactured at 46,225 mm² and containing 4 trillion transistors. It is designed for organizations training and deploying large language models (LLMs) and other memory-bandwidth-intensive AI workloads, su

Trainium

AWS Trainium is a purpose-built machine learning accelerator designed specifically for training large deep learning models, targeting organizations that need cost-effective, high-throughput training for transformer-based architectures like LLMs and vision models. Developed by Amazon Web Services, it

MLU370

The MLU370-S4/S8 is a family of cloud inference accelerators from Cambricon, a Beijing-based AI chip company that has become China’s most valuable listed stock (market cap ~$81B as of August 2025). Designed for high-density deployment in servers, these half-height, half-length, single-slot cards tar

Metis AIPU

The Metis AIPU is a quad-core AI inference chip built on a proprietary RISC-V architecture with Digital In-Memory Computing (D-IMC), designed for computer vision workloads at the edge. It targets developers and system integrators deploying multi-camera surveillance, robotics, medical imaging, agrite

MLSoC

SiMa.ai’s MLSoC (Machine Learning System-on-Chip) is a purpose-built hardware and software platform for deploying AI at the embedded edge, targeting applications in computer vision, transformers, generative AI, and multimodal processing. It is designed for enterprises in industrial automation, auton