Axelera AI is shipping Europa, the second-generation AI Processing Unit (AIPU) it has been previewing since last year, and it’s launching with validated servers from Dell and Supermicro attached. The Eindhoven company’s pitch is inference on infrastructure the customer controls: agentic systems, vision-language models, generative AI, and computer vision running in a standard rackmount server on premises, for the financial services, healthcare, legal, defense, and government buyers whose compliance or sovereignty rules keep them off public cloud AI. Europa comes three ways: as a bare chip for customers designing their own boards, as the half-height, half-length Axelera Edge 232p PCIe card, and as the full-height, full-length Axelera Server 250p.
629 TOPS at 45W, With the Pre- and Post-Processing On Board
Europa’s headline number, from Axelera’s product page, is 629 TOPS at INT4, INT8, or INT16 inside a 45W TDP, from eight second-generation AI processing cores, double the count in the first-generation Metis part. Alongside them sit 16 RISC-V vector cores that handle pre- and post-processing on the chip, so the host CPU stays free for application logic, plus an onboard video decoder that keeps vision pipelines from bouncing frames through system memory. Memory is 128MB of L2 SRAM backed by 200GB/s of DRAM bandwidth, and the silicon is built on Samsung’s 5nm process. Axelera’s comparison charts claim 3x to 5x performance per dollar and 2x to 3x performance per watt against unnamed competitors on Llama 3 8B, Llama 3 70B, and Llama 3.2 11B Vision, with the four-chip configuration doing most of the winning; those figures combine Axelera internal testing with competitors’ published NIM benchmark data.
The card photos tell you how the two form factors divide the work. With the heatsink off, the Server 250p carries four Europa AIPUs in a row along a full-length board, which lines up with the “Europa 4Chip” column in Axelera’s charts, while the Edge 232p mounts a single AIPU on a half-length board. Both are standard PCIe cards that Axelera says drop into existing servers without rebuilding the environment around a new platform, and the Edge 232p is the one shipping in validated systems today: Dell’s XE5 and Supermicro’s 111AD, which join a list of validated OEM platforms from Advantech, Axiomtek, HPE, Lenovo, and Seco, with more to be announced. “We built our architecture around some of the hardest constraints in computing: power, energy, cost and the need to process data locally,” said Fabrizio Del Maffeo, CEO and co-founder of Axelera AI. “Europa applies those same principles, expanding from Physical to Enterprise AI, giving organizations the performance they need for increasingly sophisticated workloads while keeping control of their data, infrastructure and economics.”
One Toolchain From Metis to Europa
The software side is the Voyager Toolchain, which compiles and optimizes existing models for every Axelera part from the embedded Metis modules up to the Server 250p, using YAML pipeline definitions so a vision or language application built on Metis moves to Europa without a rewrite. On top of it sits Voyager Wingman, the agentic development layer Axelera introduced in July that ports existing inference pipelines, builds new ones, and optimizes code. The newest piece is AxeleraScript (AxScript), the model compilation approach Axelera disclosed last week to widen model support, which the company calls out as the limiting factor for many customers and silicon vendors alike.
Del Maffeo delivers a keynote titled “The Physical AI Inflection Point: Industry Shifts, Adoption Barriers, and the Infrastructure Imperative” at 11:00 a.m. local time today at AI Infra Summit in Santa Clara, the same show where Lightbits is debuting its Inferra KV cache engine, and Europa is running at booth 930. Axelera says it has deployed across more than 600 customers to date; Europa is the part that has to prove the same edge-first design holds up when the workload is a 70B-parameter model serving multiple users from one server.




아마존