Marvell Brings Radix, Low Latency, and Bandwidth to Bear with Teralynx T100

rbanffy1 pts0 comments

Marvell Brings Radix, Low Latency, And Bandwidth To Bear With Teralynx T100

Jump to main content

Search

NEXTPLATFORM AD

Marvell Brings Radix, Low Latency, And Bandwidth To Bear With Teralynx T100

Timothy Prickett Morgan

Timothy Prickett<br>Morgan

Co-Editor, Co-Founder, The Next Platform

Published<br>fri 17 Jul 2026 // 20:40 UTC

When it comes to modern HPC and AI systems, high switch radix – meaning the number of ports you can hang off of a given switch ASIC – has become as important as low latency and high bandwidth.<br>Ideally, you would have a megaswitch that delivered all of this, one with an ASIC socket the size of a half dozen squared off 12-inch silicon wafers. But such a monster switch would present its own problems, which is why we are still breaking up the network just like we do compute and storage.

NEXTPLATFORM AD

But just the same, getting a balance of these three different vectors – radix, latency, and bandwidth – is the job that switch chip designers are tasked with, and Innovium, the hyperscale Ethernet chip startup that was acquired by Marvell in August 2021, has been pushing the latency and bandwidth barriers since it was founded in 2015 and dropped out of stealth mode two years later. And with the new Teralynx T100, which will start sampling any day now, Marvell’s switch ASIC team is pushing the limits to the hilt. By doing so, Marvell is likely to expand the use of the Teralynx chips and switches by the hyperscalers and cloud builders, and may even get some traction at the AI model builders looking for high bandwidth, high radix, and low latency Ethernet for their scale out and scale up networks.<br>The Teralynx chips have been on a tear since the coronavirus pandemic hit in 2020. The formerly independent Innovium shipped over 1 million 400 Gb/sec Ethernet ports with a mix of its Teralynx 7 and 8 chips in 2020, and two years later Marvell pushed the Teralynx ASICs into mass production, with ASICs ranging from 12.8 Tb/sec to 51.2 Tb/sec of aggregate bandwidth.

By 2023, Marvell had shipped over 5 million 400 Gb/sec ports, which works out to somewhere well north of 300,000 switches. By last year, Marvell had sold more than 20 million 400 Gb/sec ports, which is probably somewhere around 1.25 million switches, thanks in large part to the delivery of the 51.2 Tb/sec Teralynx 10 ASIC that ramped in 2024 using 5 nanometer processes from Taiwan Semiconductor Manufacturing Co.<br>This time around with the Teralynx T100, Rishi Chugh, general manager of the Network Switching Division of Marvell, tells The Next Platform that it is pushing the limits hard, making a monolithic switch ASIC that is up against the reticle limits and that is employing TSMC’s 3 nanometer process, and probably its N3E refined variant if we had to guess.

NEXTPLATFORM AD

We very rarely get a die shot of a switch ASIC these days, but Chugh made sure we got one for the T100:

Pretty, isn’t it?<br>At 100.4 Tb/sec of aggregate bandwidth and using 224 Gb/sec SerDes (which is cut down to 200 Gb/sec after signaling overhead is taken off), there are 512 SerDes wrapped around the outside of the packet processing engines and SRAM cache used in the T100. That red area is the SRAM, and there are clearly four blocks of packet processing engines.<br>I count 76 SerDes blocks – 18 on each side left and right, and 18 on the top and bottom plus four in each corner. If each block is four SerDes, then that works out to 608 SerDes, and if you assume an 84 percent yield on these blocks, you get 512 working SerDes.

NEXTPLATFORM AD

Chugh says that by sticking with a monolithic die and not breaking the SerDes into chiplets, the T100 has 25 percent lower power consumption. There is that yield issue with a big monolithic chip on a fairly new N3 process, of course, which is why the actual SerDes count on the chip is 19 percent higher than what the spec sheet says. But if you want to drive performance and lower power, you make a more expensive chip.<br>Clearly in this AI era, money is not yet a limiting factor. . . .

The typical power consumption of the T100 chip will be under 1,000 watts and significantly the chip has a port to port hop latency of only 420 nanoseconds. That is not as low as InfiniBand can get at the same port speed – about 130 nanoseconds – but with some extra engineering Marvell hopes to get the latency down to the middle 300 nanoseconds range, which is low enough that the T100 can support the ESUN scale up memory sharing protocol for AI accelerators, which is going to become an alternative to Nvidia’s NVLink protocol and its NVSwitch memory fabric among the XPUs of the world.<br>The T100 also supports the Ultra Ethernet Consortium protocols, importantly the adaptive routing functions that make scale out networks work well with tens to hundreds of thousands of XPUs. And like the Teralynx 7 and Teralynx 10 chips before it, it can run the open source SONiC Linux-derived network operating system that was originally created by Microsoft as well as...

teralynx marvell t100 latency bandwidth serdes

Related Articles