RoCE: An Ethernet-InfiniBand Love Story

By Michael Feldman

April 22, 2010

To go along with the low-latency theme of this week’s High Performance Computing Linux Financial Markets confab in New York City, the InfiniBand Trade Association (IBTA) announced the release of the RDMA over Converged Ethernet standard that brings InfiniBand-like performance and efficiency into the Ethernet realm.

Abbreviated RoCE (and pronounced “Rocky”), the new standard allows the RDMA guts of InfiniBand to run over Ethernet. Basically the IBTA has taken the InfiniBand stack, left the IB transport and network layers intact, and swapped the IB link layer for Ethernet. Or as OpenFabrics Alliance Executive Director Bill Boas put it: “The only change here is that the verbs in the InfiniBand standard have been implemented over Ethernet.”

This is a much simpler solution than iWARP (Internet Wide Area RDMA Protocol), which also uses RDMA, but incorporates TCP/IP into the stack. In a sense, iWARP tried to unify InfiniBand and IP, but that model has garnered limited appeal. Supporting the TCP/IP stack meant latency could only get into the 10 microsecond range. Freed of that extra processing burden, RoCE latency can approach 1-3 microsecond territory. And it can be implemented more cheaply and with less power consumption. Yes, IP support is missing, but in a closed cluster environment, you would normally just use a gateway node to talk to the outside world.

In general, RoCE is aimed at users of clustered computing setups who might otherwise have opted for InfiniBand because of its speed and agility, but who are already married to Ethernet — either to maintain compatibility with existing storage networks and compute infrastructure or because their local datacenter already has a big investment in Ethernet technology, expertise and management tools. Mellanox has been talking about this technology for a year or so, under the moniker low-latency Ethernet.

“Essentially what you’re able to do now is run close to InfiniBand-like latency over 10 Gigabit Ethernet,” says Brian Sparks, IBTA marketing working group co-chair and director of marketing communications at Mellanox . “But you don’t have the InfiniBand barrier and the learning curve that goes with that.”

RoCE isn’t quite InfiniBand-strength, though. QDR IB nets 32 Gbps and sub-microsecond latencies, while RoCE is currently limited to 10 Gig and latencies closer to single-digit microseconds. For most apps, though, 10 Gig is plenty of bandwidth (and there’s a clean path to 40 and 100 Gig when Ethernet catches up). The real hurt is on the latency side.

Financial services, database warehousing, cloud computing and related virtualization apps are all potential targets of this technology. One of the tastiest low-hanging fruits for RoCE is high frequency trading (HFT), an Ethernet-based application that is all about latency. HFT is a highly lucrative class of algorithmic trading that relies far more on network performance than compute muscle. The object of the game is to turn reams of market data coming in from Ethernet-based ticker feeds into split-second arbitrage opportunities. One person I recently spoke with characterized it as “picking up a nickel in front of a freight train.” RoCE seems tailor-made for this type of application.

In more traditional HPC, RoCE could have plenty of takers. Again, the real draw here is the ubiquity of the Ethernet ecosystem and the promise of near-InfiniBand performance. It’s worth noting that more than half the systems on the TOP500 list are still employing Ethernet interconnects. That’s because there are plenty of big cluster-based workloads (for example, data mining) that don’t require obsessively tight coupling, but would still benefit from better latency than vanilla Ethernet. As HPC makes deeper inroads into the enterprise, RoCE could look fill this role.

As of this week, RoCE is implemented in OpenFabrics Enterprise Distribution (OFED) 1.5.1. The Linux version is available today, with a Windows implementation to follow later this year. That makes it especially nice for applications already written for OFED RDMA. In these cases, there would be no need to twiddle with the code again; the apps should just auto-magically run over any RoCE fabric.

On the hardware side, basically you need an L2 Ethernet switch with IEEE DCB (Data Center Bridging, aka Converged Enhanced Ethernet) with support for priority flow control. On the compute or storage server end, you need an RoCE-capable network adapter. Expect the most enthusiastic vendors to come out with products later this year. Mellanox has already declared its intentions to offer RoCE-friendly adapters. OpenFabrics will release a software-based RoCE later in the second quarter. Soft-RoCE will make a regular 10GbE NIC act like the hardware version.

One might wonder why the IBTA and its InfiniBand-loving members decided to push an Ethernet protocol at all. If RoCE is successful, there’s bound to be some cannibalization of the InfiniBand market. But that’s the wrong way to think about it. First, there are no InfiniBand vendors anymore, at least not in the strict sense. All these companies — Mellanox, Voltaire and QLogic — offer Ethernet products of one sort or another. The market decided some time ago that IB technology would only spread so far. RoCE is another way for these vendors to reach customers they couldn’t attract before. The calculation is that there’s enough daylight between RoCE and InfiniBand to support the viability of both technologies.

Subscribe to HPCwire's Weekly Update!

Be the most informed person in the room! Stay ahead of the tech trends with industry updates delivered to you every week!

Connecting with ISC 2025: Amanda Randles Keeps HPC Flowing

December 13, 2024

Amanda Randles is an accomplished computational scientist and biomedical engineer, contributing to high performance computing (HPC), bioengineering, and computational fluid dynamics (CFD). Through the innovative approach Read more…

Accelerating the Weather: Static Code Analysis with Codee

December 12, 2024

Optimization is a challenging task. Beyond the compiler optimizing your code, which can be a hit-or-miss proposition, the next best thing is often detailed code inspection. Codee, previously known as Parallelware Analyze Read more…

EDA for Quantum Algorithm and Circuit Design? Classiq Paper Shows Progress

December 12, 2024

Developing practical EDA tools for quantum software development may still seem distant but Classiq, a Israel-based quantum software specialist, has just posted an interesting paper — Design and synthesis of scalable qu Read more…

Nvidia’s Blackwell Showcases the Future of AI Is Water-Cooled – For Now

December 11, 2024

Nvidia’s Blackwell processor is a game changer. It is also incredibly dense and it runs hot. Apparently, this heat doesn’t become a big problem until you have a whopping 72 of the processors in a rack, but if you get Read more…

Nvidia, Intel, and AMD Invest in Startup Ayar Labs

December 11, 2024

AMD, Intel, and Nvidia disagree on many things but are eye to eye on one thing: investing in optical interconnect firm Ayar Labs. The company, which is creating technology for chips and systems to communicate using pulse Read more…

Microsoft Azure & AMD Solution Channel

Announcing Azure HBv5 Virtual Machines: A Breakthrough in Memory Bandwidth for HPC

The most powerful Azure virtual machine for HPC

On November 19 at Ignite 2024, Microsoft unveiled their most advanced and efficient high-performance computing infrastructure to date, Azure HBv5. Read more…

Shining a Light on AI Risks: Inside MLCommons’ AILuminate Benchmark

December 10, 2024

As the world continues to navigate new pathways brought about by generative AI, the need for tools that can illuminate the risk and reliability of these systems has never felt more urgent. MLCommons is working to shine a Read more…

Shutterstock 2283618597

Connecting with ISC 2025: Amanda Randles Keeps HPC Flowing

December 13, 2024

Amanda Randles is an accomplished computational scientist and biomedical engineer, contributing to high performance computing (HPC), bioengineering, and computa Read more…

Nvidia’s Blackwell Showcases the Future of AI Is Water-Cooled – For Now

December 11, 2024

Nvidia’s Blackwell processor is a game changer. It is also incredibly dense and it runs hot. Apparently, this heat doesn’t become a big problem until you ha Read more…

Shining a Light on AI Risks: Inside MLCommons’ AILuminate Benchmark

December 10, 2024

As the world continues to navigate new pathways brought about by generative AI, the need for tools that can illuminate the risk and reliability of these systems Read more…

Google Debuts New Quantum Chip, Error Correction Breakthrough, and Roadmap Details

December 9, 2024

Google today introduced its latest quantum chip — Willow (~100 qubits) — coinciding with two key achievements run on the new chip: breaking of the so-called Read more…

Recap: Hardware News From AWS Re:Invent

December 7, 2024

The first few days of AWS Re:Invent have shown how the company has set itself apart from rivals in the AI space. The cloud provider now has its own foundation A Read more…

Shutterstock 1179306271

Intel’s HPC Future Is Uncertain After Gelsinger’s Retirement

December 6, 2024

At this year's Supercomputing 2024 show, Intel wasn't shouting about its tech from the rooftops. The honors went to rivals AMD and Nvidia, which boasted about h Read more…

AWS Delivers the AI Heat: Project Rainier and GenAI Innovations Lead the Way

December 5, 2024

At AWS re:Invent 2024 in Las Vegas, Amazon unveiled a series of transformative AI initiatives, including the development of one of the world's largest AI superc Read more…

National Quantum Initiative Act Reauthorization Bill Calls for $2.7B and New Centers

December 4, 2024

After more than a year of delay, the National Quantum Initiative Reauthorization Act has been submitted to the U.S. Senate. The latest NQIA shifts emphasis from Read more…

CORNELL I-WAY DEMONSTRATION PITS PARASITE AGAINST VICTIM

October 6, 1995

Ithaca, NY --Visitors to this year's Supercomputing '95 (SC'95) conference will witness a life-and-death struggle between parasite and victim, using virtual Read more…

SGI POWERS VIRTUAL OPERATING ROOM USED IN SURGEON TRAINING

October 6, 1995

Surgery simulations to date have largely been created through the development of dedicated applications requiring considerable programming and computer graphi Read more…

U.S. Will Relax Export Restrictions on Supercomputers

October 6, 1995

New York, NY -- U.S. President Bill Clinton has announced that he will definitely relax restrictions on exports of high-performance computers, giving a boost Read more…

Dutch HPC Center Will Have 20 GFlop, 76-Node SP2 Online by 1996

October 6, 1995

Amsterdam, the Netherlands -- SARA, (Stichting Academisch Rekencentrum Amsterdam), Academic Computing Services of Amsterdam recently announced that it has pur Read more…

Cray Delivers J916 Compact Supercomputer to Solvay Chemical

October 6, 1995

Eagan, Minn. -- Cray Research Inc. has delivered a Cray J916 low-cost compact supercomputer and Cray's UniChem client/server computational chemistry software Read more…

NEC Laboratory Reviews First Year of Cooperative Projects

October 6, 1995

Sankt Augustin, Germany -- NEC C&C (Computers and Communication) Research Laboratory at the GMD Technopark has wrapped up its first year of operation. Read more…

Sun and Sybase Say SQL Server 11 Benchmarks at 4544.60 tpmC

October 6, 1995

Mountain View, Calif. -- Sun Microsystems, Inc. and Sybase, Inc. recently announced the first benchmark results for SQL Server 11. The result represents a n Read more…

New Study Says Parallel Processing Market Will Reach $14B in 1999

October 6, 1995

Mountain View, Calif. -- A study by the Palo Alto Management Group (PAMG) indicates the market for parallel processing systems will increase at more than 4 Read more…

Leading Solution Providers

Contributors

CORNELL I-WAY DEMONSTRATION PITS PARASITE AGAINST VICTIM

October 6, 1995

Ithaca, NY --Visitors to this year's Supercomputing '95 (SC'95) conference will witness a life-and-death struggle between parasite and victim, using virtual Read more…

SGI POWERS VIRTUAL OPERATING ROOM USED IN SURGEON TRAINING

October 6, 1995

Surgery simulations to date have largely been created through the development of dedicated applications requiring considerable programming and computer graphi Read more…

U.S. Will Relax Export Restrictions on Supercomputers

October 6, 1995

New York, NY -- U.S. President Bill Clinton has announced that he will definitely relax restrictions on exports of high-performance computers, giving a boost Read more…

Dutch HPC Center Will Have 20 GFlop, 76-Node SP2 Online by 1996

October 6, 1995

Amsterdam, the Netherlands -- SARA, (Stichting Academisch Rekencentrum Amsterdam), Academic Computing Services of Amsterdam recently announced that it has pur Read more…

Cray Delivers J916 Compact Supercomputer to Solvay Chemical

October 6, 1995

Eagan, Minn. -- Cray Research Inc. has delivered a Cray J916 low-cost compact supercomputer and Cray's UniChem client/server computational chemistry software Read more…

NEC Laboratory Reviews First Year of Cooperative Projects

October 6, 1995

Sankt Augustin, Germany -- NEC C&C (Computers and Communication) Research Laboratory at the GMD Technopark has wrapped up its first year of operation. Read more…

Sun and Sybase Say SQL Server 11 Benchmarks at 4544.60 tpmC

October 6, 1995

Mountain View, Calif. -- Sun Microsystems, Inc. and Sybase, Inc. recently announced the first benchmark results for SQL Server 11. The result represents a n Read more…

New Study Says Parallel Processing Market Will Reach $14B in 1999

October 6, 1995

Mountain View, Calif. -- A study by the Palo Alto Management Group (PAMG) indicates the market for parallel processing systems will increase at more than 4 Read more…

  • arrow
  • Click Here for More Headlines
  • arrow
HPCwire