Chinese GPU Firm Biren Plans IPO to Better Compete Against Nvidia

(Image credit: Biren Technology)

Biren, a Chinese developer of compute GPUs, is mulling an initial public offering (IPO) in Hong Kong this year. This comes as domestic clients are increasingly favoring its AI chips over those from Nvidia, which are expensive and in short supply, according to a report from Bloomberg. Hoping to seize on this opportunity, the tech startup is positioning itself to capitalize on the increased demand for its products

Biren is expected to apply for its maiden share sale in the next few week, according to the report that cites anonymous sources. Concurrently, Biren is negotiating with potential investors, including government-supported funds in Guangzhou. These discussions are focused on another independent round of funding that could garner about 2 billion yuan ($279 million). Biren was seeking to raise funds last year at a valuation of 17 billion yuan, which is approximately $2.4 billion. For now, Biren has yet to determine the scope of the IPO, along with exact timeframe.

The reason why Biren is so confident of its valuation is that the company's products look competitive compared to compute GPUs from Nvidia (at least on paper) and the market of AI-capable compute GPUs is booming these days.

Biren's debut family of compute GPUs consists of two options: the BR100 and the BR104. The 'baseline' BR104 delivers performance up to 128 FP32 TFLOPS or 1 INT8 PetaFLOPS, whereas the higher-end BR100 — which is essentially two BR104s on one silicon interposer — offers performance up to 256 FP32 TFLOPS or 2 INT8 PetaFLOPS. The mid-tier BR104 comes with 32GB of HBM2E memory, using a 2048-bit interface that provides bandwidth of 819 GB/s. By contrast, the premium BR100 is equipped with 64GB of HBM2E memory, featuring a 4096-bit interface with bandwidth of 1.64 TB/s.

Header Cell - Column 0 Biren BR104Biren BR100Nvidia A100Nvidia H100
Form-FactorFHFL CardOAM ModuleSXM4SXM5
Transistor Count?77 billion54.2 billion80 billion
NodeN7N7N74N
Power300W550W400W700W
FP32 TFLOPS12825619.560
TF32+ TFLOPS256512??
TF32 TFLOPS??156/312*500/1000*
FP16 TFLOPS??78120
FP16 TFLOPS Tensor??312/624*1000/2000*
BF16 TFLOPS512102439120
BF16 TFLOPS Tensor??312/624*1000/2000*
INT810242048??
INT8 TFLOPS Tensor??624/1248*2000/4000*

There is another reason for the Biren's optimism. Its fundraising efforts coincide with the Chinese government's vigorous push to advance its domestic semiconductor industry. This move is a response to a U.S.-led campaign that blocked Chinese companies from acquiring numerous compute GPUs from AMD, Intel, and Nvidia, all of which compete against Biren's products. Since Nvidia's products are expensive and in short supply, according to media reports, Biren can sell more of its GPUs, at least to companies that do not use Nvidia's CUDA software stack for their AI workloads.

But Biren is facing numerous challenges too. Last year, TSMC temporarily halted shipments of compute GPUs to Biren in a bid to make sure that they meet U.S. export rules in terms of performance and capabilities. This forced the company to slash its headcount to cut costs. Apparently, Biren can procure enough silicon for now, so its main job at the moment is to ensure that its software stack is competitive when compared to those of Nvidia, Intel, and AMD.

In this field, Nvidia is extremely hard to beat. The company has spent nearly two decades refining CUDA and in recent years invested hundreds of millions in making CUDA platform of choice for AI development. For now, numerous Chinese hyperscalers prefer to use Nvidia's GPUs for their AI products due to the superiority of CUDA and the amount of money they have already invested in this ecosystem.

Stay on the Cutting Edge

Join the experts who read Tom's Hardware for the inside track on enthusiast PC tech news — and have for over 25 years. We'll send breaking news and in-depth reviews of CPUs, GPUs, AI, maker hardware and more straight to your inbox.

Anton Shilov is a Freelance News Writer at Tom’s Hardware US. Over the past couple of decades, he has covered everything from CPUs and GPUs to supercomputers and from modern process technologies and latest fab tools to high-tech industry trends.