Skip to main content

Iluvatar CoreX BI-V150

Product Overview​

Iluvatar CoreX BI-V150 is a general-purpose GPU accelerator card from Iluvatar CoreX for the cloud training market, released in 2023 and in mass production in 2024. Based on Iluvatar CoreX's self-developed ivcore11 general-purpose GPU architecture, built on a 7nm process with 2.5D CoWoS packaging technology, it aims to provide Chinese compute solutions for AI training, high-performance computing, and other scenarios.

Product positioning: a Chinese general-purpose GPU training card, CUDA-ecosystem compatible, supporting FP32, FP16, and INT8 multi-precision computing.


Core Specifications​

ParameterValue
Architectureivcore11 (second-generation general-purpose GPU architecture)
ProcessTSMC 7nm
Packaging2.5D CoWoS
FP3248 TFLOPS
FP16192 TFLOPS (confirmed by official/procurement specifications)
INT8384 TOPS (confirmed by official/procurement specifications)
Memory64 GB HBM2e
Memory Bandwidth1.2 TB/s (HBM2e, confirmed by official/procurement specifications)
TDP350 W
InterfacePCIe 4.0 x16
Release2023
Mass Production2024
Form FactorAir-cooled PCIe accelerator card / OAM module

Data notes:

  • ✅ FP32, memory, TDP, FP16, INT8, and memory bandwidth have all been confirmed by official/procurement specifications, including Nankai University's 2025 accelerator card procurement announcement
  • Memory bandwidth 1.2 TB/s (HBM2e), FP16 192 TFLOPS, INT8 384 TOPS

Product Highlights​

1. Fully Self-Developed Architecture​

  • ivcore11 architecture: Iluvatar CoreX's second-generation general-purpose GPU architecture, with a complete instruction set system
  • CUDA-ecosystem compatible: supports CUDA C++ programming, low migration cost
  • Multi-precision support: FP32, FP16, INT8, FP8 (requires the ixTE library)

2. Large Memory Capacity​

  • 64GB HBM2e: supports large-scale model training
  • High bandwidth: 1.2 TB/s memory bandwidth (HBM2e, officially confirmed)

3. IXUCA Software Stack​

  • Mainstream framework compatibility: TensorFlow, PyTorch, PaddlePaddle
  • Complete toolchain: compiler, math libraries, communication libraries, management tools
  • Seamless migration: highly compatible with the CUDA ecosystem, migration time reduced by over 50%

Software Stack: IXUCA​

IXUCA (Iluvatar Unified Computing Architecture) is the unified computing architecture software stack self-developed by Iluvatar CoreX.

ComponentNameFunctionCounterpart
Deep learning frameworksPyTorch-Cambricon, TensorFlow-CambriconAdapted deep learning frameworksPyTorch, TensorFlow
Inference frameworkIGIEHigh-performance inference frameworkTensorRT
Inference engineIxRTDedicated inference acceleration engineTensorRT
LLM inference frameworkIxFormerLarge-model inference and training optimizationvLLM
CompilerIXUCA CompilerCompilernvcc
Math librariesixDNN, ixBLASFundamental deep learning operatorscuDNN, cuBLAS
Communication libraryixCCLMulti-card communication libraryNCCL
Management toolixsmiGPU management toolnvidia-smi

Use Cases​

  • ✅ AI model training (CNN, RNN, Transformer, etc.)
  • ✅ High-performance computing (HPC)
  • ✅ Large-model pretraining (requires multi-card parallelism)
  • ✅ Domestic substitution projects (government, state-owned enterprises, defense industry)
  • ❌ Top-tier frontier model training (compute limitations)
  • ❌ International markets (constrained by US export controls)

Performance Comparison​

MetricBI-V150A100 80GBGap
FP3248 TFLOPS19.5 TFLOPS+146%
Memory64 GB80 GB-20%
TDP350W400W-12.5%

Note: the BI-V150's FP32 compute is higher than the A100's (possibly due to different precision definitions or test conditions), but actual training performance also depends on the degree of software stack optimization.


Vendor Information​

ItemDetails
CompanyShanghai Iluvatar CoreX Semiconductor Co., Ltd.
English NameIluvatar CoreX
Founded2015
FounderDiao Shijing
HeadquartersShanghai
PositioningChinese general-purpose GPU chip design company
Official Websitehttps://www.iluvatar.com
Software Stackhttps://support.iluvatar.com

Market Progress: ByteDance Shipments Double (September 2026 Update)​

According to a Reuters report on 2026-09-10, Iluvatar CoreX's GPU shipments to ByteDance this year have nearly doubled to about 100,000 units, and the company has reallocated GPUs originally reserved for internal use to prioritize this order.

In ByteDance's domestic chip supplier hierarchy, Huawei remains the largest supplier, followed by Cambricon, with Iluvatar CoreX third.

During the same period, Iluvatar CoreX, like other domestic vendors, raised its prices (+20~30%), mainly due to rising HBM costs and tight domestic HBM capacity.

Key insight: a shipment scale of 100,000 units shows Iluvatar CoreX has moved from "validation purchases" into the "volume supply" stage. However, note that these orders are mainly for inference; the company's software ecosystem maturity on the training side still lags the leading vendors.


Information To Be Added​

  • Official FP16/INT8 compute figures (192 TFLOPS / 384 TOPS, confirmed by procurement specifications)
  • Official memory bandwidth figure (1.2 TB/s, HBM2e)
  • Actual training performance tests (ResNet, BERT, LLM, etc.)
  • Multi-card scaling performance (ixCCL)
  • Energy efficiency tests

Data sources:

  • SMZDM unboxing review (2026-01-08)
  • Gitee AI product documentation
  • Iluvatar CoreX official materials

Last updated: 2026-06-28