TCO Analysis →
Why MDCX · Our Approach为何 MDCX · 我们的方法

Idle GPUs are sunk capital.
We optimize for time-to-revenue.
闲置 GPU 是沉没资本。
我们优化投产时间

You've seen the economics. This is the architecture behind them — and why every choice here exists to protect that return.你已经看过经济性测算。这里是支撑它的架构——以及每一个设计取舍为何都在守护那份回报。

Computing Unit计算单元 + BESS + Hybrid Chiller混合冷机
01
The Solution整体方案

Every AI factory needs three things. We made each one a block.AI 工厂需要三样东西。我们把每一样都做成了模块。

Compute, power, cooling — each delivered as a factory-integrated module. On site you only connect power, cooling, and network. No on-site system commissioning.算力、电力、冷却——每一项都以工厂集成模块交付。现场只需接入电、冷、网,无需现场系统调试。

Compute算力
Compute module

Computing Unit计算单元

The compute engine — pick a liquid-cooled path by workload, or run both on one shared 28–35 °C loop.算力引擎——按负载选择液冷路线,或让两者共用一条 28–35 °C 回路。

  • Liquid Cooling — cold-plate training & high-performance inference (L1240C45)液冷 — 冷板式训练与高性能推理(L1240C45)
  • Immersion Cooling — PCIe inference, simple mechanics, low OpEx (I400C45 / I400C40)浸没液冷 — PCIe 推理,机械结构简单,OpEx 低(I400C45 / I400C40)
+
Power电力
Power and BESS

BESS

A grid-tied power asset — not just backup. 120 min ride-through standard, plus peak-shaving and demand response.并网电力资产——不只是备电。标准 120 分钟穿越能力,兼顾削峰填谷与需求响应。

  • Power quality: voltage / harmonics / reactive电能质量:电压 / 谐波 / 无功
  • Curtailment-ready reverse feed支持弃电消纳的反向馈电
  • DC-Attach ready预留 DC-Attach for a future direct-DC bus预留 DC-Attach,面向未来直流母线
+
Cooling冷却
Cooling systems

Hybrid Chiller混合冷机

Free cooling whenever the climate allows, mechanical trim only when it doesn't — across all climates.气候允许时全程自然冷却,不允许时才启动机械补冷——适配各类气候。

  • Immersion PUE ≈ 1.05 · Liquid Cooling PUE ≈ 1.15–1.20浸没液冷 PUE ≈ 1.05 · 液冷 PUE ≈ 1.15–1.20
×
Software软件
CIOS · the operating layerCIOS · 运营层
= AI Factory, revenue-ready · standard power / cooling / network interfaces= 可创收的 AI 工厂 · 标准电力 / 冷却 / 网络接口
02
The Shift范式转变

Compute is now gated by power — not construction算力的瓶颈已是电力,而非土建

AI racks draw 80–150 kW and climb; GPU generations turn over every 1–2 years. The 12–24-month build model can't keep up. The real constraint is no longer how fast you can build — it's how fast you can energize, and how well you work with the grid.AI 机柜功率已达 80–150 kW 并持续上升;GPU 每 1–2 年换代。12–24 个月的土建模式跟不上节奏。真正的约束不再是建得多快,而是通电多快、以及与电网协同得多好。

Traditional Hyperscale传统超大规模模式
Delivery model交付模式Custom building project定制土建项目
Time to online上线周期12–24 months12–24 个月
Rack density机柜密度5–15 kW
Scaling扩容方式Pre-build everything一次性全部预建
Redundancy冗余Full-system 2N (Tier IV)全系统 2N(Tier IV)
Backup power备用电源Diesel generators柴油发电机
Grid role电网角色Passive load被动负荷
MDCX — AI FactoryMDCX — AI 工厂
Delivery model交付模式Manufactured product制造化产品
Time to online上线周期~4–6 mo → 75–90 d (repeat)约 4–6 个月 → 复制项目 75–90 天
Rack density机柜密度80–150 kW+
Scaling扩容方式Linear, on-demand modules按需线性增加模块
Redundancy冗余Differentiated, workload-matched分级冗余,匹配负载
Backup power备用电源BESS-firstBESS 优先
Grid role电网角色Flexible, grid-cooperative asset柔性、与电网协同的资产
03
Why · Power为何 · 电力

Turn power from your biggest cost into your biggest edge把电力从最大成本变成最大优势

Where you can build is decided by power. MDCX is engineered to be a grid-cooperative load — so it clears interconnection faster, and earns on power others can't use.能在哪里建,由电力决定。MDCX 被设计成与电网协同的负荷——因此并网审批更快,并能靠别人用不上的电赚钱。

Europe · Interconnection欧洲 · 并网
A flexible, zero-emission load clears grid approval faster — and can unlock more capacity.柔性、零排放的负荷通过电网审批更快——还能释放更多容量。
US · Curtailment美国 · 弃电
Run as an interruptible load — turn curtailed, near-free power into your cheapest compute.作为可中断负荷运行——把被弃掉的近乎免费的电,变成你最便宜的算力。

Powered by BESS, not diesel —以 BESS 供电,而非柴发 —

Lower OpEx更低 OpEx Simpler to ship运输更简单 DC-Attach ready预留 DC-Attach Cleaner approvals审批更顺畅

Myth: diesel isn't faster — UPS bridges the gap in both cases. So BESS wins on cost, logistics, and approvals.常见误解:柴发并不更快——两种方案都由 UPS 承担切换间隙。因此 BESS 在成本、物流与审批上均占优。

So: BESS by default. Diesel must justify itself — not the other way around.所以:默认选 BESS。需要论证的是柴发,而不是反过来。
04
Why · Reliability为何 · 可靠性

Tier III is the wrong yardstick for an AI factoryTier III 不是衡量 AI 工厂的正确标尺

An AI factory runs ~10× the scale of an enterprise data center. Full 2N / Tier III doubles your infrastructure CapEx on a massive base — straight against the IRR you just modeled.AI 工厂的规模约为企业数据中心的 10 倍。在如此大的基数上做全 2N / Tier III,会让基础设施 CapEx 翻倍——直接冲击你刚测算出的 IRR。

And AI training isn't a financial-transaction workload. It checkpoints, restarts, and fails over at the node level — it carries no per-second SLA penalty. Buying enterprise-grade redundancy for it is capital spent to protect against a risk that doesn't cost you.而且 AI 训练不是金融交易类负载。它有检查点、可重启、在节点级做故障切换——不存在按秒计的 SLA 罚则。为它购买企业级冗余,是花钱防一个本来不会让你损失的风险。

Redundancy exists to enable maintenance — not to mask low reliability.冗余的意义是支撑在线维护——不是用来掩盖低可靠性。

So redundancy is matched to each component's replaceability and blast radius:因此,冗余按各部件的可更换性与故障影响范围来匹配:

System系统Strategy策略Logic逻辑
Busbar / copper rails母线 / 铜排Single-path, high-integrity单路径、高完整性Not field-replaceable — redundant paths only add connection points and failure modes.无法现场更换——冗余路径只会增加连接点与故障模式。
Cooling pumps / CDU冷却泵 / CDU2NSupports online maintenance — swap without stopping compute.支持在线维护——更换无需停止算力。
Fans / compressors风机 / 压缩机N+1Degrade-and-continue on failure, not an immediate stop.故障时降级运行而非立即停机。
Compute nodes计算节点Node-level failover节点级故障切换Local fault containment — failures don't propagate across the cluster.故障就地隔离——不会在集群内扩散。
Control plane控制平面Redundant冗余配置No single point of control can take down the whole system.不存在能拖垮整个系统的单点控制。
So: skip the Tier badge. Match reliability to the workload — and put the saved capital into compute that earns.所以:不必追逐 Tier 认证徽章。让可靠性匹配负载——把省下的资本投入能创收的算力。
05
Why · Speed为何 · 速度

Modular & fast — monetization is the moat模块化且快——变现速度就是护城河

MDCX is a manufactured product, not a construction project: fully integrated and tested at the factory, online on arrival. The first cluster ships in ~4–6 months — and because the design is proven and the supply chain is warm, every batch after lands in 75–90 days.MDCX 是制造出来的产品,不是工程项目:在工厂完成全部集成与测试,到场即可上线。首个集群约 4–6 个月交付——由于设计已验证、供应链处于热状态,后续每批次 75–90 天到位。

First batch · ~4–6 months首批 · 约 4–6 个月
90–120
days
Factory build & integration工厂制造与集成
Complete system integration and testing — no on-site commissioning of core systems.完成整机集成与测试——核心系统无需现场调试。
30–45
days
Transport运输
Standardized form factors and logistics routing.标准化外形尺寸与物流路径。
15–30
days
On-site deploy现场部署
Connect power, cooling, network — once all parties' site conditions (power / cooling / network) are ready.接入电、冷、网——前提是各方现场条件(电力 / 冷却 / 网络)已就绪。
2nd batch & beyond第二批及以后
Design proven, supply chain warm — no re-engineering, just replication.设计已验证、供应链处于热状态——无需重新设计,只需复制。
75–90days
Linear scaling — 1 → 4 → 16 → 64 modules — adds capacity without touching what's already running.线性扩容——1 → 4 → 16 → 64 个模块——新增容量不影响已在运行的部分。
Every month saved is a month of GPU revenue you never get back. Speed isn't a feature — it's the return.省下的每一个月,都是再也拿不回来的一个月 GPU 收入。速度不是功能——它就是回报。
06
Why · Operations为何 · 运营

Hardware without an operating system is just metal.没有操作系统的硬件只是一堆金属。

CIOS ships with every block: see every sensor, turn every alarm into a ticket, meter every kilowatt-hour and GPU-hour.CIOS 随每套模块交付:看见每一个传感器,把每一条告警变成工单,计量每一度电与每一个 GPU 小时。

Visibility

See it看得见

Path-addressed telemetry:路径寻址遥测: sgp01.pod002.cdu000.fws.supply.flow

Control

Control it控得住

Alarms → tickets → SLA, policy-gated setpoints.告警 → 工单 → SLA,设定点受策略约束。

Monetize

Monetize it变得现

Usage metering, capacity headroom, ops reports.用量计量、容量余量、运维报表。

How CIOS works →CIOS 如何运作 →

07
Built to last为长期运行而造

First-class components, fully certified一流部件,完整认证

Fast deployment and lean redundancy don't mean cutting corners. Every block is built from tier-one industrial hardware and certified to recognized standards.快速部署与精简冗余不等于偷工减料。每套模块都采用一线工业级硬件,并按公认标准取得认证。

Quality

First-class only只用一流部件

No compromise components. Every block is assembled from proven, tier-one industrial hardware.部件不做妥协。每套模块均由经过验证的一线工业级硬件装配而成。

Compliance

Certified components部件认证

All components meet UL / CE / CSA and the applicable regional standards.所有部件符合 UL / CE / CSA 及适用的地区标准。

ULCECSA
Full-system certification

Full-system certification整机系统认证

End-to-end system-level certification targeted for completion in 2027.端到端系统级认证目标于 2027 年完成。

Our System Suppliers我们的系统供应商
EATON VERTIV SIEMENS ABB SCHNEIDER SHELL CASTROL
08
Choose your path选择你的路线

You've run the numbers. Now pick your GPU.数字已经算过了。现在选你的 GPU。

Two container profiles — chosen by the GPU you want to run and the business you're building. Frontier training on Blackwell, or a lean inference business at scale.两种集装箱配置——取决于你要跑的 GPU 与你要做的生意。基于 Blackwell 的前沿训练,或规模化的精简推理业务。

Running any of the above →运行上述任一方案 → CIOS · included →CIOS · 已包含 →

Numbers don't match your plan?
Adjust your assumptions and re-run the TCO before you commit.
数字和你的计划对不上?
在做决定前,调整假设并重新运行 TCO。