智算中心论文专站

AIDC Research Papers

Liquid Cooling AI Data Center Power & Thermal Systems
Current Issue

Volume 2026 · Issue 09-06

按期刊卷期页方式整理本期论文。每条仅使用日报已列出的可追溯公开来源,不新增未经核验事实。

Research Article算电协同

Hosting Capacity Assessment of Data Centers with Voltage Ride-Through Capability in Power Systems

Pengyu Ren、Wei Sun、Fei Teng

Published 2026-09-03 · arXiv · Credibility S

Large data centers are emerging as concentrated, power-electronic grid loads whose abrupt disconnection or transfer to on-site backup supply during voltage disturbances can remove large demand from the power system, and may create a system-level stability problem. Their interconnection feasibility therefore depends not only on steady-state thermal and voltage limits, but also on whether internal power-conditioning s…

Abstract, interpretation and reference

Abstract

Large data centers are emerging as concentrated, power-electronic grid loads whose abrupt disconnection or transfer to on-site backup supply during voltage disturbances can remove large demand from the power system, and may create a system-level stability problem. Their interconnection feasibility therefore depends not only on steady-state thermal and voltage limits, but also on whether internal power-conditioning systems can maintain IT service while limiting customer-initiated load reduction. This paper presents a voltage ride-through (VRT)-aware data center and grid co-planning framework that couples transmission-level fault simulation with an internal data center ride-through model. Python-based dynamic simulations generate point-of-interconnection (POI) voltage trajectories under selected network faults, and the resulting waveforms drive an internal model incorporating IT and cooling-load dynamics, DC-link, Uninterruptible Power Supply (UPS) response, and converter apparent power limits. The IEEE 118-bus case study shows that internal VRT capability can become a binding interconnection constraint: steady-state planning alone can overestimate feasible data center capacity, whereas increased UPS converter headroom progressively restores hosting capacity. Under the reduced-order response models studied, the grid-forming mode provides greater ride-through margin than the current-limited grid-following mode under the same network fault conditions. The results further show that VRT constraints can materially change both the total hosting capacity of data centers and its spatial allocation across candidate interconnection buses.

中文解读

背景:AI 数据中心负载、功率密度和能源约束同步上升,算力负载与电网侧资源的协同调度正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用框架构建和频域/系统级分析,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向AI 负载波动对电网设备寿命和调频边界的影响。意义:对日报读者而言,它可用于判断智算中心建设是否受电网容量、负载波动和调度机制约束。仍需结合全文实验条件、样本范围和成本假设核验。

参考文献

Pengyu Ren, Wei Sun, Fei Teng. Hosting Capacity Assessment of Data Centers with Voltage Ride-Through Capability in Power Systems[J/OL]. (2026-09-03)[2026-09-06]. http://arxiv.org/abs/2609.03030v1.

Full text 中文海报
算电协同 论文图示
Research Article余热回收

Real-Time Control of Sustainable Data Centers: A Two-Layer Model Predictive Control Framework with Workload Flexibility and Heat Recovery

Wenyu Liu、Enea Figini、Mario Paolone

Published 2026-08-17 · arXiv · Credibility S

This paper proposes a two-layer model predictive control (MPC) framework for the real-time operation of data centers integrated with on-site photovoltaic generation, battery energy storage, waste heat recovery, and district heating. The upper layer employs scenario-based stochastic optimization to jointly optimize intraday market participation, workload scheduling, and energy management under uncertainty. The lower …

Abstract, interpretation and reference

Abstract

This paper proposes a two-layer model predictive control (MPC) framework for the real-time operation of data centers integrated with on-site photovoltaic generation, battery energy storage, waste heat recovery, and district heating. The upper layer employs scenario-based stochastic optimization to jointly optimize intraday market participation, workload scheduling, and energy management under uncertainty. The lower layer adopts an adaptive tube-based MPC strategy that compensates short-term disturbances while tracking the dispatch references given by the upper layer. The framework further integrates multi-horizon forecasting to support real-time decision making. Microservice-based simulation studies under representative clear-sky and overcast operating conditions demonstrate that the proposed framework accurately tracks dispatch plans despite fast photovoltaic and workload fluctuations. Compared with single-layer control strategies, the adaptive lower-layer controller substantially reduces real-time dispatch deviations and the associated imbalance costs. In addition, the proposed framework naturally adapts to seasonal operating conditions and responds to carbon-aware operating signals, offering a practical approach for economically efficient, sustainable, and grid-supportive operation of future data centers.

中文解读

背景:AI 数据中心负载、功率密度和能源约束同步上升,余热回收、热泵耦合和二次能源利用正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用建模优化、调度分析或算法评估,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向AI 负载波动对电网设备寿命和调频边界的影响。意义:对日报读者而言,它可用于判断数据中心余热能否从成本项转化为能源资产。仍需结合全文实验条件、样本范围和成本假设核验。

参考文献

Wenyu Liu, Enea Figini, Mario Paolone. Real-Time Control of Sustainable Data Centers: A Two-Layer Model Predictive Control Framework with Workload Flexibility and Heat Recovery[J/OL]. (2026-08-17)[2026-09-06]. http://arxiv.org/abs/2608.16432v1.

Full text 中文海报
余热回收 论文图示
Research ArticleAI 运维优化

Quantifying AI data center nitrogen oxide (NO$_x$) emissions from space

Kevin D. Gauld、Daniel J. Varon、Nicholas Balasus、Daniel H. Cusworth

Published 2026-08-23 · arXiv · Credibility S

AI data center power demand is spurring rapid deployment of on- and near-site natural gas turbines. Nitrogen oxide (NO$_x$) pollution from this equipment is a growing concern but has not previously been quantified with atmospheric observations. Here we demonstrate space-based detection and quantification of NO$_x$ emissions from the SpaceXAI Colossus 2 power plant in Southaven, Mississippi. Using observations from t…

Abstract, interpretation and reference

Abstract

AI data center power demand is spurring rapid deployment of on- and near-site natural gas turbines. Nitrogen oxide (NO$_x$) pollution from this equipment is a growing concern but has not previously been quantified with atmospheric observations. Here we demonstrate space-based detection and quantification of NO$_x$ emissions from the SpaceXAI Colossus 2 power plant in Southaven, Mississippi. Using observations from the geostationary TEMPO satellite instrument, we detect a strong increase in local mean NO$_2$ column concentrations after the plant began operations in late 2025. We then use TEMPO to estimate two-week-average NO$_x$ source rates from August 2025 to mid-August 2026, calibrating against continuous emission monitoring system (CEMS) data from US power plants. TEMPO first detected NO$_x$ emissions in December 2025 at 460$\pm$180 kg h$^{-1}$. We find that emissions increased through August 2026, averaging 730$\pm$185 kg h$^{-1}$ after February 2026, roughly 16 times higher than expected from the facility's March 2026 permit for 41 turbines operating under best available control technology (BACT) requirements ($\sim$47 kg h$^{-1}$). Emissions at the expected level would be undetectable by our TEMPO analysis.

中文解读

背景:AI 数据中心负载、功率密度和能源约束同步上升,AI 运维、负载预测和设施调优正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用仿真建模和情景分析,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向AI 负载波动对电网设备寿命和调频边界的影响。意义:对日报读者而言,它可用于判断AI 工具是否能降低运维复杂度并提升可用性。仍需结合全文实验条件、样本范围和成本假设核验。

参考文献

Kevin D. Gauld, Daniel J. Varon, Nicholas Balasus, 等. Quantifying AI data center nitrogen oxide (NO$_x$) emissions from space[J/OL]. (2026-08-23)[2026-09-06]. http://arxiv.org/abs/2608.22153v1.

Full text 论文栏目封面
智算中心论文栏目通用封面
Research Article芯片与算力

Towards Terabit/$λ$/s Multidimensional Silicon Photonic Engine

Hao Chen、Zengqi Chen、Wu Zhou、Kaihang Lu、Mingyuan Zhang、Yuxiang Yin、Yiou Cui、Chaoran Huang

Published 2026-08-12 · arXiv · Credibility S

Increasing artificial intelligence (AI) workloads drive co-packaged optics (CPO), which integrates optical engines with electronic components. Optical interconnects can extend transmission distances and reduce latency, allowing distributed clusters in AI factories to operate as a unified computational unit. However, escalating data throughput necessitates greater parallelization of light within ultracompact form fac…

Abstract, interpretation and reference

Abstract

Increasing artificial intelligence (AI) workloads drive co-packaged optics (CPO), which integrates optical engines with electronic components. Optical interconnects can extend transmission distances and reduce latency, allowing distributed clusters in AI factories to operate as a unified computational unit. However, escalating data throughput necessitates greater parallelization of light within ultracompact form factors while maintaining stringent energy efficiency and latency constraints. Here, we present a multidimensional silicon photonic engine that achieves a communication capacity exceeding 1.8 terabit/s/lambda/s. By monolithically integrating transceivers, spatial and polarization (de)multiplexers, and optical signal processors on a single chip, we eliminate bulky discrete (de)multiplexers and power-hungry digital signal processing (DSP). In experiments, the photonic engine can be self-configured to identify two, four, or six concurrent spatial and polarization channels per fiber while mitigating dynamic channel crosstalk. Compared with the state-of-art DSP, our approach achieves >5,000-fold reductions in both power consumption and processing latency at a MIMO processing order of six. Furthermore, we demonstrate full-duplex, modulation-format-transparent inter-chip communication over 300-meter fiber. These results represent a paradigm shift for optical engines in future high-performance computing and AI-driven data centers.

中文解读

背景:AI 数据中心负载、功率密度和能源约束同步上升,芯片、服务器和高密度算力部署正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用框架构建和频域/系统级分析,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向能效评价口径、运营指标和优化目标的系统化梳理。意义:对日报读者而言,它可用于判断芯片路线和服务器密度变化如何传导到机房设计。仍需结合全文实验条件、样本范围和成本假设核验。

参考文献

Hao Chen, Zengqi Chen, Wu Zhou, 等. Towards Terabit/$λ$/s Multidimensional Silicon Photonic Engine[J/OL]. (2026-08-12)[2026-09-06]. http://arxiv.org/abs/2608.11639v1.

Full text 中文海报
芯片与算力 论文图示
Research Article算电协同

Exploiting the Benefits of V2B Application on Peak Shaving of Data Center Loads

Arya Joshi、Hamed Haggi、Chinmay Morankar

Published 2026-09-01 · arXiv · Credibility S

The accelerated growth in data center projects has introduced a demand-driven bottleneck throughout power grids and contributed to a substantial increase in carbon emissions. These concerns are fueling discussions on methods to use existing energy assets to drive operational efficiency. To this end, this paper explores the benefits of Vehicle-to-Building (V2B) applications to support peak shaving of data center cool…

Abstract, interpretation and reference

Abstract

The accelerated growth in data center projects has introduced a demand-driven bottleneck throughout power grids and contributed to a substantial increase in carbon emissions. These concerns are fueling discussions on methods to use existing energy assets to drive operational efficiency. To this end, this paper explores the benefits of Vehicle-to-Building (V2B) applications to support peak shaving of data center cooling loads. Initially, a literature review was conducted considering V2B constraints and optimization methods including SoC limitations, EV participation, tariffs, and building loads. This analysis was then used to develop a conceptual case study of a 10 MW data center in Loudoun County, VA by simulating a temperature-dependent load profile and adjusting the V2B participation of 40 commercial and passenger EVs. Simulation results indicate that, depending on seasonal variations in cooling load demands, strategic deployment of V2B assets between 12-5pm can offset gross cooling loads by 13-36%.

中文解读

背景:AI 数据中心负载、功率密度和能源约束同步上升,算力负载与电网侧资源的协同调度正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用综述归纳和指标比较,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向AI 负载波动对电网设备寿命和调频边界的影响。意义:对日报读者而言,它可用于判断智算中心建设是否受电网容量、负载波动和调度机制约束。仍需结合全文实验条件、样本范围和成本假设核验。

参考文献

Arya Joshi, Hamed Haggi, Chinmay Morankar. Exploiting the Benefits of V2B Application on Peak Shaving of Data Center Loads[J/OL]. (2026-09-01)[2026-09-06]. http://arxiv.org/abs/2609.00204v1.

Full text 中文海报
算电协同 论文图示
Research Article算电协同

Flexible Training Workloads in Large-Scale AI Data Centers for Transient-Stability Support in Transmission-Constrained Power Systems

Jae-Kyeong Kim

Published 2026-08-31 · arXiv · Credibility S

The rapid expansion of large-scale artificial intelligence (AI) data centers is adding substantial, concentrated, and rapidly varying loads to transmission-constrained power systems. Although such load variations are generally regarded as operational challenges, this paper presents an alternative perspective in which the upward load flexibility of AI data centers could be coordinated for transient-stability support.…

Abstract, interpretation and reference

Abstract

The rapid expansion of large-scale artificial intelligence (AI) data centers is adding substantial, concentrated, and rapidly varying loads to transmission-constrained power systems. Although such load variations are generally regarded as operational challenges, this paper presents an alternative perspective in which the upward load flexibility of AI data centers could be coordinated for transient-stability support. To this end, this paper proposes training-induced load surge (TILS), a fast demand-side strategy that initiates or resumes flexible AI training workloads after fault clearing to increase active-power demand at electrically effective locations. The resulting load increase allows accelerating generators to supply additional electrical power, thereby reducing the accelerating-power imbalance and limiting the first-swing rotor-angle excursion. The underlying mechanism is first clarified in a single-machine infinite-bus (SMIB) system and then evaluated in the IEEE 39-bus system and a large-scale Korean power system. Results across all three systems demonstrate that TILS can increase the transient-stability-constrained generation limit. Larger responses, earlier activation, and siting at buses with a stronger electrical influence on the critical generators provide greater generation-limit increases. These results suggest that the upward load-response capability of AI data centers can provide complementary transient-stability support when sufficient electrical headroom, flexible workloads, and reliable grid-triggered activation are available.

中文解读

背景:AI 数据中心负载、功率密度和能源约束同步上升,算力负载与电网侧资源的协同调度正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用文献摘要中的模型、实验或案例分析,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向AI 负载波动对电网设备寿命和调频边界的影响。意义:对日报读者而言,它可用于判断智算中心建设是否受电网容量、负载波动和调度机制约束。仍需结合全文实验条件、样本范围和成本假设核验。

参考文献

Jae-Kyeong Kim. Flexible Training Workloads in Large-Scale AI Data Centers for Transient-Stability Support in Transmission-Constrained Power Systems[J/OL]. (2026-08-31)[2026-09-06]. http://arxiv.org/abs/2608.30901v1.

Full text 中文海报
算电协同 论文图示
Research Article算电协同

Techno-Economic Boundary Analysis of Small Modular Reactor Cogeneration for Hyperscale Data Center IT and Cooling Loads

Honglin Li、Buxin She、Jie Zhang

Published 2026-08-11 · arXiv · Credibility S

Hyperscale data centers are adding firm, high-utilization demand faster than grids can serve it, renewing interest in colocating them with small modular reactors. Such a plant could earn revenue in two ways, selling low-carbon power and diverting steam to absorption chillers that serve a cooling load accounting for 20-40% of facility electricity use, but neither revenue stream has been priced across the conditions t…

Abstract, interpretation and reference

Abstract

Hyperscale data centers are adding firm, high-utilization demand faster than grids can serve it, renewing interest in colocating them with small modular reactors. Such a plant could earn revenue in two ways, selling low-carbon power and diverting steam to absorption chillers that serve a cooling load accounting for 20-40% of facility electricity use, but neither revenue stream has been priced across the conditions that must coincide. Here we co-optimize reactor dispatch, steam extraction, absorption cooling and grid exchange hourly for a 200 MW$_\mathrm{e}$ data center in the Electric Reliability Council of Texas (ERCOT) region, across 109 runs spanning capital, market, policy, financing and cooling efficiency. At 2023 mid-range reactor capital, the nuclear configurations cost 49-62% more than grid supply even with the Section 45Y production tax credit. The viable region opens near \$5,000 kW$_\mathrm{e}^{-1}$, and nth-of-a-kind capital makes them 77-89% cheaper in 2023, though between parity and 34% more expensive in the low-price 2024 market. A carbon price of \$53-64 tCO$_2^{-1}$ closes the mid-range gap under hourly export crediting. Absorption cooling is dispatched in response to hourly electricity prices and supplies 38% of annual cooling, at an added cost of \$9.2 million yr$^{-1}$ relative to the reactor-only plant; that gap closes at an installed absorption cost of \$60 kW$_\mathrm{c}^{-1}$ at baseline efficiency and \$570 kW$_\mathrm{c}^{-1}$ on a legacy-efficiency campus, against surveyed commercial prices of \$450-1,200 kW$_\mathrm{c}^{-1}$. Together these results delineate the capital, market and policy conditions under which colocated reactor cogeneration is competitive with grid procurement, and the range over which each condition moves the outcome.

中文解读

背景:AI 数据中心负载、功率密度和能源约束同步上升,算力负载与电网侧资源的协同调度正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用综述归纳和指标比较,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向AI 负载波动对电网设备寿命和调频边界的影响。意义:对日报读者而言,它可用于判断智算中心建设是否受电网容量、负载波动和调度机制约束。仍需结合全文实验条件、样本范围和成本假设核验。

参考文献

Honglin Li, Buxin She, Jie Zhang. Techno-Economic Boundary Analysis of Small Modular Reactor Cogeneration for Hyperscale Data Center IT and Cooling Loads[J/OL]. (2026-08-11)[2026-09-06]. http://arxiv.org/abs/2608.10999v1.

Full text 中文海报
算电协同 论文图示
Research Article算电协同

Beyond the Grid: Cost, Carbon, and Capital Requirements of On-Site Power Technologies for AI Data Centers

Eliseo Curcio

Published 2026-08-08 · arXiv · Credibility S

Interconnection queues, not electricity prices, now govern where data centers can be built, and the standard levelized-cost comparison answers a question no developer faces: it assumes a load profile, freezes the grid price while modeling the demand that moves it, and quotes busbar costs a facility cannot buy. This paper evaluates nine on-site supply technologies against a delivered grid whose price is endogenous to…

Abstract, interpretation and reference

Abstract

Interconnection queues, not electricity prices, now govern where data centers can be built, and the standard levelized-cost comparison answers a question no developer faces: it assumes a load profile, freezes the grid price while modeling the demand that moves it, and quotes busbar costs a facility cannot buy. This paper evaluates nine on-site supply technologies against a delivered grid whose price is endogenous to projected data-center demand, on a complete-site basis that retains standby charges, with measured GPU training load, delivered fuel prices, production-pathway carbon, and statutory 45V and 48E incentive mechanics. Nothing beats the wire: gas combined cycle produces at 47 USD/MWh but costs about 114 USD per megawatt-hour of complete site energy against a 92 USD grid; four-hour storage is physically capped near 18 percent of annual energy and, charged at the margin, dirtier than the grid; hydrogen from grid-priced power fails on cost and carbon together. An investment inversion converts these findings into capital terms: conversion-hardware learning buys nothing, because free hardware still exceeds the grid for every low-carbon arm, while global electrolyser deployment on sited sub-20 USD/MWh power brings PEM hydrogen power to about 2.2 times the grid at 300 billion USD and 1.9 times at 1 trillion USD (2.7 and 2.3 for the hydrogen engine), with a carbon reduction of roughly 85 percent (6.8-fold) against grid-power production. Grid parity is not purchasable at any budget. On-site supply is an access and depth product; most current investment targets the wrong term.

中文解读

背景:AI 数据中心负载、功率密度和能源约束同步上升,算力负载与电网侧资源的协同调度正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用仿真建模和情景分析,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向AI 负载波动对电网设备寿命和调频边界的影响。意义:对日报读者而言,它可用于判断智算中心建设是否受电网容量、负载波动和调度机制约束。仍需结合全文实验条件、样本范围和成本假设核验。

参考文献

Eliseo Curcio. Beyond the Grid: Cost, Carbon, and Capital Requirements of On-Site Power Technologies for AI Data Centers[J/OL]. (2026-08-08)[2026-09-06]. http://arxiv.org/abs/2608.08170v1.

Full text 中文海报
算电协同 论文图示