Research Article算电协同
Jae-Kyeong Kim
Published 2026-08-31 · arXiv · Credibility S
The rapid expansion of large-scale artificial intelligence (AI) data centers is adding substantial, concentrated, and rapidly varying loads to transmission-constrained power systems. Although such load variations are generally regarded as operational challenges, this paper presents an alternative perspective in which the upward load flexibility of AI data centers could be coordinated for transient-stability support.…
Abstract, interpretation and reference
Abstract
The rapid expansion of large-scale artificial intelligence (AI) data centers is adding substantial, concentrated, and rapidly varying loads to transmission-constrained power systems. Although such load variations are generally regarded as operational challenges, this paper presents an alternative perspective in which the upward load flexibility of AI data centers could be coordinated for transient-stability support. To this end, this paper proposes training-induced load surge (TILS), a fast demand-side strategy that initiates or resumes flexible AI training workloads after fault clearing to increase active-power demand at electrically effective locations. The resulting load increase allows accelerating generators to supply additional electrical power, thereby reducing the accelerating-power imbalance and limiting the first-swing rotor-angle excursion. The underlying mechanism is first clarified in a single-machine infinite-bus (SMIB) system and then evaluated in the IEEE 39-bus system and a large-scale Korean power system. Results across all three systems demonstrate that TILS can increase the transient-stability-constrained generation limit. Larger responses, earlier activation, and siting at buses with a stronger electrical influence on the critical generators provide greater generation-limit increases. These results suggest that the upward load-response capability of AI data centers can provide complementary transient-stability support when sufficient electrical headroom, flexible workloads, and reliable grid-triggered activation are available.
中文解读
背景:AI 数据中心负载、功率密度和能源约束同步上升,算力负载与电网侧资源的协同调度正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用文献摘要中的模型、实验或案例分析,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向AI 负载波动对电网设备寿命和调频边界的影响。意义:对日报读者而言,它可用于判断智算中心建设是否受电网容量、负载波动和调度机制约束。仍需结合全文实验条件、样本范围和成本假设核验。
参考文献
Jae-Kyeong Kim. Flexible Training Workloads in Large-Scale AI Data Centers for Transient-Stability Support in Transmission-Constrained Power Systems[J/OL]. (2026-08-31)[2026-09-01]. http://arxiv.org/abs/2608.30901v1.
Research ArticleAI 运维优化
Slava G. Turyshev
Published 2026-08-27 · arXiv · Credibility S
Megawatt-class orbital data centers require continuous maintenance, replacement, inventory, and service capacity in addition to spacecraft power/thermal systems. We formulate an analytical lifecycle framework for permanent/transient failures, modular orbital replacement units, robotic servicing, spare inventory, scheduled technology refresh, correlated faults, cybersecurity, optional human support. The model combine…
Abstract, interpretation and reference
Abstract
Megawatt-class orbital data centers require continuous maintenance, replacement, inventory, and service capacity in addition to spacecraft power/thermal systems. We formulate an analytical lifecycle framework for permanent/transient failures, modular orbital replacement units, robotic servicing, spare inventory, scheduled technology refresh, correlated faults, cybersecurity, optional human support. The model combines nonhomogeneous component hazards, capacity-weighted availability, multiclass robotic-service capacity, Poisson base-stock inventory, replacement-flow accounting, human-support break-even relations. For a 1 MW cluster with 10 active 100 kW nodes, 1 reserve node, ~200 5 kW compute cartridges, low, nominal, high deployed-mass allocations span ~50-75 kg/kW. Assumptions yield 70.2 random or life-limited interventions and 323-349 planned refresh operations/(MW-year), for a total of 393-419 standardized operations/(MW-year). Analysis gives a first-generation logistics of 5.3-9.0 t/(MW-year), with a nominal case of ~ 6.6 t/(MW year), 560-700 productive robot-hors/(MW-year). Planned refresh exceeds random replacement under the stated component populations, hazards, 3-15-year intervals. At ~400 standardized operations/(MW-year), the post-internal-recovery exception probability is <$10^{-3}$, with an objective near $10^{-4}$ at large scale; terminal non-recovery $p_U$ requires a smaller mission-level allocation. The target catastrophic-loss hazard for a 100 kW node is 0.01-0.03 1/yr. Parametric workload and cost cases place contingency visits at 10s of MWs, periodic campaigns at 10-100s of MWs, dedicated personnel at several 100 MWs to GWs. The reference first deployment is uncrewed, autonomously fault-managed, robotically maintainable, supported by specific inventory based on a 6-month replenishment horizon, compatible with later human access without permanent habitation.
中文解读
背景:AI 数据中心负载、功率密度和能源约束同步上升,AI 运维、负载预测和设施调优正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用框架构建和频域/系统级分析,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向能效评价口径、运营指标和优化目标的系统化梳理。意义:对日报读者而言,它可用于判断AI 工具是否能降低运维复杂度并提升可用性。仍需结合全文实验条件、样本范围和成本假设核验。
参考文献
Slava G. Turyshev. Operations, Maintenance, and Industrial Scaling of MW-Class Orbital Data Centers[J/OL]. (2026-08-27)[2026-09-01]. http://arxiv.org/abs/2608.27499v1.
Research Article热管理与液冷
Seung Hyeon Ham、Yu Min Kim、In Hyeok Choi、Jeong Woo Han
Published 2026-08-26 · arXiv · Credibility S
The rapid growth of generative AI has intensified the need for efficient heat dissipation in large-scale data centers. To control heat flow, thermal metamaterials with layered structures have been widely used, which impart the anisotropic properties of thermal conductivities. However, the conventional effective medium approximation (EMA) often fails to provide accurate predictions in systems with a high thermal cond…
Abstract, interpretation and reference
Abstract
The rapid growth of generative AI has intensified the need for efficient heat dissipation in large-scale data centers. To control heat flow, thermal metamaterials with layered structures have been widely used, which impart the anisotropic properties of thermal conductivities. However, the conventional effective medium approximation (EMA) often fails to provide accurate predictions in systems with a high thermal conductivity contrast between adjacent layers embedded in a background medium. Here, we generalize the EMA by introducing two corrective coefficients that extend its validity to regimes where the conventional EMA was previously inapplicable, i.e., high-contrast thermal metamaterials with the background medium. Notably, one of these coefficients that we proposed has the same mathematical form as the Fresnel reflection coefficient in optics. This allows us to interpret the "reflection-like" behavior of heat flow as it penetrates adjacent layers with high thermal contrast. Our findings suggest that heat diffusion, traditionally viewed as a purely dissipative process, can be understood intuitively through the framework of ray optics.
中文解读
背景:AI 数据中心负载、功率密度和能源约束同步上升,液冷、热管理和数据中心能效正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用框架构建和频域/系统级分析,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向冷却效率、能源利用或运维策略的改进方向。意义:对日报读者而言,它可用于判断液冷方案、热管理路线和高密度部署节奏。仍需结合全文实验条件、样本范围和成本假设核验。
参考文献
Seung Hyeon Ham, Yu Min Kim, In Hyeok Choi, 等. Generalizing Thermal Transport in High-Contrast Metamaterials through Interfacial Fresnel Reflection[J/OL]. (2026-08-26)[2026-09-01]. http://arxiv.org/abs/2608.25499v1.
Research Article算电协同
Hassan Zahid Butt、Rida Fatima、Xingpeng Li
Published 2026-08-30 · arXiv · Credibility S
Securing grid interconnection capacity has become a bottleneck for AI data center projects and can take longer than constructing the facilities themselves. This mismatch can delay deployment for years, making early interconnection planning essential. This paper develops ICP-AI, an interconnection capacity planning framework from a data center developer's perspective. The framework minimizes grid import capacity unde…
Abstract, interpretation and reference
Abstract
Securing grid interconnection capacity has become a bottleneck for AI data center projects and can take longer than constructing the facilities themselves. This mismatch can delay deployment for years, making early interconnection planning essential. This paper develops ICP-AI, an interconnection capacity planning framework from a data center developer's perspective. The framework minimizes grid import capacity under a prescribed onsite investment budget while jointly sizing photovoltaic (PV) and battery energy storage system (BESS) resources and scheduling deadline constrained workload flexibility. A secondary refinement fixes the minimum grid capacity and selects the minimum-investment PV-BESS portfolio among solutions that achieve that capacity. The framework is evaluated using monthly composite stress profiles across varying temporal assumptions, load shapes, flexible load fractions, and deferral windows. Results show that interconnection capacity reduction depends strongly on the planning environment: at a $100M budget, it is about 6% for the high load factor baseline, exceeds 10% under monthly average solar availability, and reaches 13.3% for a more diurnal load. At a $10M budget, 5% flexible load with a 1 h workload deferral window reduces BESS capacity from 15.30 to 4.87 MWh while increasing capacity reduction from 4.43% to 4.84%. To test sensitivity to temporal compression, the model is also solved over the full 8,760 h chronology, which preserves the main capacity and flexibility trends. Overall, ICP-AI quantifies the interconnection capacity and infrastructure substitution value of workload flexibility, providing an investment-interconnection frontier to support capital allocation and early project planning in constrained grid environments.
中文解读
背景:AI 数据中心负载、功率密度和能源约束同步上升,算力负载与电网侧资源的协同调度正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用建模优化、调度分析或算法评估,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向AI 负载波动对电网设备寿命和调频边界的影响。意义:对日报读者而言,它可用于判断智算中心建设是否受电网容量、负载波动和调度机制约束。仍需结合全文实验条件、样本范围和成本假设核验。
参考文献
Hassan Zahid Butt, Rida Fatima, Xingpeng Li. Minimizing Grid Interconnection Capacity Requirements for AI Data Centers: A Developer-Side Planning Framework with Onsite Resources and Workload Flexibility[J/OL]. (2026-08-30)[2026-09-01]. http://arxiv.org/abs/2608.29359v1.
Research Article算电协同
Muhammad Hamza Ali、Peng Sang、Hyeon Woo、Hyein Kang、Sungyun Choi、Amritanshu Pandey
Published 2026-08-18 · arXiv · Credibility S
Planners currently represent data centers as aggregate constant-PQ or ZIP loads in steady-state interconnection and contingency studies. These aggregate models are computationally convenient. However, they obscure the electrical relationship between computational workloads, server utilization, and grid-side demand. They ignore the internal power-electronic conversion stages of IT loads and assume homogeneous workloa…
Abstract, interpretation and reference
Abstract
Planners currently represent data centers as aggregate constant-PQ or ZIP loads in steady-state interconnection and contingency studies. These aggregate models are computationally convenient. However, they obscure the electrical relationship between computational workloads, server utilization, and grid-side demand. They ignore the internal power-electronic conversion stages of IT loads and assume homogeneous workload distributions across the compute clusters. This hides operating-point-dependent converter losses and efficiency variations. We propose a steady-state equivalent-circuit model (ECM) for data centers, which explicitly builds circuit models for IT loads, power supply units, cooling, and auxiliary systems. For power supply units, the equivalent circuit model explicitly represents internal power-electronic conversion stages. For IT loads, we develop a utilization-dependent server power model, and we combine it with loss-aware ECMs of power supply units. This approach captures the grid-side impact of heterogeneous workload distributions while preserving compatibility with conventional power-flow analysis. We evaluate this data center ECM in large-scale transmission power flows, using Monte Carlo simulations under heterogeneous and homogeneous cluster utilization. In comparison with the fixed-efficiency constant-PQ model, the ECM predicts that the most stressed line exceeds its thermal limit in about 30% of Monte Carlo samples. The results further show that homogeneous server utilization overstates line-loading variability by 17%-46% relative to heterogeneous server utilization, depending on the intra-cluster workload correlation.
中文解读
背景:AI 数据中心负载、功率密度和能源约束同步上升,算力负载与电网侧资源的协同调度正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用框架构建和频域/系统级分析,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向AI 负载波动对电网设备寿命和调频边界的影响。意义:对日报读者而言,它可用于判断智算中心建设是否受电网容量、负载波动和调度机制约束。仍需结合全文实验条件、样本范围和成本假设核验。
参考文献
Muhammad Hamza Ali, Peng Sang, Hyeon Woo, 等. Steady-State Equivalent Circuit Model for Data Center Loads[J/OL]. (2026-08-18)[2026-09-01]. http://arxiv.org/abs/2608.17925v1.
Research Article算电协同
Eliseo Curcio
Published 2026-08-08 · arXiv · Credibility S
Interconnection queues, not electricity prices, now govern where data centers can be built, and the standard levelized-cost comparison answers a question no developer faces: it assumes a load profile, freezes the grid price while modeling the demand that moves it, and quotes busbar costs a facility cannot buy. This paper evaluates nine on-site supply technologies against a delivered grid whose price is endogenous to…
Abstract, interpretation and reference
Abstract
Interconnection queues, not electricity prices, now govern where data centers can be built, and the standard levelized-cost comparison answers a question no developer faces: it assumes a load profile, freezes the grid price while modeling the demand that moves it, and quotes busbar costs a facility cannot buy. This paper evaluates nine on-site supply technologies against a delivered grid whose price is endogenous to projected data-center demand, on a complete-site basis that retains standby charges, with measured GPU training load, delivered fuel prices, production-pathway carbon, and statutory 45V and 48E incentive mechanics. Nothing beats the wire: gas combined cycle produces at 47 USD/MWh but costs about 114 USD per megawatt-hour of complete site energy against a 92 USD grid; four-hour storage is physically capped near 18 percent of annual energy and, charged at the margin, dirtier than the grid; hydrogen from grid-priced power fails on cost and carbon together. An investment inversion converts these findings into capital terms: conversion-hardware learning buys nothing, because free hardware still exceeds the grid for every low-carbon arm, while global electrolyser deployment on sited sub-20 USD/MWh power brings PEM hydrogen power to about 2.2 times the grid at 300 billion USD and 1.9 times at 1 trillion USD (2.7 and 2.3 for the hydrogen engine), with a carbon reduction of roughly 85 percent (6.8-fold) against grid-power production. Grid parity is not purchasable at any budget. On-site supply is an access and depth product; most current investment targets the wrong term.
中文解读
背景:AI 数据中心负载、功率密度和能源约束同步上升,算力负载与电网侧资源的协同调度正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用仿真建模和情景分析,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向AI 负载波动对电网设备寿命和调频边界的影响。意义:对日报读者而言,它可用于判断智算中心建设是否受电网容量、负载波动和调度机制约束。仍需结合全文实验条件、样本范围和成本假设核验。
参考文献
Eliseo Curcio. Beyond the Grid: Cost, Carbon, and Capital Requirements of On-Site Power Technologies for AI Data Centers[J/OL]. (2026-08-08)[2026-09-01]. http://arxiv.org/abs/2608.08170v1.
Research Article算电协同
Johanna Bolaños-Zuñiga、Alberto J. Lamadrid
Published 2026-08-11 · arXiv · Credibility S
In this study, we use electricity demand growth, cooling requirements, and backup system operation to evaluate the environmental and economic implications of artificial intelligence data centers in the United States. Our results indicate that impacts are not determined solely by facility design, but by the broader electricity, water, and land-use systems in which these facilities operate. Emissions are primarily dri…
Abstract, interpretation and reference
Abstract
In this study, we use electricity demand growth, cooling requirements, and backup system operation to evaluate the environmental and economic implications of artificial intelligence data centers in the United States. Our results indicate that impacts are not determined solely by facility design, but by the broader electricity, water, and land-use systems in which these facilities operate. Emissions are primarily driven by electricity consumption and therefore depend on marginal generation mixes, transmission constraints, and the spatial and temporal distribution of demand. Analysis further shows that local effects include pressures on water resources, increased noise exposure, and land-use changes, with outcomes varying across regions and infrastructure conditions. The assessment of technological and operational measures shows that improvements in energy efficiency, cooling configurations, and operational strategies can reduce these impacts, although their effectiveness depends on system-level conditions. Evaluation of regulatory and market structures suggests that existing frameworks may not fully account for location- and time-specific externalities. These findings support the need for integrated policy approaches that align data center deployment and operation with electricity system characteristics, water availability, and land-use planning to improve overall environmental and economic performance.
中文解读
背景:AI 数据中心负载、功率密度和能源约束同步上升,算力负载与电网侧资源的协同调度正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用框架构建和频域/系统级分析,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向能效评价口径、运营指标和优化目标的系统化梳理。意义:对日报读者而言,它可用于判断智算中心建设是否受电网容量、负载波动和调度机制约束。仍需结合全文实验条件、样本范围和成本假设核验。
参考文献
Johanna Bolaños-Zuñiga, Alberto J. Lamadrid. Environmental and Economic Implications of Artificial Intelligence Data Centers in the United States[J/OL]. (2026-08-11)[2026-09-01]. http://arxiv.org/abs/2608.09882v1.
Research Article算电协同
Saroj Khanal、Geon Roh、Boyu Yao、Abraham Silverman、Dennice Gayme、Charalambos Konstantinou、Jip Kim、Yury Dvorkin
Published 2026-08-20 · arXiv · Credibility S
Data-center growth risks overbuilding power grid infrastructure and stranding capital. Flexible data-center operation can defer infrastructure investments, but its value depends on the flexibility mechanism and the host power grid characteristics. We classify data-center load as firm, flexible or interruptible, and embed them in capacity expansion applied to market-organized, fossil-heavy PJM and carbon-capped, cent…
Abstract, interpretation and reference
Abstract
Data-center growth risks overbuilding power grid infrastructure and stranding capital. Flexible data-center operation can defer infrastructure investments, but its value depends on the flexibility mechanism and the host power grid characteristics. We classify data-center load as firm, flexible or interruptible, and embed them in capacity expansion applied to market-organized, fossil-heavy PJM and carbon-capped, centrally coordinated Korea. In PJM, the flexibility value is spatial: shifting workloads between zones reduces system cost by 6% in 2028 and 19% in 2038, avoiding 4.4 GW and 8.9 GW of gas and nuclear generation. In Korea, it is temporal: shifting load into midday solar hours makes 0.5 GW of additional solar worth building in 2028 and avoids 1.2 GW of gas and 0.3 GW of batteries in 2038. In both, realistic event-shape limits diminish the value of curtailment. The results show that flexibility procurement and its value are driven by grid characteristics and policy objectives.
中文解读
背景:AI 数据中心负载、功率密度和能源约束同步上升,算力负载与电网侧资源的协同调度正在成为智算中心设计的关键变量。问题:论文聚焦现有方案在效率、可靠性或工程协同上的瓶颈。方法:摘要显示作者采用文献摘要中的模型、实验或案例分析,把运行负载、冷却/能源系统和基础设施约束放在同一分析框架中。结果:研究重点指向AI 负载波动对电网设备寿命和调频边界的影响。意义:对日报读者而言,它可用于判断智算中心建设是否受电网容量、负载波动和调度机制约束。仍需结合全文实验条件、样本范围和成本假设核验。
参考文献
Saroj Khanal, Geon Roh, Boyu Yao, 等. Shift or curtail? How much data-center flexibility is worth depends on the host power grid[J/OL]. (2026-08-20)[2026-09-01]. http://arxiv.org/abs/2608.19622v1.