← 論文一覧に戻る

米国CCGT運転ベンチマーク:米国コンバインドサイクル発電所群の実測効率、CO2原単位、稼働率

US CCGT Operational Benchmark: measured efficiency, CO2 intensity and utilisation of the US combined-cycle fleet (原題)

Aheiev, Dmytro

Zenodoデータセット2026-08-30#エネルギー転換Origin: US経営インパクト: コスト削減対象セクター: power
DOI: 10.5281/zenodo.22172851
原典: https://zenodo.org/records/22172851

🤖 gxceed AI 要約

日本語

米国409基のCCGT発電所について、EPA eGRID2022の実測データ(CEMS)を用いて、正味熱効率(中央値46.2%)、CO2排出原単位、設備利用率などを系統的にベンチマークしたデータセット。発電所の構造特性(建設年、HRSG構成、CHP有無など)と実測値を厳密に結合し、機械学習タスク用に漏洩防止の特徴量設計と交差検証フォールドを提供する。発電効率の差異を技術・経年・運用の観点から分析する基盤となる。

English

A benchmark dataset of 409 US CCGT plants with measured annual performance (net efficiency median 46.2%, CO2 intensity, capacity factor) from EPA eGRID2022, joined to fleet structure (vintage, HRSG, CHP, etc.) with strict entity matching. Provides leakage-safe ML tasks with excluded features and CV folds, enabling analysis of efficiency drivers across the fleet. Useful for decarbonization benchmarking and energy transition research.

Unofficial AI-generated summary based on the public title and abstract. Not an official translation.

📝 gxceed 編集解説 — Why this matters

日本のGX文脈において

日本の火力発電所の効率・CO2原単位ベンチマークは、SSBJ開示やトランジション・ファイナンス評価において重要。本データセットの手法(実測値と構造データの厳密な結合、MLタスク設計)は、日本の発電所データ整備や開示インフラ構築に参考になる。

In the global GX context

This benchmark provides a rigorous, open methodology for comparing CCGT efficiency and CO2 intensity across a national fleet, directly relevant to global efforts on transition finance, TCFD/ISSB disclosure, and climate risk assessment. The entity-matching and leakage-safe ML design set a standard for similar datasets in other countries.

👥 読者別の含意

🔬研究者:Provides a high-quality, leakage-safe dataset for benchmarking CCGT efficiency and CO2 intensity, enabling cross-plant analysis and ML modeling.

🏢実務担当者:Offers a reference for benchmarking plant performance and identifying efficiency improvement opportunities, useful for sustainability reporting and asset management.

🏛政策担当者:Demonstrates a transparent, data-driven approach to monitoring power sector decarbonization, informing policy on emissions standards and transition planning.

📄 Abstract(原文)

409 US combined-cycle gas (CCGT) plants, one row per EIA plant code , with measured annual performance from US EPA eGRID2022 (CEMS/metered): net thermal efficiency (median 46.2%, range 22.4-65.4%), heat rate, CO2 output rate, capacity factor and net generation - joined to fleet structure: capacity, commissioning year, HRSG and gas-turbine unit counts, duct burners, HRSG bypass capability, CHP flag, carbon capture, planned retirement (US EIA Form 860), owner and location. Why it exists. The widely used UCI Combined Cycle Power Plant dataset contains 9,568 hourly observations of a single plant. This benchmark is the complementary view - many plants compared across the fleet on regulator-reported annual data - shifting the question from "how does ambient temperature move one plant's output" to "why is one plant more efficient than another" (vintage, technology class, utilisation, cogeneration). Entity guard. The upstream fleet inventory matched plant records to EIA plant codes by coordinate proximity; a row is kept only if its identifier embeds the joined EIA code, guaranteeing that structural features and measured values describe the same asset. This removes 9 duplicate neighbour joins and 12 single records whose capacity/vintage describe a different asset than the joined eGRID data (e.g. a 4.2 MW 1993 cogen carrying the measured values of a 1,267 MW plant 5 km away). Every join was verified against eGRID2022 names, coordinates and nameplate capacity - evidence in join_verification.csv, all 21 removals logged in dropped_mismatched_joins.csv. One plant = one row = one set of measured values. Leakage-safe ML tasks. tasks.json defines three regression targets (thermal efficiency, CO2 intensity, capacity factor) each with an explicit excluded-feature list, because several columns are deterministic transformations of one another (efficiency = 3412.14 / heat rate; CO2 intensity is approximately fuel emission factor x heat rate; capacity factor = generation / (capacity x 8760)). A deterministic cv_fold column supports reproducible 5-fold cross-validation. Caveats. 25 of the 27 plants above 55% apparent efficiency are CHP sites, where eGRID allocates part of the fuel to useful heat - filter or stratify on the chp flag. Values are annual averages for a single year (eGRID2022), efficiencies are net and HHV-basis (lower than LHV vendor figures), turbine_class is populated for only 11 of 409 plants, and coverage is US-only because measured CEMS data of this kind is not openly available for most other countries. v1.1.0 : negative numeric values (longitudes) are now plain numbers (v1.0.x escaped them with a leading apostrophe as an over-broad spreadsheet-formula guard, breaking numeric parsing); the entity guard extended to single rows (409 plants, was 421); croissant.json upgraded to conformant MLCommons Croissant 1.0. Sources: US EPA eGRID2022; US EIA Form 860; WRI Global Power Plant Database; Global Energy Monitor GOGPT; Climate TRACE. CC BY 4.0.

🔗 Provenance — このレコードを発見したソース

🔔 こうした論文の新着を逃したくない方は キーワードアラート に登録(無料・3キーワードまで)。

gxceed は公開メタデータに基づく研究支援データセットです。要約・翻訳・解説は AI 支援で生成されています。 最終的な解釈・検証は利用者が原典資料に基づいて行うことを前提とします。