{"id":"research-compute-resourcing","title":"Compute Resourcing for Atomistic Simulation and ML","subtitle":"Public allocation programs, cloud and spot pricing, and resourcing routes, compiled 2026-07-21.","category":"references","tags":["literature-review","savings-stack","allocations","cloud","funding"],"source":"articles/lit-review/research-compute-resourcing.md","lang":"en","words":1498,"readMinutes":7,"toc":[{"depth":2,"text":"1. Public allocation programs","id":"1-public-allocation-programs"},{"depth":3,"text":"US — NSF","id":"us-nsf"},{"depth":3,"text":"US — DOE (open to any researcher worldwide, no DOE funding required)","id":"us-doe-open-to-any-researcher-worldwide-no-doe-funding-required"},{"depth":3,"text":"Europe","id":"europe"},{"depth":3,"text":"UK — ARCHER2 — [access page](https://www.archer2.ac.uk/support-access/access.html)","id":"uk-archer2-access-page-https-www-archer2-ac-uk-support-access-access-html"},{"depth":3,"text":"China [partially uncertain — thin English-language primary sources]","id":"china-partially-uncertain-thin-english-language-primary-sources"},{"depth":2,"text":"2. Cloud research credit programs","id":"2-cloud-research-credit-programs"},{"depth":3,"text":"GPU-specialty clouds (verified list prices)","id":"gpu-specialty-clouds-verified-list-prices"},{"depth":2,"text":"3. Economics","id":"3-economics"},{"depth":2,"text":"4. Community / sharing models","id":"4-community-sharing-models"},{"depth":2,"text":"Suggested stack for a small atomistic-sim + ML team","id":"suggested-stack-for-a-small-atomistic-sim-ml-team"}],"html":"<blockquote>\n<p><strong>Provenance:</strong> explore agent <code>agent-12</code> (director-commissioned deep research, backdrop research, 2026-07-21) — materialized verbatim. Quantitative claims are as reported by the research agent from sources it accessed; see citations inline. Citation-verification pass pending before any external publication.</p>\n</blockquote>\n<h1 id=\"compute-resourcing-for-atomistic-simulation-ml-program-digest\">Compute Resourcing for Atomistic Simulation + ML: Program Digest</h1><p>Compiled 2026-07-21. Prices are volatile and region-dependent; all figures come from the cited pages — treat them as pointers, not quotes. Items I could not verify against a primary source are marked <strong>[uncertain]</strong>.</p>\n<h2 id=\"1-public-allocation-programs\">1. Public allocation programs</h2><h3 id=\"us-nsf\">US — NSF</h3><p><strong>ACCESS</strong> (XSEDE successor; NSF awarded $52M over 5 years) — <a href=\"https://access-ci.org/\">access-ci.org</a></p>\n<ul>\n<li>Free; no NSF award required to start. Four project types, awarded in ACCESS Credits (1 credit ≈ 1 core-hour or 1 GB storage, per the <a href=\"https://allocations.access-ci.org/exchange_calculator\">exchange calculator</a>):<ul>\n<li><strong>Explore</strong>: 400,000 credits, anytime, overview-only application, approval in ~1–2 business days</li>\n<li><strong>Discover</strong>: 1,500,000 credits, 1-page proposal</li>\n<li><strong>Accelerate</strong>: 3,000,000 credits, 3-page proposal, panel review</li>\n<li><strong>Maximize</strong>: awarded in resource units, 10-page proposal + code performance doc, semi-annual windows (next: Jun 15–Jul 31, 2026, awards start Oct 1, 2026) — <a href=\"https://allocations.access-ci.org/project-types\">project types</a>, <a href=\"https://allocations.access-ci.org/prepare-requests\">prepare requests</a></li>\n</ul>\n</li>\n<li>Graduate students can be PI with an advisor letter. Resources include GPU clusters (Delta, Expanse, Anvil, ACES) and the Jetstream2 research cloud — useful for a mixed DFT + ML workflow.</li>\n</ul>\n<p><strong>NAIRR Pilot</strong> — <a href=\"https://nairrpilot.org\">nairrpilot.org</a>. US-based researchers/educators; 3-page application; distributes HPC time (Frontera, Expanse/Voyager, Anvil, DOE AI testbeds) and cloud credits from AWS/Google/Microsoft/NVIDIA partners. 375+ projects supported. ML-for-materials fits its focus areas.</p>\n<h3 id=\"us-doe-open-to-any-researcher-worldwide-no-doe-funding-required\">US — DOE (open to any researcher worldwide, no DOE funding required)</h3><ul>\n<li><strong>INCITE</strong> — flagship; awards ~60% of ALCF/OLCF time (Aurora, Frontier). Typical awards ~0.5–2.5M node-hours, 1–3 years, renewable. Annual call Apr–Jun (2026 call: Apr 11–Jun 16, 2025); 10% reserved for an Early Career Track. <a href=\"https://www.alcf.anl.gov/science/incite-allocation-program\">ALCF INCITE page</a>, <a href=\"https://doeleadershipcomputing.org/\">doeleadershipcomputing.org</a></li>\n<li><strong>ALCC</strong> — DOE-mission-aligned, high-risk/high-payoff; 2025–26 cycle: 38M node-hours across 56 projects; now 1–3 year requests, proposals due ~late January, no pre-proposal. <a href=\"https://science.osti.gov/ascr/Facilities/Accessing-ASCR-Facilities/ALCC/FAQ-Detail\">DOE FAQ</a>, <a href=\"https://content.govdelivery.com/accounts/USDOEOS/bulletins/3e88768\">award announcement</a></li>\n<li><strong>Director&#39;s Discretionary (ALCF/OLCF)</strong> — small &quot;get started&quot; awards, anytime; the realistic entry point for a small team before INCITE/ALCC. <a href=\"https://www.alcf.anl.gov/sites/default/files/2025-12/ALCF_2025ScienceReport.pdf\">ALCF programs overview</a></li>\n<li><strong>NERSC (ERCAP)</strong> — annual call (AY2026: opened Aug 11, 2025, due Oct 2025; year runs ~Jan 21–Jan 19). Requires DOE Office of Science funding alignment; quarterly usage-based reductions. <a href=\"https://docs.nersc.gov/allocations/ercap_2026/\">ERCAP guidance</a></li>\n</ul>\n<h3 id=\"europe\">Europe</h3><p><strong>EuroHPC JU</strong> (the PRACE successor for allocation purposes) — <a href=\"https://eurohpc-ju.europa.eu/access-our-supercomputers/access-policy-and-faq_en\">access policy</a></p>\n<ul>\n<li>Free of charge for EU/associated-country researchers (academia, industry, public sector). JU manages 35% (petascale) to 50% (pre-exascale) of system capacity on LUMI, Leonardo, MareNostrum5, Jupiter, etc.</li>\n<li>Modes: <strong>Benchmark</strong> (≤3 months, monthly cut-offs), <strong>Development</strong> (≤1 year, monthly cut-offs), <strong>AI &amp; Data-Intensive</strong> (SME/industry/public sector), <strong>Regular</strong> (<del>2 cut-offs/year, 12-page proposal), <strong>Extreme Scale</strong> (</del>2 cut-offs/year, e.g. Oct 17, 2025; tracks for scientific/industry/SME). <a href=\"https://www.it4i.cz/en/about/infoservice/news/gain-access-to-european-supercomputers-including-karolina-eurohpc-ju-announced-calls-for-2025\">Call summary</a>, <a href=\"https://lumi-supercomputer.eu/current-open-calls-for-lumi-resources/\">LUMI calls</a></li>\n<li><strong>PRACE itself no longer allocates compute</strong> — it transformed into an association of users and HPC centres (&quot;PRACE 3.0&quot;); EuroHPC runs its own peer-review platform. <a href=\"https://landscape2024.esfri.eu/media/yiplqqqc/esfri-la-2024_section1-digit.pdf\">ESFRI Landscape 2024</a>, <a href=\"https://www.developmentaid.org/tenders/view/1342030/\">EuroHPC peer-review tender</a></li>\n</ul>\n<h3 id=\"uk-archer2-access-page\">UK — ARCHER2 — <a href=\"https://www.archer2.ac.uk/support-access/access.html\">access page</a></h3><ul>\n<li>Routes: <strong>Driving Test</strong> (800 CU, new users, incl. GPU test), <strong>Pump Priming</strong> (≤4,000 CU / 6 months, always open, ~2-week decision), <strong>Access to HPC</strong> (EPSRC, semi-annual), <strong>Pioneer Projects</strong> (large, ≤2 years), <strong>HEC Consortia</strong> (e.g. materials/chemistry consortia with continuous allocations — a practical route for a materials team), plus ARCHER2-on-EPSRC-grant.</li>\n<li>1 CU = 1 node-hour (128 cores). Notional cost £0.20/CU (EPSRC/NERC), £0.39 otherwise. <a href=\"https://digitalresearchservices.ed.ac.uk/resources/archer2\">Edinburgh DRS page</a></li>\n<li><strong>Caution</strong>: ARCHER2 service ends <strong>21 November 2026</strong>; successor provision is not guaranteed — flagged on the official access page.</li>\n</ul>\n<h3 id=\"china-partially-uncertain-thin-english-language-primary-sources\">China <strong>[partially uncertain — thin English-language primary sources]</strong></h3><ul>\n<li>NSFC does not itself allocate supercomputer time; access runs through the National Supercomputing Centers (Guangzhou, Tianjin, Wuxi, etc.) via institutional agreements, paid services, or project-tied allocations (e.g., National Key R&amp;D projects). <strong>[uncertain — verify per-center]</strong></li>\n<li>New: the <strong>National Supercomputing Internet (国家超算互联网, scnet)</strong>, launched April 2024 by the Ministry of Science and Technology — a marketplace aggregating 200+ providers, 25+ resource types, 6,000+ software listings, 100k+ users; a &quot;core node&quot; with 100k+ AI cards reported online in 2026. <a href=\"https://english.www.gov.cn/news/202404/12/content_WS66187785c6d0868f4e8e5f5c.html\">PRC State Council announcement</a>, <a href=\"http://qiye.chinadaily.com.cn/a/202409/26/WS66f4ed0ea310b59111d9b490.html\">China Daily</a>. Practical access for a foreign-affiliated team is <strong>uncertain</strong>; typically via Chinese institutional collaborators.</li>\n</ul>\n<h2 id=\"2-cloud-research-credit-programs\">2. Cloud research credit programs</h2><div class=\"table-wrap\"><table><thead><tr>\n<th>Program</th>\n<th>Offer</th>\n<th>How a small team qualifies</th>\n<th>Link</th>\n</tr>\n</thead><tbody><tr>\n<td data-label=\"Program\"><strong>AWS Cloud Credit for Research</strong></td>\n<td data-label=\"Offer\">Credits for on-demand + spot EC2; students ≤$5,000; faculty/staff uncapped; 1-year term</td>\n<td data-label=\"How a small team qualifies\">Rolling application, 90–120 day review; academic affiliation</td>\n<td data-label=\"Link\"><a href=\"https://aws.amazon.com/government-education/research-and-technical-computing/cloud-credit-for-research/\">aws.amazon.com</a></td>\n</tr>\n<tr>\n<td data-label=\"Program\"><strong>Google Cloud research credits</strong></td>\n<td data-label=\"Offer\">Credits expire 365 days after redemption; faculty, PhD students, postdocs at accredited institutions in approved countries; application requires a 250-word proposal + pricing-calculator estimate. Award size not published — historically ~$5,000 <strong>[uncertain]</strong></td>\n<td data-label=\"How a small team qualifies\">Online form, rolling</td>\n<td data-label=\"Link\"><a href=\"https://edu.google.com/programs/credits/research/\">edu.google.com</a></td>\n</tr>\n<tr>\n<td data-label=\"Program\"><strong>Azure Research Credits</strong></td>\n<td data-label=\"Offer\">&quot;Azure for Academic Research&quot;; contact-form intake. Third-party listings cite up to ~<span class=\"katex\"><span class=\"katex-mathml\"><math xmlns=\"http://www.w3.org/1998/Math/MathML\"><semantics><mrow><mn>60</mn><mo separator=\"true\">,</mo><mn>000</mn><mo>∗</mo><mo>∗</mo><mo stretchy=\"false\">[</mo><mi>u</mi><mi>n</mi><mi>c</mi><mi>e</mi><mi>r</mi><mi>t</mi><mi>a</mi><mi>i</mi><mi>n</mi><mo separator=\"true\">;</mo><mi>i</mi><mi>n</mi><mi>s</mi><mi>t</mi><mi>i</mi><mi>t</mi><mi>u</mi><mi>t</mi><mi>i</mi><mi>o</mi><mi>n</mi><mi>a</mi><mi>l</mi><mi>d</mi><mi>e</mi><mi>a</mi><mi>l</mi><mi>s</mi><mi>v</mi><mi>a</mi><mi>r</mi><mi>y</mi><mtext>—</mtext><mi>e</mi><mi mathvariant=\"normal\">.</mi><mi>g</mi><mi mathvariant=\"normal\">.</mi><mo separator=\"true\">,</mo><mi>S</mi><mi>t</mi><mi>a</mi><mi>n</mi><mi>f</mi><mi>o</mi><mi>r</mi><mi>d</mi><mi>H</mi><mi>A</mi><mi>I</mi><mi>o</mi><mi>f</mi><mi>f</mi><mi>e</mi><mi>r</mi><mi>s</mi><mi>u</mi><mi>p</mi><mi>t</mi><mi>o</mi></mrow><annotation encoding=\"application/x-tex\">60,000 **[uncertain; institutional deals vary — e.g., Stanford HAI offers up to</annotation></semantics></math></span><span class=\"katex-html\" aria-hidden=\"true\"><span class=\"base\"><span class=\"strut\" style=\"height:0.8389em;vertical-align:-0.1944em;\"></span><span class=\"mord\">60</span><span class=\"mpunct\">,</span><span class=\"mspace\" style=\"margin-right:0.1667em;\"></span><span class=\"mord\">000</span><span class=\"mspace\" style=\"margin-right:0.2222em;\"></span><span class=\"mbin\">∗</span><span class=\"mspace\" style=\"margin-right:0.2222em;\"></span></span><span class=\"base\"><span class=\"strut\" style=\"height:1em;vertical-align:-0.25em;\"></span><span class=\"mord\">∗</span><span class=\"mopen\">[</span><span class=\"mord mathnormal\">u</span><span class=\"mord mathnormal\">n</span><span class=\"mord mathnormal\" style=\"margin-right:0.0278em;\">cer</span><span class=\"mord mathnormal\">t</span><span class=\"mord mathnormal\">ain</span><span class=\"mpunct\">;</span><span class=\"mspace\" style=\"margin-right:0.1667em;\"></span><span class=\"mord mathnormal\">in</span><span class=\"mord mathnormal\">s</span><span class=\"mord mathnormal\">t</span><span class=\"mord mathnormal\">i</span><span class=\"mord mathnormal\">t</span><span class=\"mord mathnormal\">u</span><span class=\"mord mathnormal\">t</span><span class=\"mord mathnormal\">i</span><span class=\"mord mathnormal\">o</span><span class=\"mord mathnormal\">na</span><span class=\"mord mathnormal\" style=\"margin-right:0.0197em;\">l</span><span class=\"mord mathnormal\">d</span><span class=\"mord mathnormal\">e</span><span class=\"mord mathnormal\">a</span><span class=\"mord mathnormal\" style=\"margin-right:0.0197em;\">l</span><span class=\"mord mathnormal\">s</span><span class=\"mord mathnormal\" style=\"margin-right:0.0359em;\">v</span><span class=\"mord mathnormal\">a</span><span class=\"mord mathnormal\" style=\"margin-right:0.0278em;\">r</span><span class=\"mord mathnormal\" style=\"margin-right:0.0359em;\">y</span><span class=\"mord\">—</span><span class=\"mord mathnormal\">e</span><span class=\"mord\">.</span><span class=\"mord mathnormal\" style=\"margin-right:0.0359em;\">g</span><span class=\"mord\">.</span><span class=\"mpunct\">,</span><span class=\"mspace\" style=\"margin-right:0.1667em;\"></span><span class=\"mord mathnormal\" style=\"margin-right:0.0576em;\">S</span><span class=\"mord mathnormal\">t</span><span class=\"mord mathnormal\">an</span><span class=\"mord mathnormal\" style=\"margin-right:0.1076em;\">f</span><span class=\"mord mathnormal\" style=\"margin-right:0.0278em;\">or</span><span class=\"mord mathnormal\">d</span><span class=\"mord mathnormal\" style=\"margin-right:0.0813em;\">H</span><span class=\"mord mathnormal\">A</span><span class=\"mord mathnormal\" style=\"margin-right:0.0785em;\">I</span><span class=\"mord mathnormal\">o</span><span class=\"mord mathnormal\" style=\"margin-right:0.1076em;\">f</span><span class=\"mord mathnormal\" style=\"margin-right:0.1076em;\">f</span><span class=\"mord mathnormal\" style=\"margin-right:0.0278em;\">er</span><span class=\"mord mathnormal\">s</span><span class=\"mord mathnormal\">u</span><span class=\"mord mathnormal\">pt</span><span class=\"mord mathnormal\">o</span></span></span></span>50k]**</td>\n<td data-label=\"How a small team qualifies\">Rolling</td>\n<td data-label=\"Link\"><a href=\"https://www.microsoft.com/en-us/azure-academic-research/\">microsoft.com</a></td>\n</tr>\n<tr>\n<td data-label=\"Program\"><strong>Oracle for Research</strong></td>\n<td data-label=\"Offer\">12-month OCI credits (IaaS+PaaS, incl. HPC/GPU); amounts vary by proposal — UCL cites ~£50k average; a Fellows track adds cash awards</td>\n<td data-label=\"How a small team qualifies\">Application form, reviewed on impact/feasibility</td>\n<td data-label=\"Link\"><a href=\"https://www.oracle.com/a/ocom/docs/corporate/research-application-example.pdf\">program PDF</a>, <a href=\"https://www.ucl.ac.uk/health/academic-careers-office/training-portfolios/data-arcade/acooracle-cloud-computing-credits\">UCL page</a></td>\n</tr>\n<tr>\n<td data-label=\"Program\"><strong>CloudBank</strong></td>\n<td data-label=\"Offer\">NSF-funded billing/brokerage so NSF awardees can spend grant funds on AWS/GCP/Azure with negotiated discounts</td>\n<td data-label=\"How a small team qualifies\">US institution + NSF award</td>\n<td data-label=\"Link\"><a href=\"https://www.cloudbank.org/\">cloudbank.org</a></td>\n</tr>\n<tr>\n<td data-label=\"Program\"><strong>NAIRR Pilot</strong></td>\n<td data-label=\"Offer\">Distributes partner cloud credits (see §1)</td>\n<td data-label=\"How a small team qualifies\">US-based, 3-page app</td>\n<td data-label=\"Link\"><a href=\"https://nairrpilot.org\">nairrpilot.org</a></td>\n</tr>\n</tbody></table></div><h3 id=\"gpu-specialty-clouds-verified-list-prices\">GPU-specialty clouds (verified list prices)</h3><ul>\n<li><strong>CoreWeave</strong> (<a href=\"https://www.coreweave.com/pricing\">pricing</a>): 8×H100 HGX <span class=\"katex\"><span class=\"katex-mathml\"><math xmlns=\"http://www.w3.org/1998/Math/MathML\"><semantics><mrow><mn>49.24</mn><mi mathvariant=\"normal\">/</mi><mi>h</mi><mi>r</mi><mi>o</mi><mi>n</mi><mo>−</mo><mi>d</mi><mi>e</mi><mi>m</mi><mi>a</mi><mi>n</mi><mi>d</mi><mo stretchy=\"false\">(</mo><mtext> </mtext></mrow><annotation encoding=\"application/x-tex\">49.24/hr on-demand (~</annotation></semantics></math></span><span class=\"katex-html\" aria-hidden=\"true\"><span class=\"base\"><span class=\"strut\" style=\"height:1em;vertical-align:-0.25em;\"></span><span class=\"mord\">49.24/</span><span class=\"mord mathnormal\">h</span><span class=\"mord mathnormal\" style=\"margin-right:0.0278em;\">r</span><span class=\"mord mathnormal\">o</span><span class=\"mord mathnormal\">n</span><span class=\"mspace\" style=\"margin-right:0.2222em;\"></span><span class=\"mbin\">−</span><span class=\"mspace\" style=\"margin-right:0.2222em;\"></span></span><span class=\"base\"><span class=\"strut\" style=\"height:1em;vertical-align:-0.25em;\"></span><span class=\"mord mathnormal\">d</span><span class=\"mord mathnormal\">e</span><span class=\"mord mathnormal\">man</span><span class=\"mord mathnormal\">d</span><span class=\"mopen\">(</span><span class=\"mspace nobreak\"> </span></span></span></span>6.16/GPU-hr), spot <span class=\"katex\"><span class=\"katex-mathml\"><math xmlns=\"http://www.w3.org/1998/Math/MathML\"><semantics><mrow><mn>19.71</mn><mi mathvariant=\"normal\">/</mi><mi>h</mi><mi>r</mi><mo separator=\"true\">;</mo><mn>8</mn><mo>×</mo><mi>H</mi><mn>200</mn></mrow><annotation encoding=\"application/x-tex\">19.71/hr; 8×H200</annotation></semantics></math></span><span class=\"katex-html\" aria-hidden=\"true\"><span class=\"base\"><span class=\"strut\" style=\"height:1em;vertical-align:-0.25em;\"></span><span class=\"mord\">19.71/</span><span class=\"mord mathnormal\">h</span><span class=\"mord mathnormal\" style=\"margin-right:0.0278em;\">r</span><span class=\"mpunct\">;</span><span class=\"mspace\" style=\"margin-right:0.1667em;\"></span><span class=\"mord\">8</span><span class=\"mspace\" style=\"margin-right:0.2222em;\"></span><span class=\"mbin\">×</span><span class=\"mspace\" style=\"margin-right:0.2222em;\"></span></span><span class=\"base\"><span class=\"strut\" style=\"height:0.6833em;\"></span><span class=\"mord mathnormal\" style=\"margin-right:0.0813em;\">H</span><span class=\"mord\">200</span></span></span></span>50.44/hr; 8×A100 <span class=\"katex\"><span class=\"katex-mathml\"><math xmlns=\"http://www.w3.org/1998/Math/MathML\"><semantics><mrow><mn>21.60</mn><mi mathvariant=\"normal\">/</mi><mi>h</mi><mi>r</mi><mo separator=\"true\">;</mo><mi>G</mi><mi>B</mi><mn>200</mn><mi>N</mi><mi>V</mi><mi>L</mi><mn>72</mn></mrow><annotation encoding=\"application/x-tex\">21.60/hr; GB200 NVL72</annotation></semantics></math></span><span class=\"katex-html\" aria-hidden=\"true\"><span class=\"base\"><span class=\"strut\" style=\"height:1em;vertical-align:-0.25em;\"></span><span class=\"mord\">21.60/</span><span class=\"mord mathnormal\">h</span><span class=\"mord mathnormal\" style=\"margin-right:0.0278em;\">r</span><span class=\"mpunct\">;</span><span class=\"mspace\" style=\"margin-right:0.1667em;\"></span><span class=\"mord mathnormal\" style=\"margin-right:0.0502em;\">GB</span><span class=\"mord\">200</span><span class=\"mord mathnormal\" style=\"margin-right:0.109em;\">N</span><span class=\"mord mathnormal\" style=\"margin-right:0.2222em;\">V</span><span class=\"mord mathnormal\">L</span><span class=\"mord\">72</span></span></span></span>42/hr/GPU; CPU-only AMD Genoa 192 vCPU <span class=\"katex\"><span class=\"katex-mathml\"><math xmlns=\"http://www.w3.org/1998/Math/MathML\"><semantics><mrow><mn>7.78</mn><mi mathvariant=\"normal\">/</mi><mi>h</mi><mi>r</mi><mo stretchy=\"false\">(</mo><mtext> </mtext></mrow><annotation encoding=\"application/x-tex\">7.78/hr (~</annotation></semantics></math></span><span class=\"katex-html\" aria-hidden=\"true\"><span class=\"base\"><span class=\"strut\" style=\"height:1em;vertical-align:-0.25em;\"></span><span class=\"mord\">7.78/</span><span class=\"mord mathnormal\">h</span><span class=\"mord mathnormal\" style=\"margin-right:0.0278em;\">r</span><span class=\"mopen\">(</span><span class=\"mspace nobreak\"> </span></span></span></span>0.04/vCPU-hr); reserved up to 60% off; <strong>free egress</strong>.</li>\n<li><strong>RunPod</strong> (<a href=\"https://www.runpod.io/gpu-instance/pricing\">pricing</a>, secure cloud, per-GPU): H100 <span class=\"katex\"><span class=\"katex-mathml\"><math xmlns=\"http://www.w3.org/1998/Math/MathML\"><semantics><mrow><mn>4.55</mn><mi mathvariant=\"normal\">/</mi><mi>h</mi><mi>r</mi><mo separator=\"true\">,</mo><mi>H</mi><mn>200</mn></mrow><annotation encoding=\"application/x-tex\">4.55/hr, H200</annotation></semantics></math></span><span class=\"katex-html\" aria-hidden=\"true\"><span class=\"base\"><span class=\"strut\" style=\"height:1em;vertical-align:-0.25em;\"></span><span class=\"mord\">4.55/</span><span class=\"mord mathnormal\">h</span><span class=\"mord mathnormal\" style=\"margin-right:0.0278em;\">r</span><span class=\"mpunct\">,</span><span class=\"mspace\" style=\"margin-right:0.1667em;\"></span><span class=\"mord mathnormal\" style=\"margin-right:0.0813em;\">H</span><span class=\"mord\">200</span></span></span></span>5.93, A100 <span class=\"katex\"><span class=\"katex-mathml\"><math xmlns=\"http://www.w3.org/1998/Math/MathML\"><semantics><mrow><mn>2.72</mn><mo separator=\"true\">,</mo><mi>R</mi><mi>T</mi><mi>X</mi><mn>4090</mn></mrow><annotation encoding=\"application/x-tex\">2.72, RTX 4090</annotation></semantics></math></span><span class=\"katex-html\" aria-hidden=\"true\"><span class=\"base\"><span class=\"strut\" style=\"height:0.8778em;vertical-align:-0.1944em;\"></span><span class=\"mord\">2.72</span><span class=\"mpunct\">,</span><span class=\"mspace\" style=\"margin-right:0.1667em;\"></span><span class=\"mord mathnormal\" style=\"margin-right:0.0077em;\">R</span><span class=\"mord mathnormal\" style=\"margin-right:0.1389em;\">T</span><span class=\"mord mathnormal\" style=\"margin-right:0.0785em;\">X</span><span class=\"mord\">4090</span></span></span></span>1.10, L4-class <span class=\"katex\"><span class=\"katex-mathml\"><math xmlns=\"http://www.w3.org/1998/Math/MathML\"><semantics><mrow><mn>0.69</mn><mo separator=\"true\">;</mo><mi>B</mi><mn>200</mn></mrow><annotation encoding=\"application/x-tex\">0.69; B200</annotation></semantics></math></span><span class=\"katex-html\" aria-hidden=\"true\"><span class=\"base\"><span class=\"strut\" style=\"height:0.8778em;vertical-align:-0.1944em;\"></span><span class=\"mord\">0.69</span><span class=\"mpunct\">;</span><span class=\"mspace\" style=\"margin-right:0.1667em;\"></span><span class=\"mord mathnormal\" style=\"margin-right:0.0502em;\">B</span><span class=\"mord\">200</span></span></span></span>8.64; community cloud is cheaper.</li>\n<li><strong>Lambda</strong> (<a href=\"https://lambda.ai/service/gpu-cloud\">pricing</a>): page is JS-rendered; third-party trackers (Apr 2026) put 1×H100 PCIe at ~<span class=\"katex\"><span class=\"katex-mathml\"><math xmlns=\"http://www.w3.org/1998/Math/MathML\"><semantics><mrow><mn>3.29</mn><mi mathvariant=\"normal\">/</mi><mi>h</mi><mi>r</mi><mo separator=\"true\">,</mo><mi>S</mi><mi>X</mi><mi>M</mi><mtext> </mtext></mrow><annotation encoding=\"application/x-tex\">3.29/hr, SXM ~</annotation></semantics></math></span><span class=\"katex-html\" aria-hidden=\"true\"><span class=\"base\"><span class=\"strut\" style=\"height:1em;vertical-align:-0.25em;\"></span><span class=\"mord\">3.29/</span><span class=\"mord mathnormal\">h</span><span class=\"mord mathnormal\" style=\"margin-right:0.0278em;\">r</span><span class=\"mpunct\">,</span><span class=\"mspace\" style=\"margin-right:0.1667em;\"></span><span class=\"mord mathnormal\" style=\"margin-right:0.0576em;\">S</span><span class=\"mord mathnormal\" style=\"margin-right:0.0785em;\">X</span><span class=\"mord mathnormal\" style=\"margin-right:0.109em;\">M</span><span class=\"mspace nobreak\"> </span></span></span></span>4.29/hr <strong>[secondary source]</strong>.</li>\n<li><strong>Vast.ai</strong> (<a href=\"https://vast.ai/pricing\">pricing</a>): marketplace; live prices (page didn&#39;t render numbers) but interruptible tier is 50%+ below on-demand; reserved (1/3/6-mo) up to 50% off. Typically the cheapest GPU source <strong>[typical rates uncertain — check live]</strong>.</li>\n<li>Big-3 comparison anchor: AWS p4d.24xlarge (8×A100) ≈ $32.77/hr on-demand (<a href=\"https://www.usage.ai/blogs/aws/ec2/instance-types/what-are-ec2-instances/\">usage.ai</a>); p5 (8×H100) is substantially higher <strong>[exact current price uncertain]</strong>. GPU-specialty clouds undercut big-3 GPU on-demand by roughly 2–5×, but lack InfiniBand-class fabrics except CoreWeave (and big-3 ND/p5/A3 tiers).</li>\n</ul>\n<h2 id=\"3-economics\">3. Economics</h2><p><strong>Purchase models.</strong> Spot/preemptible: AWS Spot up to 90% off, GCP Spot VMs 60–91% off, Azure Spot up to 90% off (<a href=\"https://www.prosperops.com/blog/spot-instances/\">ProsperOps comparison</a>). Verified example: GCP h3-standard-88 <span class=\"katex\"><span class=\"katex-mathml\"><math xmlns=\"http://www.w3.org/1998/Math/MathML\"><semantics><mrow><mn>4.92</mn><mtext>–</mtext><mn>5.97</mn><mi mathvariant=\"normal\">/</mi><mi>h</mi><mi>r</mi><mi>o</mi><mi>n</mi><mo>−</mo><mi>d</mi><mi>e</mi><mi>m</mi><mi>a</mi><mi>n</mi><mi>d</mi><mi>v</mi><mi>s</mi></mrow><annotation encoding=\"application/x-tex\">4.92–5.97/hr on-demand vs</annotation></semantics></math></span><span class=\"katex-html\" aria-hidden=\"true\"><span class=\"base\"><span class=\"strut\" style=\"height:1em;vertical-align:-0.25em;\"></span><span class=\"mord\">4.92–5.97/</span><span class=\"mord mathnormal\">h</span><span class=\"mord mathnormal\" style=\"margin-right:0.0278em;\">r</span><span class=\"mord mathnormal\">o</span><span class=\"mord mathnormal\">n</span><span class=\"mspace\" style=\"margin-right:0.2222em;\"></span><span class=\"mbin\">−</span><span class=\"mspace\" style=\"margin-right:0.2222em;\"></span></span><span class=\"base\"><span class=\"strut\" style=\"height:0.6944em;\"></span><span class=\"mord mathnormal\">d</span><span class=\"mord mathnormal\">e</span><span class=\"mord mathnormal\">man</span><span class=\"mord mathnormal\">d</span><span class=\"mord mathnormal\" style=\"margin-right:0.0359em;\">v</span><span class=\"mord mathnormal\">s</span></span></span></span>2.45/hr spot (<a href=\"https://www.devzero.io/instances/gcp/h3-standard-88\">devzero</a>). Reserved/committed: ~37–57% off 1–3 yr (AWS RI), up to 60% off (CoreWeave), up to 50% off (Vast). Batch DFT/MD is checkpointable and embarrassingly parallel → spot-friendly; long MD trajectories need fault-tolerant restart logic.</p>\n<p><strong>Cloud HPC instances, verified prices (on-demand, cheapest region):</strong></p>\n<ul>\n<li>AWS <strong>hpc6a.48xlarge</strong> (96 AMD cores): <span class=\"katex\"><span class=\"katex-mathml\"><math xmlns=\"http://www.w3.org/1998/Math/MathML\"><semantics><mrow><mn>2.88</mn><mi mathvariant=\"normal\">/</mi><mi>h</mi><mi>r</mi><mo>→</mo><mo>∗</mo><mo>∗</mo></mrow><annotation encoding=\"application/x-tex\">2.88/hr → **</annotation></semantics></math></span><span class=\"katex-html\" aria-hidden=\"true\"><span class=\"base\"><span class=\"strut\" style=\"height:1em;vertical-align:-0.25em;\"></span><span class=\"mord\">2.88/</span><span class=\"mord mathnormal\">h</span><span class=\"mord mathnormal\" style=\"margin-right:0.0278em;\">r</span><span class=\"mspace\" style=\"margin-right:0.2778em;\"></span><span class=\"mrel\">→</span><span class=\"mspace\" style=\"margin-right:0.2778em;\"></span></span><span class=\"base\"><span class=\"strut\" style=\"height:0.4653em;\"></span><span class=\"mord\">∗</span><span class=\"mord\">∗</span></span></span></span>0.030/core-hr** (<a href=\"https://cloudprice.net/aws/ec2/instances/hpc6a.48xlarge\">cloudprice.net</a>); hpc7g (Graviton3E ARM) exists and is cheaper <strong>[exact price unverified]</strong>.</li>\n<li>GCP <strong>h3-standard-88</strong> (Sapphire Rapids, HPC-only, 2 regions): ~<span class=\"katex\"><span class=\"katex-mathml\"><math xmlns=\"http://www.w3.org/1998/Math/MathML\"><semantics><mrow><mn>4.92</mn><mi mathvariant=\"normal\">/</mi><mi>h</mi><mi>r</mi><mo>→</mo><mtext> </mtext></mrow><annotation encoding=\"application/x-tex\">4.92/hr → ~</annotation></semantics></math></span><span class=\"katex-html\" aria-hidden=\"true\"><span class=\"base\"><span class=\"strut\" style=\"height:1em;vertical-align:-0.25em;\"></span><span class=\"mord\">4.92/</span><span class=\"mord mathnormal\">h</span><span class=\"mord mathnormal\" style=\"margin-right:0.0278em;\">r</span><span class=\"mspace\" style=\"margin-right:0.2778em;\"></span><span class=\"mrel\">→</span><span class=\"mspace nobreak\"> </span></span></span></span>0.056/vCPU-hr; <strong>c3-standard-88</strong> $4.44/hr us-central1 (<a href=\"https://www.devzero.io/instances/gcp/c3-standard-88\">devzero</a>).</li>\n<li>Azure <strong>HB176rs_v4</strong> (176 Zen4 cores): from <span class=\"katex\"><span class=\"katex-mathml\"><math xmlns=\"http://www.w3.org/1998/Math/MathML\"><semantics><mrow><mn>5</mn><mo separator=\"true\">,</mo><mn>256</mn><mi mathvariant=\"normal\">/</mi><mi>m</mi><mi>o</mi><mo>≈</mo></mrow><annotation encoding=\"application/x-tex\">5,256/mo ≈</annotation></semantics></math></span><span class=\"katex-html\" aria-hidden=\"true\"><span class=\"base\"><span class=\"strut\" style=\"height:1em;vertical-align:-0.25em;\"></span><span class=\"mord\">5</span><span class=\"mpunct\">,</span><span class=\"mspace\" style=\"margin-right:0.1667em;\"></span><span class=\"mord\">256/</span><span class=\"mord mathnormal\">m</span><span class=\"mord mathnormal\">o</span><span class=\"mspace\" style=\"margin-right:0.2778em;\"></span><span class=\"mrel\">≈</span></span></span></span>7.20/hr → ~$0.041/vCPU-hr (<a href=\"https://cloudprice.net/vm/Standard_HB176rs_v4\">cloudprice.net</a>). <strong>HBv5</strong> (custom EPYC 9V64H, up to 352 cores, 6.7 TB/s HBM3) reached <strong>GA in November 2025</strong> (<a href=\"https://www.phoronix.com/review/azure-hbv5-amd-epyc-9v64h\">Phoronix</a>); pricing <strong>[uncertain — portal only]</strong>.</li>\n</ul>\n<p><strong>When cloud beats an allocation.</strong> Reference point: ARCHER2&#39;s notional £0.20/node-hour ≈ £0.0016/core-hour — allocations are <del>10–20× cheaper per core-hour than any cloud; ACCESS/EuroHPC/DOE are free outright. Cloud wins when: (a) the workload is bursty or small (&lt;</del>100k core-hours), (b) you can&#39;t wait 3–9 months of review cycles, (c) you have credits making marginal cost $0, (d) you need elasticity (thousands of short jobs now, not queued), or (e) GPU inference/fine-tuning where big-3 managed tooling matters. Steady, large MPI workloads belong on allocations.</p>\n<h2 id=\"4-community-sharing-models\">4. Community / sharing models</h2><ul>\n<li><strong>OSG OSPool (PATh, NSF-funded)</strong> — the strongest current federated option for a US-affiliated team: free, fair-share, no allocation review; account + short consultation; HTCondor high-throughput model ideal for DFT screening ensembles and ML dataset generation (jobs &lt;~10 hrs, &lt;8 cores each); OSG-wide delivers &gt;2B core-hours/year; documented computational-chemistry use (70M-molecule dataset regeneration). <a href=\"https://osg-htc.org/services/ospool-registration.html\">OSPool registration</a>, <a href=\"https://osg-htc.org/spotlights/ospool-computation.html\">chemistry spotlight</a></li>\n<li><strong>Volunteer computing</strong>: BOINC ecosystem still active (<a href=\"https://arstechnica.com/civis/threads/formula-boinc-2025.1507290/\">project list/competition</a>); <strong>Folding@home</strong> (volunteer GPU molecular dynamics — same method family as MD, but run by the project, not usable as your cluster) and <strong>GPUGRID</strong> (BOINC, MD). Running your own BOINC project is possible but only pays off with a large volunteer base — not realistic for a small team.</li>\n<li><strong>World Community Grid</strong> — corporate-philanthropic volunteer grid; materials/chemistry projects have run on it historically (e.g., Harvard Clean Energy Project). Transferred from IBM to Krembil Research Institute in 2021 <strong>[current submission process uncertain]</strong>.</li>\n<li><strong>WLCG</strong> — heritage grid (~1M+ cores) but restricted to LHC collaborations; not accessible for materials work. <strong>EGI</strong> remains Europe&#39;s general research grid/cloud federation <strong>[current onboarding terms not verified this session]</strong>.</li>\n<li>Materials-specific federated compute doesn&#39;t exist at scale today; nearest things are ACCESS science gateways (<a href=\"https://sciencegateways.org/\">SGX3</a>) and data-sharing infrastructures (Materials Cloud/NOMAD), which share data, not cycles.</li>\n</ul>\n<h2 id=\"suggested-stack-for-a-small-atomistic-sim-ml-team\">Suggested stack for a small atomistic-sim + ML team</h2><ol>\n<li><strong>Today</strong>: ACCESS Explore (400k credits, ~2 days) + AWS/Google/Oracle research credits for GPU work.</li>\n<li><strong>Throughput</strong>: OSPool for screening ensembles; institutional cluster if available.</li>\n<li><strong>Scale-up</strong>: ACCESS Discover→Accelerate; EuroHPC Development→Regular if EU-affiliated; DOE Director&#39;s Discretionary → ALCC for exascale ambitions.</li>\n<li><strong>Burst/ML training</strong>: spot GPUs on CoreWeave/RunPod/Vast with checkpointing; reserved only for sustained GPU needs.</li>\n</ol>\n<p><strong>Uncertain items to verify before publishing</strong>: Google/Azure credit award amounts, Lambda and Vast live prices, Azure HBv5 pricing, AWS hpc7g price, China scnet practical access, WCG submission process, EGI onboarding, AWS p5 current on-demand price.</p>\n"}