{"id":138740,"date":"2026-10-06T14:17:04","date_gmt":"2026-10-06T14:17:04","guid":{"rendered":"\/uk\/tutorials\/cloud-gpu-pricing"},"modified":"2026-10-06T14:17:04","modified_gmt":"2026-10-06T14:17:04","slug":"cloud-gpu-pricing","status":"publish","type":"post","link":"\/uk\/tutorials\/cloud-gpu-pricing\/","title":{"rendered":"Cloud GPU pricing compared: what GPUs cost per hour in 2026"},"content":{"rendered":"<p class=\"wp-block-paragraph\">Cloud GPU pricing ranges from under $0.50 per hour for an entry-level card to over $12 for the most powerful ones. For the same GPU, the price can still vary by as much as sevenfold depending on which provider you rent from. That gap comes from how providers package the hardware, not from the chip itself.<\/p><p class=\"wp-block-paragraph\">What you pay depends heavily on how that access is structured: some providers rent you just the GPU; others bundle it inside a full server with dozens of processors and terabytes of memory, charge for everything, and require you to rent several GPUs at once.<\/p><p class=\"wp-block-paragraph\">The pricing model adds a second layer: on-demand, interruptible, or reserved access can halve or double the rate for the same card.<\/p><p class=\"wp-block-paragraph\">The advertised rate also doesn&rsquo;t capture the full cost: storage fees, data transfer charges, and idle GPUs running between jobs can push a real monthly bill well above what the hourly rate suggests.<\/p><h2 class=\"wp-block-heading h-t-title-2\" id=\"h-how-much-do-cloud-gpus-cost-per-hour\">How much do cloud GPUs cost per hour?<\/h2><p class=\"wp-block-paragraph\">Cloud GPUs cost between $0.14 and $12.29 per GPU-hour (the cost of renting one GPU for one hour) at on-demand (pay-as-you-go) list prices, as of September 2026.<\/p><p class=\"wp-block-paragraph\">The comparison table covers seven GPU models from seven providers across US regions. The VRAM column shows each GPU&rsquo;s on-chip memory, which determines how large a model the GPU can run.<\/p><figure tabindex=\"0\" class=\"wp-block-table\"><table><tbody><tr><td colspan=\"1\" rowspan=\"1\"><p><strong>GPU<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>VRAM<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Hostinger<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>RunPod<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Vast.ai<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Lambda<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>AWS<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>GCP<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Azure<\/strong><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>RTX 4090<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>24 GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0.38<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0.74<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>from <\/span><strong>$0.14<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>RTX A6000<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>48 GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0.53<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>from <\/span><strong>$0.28<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$1.09<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>RTX PRO 6000<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>96 GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0.60<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$2.09<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>from <\/span><strong>$0.89<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$4.50<\/strong><span>&dagger;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>L40S<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>48 GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0.92*<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0.99<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>from <\/span><strong>$0.47<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>A100 80GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>80 GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$1.43<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$1.39<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>from <\/span><strong>$0.44<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>from <\/span><strong>$1.99<\/strong><span>&sect;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>$3.43&dagger;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$3.67<\/strong><span>&sect;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$3.67<\/strong><span>&dagger;<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>H100<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>80 GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$2.89<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>from <\/span><strong>$1.33<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$3.29<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>$6.88<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$11.06<\/strong><span>&dagger;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$12.29<\/strong><span>&dagger;<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>B200<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>192 GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$4.50<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$6.79<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>from <\/span><strong>$4.38<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>from <\/span><strong>$6.69<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$16.11<\/strong><span>&Dagger;<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>&ndash;<\/span><\/p><\/td><\/tr><\/tbody><\/table><\/figure><p class=\"has-small-font-size wp-block-paragraph\"><em>On-demand rates, normalized per GPU-hour, US regions, as of September 2026. RunPod: Secure Cloud tier. Vast.ai: marketplace range; prices change daily. (&ndash;) = not offered.<\/em><\/p><p class=\"has-small-font-size wp-block-paragraph\"><em>&dagger; Whole-VM or minimum 8-GPU node; per-GPU rate derived from instance total.<\/em><\/p><p class=\"has-small-font-size wp-block-paragraph\"><em>&Dagger; Reservation required; no standard on-demand.<\/em><\/p><p class=\"has-small-font-size wp-block-paragraph\"><em>* Availability varies.<\/em><\/p><p class=\"has-small-font-size wp-block-paragraph\"><em>&sect; Lambda $1.99 and GCP $3.67 reflect A100 40GB variants (cheapest option). Lambda 80GB SXM is approximately $2.79; GCP a2-ultragpu 80GB is approximately $5.03.<\/em><\/p><p class=\"wp-block-paragraph\">Beyond the per-GPU rate, the tier you pick &ndash; shared community hardware or dedicated data-center capacity &ndash;, regional availability, and uptime commitments all affect the real cost of a workload.<a data-wpel-link=\"external\" href=\"\/gb\/tutorials\/best-gpu-cloud-providers\/\" target=\"_blank\" rel=\"nofollow noopener noreferrer\"> On those dimensions, <\/a><a href=\"\/gb\/tutorials\/best-gpu-cloud-providers\/\" target=\"_blank\" rel=\"noreferrer noopener nofollow\">comparing GPU cloud providers<\/a> often changes the final choice more than a small rate difference would.<\/p><p class=\"wp-block-paragraph\"><div class=\"announcement-block announcement-block--important\">\n            <span class=\"announcement-block__heading\">\n                <svg width=\"24\" height=\"24\" viewBox=\"0 0 24 24\" fill=\"none\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\">\n                    <path fill-rule=\"evenodd\" clip-rule=\"evenodd\"\n                          d=\"M12 22.5C17.799 22.5 22.5 17.799 22.5 12C22.5 6.20101 17.799 1.5 12 1.5C6.20101 1.5 1.5 6.20101 1.5 12C1.5 17.799 6.20101 22.5 12 22.5ZM13.637 7.65198C13.637 6.74791 12.9041 6.01501 12 6.01501C11.0959 6.01501 10.363 6.74791 10.363 7.65198C10.5335 9.53749 10.875 13.383 10.875 13.383C10.875 14.0043 11.3787 14.508 12 14.508C12.6213 14.508 13.125 14.0043 13.125 13.383V13.38L13.637 7.65198ZM11.9927 15.714C11.3714 15.714 10.8677 16.2177 10.8677 16.839C10.8677 17.4603 11.3714 17.964 11.9927 17.964H12.0073C12.6286 17.964 13.1323 17.4603 13.1323 16.839C13.1323 16.2177 12.6286 15.714 12.0073 15.714H11.9927Z\"\n                          fill=\"#FEA419\"\/>\n                <\/svg>\n                Important\n            <\/span>\n            <p class=\"announcement-block__content\">\n                 GPU cloud pricing updates monthly and sometimes weekly; marketplace platforms like Vast.ai can change hourly. The rates above were verified in September 2026 &ndash; re-check primary pricing pages before you deploy anything. \n            <\/p><\/div><\/p><h2 class=\"wp-block-heading h-t-title-2\" id=\"h-how-does-cloud-gpu-pricing-work\">How does cloud GPU pricing work?<\/h2><p class=\"wp-block-paragraph\">Cloud GPU pricing runs on three models: on-demand pay-as-you-go, interruptible spot capacity at a significant discount, and reserved terms for predictable long-running workloads.<\/p><h3 class=\"wp-block-heading h-t-title-3\">On-demand<\/h3><p class=\"wp-block-paragraph\">On-demand gives you immediate access at a fixed rate with no commitment. You pay for the whole time the instance (your running virtual server) is running, whether the GPU is processing a job or sitting idle. It&rsquo;s the default model for variable or short-term workloads and the basis for every rate in the comparison table.<\/p><h3 class=\"wp-block-heading h-t-title-3\">Spot and interruptible<\/h3><p class=\"wp-block-paragraph\">Spot pricing cuts costs from 50 to 91 percent depending on the provider, but the instance can be reclaimed with little notice when the provider needs the capacity back. Google Cloud Platform (GCP) gives 30 seconds of warning; Amazon Web Services (AWS) gives two minutes.<\/p><p class=\"wp-block-paragraph\">Spot only makes sense with proper checkpointing: saving your job&rsquo;s progress at intervals so it can resume from the last save point after an interruption, rather than from scratch. A job with no checkpoints that gets evicted mid-run can end up costing more than an on-demand run would have.<\/p><p class=\"wp-block-paragraph\">        <div class=\"protip\">\n            <div class=\"protip__heading\">\n                <svg width=\"24\" height=\"24\" viewBox=\"0 0 24 24\" fill=\"none\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\">\n                    <path d=\"M1.49234 23.5024C1.23229 23.5024 0.972242 23.4024 0.782206 23.2123C0.562165 22.9923 0.452144 22.6822 0.502153 22.3722C0.562165 21.9221 1.14227 17.9113 3.00262 16.351C3.63274 15.8209 4.43289 15.5509 5.26305 15.5609C6.09321 15.5909 6.87335 15.9109 7.47347 16.4911C8.6937 17.6913 8.76371 19.6717 7.6435 20.9919C6.0832 22.8523 2.08245 23.4324 1.63237 23.4924C1.59236 23.4924 1.54235 23.4924 1.50234 23.4924L1.49234 23.5024ZM5.16303 17.5613C4.84297 17.5613 4.53291 17.6713 4.29287 17.8813C3.60274 18.4614 3.07264 19.9317 2.75258 21.242C4.06282 20.9219 5.5331 20.3918 6.11321 19.7017C6.55329 19.1716 6.54329 18.3814 6.0832 17.9213C5.85316 17.7013 5.5431 17.5713 5.20304 17.5613C5.19304 17.5613 5.17303 17.5613 5.16303 17.5613ZM11.7243 21.8821C11.4942 21.8821 11.2642 21.8021 11.0841 21.652C10.8541 21.462 10.7241 21.1819 10.7241 20.8819V15.9109L8.08358 13.2705H3.11264C2.81259 13.2705 2.53254 13.1404 2.3425 12.9104C2.15246 12.6803 2.07245 12.3803 2.12246 12.0902C2.19247 11.7102 2.84259 8.36953 4.70294 7.12929C6.33325 6.04909 8.96375 6.49918 10.244 6.80923C11.5442 4.96889 13.2546 3.4286 15.2349 2.33839C17.4553 1.11816 19.9858 0.518051 22.4963 0.498047C23.0464 0.498047 23.4865 0.948132 23.4865 1.49824C23.4865 5.0389 22.3763 9.97983 17.1753 13.7605C17.4853 15.0408 17.9354 17.6613 16.8552 19.2816C15.615 21.1419 12.2744 21.7921 11.8943 21.8621C11.8343 21.8721 11.7743 21.8821 11.7143 21.8821H11.7243ZM12.7245 16.181V19.6016C13.7146 19.2916 14.7948 18.7915 15.2049 18.1814C15.675 17.4812 15.605 16.091 15.385 14.9008C14.5248 15.3808 13.6346 15.8109 12.7245 16.181ZM9.66388 12.0302L11.9643 14.3307C13.1845 13.8306 14.3648 13.2204 15.485 12.5103C19.9358 9.51974 21.2361 5.60901 21.4561 2.53843C19.6157 2.67846 17.8254 3.20856 16.2051 4.09872C14.2847 5.14892 12.6544 6.68921 11.4942 8.54956C10.7841 9.65977 10.174 10.82 9.66388 12.0302ZM4.39289 11.2701H7.81353C8.1936 10.3599 8.63368 9.46974 9.11377 8.60957C7.92355 8.38953 6.51329 8.31952 5.81315 8.78961C5.19304 9.19968 4.70294 10.3099 4.39289 11.2701Z\" fill=\"#673DE6\"\/>\n                <\/svg>\n                <p class=\"protip__title\">\n                    Pro tip                <\/p>\n            <\/div>\n            <p class=\"protip__content\"> Spot prices are market rates, not guaranteed discounts. On bid-based platforms, a supply crunch can push the spot rate close to or above on-demand. Always check the live spot rate before assuming it's cheaper than on-demand. <\/p>\n                    <\/div>\n        <\/p><figure tabindex=\"0\" class=\"wp-block-table\"><table><tbody><tr><td colspan=\"1\" rowspan=\"1\"><p><strong>Provider<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Spot\/interruptible model<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Typical discount<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Notes<\/strong><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>AWS<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Spot Instances<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>~57%<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>~$2.96\/GPU-hr for H100 (p5.48xlarge); fluctuates<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>GCP<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Spot VMs<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>60&ndash;91%<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>30-second eviction notice<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Azure<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Spot VMs<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>~80%<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>30-second eviction notice<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>RunPod<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Community Cloud<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>~50%<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>No published spot rate card<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Vast.ai<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Interruptible<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>50%+<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Host-set; varies daily<\/span><\/p><\/td><\/tr><\/tbody><\/table><\/figure><h3 class=\"wp-block-heading h-t-title-3\">Reserved and committed<\/h3><p class=\"wp-block-paragraph\">Reserved pricing locks in a GPU for a fixed term in exchange for a lower effective hourly rate. Discounts range from 25 to 65 percent, depending on the provider and commitment length.<\/p><figure tabindex=\"0\" class=\"wp-block-table\"><table><tbody><tr><td colspan=\"1\" rowspan=\"1\"><p><strong>Provider<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Discount<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Term options<\/strong><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>AWS<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>25&ndash;45%<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Capacity Blocks (variable), Savings Plans<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>GCP<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>up to 65%<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>1 or 3 years (Committed Use Discounts)<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Azure<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>up to 63%<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>1 or 3 years; 5 years on ND H100 v5 (Reserved Instances)<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Vast.ai<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Up to 50%<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>1, 3, or 6 months<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Lambda<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>1-Click Clusters from $5.54\/GPU\/hr<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>2 weeks &ndash; 1 year; 16+ GPUs minimum; 1yr+ contact sales<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>RunPod<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Contact sales<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Enterprise agreements<\/span><\/p><\/td><\/tr><\/tbody><\/table><\/figure><h3 class=\"wp-block-heading h-t-title-3\">Billing granularity<\/h3><p class=\"wp-block-paragraph\">Providers charge in different units, and the difference matters most for short or interrupted jobs. RunPod and Vast.ai bill per second; Lambda bills per-minute; AWS, GCP, and Azure bill per second with a one-minute minimum. A 16-minute job on RunPod costs roughly 27 percent of the hourly rate; on Lambda it costs the full hour.<\/p><p class=\"wp-block-paragraph\">Hostinger bills hourly using a prepaid credit system (1 credit = $0.01), with the first full hour deducted at the moment of deployment. Destroying an instance mid-hour does not return unused time.<\/p><h2 class=\"wp-block-heading h-t-title-2\" id=\"h-cloud-gpu-pricing-by-model\">Cloud GPU pricing by model<\/h2><p class=\"wp-block-paragraph\">Cloud GPU models fall into several pricing tiers, from under $0.14\/hr for the RTX 4090 to $6.79\/hr for the B200 on standard on-demand &ndash; and higher still at hyperscalers, where a reserved B200 runs $16.11\/hr. Each tier offers different VRAM capacity and suits different jobs.<\/p><h3 class=\"wp-block-heading h-t-title-3\">RTX 4090 pricing per hour<\/h3><p class=\"wp-block-paragraph\">The RTX 4090 costs from <strong>$0.14<\/strong> per hour on Vast.ai and <strong>$0.38<\/strong> per hour on Hostinger, as of September 2026. RunPod lists it at <strong>$0.34<\/strong> on Community Cloud and <strong>$0.74<\/strong> on Secure Cloud.<\/p><p class=\"wp-block-paragraph\">Community Cloud uses hardware contributed by independent hosts that RunPod has screened; Secure Cloud uses data-center-grade infrastructure with higher reliability guarantees.<\/p><p class=\"wp-block-paragraph\">No hyperscaler carries the RTX 4090, since GPUs originally built for gaming and workstation use aren&rsquo;t available on AWS, GCP, or Azure.<\/p><p class=\"wp-block-paragraph\">With 24 GB of VRAM, the 4090 handles most 7-billion-parameter models at full 16-bit precision and up to roughly 30 billion parameters at 4-bit quantization, a compression method that reduces memory usage at a small cost to accuracy.<\/p><p class=\"wp-block-paragraph\">Ollama, a free tool for running open-source AI models on cloud or local hardware, is a natural fit for this card.<a data-wpel-link=\"external\" href=\"\/gb\/tutorials\/ollama-gpu-requirements\/\" target=\"_blank\" rel=\"nofollow noopener noreferrer\"> A<\/a>t the 24 GB tier, <a href=\"\/gb\/tutorials\/ollama-gpu-requirements\/\" target=\"_blank\" rel=\"noreferrer noopener nofollow\">choosing a GPU for Ollama<\/a> means most popular open-weight models run without additional configuration.<\/p><p class=\"wp-block-paragraph\">The 4090 is also a popular choice for Stable Diffusion (an open-source image generation model) and ComfyUI (a visual workflow tool for running it), where generating images quickly matters more than raw training speed.<\/p><h3 class=\"wp-block-heading h-t-title-3\">RTX A6000 and RTX PRO 6000 pricing<\/h3><p class=\"wp-block-paragraph\">The RTX A6000 starts at <strong>$0.28<\/strong> per hour on Vast.ai, <strong>$0.53<\/strong> on RunPod, and <strong>$1.09<\/strong> on Lambda, as of September 2026. Built on NVIDIA&rsquo;s older Ampere architecture with 48 GB of VRAM, it keeps NVLink, which lets multiple GPUs share memory directly &ndash; a feature the newer L40S drops.<\/p><p class=\"wp-block-paragraph\">The RTX PRO 6000 (96 GB, built on NVIDIA&rsquo;s newer Blackwell architecture) is its successor. Hostinger lists it at <strong>$0.60<\/strong> per hour, while RunPod charges $2.09. GCP&rsquo;s G4 series bundles the same card with 48 vCPUs (virtual CPU cores) and 180 GB of RAM into a single virtual machine from approximately <strong>$4.50<\/strong> per hour.<\/p><p class=\"wp-block-paragraph\">The memory difference is the decision point. At 48 GB, the A6000 can hold a model of roughly 20 billion parameters at full FP16 precision (16-bit floating point, the standard format for full-accuracy AI model weights) once you leave room for the context window; larger models require quantization to fit.<\/p><p class=\"wp-block-paragraph\">With 96 GB, the PRO 6000 can handle most 70-billion-parameter models on a single card at 4-bit quantization, without needing to link multiple GPUs.<\/p><h3 class=\"wp-block-heading h-t-title-3\">L40S pricing per hour<\/h3><p class=\"wp-block-paragraph\">The L40S costs <strong>$0.92<\/strong> per hour on Hostinger and <strong>$0.99<\/strong> on RunPod Secure Cloud, as of September 2026. It&rsquo;s built on NVIDIA&rsquo;s Ada Lovelace architecture with 48 GB of VRAM and native FP8 support &ndash; FP8 is an 8-bit number format that reduces memory usage and can speed up AI inference workloads.<\/p><p class=\"wp-block-paragraph\">That FP8 support makes it the better option for inference (running a trained AI model to generate outputs) in the 48 GB tier compared to the older A6000.<\/p><p class=\"wp-block-paragraph\">GCP&rsquo;s G2 series offers the L4 (24 GB) from <strong>$0.71<\/strong> per hour. The L4 and L40S share the same Ada Lovelace architecture, but the L4 is a low-power inference card with half the VRAM and a fraction of the throughput &ndash; it draws 72 W against the L40S&rsquo;s 350 W. If both appear in GCP&rsquo;s pricing for an inference workload, the L4 is not a cheaper L40S: it is a much smaller card.<\/p><p class=\"wp-block-paragraph\"><div class=\"announcement-block announcement-block--important\">\n            <span class=\"announcement-block__heading\">\n                <svg width=\"24\" height=\"24\" viewBox=\"0 0 24 24\" fill=\"none\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\">\n                    <path fill-rule=\"evenodd\" clip-rule=\"evenodd\"\n                          d=\"M12 22.5C17.799 22.5 22.5 17.799 22.5 12C22.5 6.20101 17.799 1.5 12 1.5C6.20101 1.5 1.5 6.20101 1.5 12C1.5 17.799 6.20101 22.5 12 22.5ZM13.637 7.65198C13.637 6.74791 12.9041 6.01501 12 6.01501C11.0959 6.01501 10.363 6.74791 10.363 7.65198C10.5335 9.53749 10.875 13.383 10.875 13.383C10.875 14.0043 11.3787 14.508 12 14.508C12.6213 14.508 13.125 14.0043 13.125 13.383V13.38L13.637 7.65198ZM11.9927 15.714C11.3714 15.714 10.8677 16.2177 10.8677 16.839C10.8677 17.4603 11.3714 17.964 11.9927 17.964H12.0073C12.6286 17.964 13.1323 17.4603 13.1323 16.839C13.1323 16.2177 12.6286 15.714 12.0073 15.714H11.9927Z\"\n                          fill=\"#FEA419\"\/>\n                <\/svg>\n                Important\n            <\/span>\n            <p class=\"announcement-block__content\">\n                 L40S availability on Hostinger fluctuates. If the card doesn&rsquo;t appear in the deploy picker, it&rsquo;s simply out of stock, not unavailable in your region. Hostinger shows live inventory before you commit to a deploy, so what you see is what&rsquo;s actually available. \n            <\/p><\/div><\/p><h3 class=\"wp-block-heading h-t-title-3\">A100 80GB pricing per hour<\/h3><p class=\"wp-block-paragraph\">The A100 80GB starts at <strong>$1.39<\/strong> per hour on RunPod and <strong>$1.43<\/strong> per hour on Hostinger, rising to <strong>$3.67<\/strong>&ndash;<strong>$5.03<\/strong> per hour at hyperscalers.<\/p><p class=\"wp-block-paragraph\">RunPod&rsquo;s $1.39 rate is the PCIe variant; the SXM variant (faster for multi-GPU training) is <strong>$1.59<\/strong>. At hyperscalers, the gap widens sharply: Azure&rsquo;s NC24ads A100 v4 is <strong>$3.67<\/strong> per hour for a single-GPU VM; the per-GPU rate on AWS&rsquo;s p4de (8&times; A100 80GB) works out to roughly <strong>$3.43<\/strong>, and GCP&rsquo;s a2-highgpu-1g (A100 40GB) runs <strong>$3.67<\/strong>.<\/p><p class=\"wp-block-paragraph\">The A100&rsquo;s 80 GB of high-bandwidth memory (HBM2e) handles fine-tuning (adapting an existing model for a specific task) and inference for most models up to 70 billion parameters in quantized form. At the specialist-cloud rate of $1.39&ndash;1.43 per hour, it&rsquo;s the default step up from the 48 GB tier for training runs that need more memory.<\/p><h3 class=\"wp-block-heading h-t-title-3\">H100 pricing per hour<\/h3><p class=\"wp-block-paragraph\">An H100 costs <strong>$2.89<\/strong> per hour on RunPod and up to <strong>$12.29<\/strong> on Azure, as of September 2026. That is a more than 4x spread between the cheapest specialist cloud and the most expensive hyperscaler for the same chip. Hostinger doesn&rsquo;t offer an H100; the B200 covers its top tier.<\/p><ul class=\"wp-block-list\">\n<li><strong>RunPod:<\/strong> <strong>$2.89<\/strong> (PCIe); <strong>$3.29<\/strong> (SXM); <strong>$3.19<\/strong> (NVL).<\/li>\n\n\n\n<li><strong>Lambda:<\/strong> <strong>$3.29<\/strong> (PCIe); from <strong>$3.99<\/strong> (SXM; cheapest per-GPU at the 8-GPU config, single GPU available at $4.29).<\/li>\n\n\n\n<li><strong>Nebius:<\/strong> <strong>$3.85.<\/strong><\/li>\n\n\n\n<li><strong>DigitalOcean:<\/strong> <strong>$4.41.<\/strong><\/li>\n\n\n\n<li><strong>AWS p5.4xlarge:<\/strong> <strong>$6.88<\/strong> (single H100 SXM5; no minimum node).<\/li>\n\n\n\n<li><strong>Azure NC H100 v5:<\/strong> <strong>$6.98<\/strong> (NVL form factor, single GPU).<\/li>\n\n\n\n<li><strong>GCP A3 (a3-highgpu-8g):<\/strong> from <strong>$11.06<\/strong> per GPU (8-GPU minimum on-demand).<\/li>\n\n\n\n<li><strong>Azure ND H100 v5:<\/strong> <strong>$12.29<\/strong> per GPU (<strong>$98.32<\/strong> for the 8-GPU node).<\/li>\n<\/ul><p class=\"wp-block-paragraph\">PCIe, SXM, and NVL are the form factors the H100 ships in. PCIe fits a standard server slot; SXM uses NVIDIA&rsquo;s proprietary high-power socket with NVLink interconnects that let multiple GPUs share memory directly, which matters for distributed training (splitting a job across several GPUs in parallel). NVL pairs two cards and carries more memory per GPU.<\/p><p class=\"wp-block-paragraph\">The SXM premium at specialist clouds is roughly 3&ndash;10%. For single-card inference, PCIe is the cheaper option with no performance difference at that scale.<\/p><p class=\"wp-block-paragraph\">GCP&rsquo;s A3 on-demand requires a minimum of 8 GPUs per node &ndash; if you need two H100s, you still pay for eight. Lambda sells H100 SXM in 1&times;, 2&times;, 4&times;, and 8&times; configurations.<\/p><h3 class=\"wp-block-heading h-t-title-3\">B200 and H200 pricing<\/h3><p class=\"wp-block-paragraph\">The B200 costs <strong>$4.50<\/strong> per hour on Hostinger, <strong>$6.79<\/strong> on RunPod, from <strong>$4.38<\/strong> on Vast.ai, and from <strong>$6.69<\/strong> on Lambda. The H200 starts at <strong>$3.87<\/strong> on Vast.ai and <strong>$4.59<\/strong> on RunPod. GCP&rsquo;s A4 series starts around <strong>$16.11<\/strong> per GPU-hour but requires a reservation. There&rsquo;s no standard on-demand option.<\/p><p class=\"wp-block-paragraph\">The H200 (141 GB HBM3e, a high-bandwidth memory type) carries roughly 75% more VRAM than the H100; RunPod lists it at <strong>$4.59<\/strong>. GCP&rsquo;s A3 Ultra and Azure&rsquo;s ND H200 v5 also require reservations and aren&rsquo;t available on standard on-demand.<\/p><p class=\"wp-block-paragraph\">Both cards sit above the H100, but in different ways: the B200 is a new architecture with higher throughput, while the H200 keeps the H100&rsquo;s compute and adds memory. As these cards become more widely available, H100 on-demand rates are likely to keep falling. A100 pricing has already dropped below $1 per hour on some marketplace providers as the market shifts toward newer GPU generations.<\/p><h2 class=\"wp-block-heading h-t-title-2\" id=\"h-why-do-hyperscalers-charge-more-for-the-same-gpu\">Why do hyperscalers charge more for the same GPU?<\/h2><p class=\"wp-block-paragraph\">Hyperscalers (AWS, Google Cloud and Azure) charge significantly more than specialist GPU clouds for on-demand access because they rent you a whole virtual machine, not just the GPU. That package bundles compute, memory, and storage, often with a minimum number of GPUs per rental, plus the compliance certifications and support tooling most AI workloads don&rsquo;t need.<\/p><p class=\"wp-block-paragraph\"><strong>Forced bundling.<\/strong> AWS&rsquo;s p5.4xlarge pairs one H100 with 16 vCPUs (virtual CPU cores) and 256 GB of system RAM. Azure&rsquo;s ND H100 v5 bundles eight GPUs with 96 vCPUs and 1.9 TB of RAM. You can&rsquo;t rent the GPU on its own &ndash; the processors and memory come attached in fixed ratios, whether you need them or not.<\/p><p class=\"wp-block-paragraph\">On Hostinger, the same B200 GPU is available in tiers from $4.50 to $6.00\/hr, a 33% spread from smallest to largest.<\/p><p class=\"wp-block-paragraph\"><strong>Quota lead time.<\/strong> AWS and GCP both cap how many high-end GPU instances a new account can launch &ndash; these limits are called quotas, and they default to zero for GPU instances.<\/p><p class=\"wp-block-paragraph\">Getting AWS P5 access requires a written justification and typically takes 3&ndash;7 business days for approval; GCP&rsquo;s A3 H100 quota can take longer. That wait is a real cost: teams blocked on quota often end up renting more capacity than they need on whatever they can access immediately..<\/p><p class=\"wp-block-paragraph\"><strong>What hyperscalers give you in return.<\/strong> They run high-speed networking between GPU nodes, up to 3.2 Tbps (terabits per second) on AWS P5 instances, which matters for distributed training across dozens of GPUs. Enterprise compliance certifications are built in.<\/p><p class=\"wp-block-paragraph\">Regional coverage spans every major geography. For regulated industries or workloads that need 64+ GPU clusters, the hyperscaler premium buys infrastructure you can&rsquo;t assemble elsewhere.<\/p><p class=\"wp-block-paragraph\">For developers running inference, fine-tuning, or image generation, most of that bundle is unnecessary. Specialist GPU clouds strip it out, and<a aria-label=\"Runpod alternatives\" data-wpel-link=\"external\" href=\"\/gb\/tutorials\/runpod-alternatives\/\" target=\"_blank\" rel=\"nofollow noopener noreferrer\"> <\/a><a href=\"\/gb\/tutorials\/runpod-alternatives\/\" target=\"_blank\" rel=\"noreferrer noopener nofollow\">alternatives to RunPod for GPU workloads<\/a> in the same tier typically come without quota approval or minimum node sizes.<\/p><p class=\"wp-block-paragraph\">The price gap is most visible when you only need one GPU, where specialist clouds consistently undercut hyperscaler rates.<\/p><h2 class=\"wp-block-heading h-t-title-2\" id=\"h-what-you-actually-pay-for-cloud-gpus\">What you actually pay for cloud GPUs<\/h2><p class=\"wp-block-paragraph\">The hourly GPU rate is only part of the cost: storage, data transfer fees, and idle time can push the real bill up by 20&ndash;40%.<\/p><h3 class=\"wp-block-heading h-t-title-3\">Data transfer costs<\/h3><p class=\"wp-block-paragraph\">Moving data out of a provider&rsquo;s network is free on some platforms and costs up to $0.12 per GB on others, which adds up fast on large workloads. Providers call this egress rate.<\/p><figure tabindex=\"0\" class=\"wp-block-table\"><table><tbody><tr><td colspan=\"1\" rowspan=\"1\"><p><strong>Provider<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Egress rate<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Notes<\/strong><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>AWS<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0.09\/GB<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>First 100 GB free per month<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>GCP<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0.08&ndash;0.12\/GB<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Varies by destination and volume<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Azure<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>~$0.087\/GB<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>First 5 TB from East US<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Vast.ai<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Host-set<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Charged per byte, both directions<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Hostinger<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>No egress fees<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>RunPod (Pods)<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Zero egress on Pod workloads<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Lambda<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>No egress fees<\/span><\/p><\/td><\/tr><\/tbody><\/table><\/figure><p class=\"wp-block-paragraph\">The difference scales with volume: moving a 140 GB Llama-70B model checkpoint out of AWS costs around $12.60, while moving a 10 TB training dataset costs roughly $900. On Hostinger, RunPod Pods, or Lambda, the same transfers cost nothing.<\/p><h3 class=\"wp-block-heading h-t-title-3\">Storage costs<\/h3><p class=\"wp-block-paragraph\">Cloud GPU storage typically runs $0.04&ndash;0.20\/GB\/month, and on most platforms it keeps billing even after you stop your instance.<\/p><ul class=\"wp-block-list\">\n<li><strong>RunPod.<\/strong> <strong>$0.10\/GB\/month<\/strong> while running, $0.20\/GB\/month when stopped. A 200 GB volume left idle for a month costs $40.<\/li>\n\n\n\n<li><strong>Vast.ai.<\/strong> Rates are host-set and vary by listing. Storage bills continuously even when stopped, typically at a higher rate than when running. Deleting the instance is the only way to stop the charge.<\/li>\n\n\n\n<li><strong>Lambda.<\/strong> <strong>$0.20\/GB\/month<\/strong> for persistent file storage, which bills even with no GPU instance attached.<\/li>\n\n\n\n<li><strong>AWS.<\/strong> <strong>$0.08\/GB\/month<\/strong> (EBS), billed separately on top of compute.<\/li>\n\n\n\n<li><strong>GCP.<\/strong> <strong>$0.04&ndash;0.17\/GB\/month<\/strong> (standard to SSD persistent disk), billed separately on top of compute.<\/li>\n\n\n\n<li><strong>Azure.<\/strong> <strong>$0.17\/GB\/month<\/strong> (Premium SSD), billed separately on top of compute.<\/li>\n\n\n\n<li><strong>Hostinger.<\/strong> Included with each instance tier. No separate storage charge.<\/li>\n<\/ul><h3 class=\"wp-block-heading h-t-title-3\">Idle GPU costs<\/h3><p class=\"wp-block-paragraph\">Idle GPUs bill at the full hourly rate, the same as when they are running a job. A GPU left running after a job finishes costs exactly as much as one that was computing the whole time.<\/p><p class=\"wp-block-paragraph\">This is true of every provider: a session forgotten overnight costs a full night&rsquo;s compute everywhere. Billing granularity only changes the rounding at the edges. Hostinger deducts the full first hour at deploy, so a 10-minute session costs the same as a 60-minute one. On RunPod or Vast.ai, that same 10-minute session costs exactly 10 minutes.<\/p><h2 class=\"wp-block-heading h-t-title-2\" id=\"h-how-much-does-hostinger-gpu-hosting-cost\">How much does Hostinger GPU hosting cost?<\/h2><p class=\"wp-block-paragraph\">Hostinger GPU hosting starts at <strong>$0.38<\/strong> per hour for an RTX 4090 and reaches <strong>$7.08<\/strong> per hour for a dedicated B200, billed hourly in credits with no egress fees, as of September 2026.<\/p><figure tabindex=\"0\" class=\"wp-block-table\"><table><tbody><tr><td colspan=\"1\" rowspan=\"1\"><p><strong>GPU<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>VRAM<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>From<\/strong><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>RTX 4090<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>24 GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0.38<\/strong><span>\/hr<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>RTX PRO 6000 (Server)<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>96 GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0.60<\/strong><span>\/hr<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>L40S<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>48 GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$0.92<\/strong><span>\/hr<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>A100 80GB PCIe<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>80 GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$1.43<\/strong><span>\/hr<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>B200<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>192 GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$4.50<\/strong><span>\/hr<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>B200 (Dedicated)<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>192 GB<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$7.08<\/strong><span>\/hr<\/span><\/p><\/td><\/tr><\/tbody><\/table><\/figure><p class=\"wp-block-paragraph\">Every instance gets a dedicated NVIDIA GPU &ndash; the card isn&rsquo;t shared between customers &ndash; with full root and SSH access. The separate B200 (Dedicated) tier goes further, giving you the whole underlying server rather than a share of it. One-click apps for Ollama, ComfyUI, and Jupyter are available at deploy time alongside a base Ubuntu 24.04 option. Deployment takes a few minutes once you select a GPU and instance size.<\/p><p class=\"wp-block-paragraph\">Each GPU comes in multiple size tiers. The B200&rsquo;s base tier (Nano) gives you 1 core, 1 GB RAM, and 10 GB storage at 450 credits per hour ($4.50). The largest shared tier (32 cores, 64 GB RAM, 500 GB storage) costs 600 credits per hour ($6.00). The spread across the shared tiers is 33 percent.<\/p><p class=\"wp-block-paragraph\">Hostinger GPU hosting uses credits purchased in advance. Three pack sizes are available:<\/p><ul class=\"wp-block-list\">\n<li>500 credits. <strong>$5.00<\/strong> ($0.0100\/credit, base rate).<\/li>\n\n\n\n<li>2,200 credits. <strong>$20.00<\/strong> ($0.0091\/credit).<\/li>\n\n\n\n<li>6,000 credits. <strong>$50.00<\/strong> ($0.0083\/credit).<\/li>\n<\/ul><p class=\"wp-block-paragraph\">The full first hour is deducted at deploy. Destroying an instance mid-hour does not return the unused portion.<\/p><p class=\"wp-block-paragraph\">Hostinger GPU hosting is US-only (Los Angeles and Dallas, region depending on the GPU) and on-demand only. No spot or reserved tiers exist. There is no GPU switching after deploy; changing to a different card requires a new instance and a data migration.<\/p><figure class=\"wp-block-image size-large\"><a aria-label=\"Vps hosting\" href=\"\/uk\/vps-hosting\" target=\"_blank\" rel=\"noreferrer noopener\"><img decoding=\"async\" width=\"1024\" height=\"300\" src=\"https:\/\/imagedelivery.net\/LqiWLm-3MGbYHtFuUbcBtA\/wp-content\/uploads\/sites\/2\/2023\/02\/VPS-hosting-banner.png\/w=1024,h=1024,fit=scale-down\" alt=\"\" class=\"wp-image-77934\" srcset=\"https:\/\/www.hostinger.com\/uk\/tutorials\/wp-content\/uploads\/sites\/51\/2023\/02\/VPS-hosting-banner.png 1024w, https:\/\/www.hostinger.com\/uk\/tutorials\/wp-content\/uploads\/sites\/51\/2023\/02\/VPS-hosting-banner-300x88.png 300w, https:\/\/www.hostinger.com\/uk\/tutorials\/wp-content\/uploads\/sites\/51\/2023\/02\/VPS-hosting-banner-150x44.png 150w, https:\/\/www.hostinger.com\/uk\/tutorials\/wp-content\/uploads\/sites\/51\/2023\/02\/VPS-hosting-banner-768x225.png 768w\" sizes=\"(max-width: 1024px) 100vw, 1024px\" \/><\/a><\/figure><h2 class=\"wp-block-heading h-t-title-2\" id=\"h-how-to-estimate-your-monthly-cloud-gpu-bill\">How to estimate your monthly cloud GPU bill<\/h2><p class=\"wp-block-paragraph\">To estimate a monthly cloud GPU bill, multiply your planned hours by the hourly rate, then add storage costs and egress fees. On hyperscalers both are billed separately and grow with the amount of data you move; on Hostinger they add nothing, while RunPod and Lambda charge for storage but not egress.<\/p><p class=\"wp-block-paragraph\">This example uses a single A100 80GB for LLM (Large Language Model) inference using a model such as Llama or Mistral, 8 hours per day over 30 days (240 GPU-hours), with 10 GB of model weights in storage and 200 GB of data transferred out per month.<\/p><figure tabindex=\"0\" class=\"wp-block-table\"><table><tbody><tr><td colspan=\"1\" rowspan=\"1\"><p><strong>Cost component<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Hostinger<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>RunPod (Secure)<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Azure (NC24ads)<\/strong><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Compute (240 hr)<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>240 &times; $1.43 = <\/span><strong>$343.20<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>240 &times; $1.39 = <\/span><strong>$333.60<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>240 &times; $3.67 = <\/span><strong>$880.80<\/strong><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Storage (10 GB\/mo)<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Included<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>10 GB &times; $0.07 = <\/span><strong>$0.70<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>10 GB &times; $0.17 = <\/span><strong>$1.70<\/strong><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Egress (200 GB\/mo)<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>$0<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>$0<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>200 GB &times; $0.087 = <\/span><strong>$17.40<\/strong><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><strong>Monthly total<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$343.20<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$334.30<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>$899.90<\/strong><\/p><\/td><\/tr><\/tbody><\/table><\/figure><p class=\"wp-block-paragraph\">At 8 hours per day, Hostinger and RunPod land within $9 of each other. Azure costs 2.7 times more, with the compute rate driving almost all the difference.<\/p><p class=\"wp-block-paragraph\">Applying Hostinger&rsquo;s $50 credit pack (effective $0.0083\/credit) reduces the Hostinger total to approximately $285. At that rate, the gap over RunPod widens in Hostinger&rsquo;s favor for this workload.<\/p><p class=\"wp-block-paragraph\">For production LLM inference, how you size the GPU to the model matters as much as which provider you pick.<\/p><p class=\"wp-block-paragraph\"><a href=\"\/gb\/tutorials\/deploy-llm-with-vllm\/\" target=\"_blank\" rel=\"noreferrer noopener nofollow\">Deploying LLMs on GPU infrastructure<\/a> with vLLM requires matching GPU memory to model size, which determines whether an A100 or a smaller card is the right choice for a given workload.<\/p><p class=\"wp-block-paragraph\">        <div class=\"protip\">\n            <div class=\"protip__heading\">\n                <svg width=\"24\" height=\"24\" viewBox=\"0 0 24 24\" fill=\"none\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\">\n                    <path d=\"M1.49234 23.5024C1.23229 23.5024 0.972242 23.4024 0.782206 23.2123C0.562165 22.9923 0.452144 22.6822 0.502153 22.3722C0.562165 21.9221 1.14227 17.9113 3.00262 16.351C3.63274 15.8209 4.43289 15.5509 5.26305 15.5609C6.09321 15.5909 6.87335 15.9109 7.47347 16.4911C8.6937 17.6913 8.76371 19.6717 7.6435 20.9919C6.0832 22.8523 2.08245 23.4324 1.63237 23.4924C1.59236 23.4924 1.54235 23.4924 1.50234 23.4924L1.49234 23.5024ZM5.16303 17.5613C4.84297 17.5613 4.53291 17.6713 4.29287 17.8813C3.60274 18.4614 3.07264 19.9317 2.75258 21.242C4.06282 20.9219 5.5331 20.3918 6.11321 19.7017C6.55329 19.1716 6.54329 18.3814 6.0832 17.9213C5.85316 17.7013 5.5431 17.5713 5.20304 17.5613C5.19304 17.5613 5.17303 17.5613 5.16303 17.5613ZM11.7243 21.8821C11.4942 21.8821 11.2642 21.8021 11.0841 21.652C10.8541 21.462 10.7241 21.1819 10.7241 20.8819V15.9109L8.08358 13.2705H3.11264C2.81259 13.2705 2.53254 13.1404 2.3425 12.9104C2.15246 12.6803 2.07245 12.3803 2.12246 12.0902C2.19247 11.7102 2.84259 8.36953 4.70294 7.12929C6.33325 6.04909 8.96375 6.49918 10.244 6.80923C11.5442 4.96889 13.2546 3.4286 15.2349 2.33839C17.4553 1.11816 19.9858 0.518051 22.4963 0.498047C23.0464 0.498047 23.4865 0.948132 23.4865 1.49824C23.4865 5.0389 22.3763 9.97983 17.1753 13.7605C17.4853 15.0408 17.9354 17.6613 16.8552 19.2816C15.615 21.1419 12.2744 21.7921 11.8943 21.8621C11.8343 21.8721 11.7743 21.8821 11.7143 21.8821H11.7243ZM12.7245 16.181V19.6016C13.7146 19.2916 14.7948 18.7915 15.2049 18.1814C15.675 17.4812 15.605 16.091 15.385 14.9008C14.5248 15.3808 13.6346 15.8109 12.7245 16.181ZM9.66388 12.0302L11.9643 14.3307C13.1845 13.8306 14.3648 13.2204 15.485 12.5103C19.9358 9.51974 21.2361 5.60901 21.4561 2.53843C19.6157 2.67846 17.8254 3.20856 16.2051 4.09872C14.2847 5.14892 12.6544 6.68921 11.4942 8.54956C10.7841 9.65977 10.174 10.82 9.66388 12.0302ZM4.39289 11.2701H7.81353C8.1936 10.3599 8.63368 9.46974 9.11377 8.60957C7.92355 8.38953 6.51329 8.31952 5.81315 8.78961C5.19304 9.19968 4.70294 10.3099 4.39289 11.2701Z\" fill=\"#673DE6\"\/>\n                <\/svg>\n                <p class=\"protip__title\">\n                    Pro tip                <\/p>\n            <\/div>\n            <p class=\"protip__content\"> Set a shutdown reminder or a budget alert for any long-running instance. On hourly billing, a two-day accidental run costs the same as two days of actual work. <\/p>\n                    <\/div>\n        <\/p><h2 class=\"wp-block-heading h-t-title-2\" id=\"h-how-to-choose-the-right-cloud-gpu-pricing-model-for-your-workload\">How to choose the right cloud GPU pricing model for your workload<\/h2><p class=\"wp-block-paragraph\">The right cloud GPU pricing model depends on three factors: how long your jobs run, whether they can be interrupted, and how often you need the GPU. For most workloads, that means on-demand access at a specialist cloud.<\/p><figure tabindex=\"0\" class=\"wp-block-table\"><table><tbody><tr><td colspan=\"1\" rowspan=\"1\"><p><strong>Workload<\/strong><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><strong>Recommended model<\/strong><\/td><td colspan=\"1\" rowspan=\"1\"><p><strong>Typical cost<\/strong><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Experiments and development<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>On-demand<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>From $0.38\/hr (RTX 4090)<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Multi-hour batch jobs<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Spot \/ interruptible<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>From $0.34\/hr (RTX 4090)<\/span><\/p><\/td><\/tr><tr><td colspan=\"1\" rowspan=\"1\"><p><span>Always-on production service<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>Reserved<\/span><\/p><\/td><td colspan=\"1\" rowspan=\"1\"><p><span>From $1.36\/hr (A100, Azure 3-year)<\/span><\/p><\/td><\/tr><\/tbody><\/table><\/figure><p class=\"wp-block-paragraph\">If your jobs take several hours and you don&rsquo;t need to watch them run, spot pricing cuts costs by 50% or more. RunPod&rsquo;s Spot Pods and Vast.ai&rsquo;s interruptible tier offer the steepest discounts, though the provider can reclaim the GPU mid-job, so save your progress regularly.<\/p><p class=\"wp-block-paragraph\">If you need a GPU running continuously for a production AI service with consistent user traffic, on-demand pricing stops making sense. Azure&rsquo;s 3-year A100 reservation drops from $3.67 to $1.36 per hour, a 63% reduction for a workload that was going to run anyway. Most teams reach this point when they&rsquo;re serving a live product.<\/p><p class=\"wp-block-paragraph\">Before picking a provider, check the data transfer rate: moving 10 TB out of GCP or AWS can add $800&ndash;1,000 per month. The same transfer is free on Hostinger, RunPod Pods, and Lambda. On data-heavy workloads, the transfer fee can outweigh any savings from a cheaper hourly rate.<\/p><p class=\"wp-block-paragraph\">One approach worth trying: keep your application and data on a VPS and spin up a GPU only when you need to run something. You pay for the GPU only while it runs.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>Cloud GPU pricing ranges from under $0.50 per hour for an entry-level card to over $12 for the most powerful ones. For the same GPU, the price can still vary by as much as sevenfold depending on which provider you rent from. That gap comes from how providers package the hardware, not from the chip [&#8230;]<\/p>\n<p><a class=\"btn btn-secondary understrap-read-more-link\" href=\"\/uk\/tutorials\/cloud-gpu-pricing\/\">Read More&#8230;<\/a><\/p>\n","protected":false},"author":356,"featured_media":0,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"rank_math_title":"Cloud GPU pricing: rates compared across 8 providers","rank_math_description":"From $0.14\/hr on a marketplace GPU to $12\/hr on a hyperscaler: here's what cloud GPU pricing actually looks like, including hidden costs.","rank_math_focus_keyword":"cloud gpu pricing","footnotes":""},"categories":[22661],"tags":[],"class_list":["post-138740","post","type-post","status-publish","format-standard","hentry","category-hosting"],"hreflangs":[{"locale":"en-US","link":"https:\/\/www.hostinger.com\/tutorials\/cloud-gpu-pricing","default":1},{"locale":"en-PH","link":"https:\/\/www.hostinger.com\/ph\/tutorials\/cloud-gpu-pricing","default":0},{"locale":"en-MY","link":"https:\/\/www.hostinger.com\/my\/tutorials\/cloud-gpu-pricing","default":0},{"locale":"en-GB","link":"https:\/\/www.hostinger.com\/uk\/tutorials\/cloud-gpu-pricing","default":0},{"locale":"en-IN","link":"https:\/\/www.hostinger.com\/in\/tutorials\/cloud-gpu-pricing","default":0},{"locale":"en-CA","link":"https:\/\/www.hostinger.com\/ca\/tutorials\/cloud-gpu-pricing","default":0},{"locale":"en-AU","link":"https:\/\/www.hostinger.com\/au\/tutorials\/cloud-gpu-pricing","default":0},{"locale":"en-NG","link":"https:\/\/www.hostinger.com\/ng\/tutorials\/cloud-gpu-pricing","default":0}],"_links":{"self":[{"href":"https:\/\/www.hostinger.com\/uk\/tutorials\/wp-json\/wp\/v2\/posts\/138740","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.hostinger.com\/uk\/tutorials\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.hostinger.com\/uk\/tutorials\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.hostinger.com\/uk\/tutorials\/wp-json\/wp\/v2\/users\/356"}],"replies":[{"embeddable":true,"href":"https:\/\/www.hostinger.com\/uk\/tutorials\/wp-json\/wp\/v2\/comments?post=138740"}],"version-history":[{"count":0,"href":"https:\/\/www.hostinger.com\/uk\/tutorials\/wp-json\/wp\/v2\/posts\/138740\/revisions"}],"wp:attachment":[{"href":"https:\/\/www.hostinger.com\/uk\/tutorials\/wp-json\/wp\/v2\/media?parent=138740"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.hostinger.com\/uk\/tutorials\/wp-json\/wp\/v2\/categories?post=138740"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.hostinger.com\/uk\/tutorials\/wp-json\/wp\/v2\/tags?post=138740"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}