{"id":10841,"date":"2026-02-25T04:19:47","date_gmt":"2026-02-25T08:19:47","guid":{"rendered":"https:\/\/wp.glbgpt.com\/?p=10841"},"modified":"2026-09-26T04:58:13","modified_gmt":"2026-09-26T08:58:13","slug":"gemini-3-1-pro-cost-complete-2026-pricing-guide","status":"publish","type":"post","link":"https:\/\/wp.glbgpt.com\/hub\/gemini-3-1-pro-cost-complete-2026-pricing-guide","title":{"rendered":"Gemini 3.1 Pro Pricing 2026: API Costs and AI Plans"},"content":{"rendered":"\n<p class=\"wp-block-paragraph\">Gemini 3.1 Pro Preview costs <strong>$2 per million input tokens and $12 per million output tokens<\/strong> on the standard Gemini API tier for prompts of up to 200,000 tokens. Above that threshold, the rates are $4 and $18. Output billing includes thinking tokens. For the Gemini app, Google AI Pro is a separate <strong>$19.99 monthly subscription in the United States<\/strong>.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">The price you should budget depends on where you use the model. An app subscription, an API project and a multi-model workspace have different billing rules. A long document, repeated conversation history or a large thinking budget can change an API bill even when the number of requests stays the same.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">For work that moves from research and writing to coding and media creation, <a href=\"https:\/\/www.glbgpt.com\/home?inviter=hub_content_home&amp;login=1\">GlobalGPT<\/a> brings Gemini, Claude, GPT and other models into one subscription and one workspace. It is a practical way to organize a broader AI workflow and manage subscription spending without opening a separate account for every tool. GlobalGPT publishes this guide; the provider prices below come from their official pages.<\/p>\n\n\n\n<div class=\"wp-block-group is-layout-constrained wp-block-group-is-layout-constrained\">\n<figure class=\"wp-block-image size-large\"><img alt=\"\" fetchpriority=\"high\" decoding=\"async\" width=\"1024\" height=\"714\" src=\"https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/08\/image-21-1024x714.png\" alt=\"\" class=\"wp-image-18357\" srcset=\"https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/08\/image-21-1024x714.png 1024w, https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/08\/image-21-300x209.png 300w, https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/08\/image-21-767x535.png 767w, https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/08\/image-21-1536x1072.png 1536w, https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/08\/image-21-17x12.png 17w, https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/08\/image-21.png 1657w\" sizes=\"(max-width: 1024px) 100vw, 1024px\" \/><\/figure>\n\n\n\n<div class=\"wp-block-buttons is-content-justification-center is-layout-flex wp-container-core-buttons-is-layout-3e41869c wp-block-buttons-is-layout-flex\">\n<div class=\"wp-block-button is-style-fill\"><a class=\"wp-block-button__link has-luminous-vivid-amber-background-color has-text-color has-background has-link-color wp-element-button\" href=\"https:\/\/www.glbgpt.com\/home\/gemini-3-7-flash?inviter=hub_content_37flash&amp;login=1\" style=\"color:#0d5a87\"><strong>Try Gemini 3.7 Flash on GlobalGPT Now<\/strong><\/a><\/div>\n<\/div>\n<\/div>\n\n\n\n<aside style=\"margin:24px 0;padding:22px;border:1px solid #abd0cf;border-left:5px solid #197c80;border-radius:12px;background:#eff8f7;color:#173d46;\"><p style=\"margin:0 0 8px;color:#173d46;font-weight:700;\">The three numbers that matter<\/p><p style=\"margin:0;color:#173d46;line-height:1.65;\">Standard API: <strong>$2 input \/ $12 output per 1M tokens<\/strong> up to 200k prompt tokens. Long prompts: <strong>$4 \/ $18<\/strong>. Google AI Pro: <strong>$19.99\/month in the US<\/strong>. These are different purchases, not interchangeable allowances.<\/p><p style=\"margin:12px 0 0;font-size:.9em;color:#35565d;\">Prices checked September 15, 2026. USD API rates and US advertised monthly plans; promotions, taxes and regional checkout totals are separate.<\/p><\/aside>\n\n\n\n<nav aria-label=\"Article contents\" style=\"margin:24px 0;padding:22px;background:#f3f5f8;border:1px solid #d4dce5;border-radius:12px;color:#173d46;\"><p style=\"margin:0 0 12px;font-weight:700;color:#173d46;\">On this page<\/p><ol style=\"margin:0;padding-left:22px;line-height:1.9;color:#173d46;\"><li><a style=\"color:#175c68;\" href=\"#model-context\">What is Gemini 3.1 Pro? Model and context window<\/a><\/li><li><a style=\"color:#175c68;\" href=\"#api-pricing\">Gemini 3.1 Pro API pricing per million tokens<\/a><\/li><li><a style=\"color:#175c68;\" href=\"#context-caching\">How context caching changes the bill<\/a><\/li><li><a style=\"color:#175c68;\" href=\"#google-ai-plans\">Google AI Pro and Ultra: US monthly subscription prices<\/a><\/li><li><a style=\"color:#175c68;\" href=\"#api-comparison\">API cost comparison: Gemini, GPT and Claude<\/a><\/li><li><a style=\"color:#175c68;\" href=\"#globalgpt-workflow\">When a multi-model subscription fits your budget<\/a><\/li><li><a style=\"color:#175c68;\" href=\"#faq\">Frequently asked questions<\/a><\/li><\/ol><\/nav>\n\n\n\n<h2 id=\"model-context\" class=\"wp-block-heading\">What is Gemini 3.1 Pro? Model and context window<\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">Google announced Gemini 3.1 Pro on February 19, 2026. The API model covered here is <code>gemini-3.1-pro-preview<\/code>. Its <a href=\"https:\/\/ai.google.dev\/gemini-api\/docs\/models\/gemini-3.1-pro-preview\">official model specification<\/a> lists a 1,048,576-token input limit and a 65,536-token output limit. It accepts text, images, video, audio and PDFs, and produces text. Those input capabilities do not make it an image- or video-generation model.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">For users comparing the <a href=\"https:\/\/www.glbgpt.com\/hub\/gemini-3-pro-token-limit\/\">Gemini Pro token limit<\/a>, capacity and cost answer different questions: the model may accept a large file, but a prompt above 200k tokens moves into the higher price bracket. The full conversation and repeated source material also contribute to the prompt you send.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">In its <a href=\"https:\/\/blog.google\/innovation-and-ai\/models-and-research\/gemini-models\/gemini-3-1-pro\/\">launch announcement<\/a>, Google reported a 77.1% ARC-AGI-2 score. That is a vendor-reported reasoning result, not a measure of your cost per useful answer. A more valuable budget check is whether the model finishes your actual document, coding or analysis task with fewer corrections. The <a href=\"https:\/\/www.glbgpt.com\/hub\/gemini-3-1-pro-coding-guide-tutorial\/\">Gemini 3.1 Pro coding guide<\/a> covers that workflow in more detail.<\/p>\n\n\n\n<h2 id=\"api-pricing\" class=\"wp-block-heading\">Gemini 3.1 Pro API pricing per million tokens<\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">The <a href=\"https:\/\/ai.google.dev\/gemini-api\/docs\/pricing#gemini-3.1-pro-preview\">Gemini API price list<\/a> separates standard requests from Batch, Flex and Priority inference. All rates in the following table are in US dollars per one million tokens. Select the bracket using the <strong>prompt token count<\/strong>; then apply the corresponding input and output rates.<\/p>\n\n\n\n<div style=\"margin:24px 0;overflow-x:auto;border:1px solid #cbd5e1;border-radius:10px;background:#ffffff;color:#172b3a;\"><table style=\"width:100%;border-collapse:collapse;font-size:0.95em;line-height:1.5;\"><caption style=\"padding:12px;text-align:left;font-weight:600;color:#153e46;\">Gemini 3.1 Pro Preview token rates \u2014 checked September 15, 2026<\/caption><thead><tr><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Mode<\/th><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Prompt \u2264200k: input \/ output<\/th><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Prompt >200k: input \/ output<\/th><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">When it fits<\/th><\/tr><\/thead><tbody><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Standard<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$2 \/ $12<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$4 \/ $18<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Regular interactive use<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Batch<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$1 \/ $6<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$2 \/ $9<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Asynchronous work; target turnaround of 24 hours<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Flex<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$1 \/ $6<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$2 \/ $9<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Latency-tolerant synchronous work; variable latency and best-effort availability<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Priority<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$3.60 \/ $21.60<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$7.20 \/ $32.40<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">A separately priced priority inference tier<\/td><\/tr><\/tbody><\/table><\/div>\n\n\n\n<p class=\"wp-block-paragraph\"><a href=\"https:\/\/ai.google.dev\/gemini-api\/docs\/batch-api\">Batch<\/a> and <a href=\"https:\/\/ai.google.dev\/gemini-api\/docs\/flex-inference\">Flex<\/a> halve the listed standard input and output rates, with different timing and availability trade-offs. They are useful for jobs such as classifying a document collection or preparing a report you do not need immediately. Their token discount does not automatically halve cache storage or every tool charge.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\">The 200k boundary: input doubles, output rises by 50%<\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">A prompt of exactly 200,000 tokens remains in the lower bracket. Above 200,000, the standard input rate rises from $2 to $4, while the output rate rises from $12 to $18. The higher rates apply to the request\u2019s applicable token counts; this is not a surcharge on only the tokens beyond 200k.<\/p>\n\n\n\n<figure style=\"margin:28px 0;padding:22px;background:#f5faf9;border:1px solid #b4d6d4;border-radius:14px;color:#143540;\"><figcaption style=\"font-size:1.15em;font-weight:700;margin-bottom:12px;color:#143540;\">How a long prompt changes the token rate<\/figcaption><div role=\"img\" aria-label=\"Standard Gemini 3.1 Pro Preview USD per million tokens. Input costs 2 below or at 200k prompt tokens and 4 above. Output including thinking costs 12 and 18 respectively.\" style=\"color:#143540;\"><div style=\"margin:16px 0;\"><div style=\"display:flex;justify-content:space-between;gap:12px;flex-wrap:wrap;color:#143540;\"><span>Input, prompt \u2264200k<\/span><strong>$2 \/ 1M tokens<\/strong><\/div><div style=\"background:#e7eef0;border-radius:5px;margin-top:6px;overflow:hidden;\"><div style=\"width:11.1111%;height:16px;background:#197c80;border-radius:5px;\"><\/div><\/div><\/div><div style=\"margin:16px 0;\"><div style=\"display:flex;justify-content:space-between;gap:12px;flex-wrap:wrap;color:#143540;\"><span>Input, prompt >200k<\/span><strong>$4 \/ 1M tokens<\/strong><\/div><div style=\"background:#e7eef0;border-radius:5px;margin-top:6px;overflow:hidden;\"><div style=\"width:22.2222%;height:16px;background:#197c80;border-radius:5px;\"><\/div><\/div><\/div><div style=\"margin:16px 0;\"><div style=\"display:flex;justify-content:space-between;gap:12px;flex-wrap:wrap;color:#143540;\"><span>Output, prompt \u2264200k<\/span><strong>$12 \/ 1M tokens<\/strong><\/div><div style=\"background:#e7eef0;border-radius:5px;margin-top:6px;overflow:hidden;\"><div style=\"width:66.6667%;height:16px;background:#197c80;border-radius:5px;\"><\/div><\/div><\/div><div style=\"margin:16px 0;\"><div style=\"display:flex;justify-content:space-between;gap:12px;flex-wrap:wrap;color:#143540;\"><span>Output, prompt >200k<\/span><strong>$18 \/ 1M tokens<\/strong><\/div><div style=\"background:#e7eef0;border-radius:5px;margin-top:6px;overflow:hidden;\"><div style=\"width:100.0000%;height:16px;background:#197c80;border-radius:5px;\"><\/div><\/div><\/div><\/div><p style=\"font-size:.9em;line-height:1.6;color:#35565d;margin:16px 0 8px;\">Source: <a href=\"https:\/\/ai.google.dev\/gemini-api\/docs\/pricing#gemini-3.1-pro-preview\">Google Gemini API pricing<\/a>. Standard tier; checked September 15, 2026. All bars use the same dollar scale.<\/p><p style=\"color:#143540;margin:8px 0 0;\">A larger input can raise the output rate too, even if the answer is short.<\/p><\/figure>\n\n\n\n<h3 class=\"wp-block-heading\">A worked monthly API estimate<\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">Suppose you make 100 standard requests, each with 10,000 input tokens and 2,000 billable output tokens. The estimate is <strong>100 \u00d7 [(10,000 \u00f7 1,000,000 \u00d7 $2) + (2,000 \u00f7 1,000,000 \u00d7 $12)] = $4.40<\/strong>. This is an illustrative calculation, not a measured bill; it excludes tools, caching, storage, taxes and platform fees.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">For one larger request with 250,000 input tokens and 10,000 billable output tokens, the same method gives <strong>$1.18<\/strong>: $1.00 for input plus $0.18 for output. Use billable output, including thinking, rather than counting only the words you can see in the answer.<\/p>\n\n\n\n<ol class=\"wp-block-list\">\n<li>Count the complete prompt, including conversation history and source files.<\/li>\n\n\n\n<li>Identify the request mode and the correct prompt-size bracket.<\/li>\n\n\n\n<li>Add billable output, then any caching, storage and tool charges.<\/li>\n\n\n\n<li>Check the usage report against the estimate before scaling up.<\/li>\n<\/ol>\n\n\n\n<p class=\"wp-block-paragraph\">The <a href=\"https:\/\/ai.google.dev\/gemini-api\/docs\/tokens\">official token guide<\/a> explains token counting. For setup and implementation, use the <a href=\"https:\/\/www.glbgpt.com\/hub\/gemini-3-1-pro-api-pricing-performance-the-complete-guide-for-developers\/\">Gemini 3.1 Pro API developer guide<\/a>.<\/p>\n\n\n\n<h2 id=\"context-caching\" class=\"wp-block-heading\">How context caching changes the bill<\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">Context caching is useful when the same document, repository or system instructions appear in repeated requests. It discounts the reused input; it does not make the whole request free. On the standard tier, cached input costs $0.20 per million tokens for prompts up to 200k, or $0.40 above 200k. Explicit cache storage adds <strong>$4.50 per million cached tokens per hour<\/strong>.<\/p>\n\n\n\n<figure class=\"wp-block-image size-full\"><img decoding=\"async\" src=\"https:\/\/static.futureshareai.com\/glb_features\/mcp\/3\/gemini-cost-original-api-rates-crop-20260915_6cb23f4ed1ea421196bc9839e240d0e2.webp\" alt=\"The original Gemini 3.1 Pro pricing table shows input, output, cached input and hourly cache storage rates.\"\/><figcaption class=\"wp-element-caption\">Pricing table preserved from the original February 2026 article. The displayed token and storage rates were rechecked against Google\u2019s price list on September 15, 2026; this is not a newly captured image.<\/figcaption><\/figure>\n\n\n\n<p class=\"wp-block-paragraph\">Google describes two approaches in its <a href=\"https:\/\/ai.google.dev\/gemini-api\/docs\/generate-content\/caching\">context caching documentation<\/a>: implicit caching automatically passes on savings when a request hits a cache; explicit caching lets you retain a specified prefix for a chosen time to live. Explicit caching defaults to one hour if you do not set a duration. Cached tokens still count toward token limits, so caching does not move a long prompt into a smaller price bracket.<\/p>\n\n\n\n<div style=\"margin:24px 0;overflow-x:auto;border:1px solid #cbd5e1;border-radius:10px;background:#ffffff;color:#172b3a;\"><table style=\"width:100%;border-collapse:collapse;font-size:0.95em;line-height:1.5;\"><caption style=\"padding:12px;text-align:left;font-weight:600;color:#153e46;\">Same five input reads; different retention times \u2014 USD, September 15, 2026<\/caption><thead><tr><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Repeated-input scenario<\/th><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Five reads of 100k tokens<\/th><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Storage<\/th><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Illustrative total<\/th><\/tr><\/thead><tbody><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Uncached, standard lower bracket<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$1.00<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$0<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$1.00<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Explicit cache retained for 1 hour<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$0.10<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$0.45<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$0.55<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Explicit cache retained for 24 hours<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$0.10<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$10.80<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$10.90<\/td><\/tr><\/tbody><\/table><\/div>\n\n\n\n<p class=\"wp-block-paragraph\">These estimates compare only the recurring input and storage components. New uncached input and output are additional. A short-lived cache reused several times can be useful; keeping a large cache overnight for only a few reads can cost more than sending those inputs again. Set retention around the actual work session and inspect cache-hit usage instead of assuming every repeated prompt receives a discount.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Tool usage can add another line item. For Google Search grounding, <a href=\"https:\/\/ai.google.dev\/gemini-api\/docs\/pricing#gemini-3.1-pro-preview\">Google\u2019s price table<\/a> lists 5,000 free \u201csearch requests\u201d per month, shared across Gemini 3.x models, then $14 per 1,000 search requests. Its <a href=\"https:\/\/ai.google.dev\/gemini-api\/docs\/google-search#pricing\">grounding guide<\/a> clarifies that Gemini 3 billing counts the search queries the model executes, with empty queries excluded when counting unique queries. A single generation request that triggers two distinct searches therefore uses two billable search units. Budget around searches performed, not just answers generated; the free allowance is not a promise of 5,000 grounded answers.<\/p>\n\n\n\n<h2 id=\"google-ai-plans\" class=\"wp-block-heading\">Google AI Pro and Ultra: US monthly subscription prices<\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">If you want to use Gemini in a chat interface, compare <a href=\"https:\/\/gemini.google\/subscriptions\/\">Google\u2019s US subscription plans<\/a> rather than multiplying chat messages by API token rates. The public US page lists the following monthly options. These are advertised subscription prices, not a quote for every country or a tax-inclusive checkout total.<\/p>\n\n\n\n<div style=\"margin:24px 0;overflow-x:auto;border:1px solid #cbd5e1;border-radius:10px;background:#ffffff;color:#172b3a;\"><table style=\"width:100%;border-collapse:collapse;font-size:0.95em;line-height:1.5;\"><caption style=\"padding:12px;text-align:left;font-weight:600;color:#153e46;\">Google AI consumer plans \u2014 United States, checked September 15, 2026<\/caption><thead><tr><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">US plan<\/th><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Advertised monthly price<\/th><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Relevant access distinction<\/th><\/tr><\/thead><tbody><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Free<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$0<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Access to the Gemini app, including varying access to 3.1 Pro<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Google AI Plus<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$4.99<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Higher usage access than Free<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Google AI Pro<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$19.99<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Higher Gemini access, Google app features and 5 TB storage<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Google AI Ultra 5x<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$99.99<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">5x higher usage limits than AI Pro on the US plan page<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Google AI Ultra 20x<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$199.99<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">20x higher usage limits than AI Pro on the US plan page<\/td><\/tr><\/tbody><\/table><\/div>\n\n\n\n<figure class=\"wp-block-image size-full\"><img decoding=\"async\" src=\"https:\/\/static.futureshareai.com\/glb_features\/mcp\/3\/google-us-ai-plans-20260915_49cc37d7d8924503b79f6324e81197fe.webp\" alt=\"Google\u2019s US subscription page displays the monthly Google AI Pro and Ultra plan prices.\"\/><figcaption class=\"wp-element-caption\">US advertised Google AI subscription prices, captured September 15, 2026. Eligibility, taxes and local offers can differ.<\/figcaption><\/figure>\n\n\n\n<p class=\"wp-block-paragraph\">The 5x and 20x labels describe the published plan comparison; they are not a promise of a fixed number of API calls. Google AI Pro is worth considering if Gemini is part of your daily work and you also use the bundled Google features and storage. Free or Plus can be enough for lighter use. Ultra makes sense when its higher app usage or additional features solve a limitation you actually encounter.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\">Does Google AI Pro include API credit?<\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">Google AI Pro includes a <strong>$10 monthly developer credit<\/strong> through the <a href=\"https:\/\/developers.google.com\/program\/plans-and-pricing\">Google Developer Program<\/a>. Google\u2019s <a href=\"https:\/\/developers.google.com\/profile\/help\/faq\">current FAQ<\/a> describes the linked Pro\/Ultra benefit as monthly Google Cloud credit. Its <a href=\"https:\/\/developers.google.com\/profile\/help\/benefits#google-cloud-credits\">general Cloud credit guidance<\/a> covers products such as Firebase, Vertex AI and Google Maps, and explains how to apply a benefit to one eligible billing account through My Benefits.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Before using the credit in your budget, check the specific benefit in My Benefits, the service it covers, its expiry date and the available balance on your billing account. Confirm that it applies to the Gemini API service and project you intend to use. Until then, plan around the API costs before credits shown in this guide. Keep your Gemini app subscription and API project billing separate when tracking spending.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">The $300 Google Cloud free trial is another, separate offer. Google\u2019s API billing guide excludes Gemini API usage from that free trial starting in March 2026. A Cloud welcome balance should therefore not be treated as the monthly developer benefit attached to a Google AI plan.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Account-specific billing setup and usage are explained in <a href=\"https:\/\/ai.google.dev\/gemini-api\/docs\/billing\">Google\u2019s Gemini API billing guide<\/a>. App usage limits and API rate limits also differ; the <a href=\"https:\/\/www.glbgpt.com\/hub\/gemini-3-1-pro-limits-2026-the-ultimate-guide-to-bypassing-rate-limits-quotas\/\">Gemini 3.1 Pro limits guide<\/a> covers the distinctions. Paying for a subscription does not remove all limits.<\/p>\n\n\n\n<h2 id=\"api-comparison\" class=\"wp-block-heading\">API cost comparison: Gemini, GPT and Claude<\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">The most useful comparison starts with the same token budget. The figures below use standard, uncached first-party API rates for a prompt of 100,000 tokens and 10,000 billable output tokens. They exclude tools, storage, taxes, regional processing premiums and third-party platform fees. Equal token counts are not necessarily equal text or equally successful answers across different model tokenizers.<\/p>\n\n\n\n<div style=\"margin:24px 0;overflow-x:auto;border:1px solid #cbd5e1;border-radius:10px;background:#ffffff;color:#172b3a;\"><table style=\"width:100%;border-collapse:collapse;font-size:0.95em;line-height:1.5;\"><caption style=\"padding:12px;text-align:left;font-weight:600;color:#153e46;\">Selected API models, including Sonnet 5 and earlier Claude versions \u2014 standard rates, September 15, 2026<\/caption><thead><tr><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Selected model<\/th><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Input \/ 1M<\/th><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Output \/ 1M<\/th><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">100k input + 10k output<\/th><\/tr><\/thead><tbody><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Gemini 3.1 Pro Preview<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$2<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$12<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$0.32<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">GPT-5.5<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$5<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$30<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$0.80<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Claude Opus 4.8<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$5<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$25<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$0.75<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Claude Sonnet 4.6<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$3<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$15<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$0.45<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Claude Sonnet 5<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$2<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$10<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">$0.30<\/td><\/tr><\/tbody><\/table><\/div>\n\n\n\n<figure style=\"margin:28px 0;padding:22px;background:#f5faf9;border:1px solid #b4d6d4;border-radius:14px;color:#143540;\"><figcaption style=\"font-size:1.15em;font-weight:700;margin-bottom:12px;color:#143540;\">One token budget, five estimated costs<\/figcaption><div role=\"img\" aria-label=\"Illustrative USD token cost for 100000 input and 10000 billable output tokens: Gemini 3.1 Pro Preview 0.32, GPT-5.5 0.80, Claude Opus 4.8 0.75, Claude Sonnet 4.6 0.45, Claude Sonnet 5 0.30.\" style=\"color:#143540;\"><div style=\"margin:16px 0;\"><div style=\"display:flex;justify-content:space-between;gap:12px;flex-wrap:wrap;color:#143540;\"><span>Gemini 3.1 Pro Preview<\/span><strong>$0.32<\/strong><\/div><div style=\"background:#e7eef0;border-radius:5px;margin-top:6px;overflow:hidden;\"><div style=\"width:40.0000%;height:16px;background:#197c80;border-radius:5px;\"><\/div><\/div><\/div><div style=\"margin:16px 0;\"><div style=\"display:flex;justify-content:space-between;gap:12px;flex-wrap:wrap;color:#143540;\"><span>GPT-5.5<\/span><strong>$0.8<\/strong><\/div><div style=\"background:#e7eef0;border-radius:5px;margin-top:6px;overflow:hidden;\"><div style=\"width:100.0000%;height:16px;background:#197c80;border-radius:5px;\"><\/div><\/div><\/div><div style=\"margin:16px 0;\"><div style=\"display:flex;justify-content:space-between;gap:12px;flex-wrap:wrap;color:#143540;\"><span>Claude Opus 4.8<\/span><strong>$0.75<\/strong><\/div><div style=\"background:#e7eef0;border-radius:5px;margin-top:6px;overflow:hidden;\"><div style=\"width:93.7500%;height:16px;background:#197c80;border-radius:5px;\"><\/div><\/div><\/div><div style=\"margin:16px 0;\"><div style=\"display:flex;justify-content:space-between;gap:12px;flex-wrap:wrap;color:#143540;\"><span>Claude Sonnet 4.6<\/span><strong>$0.45<\/strong><\/div><div style=\"background:#e7eef0;border-radius:5px;margin-top:6px;overflow:hidden;\"><div style=\"width:56.2500%;height:16px;background:#197c80;border-radius:5px;\"><\/div><\/div><\/div><div style=\"margin:16px 0;\"><div style=\"display:flex;justify-content:space-between;gap:12px;flex-wrap:wrap;color:#143540;\"><span>Claude Sonnet 5<\/span><strong>$0.3<\/strong><\/div><div style=\"background:#e7eef0;border-radius:5px;margin-top:6px;overflow:hidden;\"><div style=\"width:37.5000%;height:16px;background:#197c80;border-radius:5px;\"><\/div><\/div><\/div><\/div><p style=\"font-size:.9em;line-height:1.6;color:#35565d;margin:16px 0 8px;\">Sources: <a href=\"https:\/\/ai.google.dev\/gemini-api\/docs\/pricing#gemini-3.1-pro-preview\">Google<\/a>, <a href=\"https:\/\/developers.openai.com\/api\/docs\/models\/gpt-5.5\">OpenAI GPT-5.5<\/a>, <a href=\"https:\/\/platform.claude.com\/docs\/en\/about-claude\/pricing\">Anthropic<\/a>. Checked September 15, 2026; calculated examples, not measured bills.<\/p><p style=\"color:#143540;margin:8px 0 0;\">Sonnet 5 costs $0.30 for this token budget, compared with Gemini\u2019s $0.32. This five-model sample is not a market-wide cheapest-model survey or an overall quality or return-on-investment ranking.<\/p><\/figure>\n\n\n\n<p class=\"wp-block-paragraph\">Anthropic now lists Sonnet 5\u2019s $2 input \/ $10 output rates as its standard price; the previously announced September price increase will not occur. Sonnet 4.6 and Opus 4.8 remain useful references for existing workloads, but they are earlier Claude versions rather than the full current lineup.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Longer prompts can change the size of the price gap. Gemini\u2019s higher bracket starts above 200k prompt tokens; GPT-5.5\u2019s documented threshold is above 272k input tokens. Anthropic lists standard pricing across the full 1M context window for Claude 4.6 and later. At 250k input tokens and 10k output tokens, Gemini\u2019s standard token estimate is $1.18, Sonnet 4.6\u2019s is $0.90 and Sonnet 5\u2019s is $0.60. Compare the applicable bracket for the workload you will actually send.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Choose on cost per completed task: include failed attempts, follow-up prompts, thinking and human corrections. Token prices alone cannot establish which model handles your work best. For the earlier model\u2019s pricing background, see the existing comparison of <a href=\"https:\/\/www.glbgpt.com\/hub\/gemini-3-pro-costs-gemini-3-api-costs-latest-insights-for-2025\/\">Gemini 3 Pro and Gemini 3 API costs<\/a>.<\/p>\n\n\n\n<h2 id=\"globalgpt-workflow\" class=\"wp-block-heading\">When a multi-model subscription fits your budget<\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">A direct API project fits a product that needs programmable calls and usage-based billing. A Google AI plan fits users who want Gemini and its Google ecosystem benefits. <a href=\"https:\/\/www.glbgpt.com\/home?inviter=hub_content_home&amp;login=1\">GlobalGPT<\/a> fits work that spans several models and functions: research a subject, draft and revise an article, ask for coding help, then create supporting media in one workspace. Its CLI also connects AI work with terminal and development workflows.<\/p>\n\n\n\n<div class=\"wp-block-group is-layout-constrained wp-block-group-is-layout-constrained\">\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" width=\"1024\" height=\"645\" src=\"https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/07\/glb-1024x645.png\" alt=\"globalgpt dashboard\" class=\"wp-image-19280\" srcset=\"https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/07\/glb-1024x645.png 1024w, https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/07\/glb-300x189.png 300w, https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/07\/glb-18x11.png 18w, https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/07\/glb-767x483.png 767w, https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/07\/glb-1536x967.png 1536w, https:\/\/wp.glbgpt.com\/wp-content\/uploads\/2026\/07\/glb-2048x1289.png 2048w\" sizes=\"(max-width: 1024px) 100vw, 1024px\" \/><\/figure>\n\n\n\n<div class=\"wp-block-buttons is-content-justification-center is-layout-flex wp-container-core-buttons-is-layout-3e41869c wp-block-buttons-is-layout-flex\">\n<div class=\"wp-block-button\"><a class=\"wp-block-button__link has-black-color has-text-color has-background has-link-color wp-element-button\" href=\"https:\/\/www.glbgpt.com\/home?inviter=hub_popup&amp;login=1\" style=\"background:linear-gradient(135deg,rgba(250,183,0,0.55) 0%,rgba(255,106,0,0.66) 96%)\"><strong>Try 100+ Top Models On GlobalGPT<\/strong><\/a><\/div>\n<\/div>\n<\/div>\n\n\n\n<div style=\"margin:24px 0;overflow-x:auto;border:1px solid #cbd5e1;border-radius:10px;background:#ffffff;color:#172b3a;\"><table style=\"width:100%;border-collapse:collapse;font-size:0.95em;line-height:1.5;\"><caption style=\"padding:12px;text-align:left;font-weight:600;color:#153e46;\">Choose the billing route around the job you need to finish<\/caption><thead><tr><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Purchase<\/th><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">How to budget<\/th><th scope=\"col\" style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;background:#e8f1f2;color:#153e46;\">Choose it for<\/th><\/tr><\/thead><tbody><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Gemini API<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Input, output, service tier, caching and tools<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Building software or automating calls<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Google AI subscription<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">US advertised monthly plan plus its eligibility and feature limits<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">Using Gemini and bundled Google products<\/td><\/tr><tr><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">GlobalGPT subscription<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">The selected platform plan and its own model\/function allowances<\/td><td style=\"padding:12px;border:1px solid #cbd5e1;text-align:left;vertical-align:top;\">A broader workflow across leading models and media tools<\/td><\/tr><\/tbody><\/table><\/div>\n\n\n\n<p class=\"wp-block-paragraph\">The value of a multi-model subscription is breadth and workflow convenience. Compare the tools you would actually use and the separate subscriptions they would otherwise require. Choose a plan whose allowances fit that workload; a platform subscription and an official provider\u2019s API or consumer plan are different products.<\/p>\n\n\n\n<ul class=\"wp-block-list\">\n<li><strong>Mostly chat and Google apps?<\/strong> Start with the Google plan whose usage level fits your routine. The guide to <a href=\"https:\/\/www.glbgpt.com\/hub\/is-gemini-3-pro-free\/\">what free Gemini access means<\/a> can help before you upgrade.<\/li>\n\n\n\n<li><strong>Shipping a product?<\/strong> Use an API budget with logs, usage estimates and the correct project billing setup.<\/li>\n\n\n\n<li><strong>Research, writing, coding and media in one routine?<\/strong> Explore <a href=\"https:\/\/www.glbgpt.com\/home?inviter=hub_content_home&amp;login=1\">GlobalGPT\u2019s multi-model workspace<\/a>. For more options, see <a href=\"https:\/\/www.glbgpt.com\/hub\/gemini-3-pro-alternative-we-tried-8-better-options\/\">Gemini alternatives<\/a> and match their strengths to your work.<\/li>\n<\/ul>\n\n\n\n<p class=\"wp-block-paragraph\">For a practical next step after choosing the access route, follow <a href=\"https:\/\/www.glbgpt.com\/hub\/how-to-use-gemini-3-1-pro-in-2026-from-basic-chat-to-api-integration\/\">how to use Gemini 3.1 Pro<\/a>. Keep the distinction between app subscription and API billing when following setup instructions.<\/p>\n\n\n\n<h2 id=\"faq\" class=\"wp-block-heading\">Frequently asked questions<\/h2>\n\n\n\n<h3 class=\"wp-block-heading\">How much does Gemini 3.1 Pro cost per million tokens?<\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">On the standard Gemini API tier, Gemini 3.1 Pro Preview costs $2 per million input tokens and $12 per million output tokens for prompts of up to 200,000 tokens. Above 200,000 prompt tokens, the rates are $4 and $18. Output billing includes thinking tokens.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\">Is there a free Gemini 3.1 Pro API tier?<\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">Google lists no API free tier for Gemini 3.1 Pro Preview. The Gemini app has a free plan with varying access to 3.1 Pro, and Google AI Studio provides a place to try models. Those access routes do not establish a free production API allowance for this model.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\">Does the API charge for images and video?<\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">Yes. Gemini 3.1 Pro accepts images, video, audio and PDFs as inputs, and their token usage contributes to the bill. The amount depends on the media and processing settings. It produces text; image or video generation uses separately priced models.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\">Is exactly 200,000 input tokens in the higher price bracket?<\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">No. A prompt of exactly 200,000 tokens is in the lower standard price bracket. A prompt above 200,000 tokens uses the higher input and output rates. Cached tokens still count toward the applicable token limits.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\">Is the $19.99 Google AI Pro plan the same as API billing?<\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">No. The US $19.99 monthly plan is a consumer subscription with a $10 monthly developer credit benefit. Google\u2019s general Cloud credit guidance covers products such as Vertex AI and explains applying a benefit to one eligible billing account. Check your specific credit in My Benefits, its eligible service, expiry date and available balance before applying it to your Gemini API budget. Until those details are confirmed, plan for API costs before credits.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\">What is the cheapest way to use Gemini 3.1 Pro?<\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">For occasional app use, start with the free plan\u2019s available access. For paid API work, estimate tokens and consider Batch or Flex when their timing trade-offs fit. For a broader routine across models and media tools, compare a multi-model subscription against the separate subscriptions you would otherwise use.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Budget around your actual route and workload: Google\u2019s <a href=\"https:\/\/ai.google.dev\/gemini-api\/docs\/pricing#gemini-3.1-pro-preview\">API price list<\/a> for software, its <a href=\"https:\/\/gemini.google\/subscriptions\/\">subscription page<\/a> for the Gemini app, or <a href=\"https:\/\/www.glbgpt.com\/home?inviter=hub_content_home&amp;login=1\">GlobalGPT<\/a> for a connected multi-model workflow.<\/p>\n\n\n\n<script type=\"application\/ld+json\">{\"@context\": \"https:\/\/schema.org\", \"@type\": \"FAQPage\", \"mainEntity\": [{\"@type\": \"Question\", \"name\": \"How much does Gemini 3.1 Pro cost per million tokens?\", \"acceptedAnswer\": {\"@type\": \"Answer\", \"text\": \"On the standard Gemini API tier, Gemini 3.1 Pro Preview costs $2 per million input tokens and $12 per million output tokens for prompts of up to 200,000 tokens. Above 200,000 prompt tokens, the rates are $4 and $18. Output billing includes thinking tokens.\"}}, {\"@type\": \"Question\", \"name\": \"Is there a free Gemini 3.1 Pro API tier?\", \"acceptedAnswer\": {\"@type\": \"Answer\", \"text\": \"Google lists no API free tier for Gemini 3.1 Pro Preview. The Gemini app has a free plan with varying access to 3.1 Pro, and Google AI Studio provides a place to try models. Those access routes do not establish a free production API allowance for this model.\"}}, {\"@type\": \"Question\", \"name\": \"Does the API charge for images and video?\", \"acceptedAnswer\": {\"@type\": \"Answer\", \"text\": \"Yes. Gemini 3.1 Pro accepts images, video, audio and PDFs as inputs, and their token usage contributes to the bill. The amount depends on the media and processing settings. It produces text; image or video generation uses separately priced models.\"}}, {\"@type\": \"Question\", \"name\": \"Is exactly 200,000 input tokens in the higher price bracket?\", \"acceptedAnswer\": {\"@type\": \"Answer\", \"text\": \"No. A prompt of exactly 200,000 tokens is in the lower standard price bracket. A prompt above 200,000 tokens uses the higher input and output rates. Cached tokens still count toward the applicable token limits.\"}}, {\"@type\": \"Question\", \"name\": \"Is the $19.99 Google AI Pro plan the same as API billing?\", \"acceptedAnswer\": {\"@type\": \"Answer\", \"text\": \"No. The US $19.99 monthly plan is a consumer subscription with a $10 monthly developer credit benefit. Google\u2019s general Cloud credit guidance covers products such as Vertex AI and explains applying a benefit to one eligible billing account. Check your specific credit in My Benefits, its eligible service, expiry date and available balance before applying it to your Gemini API budget. Until those details are confirmed, plan for API costs before credits.\"}}, {\"@type\": \"Question\", \"name\": \"What is the cheapest way to use Gemini 3.1 Pro?\", \"acceptedAnswer\": {\"@type\": \"Answer\", \"text\": \"For occasional app use, start with the free plan\u2019s available access. For paid API work, estimate tokens and consider Batch or Flex when their timing trade-offs fit. For a broader routine across models and media tools, compare a multi-model subscription against the separate subscriptions you would otherwise use.\"}}]}<\/script>\n","protected":false},"excerpt":{"rendered":"<p>Gemini 3.1 Pro Preview costs $2 per million input tokens and $12 per million output tokens on the standard Gemini API tier for prompts of up to 200,000 tokens. Above that threshold, the rates are $4 and $18. Output billing includes thinking tokens. For the Gemini app, Google AI Pro is a separate $19.99 monthly [&hellip;]<\/p>\n","protected":false},"author":7,"featured_media":19939,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"_acf_changed":false,"_seopress_robots_primary_cat":"","_seopress_titles_title":"Gemini 3.1 Pro Pricing 2026: API Costs and AI Plans","_seopress_titles_desc":"Gemini 3.1 Pro costs $2 per million input tokens and $12 for output at the standard tier. Compare long-context costs, caching and US monthly Google AI plans.","_seopress_robots_index":"","footnotes":""},"categories":[7],"tags":[],"class_list":["post-10841","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-ai-chat"],"acf":[],"_links":{"self":[{"href":"https:\/\/wp.glbgpt.com\/wp-json\/wp\/v2\/posts\/10841","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/wp.glbgpt.com\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/wp.glbgpt.com\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/wp.glbgpt.com\/wp-json\/wp\/v2\/users\/7"}],"replies":[{"embeddable":true,"href":"https:\/\/wp.glbgpt.com\/wp-json\/wp\/v2\/comments?post=10841"}],"version-history":[{"count":7,"href":"https:\/\/wp.glbgpt.com\/wp-json\/wp\/v2\/posts\/10841\/revisions"}],"predecessor-version":[{"id":19944,"href":"https:\/\/wp.glbgpt.com\/wp-json\/wp\/v2\/posts\/10841\/revisions\/19944"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/wp.glbgpt.com\/wp-json\/wp\/v2\/media\/19939"}],"wp:attachment":[{"href":"https:\/\/wp.glbgpt.com\/wp-json\/wp\/v2\/media?parent=10841"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/wp.glbgpt.com\/wp-json\/wp\/v2\/categories?post=10841"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/wp.glbgpt.com\/wp-json\/wp\/v2\/tags?post=10841"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}