Gemini 3 Pro Preview
Google logo

Gemini 3 Pro Preview

gemini-3-pro-preview
Google

Pricing

TierPricingCache ReadWeb SearchCache Storage
Input<=200K
$2$12
$0.5/M tokens$0.014/request$4.5/h/M tokens
200K<Input
$4$18
$1/M tokens$0.014/request$4.5/h/M tokens

Input Modalities

    Providers

    VertexAI gemini-3-pro-preview
    Pricing$2$12
    Cache Read$0.5/M tokens
    Web Search$0.014/request
    Cache Storage$4.5/h/M tokens
    Pricing$4$18
    Cache Read$1/M tokens
    Web Search$0.014/request
    Cache Storage$4.5/h/M tokens
    Context1M
    Max output65K
    Latency2.4S
    Throughput89.6TPS
    Uptime
    0.00% uptime 2 days ago
    0.00% uptime yesterday
    0.00% uptime today
    Google AI Studio gemini-3-pro-preview
    Pricing$2$12
    Cache Read$0.5/M tokens
    Web Search$0.014/request
    Cache Storage$4.5/h/M tokens
    Pricing$4$18
    Cache Read$1/M tokens
    Web Search$0.014/request
    Cache Storage$4.5/h/M tokens
    Context1M
    Max output65K
    Latency3.2S
    Throughput101.3TPS
    Uptime
    0.00% uptime 2 days ago
    0.00% uptime yesterday
    0.00% uptime today

    Performance for gemini-3-pro-preview

    Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).

    Uptime
    Loading...
    Latency
    Loading...
    Throughput
    Loading...

    Try this model

    Python
    import os
    from openai import OpenAI
    
    client = OpenAI(
        api_key=os.environ["AIHUBMIX_API_KEY"],
        base_url="https://shkq.org/v1",
    )
    
    response = client.chat.completions.create(
        model="gemini-3-pro-preview",
        messages=[
          {
            "role": "user",
            "content": "Hello, how are you?"
          }
        ],
        max_tokens=1024,
        stream=False,
    )
    
    print(response.choices[0].message.content)

    Frequently asked questions

    How much does Gemini 3 Pro Preview cost?

    On AIHubMix, Gemini 3 Pro Preview costs $2 per million input tokens and $12 per million output tokens. Cached input reads are billed at $0.5 per million tokens.