BazaarLinkBazaarLink
Sign in

Gemini 2.5 Flash API Pricing & Quick Start

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

Intelligence#346 / 633
10
Artificial Analysis Intelligence Index
Speed#35 / 331
212.7
Output tokens per second (median)
Input Price#164 / 436
$0.30NT$ 9
USD / 1M tokens
Output Price#225 / 436
$2.50NT$ 79
USD / 1M tokens
First-token latency#3 / 331
0.44s
Time to first token (AA median)
ProviderGoogle
ReleasedJune 2025
Model IDgoogle/gemini-2.5-flash

Technical specifications

Context window
1M tokens
Reasoning
Yes
Input
FileImageTextAudioVideo
Output
Text

Pricing comparison

BazaarLink retail price alongside AA's observed market standard.

BazaarLink price our site

Input$0.30NT$ 9
Output$2.50NT$ 79
Cache read$0.03NT$ 1
Cache write$0.08NT$ 3
3:1 blended (est.)$0.85NT$ 27
All prices per 1M tokens. Supports TWD credit card / official invoice.

Market reference Artificial Analysis

Input$0.30
Output$2.50
Cache readAA not tracked
Cache writeAA not tracked
3:1 blended$0.85
✓ Matches market reference — BazaarLink does not mark up.
Source: artificialanalysis.ai

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 10
Coding
Math 60
MMLU 81
GPQA 68
This modelCategory leader
Overall intelligence
#346out of 633 models
Aggregated across academic benchmarks
Time to first token
🥉
#3out of 331 models
AA 全球 p50:0.44 秒
Intelligence
10
Mid-tier
This model
10
Category leader
53
Median
11
Math
60
This model
60
Category leader
99
Median
53
MMLU Pro
81%
This model
81%
Category leader
90%
Median
75%
GPQA
68%
This model
68%
Category leader
94%
Median
69%
LiveCodeBench
50%
This model
50%
Category leader
42%
Median
42%
HLE
5%
This model
5%
Category leader
53%
Median
7%
SciCode
29%
This model
29%
Category leader
60%
Median
33%
IFBench
39%
This model
39%
Category leader
83%
Median
44%
τ²-Bench Telecom
15%
This model
15%
Category leader
99%
Median
46%
AA-LCR
46%
This model
46%
Category leader
76%
Median
40%
Terminal-Bench Hard
12%
This model
12%
Category leader
66%
Median
14%

Top 5 — CodingCoding Index leaderboard

1Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)?/claude-fable-5-182$10 / $50
2Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback)?/claude-fable-5-1-xhigh81$10 / $50
3Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)?/claude-fable-5-1-high79$10 / $50
4GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$4 / $20
5Claude Opus 5 (Adaptive Reasoning, Max Effort)anthropic/claude-opus-578$5 / $25
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →

Quick Start

Use Gemini 2.5 Flash via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://api.bazaarlink.ai/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="google/gemini-2.5-flash",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://api.bazaarlink.ai/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "google/gemini-2.5-flash",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Gemini 2.5 Flash via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Gemini 2.5 Flash?

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

How much does the Gemini 2.5 Flash API cost?

Gemini 2.5 Flash costs $0.30 per 1M input tokens and $2.50 per 1M output tokens when accessed through BazaarLink.

How do I use Gemini 2.5 Flash with the OpenAI SDK?

Set base_url to "https://api.bazaarlink.ai/v1" and use model ID "google/gemini-2.5-flash". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Gemini 2.5 Flash?

Gemini 2.5 Flash supports a context window of 1,048,576 tokens.

Is Gemini 2.5 Flash available for free?

Gemini 2.5 Flash is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

Relay integrity checks

No integrity-check data is available for this model yet.

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: Google

Gemini 3.7 FlashGemini 3 Pro Image PreviewGemma 3 12b ItGemini 2.5 Pro Preview 05 06Gemini 3.1 Pro Preview CustomtoolsGemini 3.1 Flash Image Preview
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.