BazaarLinkBazaarLink
Sign in

Deepseek V4 Flash API Pricing & Quick Start

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

Intelligence#55 / 633
35
Artificial Analysis Intelligence Index
Speed#99 / 331
127.8
Output tokens per second (median)
Input Price#214 / 436
$0.20NT$ 6
USD / 1M tokens
Output Price#183 / 436
$0.18NT$ 6
USD / 1M tokens
First-token latency#57 / 331
0.92s
Time to first token (AA median)
ProviderDeepSeek
ReleasedApril 2026
Model IDdeepseek/deepseek-v4-flash

Technical specifications

Context window
1M tokens
Reasoning
Yes
Input
Text
Output
Text

Pricing comparison

BazaarLink retail price alongside AA's observed market standard.

BazaarLink price our site

Input$0.20NT$ 6
Output$0.18NT$ 6
Cache read$0.02NT$ 1
Cache write
3:1 blended (est.)$0.20NT$ 6
All prices per 1M tokens. Supports TWD credit card / official invoice.

Market reference Artificial Analysis

Input$0.44
Output$1.32
Cache readAA not tracked
Cache writeAA not tracked
3:1 blended
Our price differs from AA reference by ~55%.
Source: artificialanalysis.ai

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 35
Coding 69
Math
MMLU
GPQA
This modelCategory leader
Coding ability
#55out of 256 models
LiveCodeBench / SciCode and similar
Overall intelligence
#55out of 633 models
Aggregated across academic benchmarks
Time to first token
#57out of 331 models
AA 全球 p50:0.92 秒
Intelligence
35
Top 9%
This model
35
Category leader
53
Median
11
Coding
69
Top 21%
This model
69
Category leader
82
Median
44

Top 5 — CodingCoding Index leaderboard

1Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)?/claude-fable-5-182$10 / $50
2Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback)?/claude-fable-5-1-xhigh81$10 / $50
3Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)?/claude-fable-5-1-high79$10 / $50
4GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$4 / $20
5Claude Opus 5 (Adaptive Reasoning, Max Effort)anthropic/claude-opus-578$5 / $25
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →

Quick Start

Use Deepseek V4 Flash via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://api.bazaarlink.ai/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="deepseek/deepseek-v4-flash",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://api.bazaarlink.ai/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "deepseek/deepseek-v4-flash",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Deepseek V4 Flash via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Deepseek V4 Flash?

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

How much does the Deepseek V4 Flash API cost?

Deepseek V4 Flash costs $0.20 per 1M input tokens and $0.18 per 1M output tokens when accessed through BazaarLink.

How do I use Deepseek V4 Flash with the OpenAI SDK?

Set base_url to "https://api.bazaarlink.ai/v1" and use model ID "deepseek/deepseek-v4-flash". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Deepseek V4 Flash?

Deepseek V4 Flash supports a context window of 1,048,576 tokens.

Is Deepseek V4 Flash available for free?

Deepseek V4 Flash is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

Relay integrity checks

Tested endpoints
16 runs
Independent hosts
1
Model-family fingerprint match
94%
Anomalies detected
7 runs
Data range
past 90 days

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: DeepSeek

Deepseek R1 Distill Llama 70bDeepseek R1Deepseek Chat V3 0324Deepseek V3.2Deepseek V3.1 TerminusDeepseek V3.2 Exp
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.