BazaarLinkBazaarLink
Sign in

Glm 4.6v API Pricing & Quick Start

GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...

Intelligence#401 / 633
8
Artificial Analysis Intelligence Index
Speed#294 / 331
44.5
Output tokens per second (median)
Input Price#164 / 436
$0.30NT$ 9
USD / 1M tokens
Output Price#145 / 436
$0.90NT$ 28
USD / 1M tokens
First-token latency#228 / 331
3.94s
Time to first token (AA median)
ProviderZhipu AI
ReleasedDecember 2025
Model IDz-ai/glm-4.6v

Technical specifications

Context window
131K tokens
Reasoning
Yes
Input
ImageTextVideo
Output
Text

Pricing comparison

BazaarLink retail price alongside AA's observed market standard.

BazaarLink price our site

Input$0.30NT$ 9
Output$0.90NT$ 28
Cache read$0.06NT$ 2
Cache write
3:1 blended (est.)$0.45NT$ 14
All prices per 1M tokens. Supports TWD credit card / official invoice.

Market reference Artificial Analysis

Input$0.30
Output$0.90
Cache readAA not tracked
Cache writeAA not tracked
3:1 blended$0.45
✓ Matches market reference — BazaarLink does not mark up.
Source: artificialanalysis.ai

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 8
Coding
Math 26
MMLU 75
GPQA 57
This modelCategory leader
Overall intelligence
#401out of 633 models
Aggregated across academic benchmarks
Time to first token
#228out of 331 models
AA 全球 p50:3.94 秒
Intelligence
8
Mid-tier
This model
8
Category leader
53
Median
11
Math
26
This model
26
Category leader
99
Median
53
MMLU Pro
75%
This model
75%
Category leader
90%
Median
75%
GPQA
57%
This model
57%
Category leader
94%
Median
69%
LiveCodeBench
41%
This model
41%
Category leader
42%
Median
42%
HLE
4%
This model
4%
Category leader
53%
Median
7%
SciCode
27%
This model
27%
Category leader
60%
Median
33%
IFBench
28%
This model
28%
Category leader
83%
Median
44%
τ²-Bench Telecom
31%
This model
31%
Category leader
99%
Median
46%
AA-LCR
12%
This model
12%
Category leader
76%
Median
40%
Terminal-Bench Hard
3%
This model
3%
Category leader
66%
Median
14%

Top 5 — CodingCoding Index leaderboard

1Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)?/claude-fable-5-182$10 / $50
2Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback)?/claude-fable-5-1-xhigh81$10 / $50
3Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)?/claude-fable-5-1-high79$10 / $50
4GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$4 / $20
5Claude Opus 5 (Adaptive Reasoning, Max Effort)anthropic/claude-opus-578$5 / $25
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →

Quick Start

Use Glm 4.6v via BazaarLink API — just change the base_url:

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://api.bazaarlink.ai/v1",
    api_key="sk-bl-YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="z-ai/glm-4.6v",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
TypeScript
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://api.bazaarlink.ai/v1",
  apiKey: "sk-bl-YOUR_API_KEY",
});

const response = await client.chat.completions.create({
  model: "z-ai/glm-4.6v",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(response.choices[0].message.content);

Why use Glm 4.6v via BazaarLink?

  • USD billing (TWD-quoted) + unified invoices — no foreign credit card needed for Taiwan teams
  • OpenAI-compatible API — zero code changes required
  • Automatic failover — multi-provider redundancy for the same model
  • Chinese-language support — local team, instant help
Try Now← All Models

Frequently Asked Questions

What is Glm 4.6v?

GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...

How much does the Glm 4.6v API cost?

Glm 4.6v costs $0.30 per 1M input tokens and $0.90 per 1M output tokens when accessed through BazaarLink.

How do I use Glm 4.6v with the OpenAI SDK?

Set base_url to "https://api.bazaarlink.ai/v1" and use model ID "z-ai/glm-4.6v". All OpenAI SDK methods (chat.completions, embeddings, streaming) work without code changes.

What is the context window for Glm 4.6v?

Glm 4.6v supports a context window of 131,072 tokens.

Is Glm 4.6v available for free?

Glm 4.6v is a paid model. BazaarLink offers free trial credits on registration so you can test it without a credit card.

Relay integrity checks

No integrity-check data is available for this model yet.

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

More from this provider: Zhipu AI

Glm 5.3Glm 4.7Glm 4.6Glm 5Glm 4.5vGlm 4.7 Flash
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.