Z.ai logo

GLM 5.1

z-ai/glm-5.1
Chat
Z.ai|
Function CallingReasoning

GLM 5.1 (z-ai/glm-5.1) is an AI model from Z.ai with a 202,752-token context window and 65,535 max output tokens, priced at $1.05/1M input and $3.50/1M output tokens. Available via the haimaker.ai OpenAI-compatible API.

Context Window
203K
tokens
Max Output
66K
tokens
Input Price
$1.05
/1M tokens
Output Price
$3.50
/1M tokens

Overview

Zhipu AI's GLM (General Language Model) with reasoning and function calling capabilities.

Features & Capabilities

Modechat
Context Window202,752 tokens
Max Output65,535 tokens
Function CallingSupported
VisionNot supported
ReasoningSupported
Web SearchNot supported
Url ContextNot supported

API Usage

from openai import OpenAI

client = OpenAI(
    base_url="https://api.haimaker.ai/v1",
    api_key="YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="z-ai/glm-5.1",
    messages=[
        {"role": "user", "content": "Hello, how are you?"}
    ],
)

print(response.choices[0].message.content)

Frequently Asked Questions

What is the context window of GLM 5.1?

GLM 5.1 (z-ai/glm-5.1) has a 202,752-token context window and supports up to 65,535 output tokens per request.

How much does GLM 5.1 cost?

GLM 5.1 is priced at $1.05 per 1M input tokens and $3.50 per 1M output tokens when accessed via the haimaker.ai OpenAI-compatible API.

What features does GLM 5.1 support?

GLM 5.1 supports function calling, reasoning.

How do I use GLM 5.1 via API?

Send requests to https://api.haimaker.ai/v1/chat/completions with model "z-ai/glm-5.1" using any OpenAI-compatible SDK. Authentication uses a Bearer API key from https://app.haimaker.ai.

Use GLM 5.1 with the haimaker API

OpenAI-compatible endpoint. Start building in minutes.

Get API Access

haimaker.ai 2026 - All Rights Reserved