Haimaker.ai Logo
Z Ai logo

Glm 4.7 Flash

z-ai/glm-4.7-flash
Chat
Z Ai|
Function CallingVisionReasoning

Glm 4.7 Flash (z-ai/glm-4.7-flash) is an AI model from Z Ai with a 131,072-token context window and 131,072 max output tokens, priced at $0.07/1M input and $0.40/1M output tokens. Available via the haimaker.ai OpenAI-compatible API.

Context Window
131K
tokens
Max Output
131K
tokens
Input Price
$0.07
/1M tokens
Output Price
$0.40
/1M tokens

Overview

Zhipu AI's GLM (General Language Model) with reasoning and function calling capabilities.

Features & Capabilities

Modechat
Context Window131,072 tokens
Max Output131,072 tokens
Function CallingSupported
VisionSupported
ReasoningSupported
Web Search-
Url Context-

API Usage

from openai import OpenAI

client = OpenAI(
    base_url="https://api.haimaker.ai/v1",
    api_key="YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="z-ai/glm-4.7-flash",
    messages=[
        {"role": "user", "content": "Hello, how are you?"}
    ],
)

print(response.choices[0].message.content)

Frequently Asked Questions

What is the context window of Glm 4.7 Flash?

Glm 4.7 Flash (z-ai/glm-4.7-flash) has a 131,072-token context window and supports up to 131,072 output tokens per request.

How much does Glm 4.7 Flash cost?

Glm 4.7 Flash is priced at $0.07 per 1M input tokens and $0.40 per 1M output tokens when accessed via the haimaker.ai OpenAI-compatible API.

What features does Glm 4.7 Flash support?

Glm 4.7 Flash supports function calling, vision, reasoning.

How do I use Glm 4.7 Flash via API?

Send requests to https://api.haimaker.ai/v1/chat/completions with model "z-ai/glm-4.7-flash" using any OpenAI-compatible SDK. Authentication uses a Bearer API key from https://app.haimaker.ai.

Use Glm 4.7 Flash with the haimaker API

OpenAI-compatible endpoint. Start building in minutes.

Get API Access

More from Z Ai