qwen/qwen3-max-thinkingQwen3 Max Thinking (qwen/qwen3-max-thinking) is an AI model from Qwen with a 262,144-token context window and 32,768 max output tokens, priced at $0.78/1M input and $3.90/1M output tokens. Available via the haimaker.ai OpenAI-compatible API.
Qwen3 Max Thinking is a chat model by Qwen. It supports a 262K token context window. Supports function calling, reasoning.
| Mode | chat |
| Context Window | 262,144 tokens |
| Max Output | 32,768 tokens |
| Function Calling | Supported |
| Vision | - |
| Reasoning | Supported |
| Web Search | - |
| Url Context | - |
from openai import OpenAI
client = OpenAI(
base_url="https://api.haimaker.ai/v1",
api_key="YOUR_API_KEY",
)
response = client.chat.completions.create(
model="qwen/qwen3-max-thinking",
messages=[
{"role": "user", "content": "Hello, how are you?"}
],
)
print(response.choices[0].message.content)Qwen3 Max Thinking (qwen/qwen3-max-thinking) has a 262,144-token context window and supports up to 32,768 output tokens per request.
Qwen3 Max Thinking is priced at $0.78 per 1M input tokens and $3.90 per 1M output tokens when accessed via the haimaker.ai OpenAI-compatible API.
Qwen3 Max Thinking supports function calling, reasoning.
Send requests to https://api.haimaker.ai/v1/chat/completions with model "qwen/qwen3-max-thinking" using any OpenAI-compatible SDK. Authentication uses a Bearer API key from https://app.haimaker.ai.
OpenAI-compatible endpoint. Start building in minutes.