Schematron V2 Turbo

inference-net/schematron-v2-turbo
Chat

Schematron V2 Turbo (inference-net/schematron-v2-turbo) is an AI model from Inference Net with a 128,000-token context window and 8,192 max output tokens, priced at $0.03/1M input and $0.15/1M output tokens. Available via the haimaker.ai OpenAI-compatible API.

Context Window
128K
tokens
Max Output
8K
tokens
Input Price
$0.03
/1M tokens
Output Price
$0.15
/1M tokens
Cached Input
$0.030
/1M tokens

Overview

Schematron V2 Turbo is a chat model by Inference Net. It supports a 128K token context window.

Features & Capabilities

Modechat
Context Window128,000 tokens
Max Output8,192 tokens
Function CallingNot supported
VisionNot supported
ReasoningNot supported
Web SearchNot supported
Url ContextNot supported

API Usage

from openai import OpenAI

client = OpenAI(
    base_url="https://api.haimaker.ai/v1",
    api_key="YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="inference-net/schematron-v2-turbo",
    messages=[
        {"role": "user", "content": "Hello, how are you?"}
    ],
)

print(response.choices[0].message.content)

Frequently Asked Questions

What is the context window of Schematron V2 Turbo?

Schematron V2 Turbo (inference-net/schematron-v2-turbo) has a 128,000-token context window and supports up to 8,192 output tokens per request.

How much does Schematron V2 Turbo cost?

Schematron V2 Turbo is priced at $0.03 per 1M input tokens and $0.15 per 1M output tokens when accessed via the haimaker.ai OpenAI-compatible API.

How do I use Schematron V2 Turbo via API?

Send requests to https://api.haimaker.ai/v1/chat/completions with model "inference-net/schematron-v2-turbo" using any OpenAI-compatible SDK. Authentication uses a Bearer API key from https://app.haimaker.ai.

Use Schematron V2 Turbo with the haimaker API

OpenAI-compatible endpoint. Start building in minutes.

Get API Access

More from Inference Net

haimaker.ai 2026 - All Rights Reserved