Schematron V2 Turbo
inference-net/schematron-v2-turboSchematron V2 Turbo (inference-net/schematron-v2-turbo) is an AI model from Inference Net with a 128,000-token context window and 8,192 max output tokens, priced at $0.03/1M input and $0.15/1M output tokens. Available via the haimaker.ai OpenAI-compatible API.
Overview
Schematron V2 Turbo is a chat model by Inference Net. It supports a 128K token context window.
Features & Capabilities
| Mode | chat |
| Context Window | 128,000 tokens |
| Max Output | 8,192 tokens |
| Function Calling | Not supported |
| Vision | Not supported |
| Reasoning | Not supported |
| Web Search | Not supported |
| Url Context | Not supported |
API Usage
from openai import OpenAI
client = OpenAI(
base_url="https://api.haimaker.ai/v1",
api_key="YOUR_API_KEY",
)
response = client.chat.completions.create(
model="inference-net/schematron-v2-turbo",
messages=[
{"role": "user", "content": "Hello, how are you?"}
],
)
print(response.choices[0].message.content)Frequently Asked Questions
What is the context window of Schematron V2 Turbo?
Schematron V2 Turbo (inference-net/schematron-v2-turbo) has a 128,000-token context window and supports up to 8,192 output tokens per request.
How much does Schematron V2 Turbo cost?
Schematron V2 Turbo is priced at $0.03 per 1M input tokens and $0.15 per 1M output tokens when accessed via the haimaker.ai OpenAI-compatible API.
How do I use Schematron V2 Turbo via API?
Send requests to https://api.haimaker.ai/v1/chat/completions with model "inference-net/schematron-v2-turbo" using any OpenAI-compatible SDK. Authentication uses a Bearer API key from https://app.haimaker.ai.
Use Schematron V2 Turbo with the haimaker API
OpenAI-compatible endpoint. Start building in minutes.