Gemma 3 12B It
google/gemma-3-12b-itGemma 3 12B It (google/gemma-3-12b-it) is an AI model from Google with a 131,072-token context window and 16,384 max output tokens, priced at $0.05/1M input and $0.15/1M output tokens. Available via the haimaker.ai OpenAI-compatible API.
Overview
Gemma 3 12B It is a chat-based AI model developed by Google with an unknown parameter count. It supports a context window of 131,072 tokens and has a maximum output length of 16,384 tokens. The model includes function calling and vision capabilities, with pricing set at $0.05 per million input tokens and $0.15 per million output tokens. Its architecture and license details are currently unspecified.
Features & Capabilities
| Mode | chat |
| Context Window | 131,072 tokens |
| Max Output | 16,384 tokens |
| Function Calling | Supported |
| Vision | Supported |
| Reasoning | Not supported |
| Web Search | Not supported |
| Url Context | Not supported |
API Usage
from openai import OpenAI
client = OpenAI(
base_url="https://api.haimaker.ai/v1",
api_key="YOUR_API_KEY",
)
response = client.chat.completions.create(
model="google/gemma-3-12b-it",
messages=[
{"role": "user", "content": "Hello, how are you?"}
],
)
print(response.choices[0].message.content)Frequently Asked Questions
What is the context window of Gemma 3 12B It?
Gemma 3 12B It (google/gemma-3-12b-it) has a 131,072-token context window and supports up to 16,384 output tokens per request.
How much does Gemma 3 12B It cost?
Gemma 3 12B It is priced at $0.05 per 1M input tokens and $0.15 per 1M output tokens when accessed via the haimaker.ai OpenAI-compatible API.
What features does Gemma 3 12B It support?
Gemma 3 12B It supports function calling, vision.
How do I use Gemma 3 12B It via API?
Send requests to https://api.haimaker.ai/v1/chat/completions with model "google/gemma-3-12b-it" using any OpenAI-compatible SDK. Authentication uses a Bearer API key from https://app.haimaker.ai.
Use Gemma 3 12B It with the haimaker API
OpenAI-compatible endpoint. Start building in minutes.