Llama 3.2 11B Vision Instruct
meta-llama/llama-3.2-11b-vision-instructLlama 3.2 11B Vision Instruct (meta-llama/llama-3.2-11b-vision-instruct) is a mllama 10.7B-parameter model from Meta Llama with a 131,072-token context window and 16,384 max output tokens, priced at $0.24/1M input and $0.24/1M output tokens. Available via the haimaker.ai OpenAI-compatible API.
Overview
Llama 3.2 11B Vision Instruct is a 10.67 billion parameter multimodal model developed by Meta Llama, capable of processing both text and visual inputs. It supports a 131,072 token context window and generates responses up to 16,384 tokens, with equal pricing for input and output tokens. The model operates under the llama3.2 license and uses the MllamaForConditionalGeneration architecture.
Features & Capabilities
| Mode | chat |
| Context Window | 131,072 tokens |
| Max Output | 16,384 tokens |
| Function Calling | Not supported |
| Vision | Supported |
| Reasoning | Not supported |
| Web Search | Not supported |
| Url Context | Not supported |
Technical Details
| Architecture | MllamaForConditionalGeneration |
| Model Type | mllama |
| Languages | en, de, fr, it, pt, hi, es, th |
| Library | transformers |
API Usage
from openai import OpenAI
client = OpenAI(
base_url="https://api.haimaker.ai/v1",
api_key="YOUR_API_KEY",
)
response = client.chat.completions.create(
model="meta-llama/llama-3.2-11b-vision-instruct",
messages=[
{"role": "user", "content": "Hello, how are you?"}
],
)
print(response.choices[0].message.content)Frequently Asked Questions
What is the context window of Llama 3.2 11B Vision Instruct?
Llama 3.2 11B Vision Instruct (meta-llama/llama-3.2-11b-vision-instruct) has a 131,072-token context window and supports up to 16,384 output tokens per request.
How much does Llama 3.2 11B Vision Instruct cost?
Llama 3.2 11B Vision Instruct is priced at $0.24 per 1M input tokens and $0.24 per 1M output tokens when accessed via the haimaker.ai OpenAI-compatible API.
What features does Llama 3.2 11B Vision Instruct support?
Llama 3.2 11B Vision Instruct supports vision.
How do I use Llama 3.2 11B Vision Instruct via API?
Send requests to https://api.haimaker.ai/v1/chat/completions with model "meta-llama/llama-3.2-11b-vision-instruct" using any OpenAI-compatible SDK. Authentication uses a Bearer API key from https://app.haimaker.ai.
Use Llama 3.2 11B Vision Instruct with the haimaker API
OpenAI-compatible endpoint. Start building in minutes.