Open-source AI infrastructure

Llama 3.2 11B Vision Instruct Turbo

Review Llama 3.2 11B Vision Instruct Turbo on Together AI: token pricing, context window, capabilities, modalities, and routing details in the Everstack model catal…

Llama 3.2 11B Vision Instruct Turbo pricing and limits on Together AI

Token pricing through Together AI: $0.900 per million input tokens, $0.900 per million output tokens. Limits: 131K token context window, 8K max output tokens.

Capabilities and modalities

Supported capabilities: chat, vision. Accepts text, image input. Returns text output. Model family: llama. Catalog status: deprecated.

Routing Llama 3.2 11B Vision Instruct Turbo through Everstack

Call Llama 3.2 11B Vision Instruct Turbo on Together AI through one OpenAI-compatible endpoint, with provider fallback, semantic caching, per-tenant rate limits, and OpenTelemetry traces for latency, tokens, and cost.