OpenRouter modele

Llama 3.2 11b Vision Instruct

Llama 3.2 11b Vision Instruct Comparez les tarifs API, le fournisseur, la longueur de contexte, les capacites, les cas d'usage, la latence et les alternatives. General chat via OpenRouter, API workloads. $0.065 / 1M tokens / $0.065 / 1M tokens. — tokens.

OpenRouterPlatform curatedCatalog
Streaming
Prix entree$0.065 / 1M tokens
Prix sortie$0.065 / 1M tokens
Longueur du contexte— tokens
Sortie max8.2k tokens

Qu'est-ce que Llama 3.2 11b Vision Instruct dans NextModel ?

Llama 3.2 11b Vision Instruct Comparez les tarifs API, le fournisseur, la longueur de contexte, les capacites, les cas d'usage, la latence et les alternatives. General chat via OpenRouter, API workloads. $0.065 / 1M tokens / $0.065 / 1M tokens. — tokens.

Meilleurs cas d'usage

  • General chat via OpenRouter
  • API workloads

Exemple compatible OpenAI

Conservez le style SDK OpenAI, pointez base_url vers NextModel et utilisez l'ID de catalogue meta-llama--llama-3.2-11b-vision-instruct.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="meta-llama/llama-3.2-11b-vision-instruct",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

Alternatives proches

OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.752 / 1M tokensInput$1.50 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.132 / 1M tokensInput$0.263 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

$0.152 / 1M tokensInput$0.302 / 1M tokensOutput32.8kContext
Best forGeneral chat via OpenRouter, API workloads
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details

FAQ

Llama 3.2 11b Vision Instruct FAQ