OpenRouter model

Llama 3.2 1B Instruct

Meta: Llama 3.2 1B Instruct So sanh gia API, nha cung cap, do dai context, kha nang, use case, do tre va cac lua chon thay the. General chat via OpenRouter, API workloads. $0.0058 / 1M tokens / $0.039 / 1M tokens. 60k tokens.

OpenRouterPlatform curatedCatalog
Streaming
Gia input$0.0058 / 1M tokens
Gia output$0.039 / 1M tokens
Do dai context60k tokens
Output toi da8.2k tokens

Llama 3.2 1B Instruct trong NextModel la gi?

Meta: Llama 3.2 1B Instruct So sanh gia API, nha cung cap, do dai context, kha nang, use case, do tre va cac lua chon thay the. General chat via OpenRouter, API workloads. $0.0058 / 1M tokens / $0.039 / 1M tokens. 60k tokens.

Use case tot nhat

  • General chat via OpenRouter
  • API workloads

Vi du tuong thich OpenAI

Giu nguyen kieu SDK OpenAI, tro base_url toi NextModel va dung ID catalog meta-llama--llama-3.2-1b-instruct.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="meta-llama/llama-3.2-1b-instruct",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

Lua chon thay the tuong tu

OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.752 / 1M tokensInput$1.50 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.132 / 1M tokensInput$0.263 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

$0.152 / 1M tokensInput$0.302 / 1M tokensOutput32.8kContext
Best forGeneral chat via OpenRouter, API workloads
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details

FAQ

Meta: Llama 3.2 1B Instruct FAQ