OpenRouter model

Llama 3.3 Nemotron Super 49b V1.5

Llama 3.3 Nemotron Super 49b V1.5 So sanh gia API, nha cung cap, do dai context, kha nang, use case, do tre va cac lua chon thay the. General chat via OpenRouter, API workloads. $0.075 / 1M tokens / $0.075 / 1M tokens. — tokens.

OpenRouterPlatform curatedCatalog
Streaming
Gia input$0.075 / 1M tokens
Gia output$0.075 / 1M tokens
Do dai context— tokens
Output toi da8.2k tokens

Llama 3.3 Nemotron Super 49b V1.5 trong NextModel la gi?

Llama 3.3 Nemotron Super 49b V1.5 So sanh gia API, nha cung cap, do dai context, kha nang, use case, do tre va cac lua chon thay the. General chat via OpenRouter, API workloads. $0.075 / 1M tokens / $0.075 / 1M tokens. — tokens.

Use case tot nhat

  • General chat via OpenRouter
  • API workloads

Vi du tuong thich OpenAI

Giu nguyen kieu SDK OpenAI, tro base_url toi NextModel va dung ID catalog nvidia--llama-3.3-nemotron-super-49b-v1.5.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="nvidia/llama-3.3-nemotron-super-49b-v1.5",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

Lua chon thay the tuong tu

OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.752 / 1M tokensInput$1.50 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.132 / 1M tokensInput$0.263 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

$0.152 / 1M tokensInput$0.302 / 1M tokensOutput32.8kContext
Best forGeneral chat via OpenRouter, API workloads
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details

FAQ

Llama 3.3 Nemotron Super 49b V1.5 FAQ