OpenRouter modelo

Llama 3.3 Nemotron Super 49b V1.5

Llama 3.3 Nemotron Super 49b V1.5 Compare precos de API, provedor, comprimento de contexto, capacidades, casos de uso, latencia e alternativas. General chat via OpenRouter, API workloads. $0.075 / 1M tokens / $0.075 / 1M tokens. — tokens.

OpenRouterPlatform curatedCatalog
Streaming
Preco de entrada$0.075 / 1M tokens
Preco de saida$0.075 / 1M tokens
Comprimento de contexto— tokens
Saida maxima8.2k tokens

O que e Llama 3.3 Nemotron Super 49b V1.5 no NextModel?

Llama 3.3 Nemotron Super 49b V1.5 Compare precos de API, provedor, comprimento de contexto, capacidades, casos de uso, latencia e alternativas. General chat via OpenRouter, API workloads. $0.075 / 1M tokens / $0.075 / 1M tokens. — tokens.

Melhores casos de uso

  • General chat via OpenRouter
  • API workloads

Exemplo compativel com OpenAI

Mantenha o estilo do SDK OpenAI, aponte base_url para NextModel e use o ID do catalogo nvidia--llama-3.3-nemotron-super-49b-v1.5.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="nvidia/llama-3.3-nemotron-super-49b-v1.5",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

Alternativas parecidas

OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.752 / 1M tokensInput$1.50 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.132 / 1M tokensInput$0.263 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

$0.152 / 1M tokensInput$0.302 / 1M tokensOutput32.8kContext
Best forGeneral chat via OpenRouter, API workloads
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details

FAQ

Llama 3.3 Nemotron Super 49b V1.5 FAQ