OpenRouter modelo

Llama 3.3 Nemotron Super 49b V1.5

Llama 3.3 Nemotron Super 49b V1.5 Compara precios API, proveedor, longitud de contexto, capacidades, casos de uso, latencia y alternativas. General chat via OpenRouter, API workloads. $0.075 / 1M tokens / $0.075 / 1M tokens. — tokens.

OpenRouterPlatform curatedCatalog
Streaming
Precio entrada$0.075 / 1M tokens
Precio salida$0.075 / 1M tokens
Longitud de contexto— tokens
Salida maxima8.2k tokens

Que es Llama 3.3 Nemotron Super 49b V1.5 en NextModel?

Llama 3.3 Nemotron Super 49b V1.5 Compara precios API, proveedor, longitud de contexto, capacidades, casos de uso, latencia y alternativas. General chat via OpenRouter, API workloads. $0.075 / 1M tokens / $0.075 / 1M tokens. — tokens.

Mejores casos de uso

  • General chat via OpenRouter
  • API workloads

Ejemplo compatible con OpenAI

Mantén el estilo del SDK de OpenAI, apunta base_url a NextModel y usa el ID de catalogo nvidia--llama-3.3-nemotron-super-49b-v1.5.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="nvidia/llama-3.3-nemotron-super-49b-v1.5",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

Alternativas similares

OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.752 / 1M tokensInput$1.50 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.132 / 1M tokensInput$0.263 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

$0.152 / 1M tokensInput$0.302 / 1M tokensOutput32.8kContext
Best forGeneral chat via OpenRouter, API workloads
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details

FAQ

Llama 3.3 Nemotron Super 49b V1.5 FAQ