Processing...Please wait while we secure this action
OpenRouter model

Llama 3.3 Nemotron Super 49b V1.5

Llama 3.3 Nemotron Super 49b V1.5 is a OpenRouter model listed in the NextModel catalog for General chat via OpenRouter, API workloads workloads. Its listed price is $0.075 / 1M tokens input and $0.075 / 1M tokens output per 1M tokens, with a — token context window.

OpenRouterPlatform curatedCatalog
Streaming
Input price$0.075 / 1M tokens
Output price$0.075 / 1M tokens
Context length— tokens
Max output8.2k tokens

What is Llama 3.3 Nemotron Super 49b V1.5 in NextModel?

Llama 3.3 Nemotron Super 49b V1.5 is a OpenRouter model listed in the NextModel catalog for General chat via OpenRouter, API workloads workloads. Its listed price is $0.075 / 1M tokens input and $0.075 / 1M tokens output per 1M tokens, with a — token context window.

Best use cases

  • General chat via OpenRouter
  • API workloads

OpenAI-compatible code example

Keep the OpenAI SDK style, set base_url to NextModel, and use the catalog model ID nvidia--llama-3.3-nemotron-super-49b-v1.5.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="nvidia/llama-3.3-nemotron-super-49b-v1.5",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

Similar alternatives

OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.752 / 1M tokensInput$1.50 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.132 / 1M tokensInput$0.263 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

$0.152 / 1M tokensInput$0.302 / 1M tokensOutput32.8kContext
Best forGeneral chat via OpenRouter, API workloads
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details

FAQ

Llama 3.3 Nemotron Super 49b V1.5 API questions