OpenRouter model

Llama 3.3 Nemotron Super 49b V1.5

Llama 3.3 Nemotron Super 49b V1.5 Bandingkan harga API, penyedia, panjang konteks, kemampuan, use case, latensi, dan alternatif. General chat via OpenRouter, API workloads. $0.075 / 1M tokens / $0.075 / 1M tokens. — tokens.

OpenRouterPlatform curatedCatalog
Streaming
Harga input$0.075 / 1M tokens
Harga output$0.075 / 1M tokens
Panjang konteks— tokens
Output maksimum8.2k tokens

Apa itu Llama 3.3 Nemotron Super 49b V1.5 di NextModel?

Llama 3.3 Nemotron Super 49b V1.5 Bandingkan harga API, penyedia, panjang konteks, kemampuan, use case, latensi, dan alternatif. General chat via OpenRouter, API workloads. $0.075 / 1M tokens / $0.075 / 1M tokens. — tokens.

Use case terbaik

  • General chat via OpenRouter
  • API workloads

Contoh kompatibel OpenAI

Pertahankan gaya SDK OpenAI, arahkan base_url ke NextModel, dan gunakan ID katalog nvidia--llama-3.3-nemotron-super-49b-v1.5.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="nvidia/llama-3.3-nemotron-super-49b-v1.5",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

Alternatif serupa

OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.752 / 1M tokensInput$1.50 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Available through the NextModel gateway via OpenRouter.

$0.132 / 1M tokensInput$0.263 / 1M tokensOutputContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

$0.152 / 1M tokensInput$0.302 / 1M tokensOutput32.8kContext
Best forGeneral chat via OpenRouter, API workloads
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details

FAQ

Llama 3.3 Nemotron Super 49b V1.5 FAQ