OpenRouter model

Phi 4

Microsoft: Phi 4. Vergelijk API-prijzen, provider, contextlengte, capaciteiten, use cases, latentie en alternatieven. General chat via OpenRouter, API workloads. $0.013 / 1M tokens / $0.027 / 1M tokens. 16.4k tokens.

OpenRouterPlatform curatedCatalog
StreamingJSON mode
Inputprijs$0.013 / 1M tokens
Outputprijs$0.027 / 1M tokens
Contextlengte16.4k tokens
Max output8.2k tokens

Wat is Phi 4 in NextModel?

Microsoft: Phi 4. Vergelijk API-prijzen, provider, contextlengte, capaciteiten, use cases, latentie en alternatieven. General chat via OpenRouter, API workloads. $0.013 / 1M tokens / $0.027 / 1M tokens. 16.4k tokens.

Beste toepassingen

  • General chat via OpenRouter
  • API workloads

OpenAI-compatibel voorbeeld

Houd de OpenAI SDK-stijl aan, wijs base_url naar NextModel en gebruik catalogus-ID microsoft--phi-4.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="microsoft/phi-4",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

Vergelijkbare alternatieven

OpenRouterCatalog

Olmo 3 32B Think is a large-scale, 32-billion-parameter model purpose-built for deep reasoning, complex logic chains and advanced instruction-following scenarios. Its capacity enables strong performance on demanding evaluation tasks and...

$0.029 / 1M tokensInput$0.094 / 1M tokensOutput65.5kContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
StreamingJSON mode
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).

$0.564 / 1M tokensInput$0.94 / 1M tokensOutput16.4kContext
Best forGeneral chat via OpenRouter, API workloads
RoutingConfigured
StreamingJSON mode
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

One of the highest performing and most popular fine-tunes of Llama 2 13B, with rich descriptions and roleplay. #merge

$0.012 / 1M tokensInput$0.012 / 1M tokensOutput8.2kContext
Best forGeneral chat via OpenRouter, API workloads
RoutingConfigured
StreamingJSON mode
Platform curatedNextModel gateway catalog (Go origin)
View details

FAQ

Microsoft: Phi 4 FAQ