Phi 4
Microsoft Research Phi-4 is designed to perform well in complex reasoning tasks and can operate efficiently in situations with limited memory or where quick responses are needed. At 14 billion…
Full description at the source- Input price
- per 1M tokens
- Output price
- per 1M tokens
- Context
- 16.4K14.7K max output
- Cost / request
- 8K in · 500 out
01Routes and pricing
02Capabilities
- Input
- text
- Output
- text
- Context
- 16.4K
- Max output
- 14.7K
Structured outputJSON modeSeeded sampling
03Quick start
Call it with your RProuter key
Any OpenAI-compatible SDK works. Point it at the RProuter base URL from your API keys page and use this model ID.
Catalog source: openrouter.ai · checked 2026-09-16
const response = await client.chat.completions.create({
model: "microsoft/phi-4",
messages: [{ role: "user", content: "A story begins…" }],
max_tokens: 500,
stream: true,
});
for await (const chunk of response) {
process.stdout.write(chunk.choices[0]?.delta?.content ?? "");
}