Gemma 4 26B A4B (free)
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at…
Full description at the source- Input price
- per 1M tokens
- Output price
- per 1M tokens
- Context
- 262.1K32.8K max output
- Cost / request
- 8K in · 500 out
01Routes and pricing
02Capabilities
- Input
- text
- Output
- text
- Context
- 262.1K
- Max output
- 32.8K
ReasoningTool callingJSON modeSeeded sampling
03Quick start
Call it with your RProuter key
Any OpenAI-compatible SDK works. Point it at the RProuter base URL from your API keys page and use this model ID.
Catalog source: openrouter.ai · checked 2026-09-16
const response = await client.chat.completions.create({
model: "google/gemma-4-26b-a4b-it:free",
messages: [{ role: "user", content: "A story begins…" }],
max_tokens: 500,
stream: true,
});
for await (const chunk of response) {
process.stdout.write(chunk.choices[0]?.delta?.content ?? "");
}