gemini-3.6-flash
Google's Flash model with near-Pro intelligence at Flash-tier cost and speed: strong coding, parallel agentic execution, thinking, and search grounding.
Compact per-layer-embedding variant of Gemma 4 from Google DeepMind. Text-only input, text output, with toggleable reasoning.
gemma-4-e4bCreate an account, top up in dinars and call it with your key:
from openai import OpenAI
client = OpenAI(
base_url="https://dzrouter.com/v1",
api_key="sk-dz-…",
)
reply = client.chat.completions.create(
model="gemma-4-e4b",
messages=[{"role": "user", "content": "Salam!"}],
)
print(reply.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://dzrouter.com/v1",
apiKey: process.env.DZROUTER_API_KEY,
});
const reply = await client.chat.completions.create({
model: "gemma-4-e4b",
messages: [{ role: "user", content: "Salam!" }],
});
console.log(reply.choices[0].message.content);curl https://dzrouter.com/v1/chat/completions \
-H "Authorization: Bearer $DZROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemma-4-e4b",
"messages": [{"role": "user", "content": "Salam!"}],
"stream": true
}'from anthropic import Anthropic
client = Anthropic(
base_url="https://dzrouter.com",
api_key="sk-dz-…",
)
msg = client.messages.create(
model="gemma-4-e4b",
max_tokens=1024,
messages=[{"role": "user", "content": "Salam!"}],
)
print(msg.content[0].text)gemini-3.6-flash
Google's Flash model with near-Pro intelligence at Flash-tier cost and speed: strong coding, parallel agentic execution, thinking, and search grounding.
google/gemini-3-flash-preview
Google's latest model with the Flash line's focus on latency, efficiency, and cost, in preview version.
google/gemini-2.5-flash-lite
Google's most cost-efficient and lowest-latency multimodal model in the 2.5 family, optimized for high-volume classification, simple data extraction, and low-latency tasks.
google/gemini-2.5-flash
Google's fast and cost-effective model with a 1M token context window. Best for high-volume, low-latency tasks and agentic use cases.
google/gemini-3.1-flash-lite
Google's lightweight and efficient model in the 3.1 family, optimized for low latency and cost, in preview version.
google/gemma-4-31b-it
Dense 30.7B multimodal model from Google DeepMind. Multimodal (text + image input), text output.
Create your account, top up from 500 DA and send your first request in minutes.
Create a free account