All models

DeepSeek

DeepSeek V4 Pro 0813

DeepSeek’s largest V4 model for reasoning, coding and agent work, in its general availability release of August 2026.

Start buildingOpen dashboard

Context window
1M tokens
Accepts
Text
Input per 1M tokens
$2.00
Output per 1M tokens
$4.00

Send your first request

Set LYCEUM_API_KEY to your API key, then run this request. Usage is billed to your account.

Base URL
https://api.lyceum.technology/openai/v1
Model ID
deepseek/deepseek-v4-pro-0813
Request fields
model, messages, stream, max_tokens, tools, tool_choice

Documented request fields for this endpoint. See the documentation for model-specific controls and limits.

app.py · Install openai and set LYCEUM_API_KEY

import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["LYCEUM_API_KEY"],
    base_url="https://api.lyceum.technology/openai/v1",
)

response = client.chat.completions.create(
    model="deepseek/deepseek-v4-pro-0813",
    messages=[{"role": "user", "content": "Hello!"}],
)

print(response.choices[0].message.content)

What it costs

You pay only for the tokens you use.

Token pricing

DeepSeek V4 Pro 0813: token prices in US dollars per million tokens
USD per million tokens
Input$2.00
Cached input$0.50
Output$4.00

No base fee. Caching applies only where listed. Check the model's processing region before sending data.

The model in detail

Checked against the sources below. Anything they don't state is left out.

Capabilities

Input and output
Text in, text out
Context window
1M tokens
Reasoning
Off by default
Tool calling
Supported

Architecture and licence

Total parameters
1.6T
Active parameters
49B
Architecture
Mixture of experts
Licence
MIT
Released
13 August 2026

Availability

API model ID
deepseek/deepseek-v4-pro-0813
Provider
DeepSeek
Processing region
EU-hosted
Sources

Common questions about DeepSeek V4 Pro 0813

Short answers about the model ID, price, limits and behaviour.

What is the model ID for DeepSeek V4 Pro 0813?

Use deepseek/deepseek-v4-pro-0813 as the model value. The OpenAI-compatible base URL is https://api.lyceum.technology/openai/v1. You authenticate with your Lyceum API key as a Bearer token.

How much does DeepSeek V4 Pro 0813 cost?

DeepSeek V4 Pro 0813 costs $2.00 per 1M input tokens, $0.50 per 1M cached input tokens and $4.00 per 1M output tokens. Prices are in US dollars. Billed per token. No base fee.

What is the context window of DeepSeek V4 Pro 0813?

DeepSeek V4 Pro 0813 has a 1M tokens context window. One request can generate up to 65,536 tokens and run for up to 300 seconds. Set stream: true for long outputs.

Can DeepSeek V4 Pro 0813 read images?

No, not today. Image requests are rejected with a 400 error. Pick a model that reads images if you need that.

Can I turn off reasoning for DeepSeek V4 Pro 0813?

Yes. Reasoning is on by default. Send reasoning_effort: "none" in the request. The reasoning trace counts as output tokens when it is on, so set max_tokens high enough for both the reasoning and the answer.

Does DeepSeek V4 Pro 0813 support tool calling?

Yes. Send tools in the standard OpenAI format. Automatic tool calls, parallel calls and tool-result round trips work. It does not enforce tool_choice: "required" or a named function. It can answer in plain text with no tool calls, so check message.tool_calls and fall back when it is empty.

Does DeepSeek V4 Pro 0813 support prompt caching?

Yes, and it is automatic. When requests share an identical start, the cached part is billed at $0.50 per 1M tokens instead of $2.00. Keep the stable part of your prompt at the front. A cache hit is best effort, so read usage.prompt_tokens_details.cached_tokens, which is missing on a miss.

Where does DeepSeek V4 Pro 0813 run, and is my data kept?

DeepSeek V4 Pro 0813 runs on EU-hosted infrastructure. Prompts and outputs are processed, not stored, and never used for training.

How do I call DeepSeek V4 Pro 0813 with the OpenAI SDK?

Point the OpenAI SDK at https://api.lyceum.technology/openai/v1, use your Lyceum API key and set model="deepseek/deepseek-v4-pro-0813". Chat completions, streaming, tool calling and usage accounting work as they do with OpenAI.

Behaviour notes follow the Lyceum API documentation, checked 6 October 2026.

Build with DeepSeek V4 Pro 0813