Qwen3.8 27B

A compact, dense Qwen3.8 model that reads text and images, bringing the family’s coding and agent gains to a smaller size.

Context window
256K tokens
Accepts
Text and images
Input per 1M tokens
$0.40
Output per 1M tokens
$2.40

Send your first request

Set LYCEUM_API_KEY to your API key, then run this request. Usage is billed to your account.

Base URL
https://api.lyceum.technology/openai/v1
Model ID
qwen/qwen3.8-27b
Request fields
model, messages, stream, max_tokens, tools, tool_choice

Documented request fields for this endpoint. See the documentation for model-specific controls and limits.

app.py · Install openai and set LYCEUM_API_KEY

import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["LYCEUM_API_KEY"],
    base_url="https://api.lyceum.technology/openai/v1",
)

response = client.chat.completions.create(
    model="qwen/qwen3.8-27b",
    messages=[{"role": "user", "content": "Hello!"}],
)

print(response.choices[0].message.content)

What it costs and how fast it runs

You pay only for the tokens you use. Speed figures come from our own status checks.

Token pricing

Qwen3.8 27B: token prices in US dollars per million tokens
USD per million tokens
Input$0.40
Cached input$0.10
Output$2.40

No base fee. Caching applies only where listed. Check the model's processing region before sending data.

Speed on Lyceum

Qwen3.8 27B: median speed in our status checks over the last 7 days
Last 7 days
Time to first token274 ms
Output speed45 tokens/s

Median over the last 7 days from our status checks: short requests. Measured 6 October 2026, 23:38 UTC. Reasoning tokens count towards both figures. Live status and availability

The model in detail

Checked against the sources below. Anything they don't state is left out.

Capabilities

Input and output
Text and images in, text out
Context window
256K tokens
Reasoning
On by default
Tool calling
Supported

Architecture and licence

Total parameters
27B
Architecture
Dense
Licence
Apache 2.0
Released
14 August 2026

Availability

API model ID
qwen/qwen3.8-27b
Provider
Qwen
Processing region
EU-hosted
Sources
  • Input and output types: official model card, checked 10 September 2026.
  • Capabilities and request fields: Lyceum API documentation, checked 9 September 2026.
  • Parameters, licence and release date: Qwen’s own sources, including its model card, checked 29 September 2026.

Common questions about Qwen3.8 27B

Short answers about the model ID, price, limits and behaviour.

What is the model ID for Qwen3.8 27B?

Use qwen/qwen3.8-27b as the model value. The OpenAI-compatible base URL is https://api.lyceum.technology/openai/v1. You authenticate with your Lyceum API key as a Bearer token.

How much does Qwen3.8 27B cost?

Qwen3.8 27B costs $0.40 per 1M input tokens, $0.10 per 1M cached input tokens and $2.40 per 1M output tokens. Prices are in US dollars. Billed per token. No base fee.

What is the context window of Qwen3.8 27B?

Qwen3.8 27B has a 256K tokens context window. One request can generate up to 65,536 tokens and run for up to 300 seconds. Set stream: true for long outputs.

Can Qwen3.8 27B read images?

Yes. Send images as OpenAI-style image_url content parts, from a public URL or a base64 data URL.

Can I turn off reasoning for Qwen3.8 27B?

Yes. Reasoning is on by default. Send reasoning_effort: "none" in the request, or call qwen/qwen3.8-27b-instant, the same model at the same price without a reasoning phase. The reasoning trace counts as output tokens when it is on, so set max_tokens high enough for both the reasoning and the answer.

Does Qwen3.8 27B support tool calling?

Yes. Send tools in the standard OpenAI format. Automatic tool calls, parallel calls and tool-result round trips work. It also honours tool_choice: "required" and a named function.

Does Qwen3.8 27B support prompt caching?

Yes, and it is automatic. When requests share an identical start, the cached part is billed at $0.10 per 1M tokens instead of $0.40. Keep the stable part of your prompt at the front. A cache hit is best effort, so read usage.prompt_tokens_details.cached_tokens, which is missing on a miss.

Where does Qwen3.8 27B run, and is my data kept?

Qwen3.8 27B runs on EU-hosted infrastructure. Prompts and outputs are processed, not stored, and never used for training.

How do I call Qwen3.8 27B with the OpenAI SDK?

Point the OpenAI SDK at https://api.lyceum.technology/openai/v1, use your Lyceum API key and set model="qwen/qwen3.8-27b". Chat completions, streaming, tool calling and usage accounting work as they do with OpenAI.

Behaviour notes follow the Lyceum API documentation, checked 6 October 2026.