Qwen
Qwen3 Embedding 8B
A multilingual text embedding model for search and retrieval, including code, and for classification and clustering.
- Context window
- 32K tokens
- Accepts
- Text
- Returns
- Embedding vectors
- Input per 1M tokens
- $0.02
Send your first request
Set LYCEUM_API_KEY to your API key, then run this request. Usage is billed to your account.
- Base URL
https://api.lyceum.technology/openai/v1- Model ID
qwen/qwen3-embedding-8b- Request fields
model, input
Documented request fields for this endpoint. See the documentation for model-specific controls and limits.
app.py · Install openai and set LYCEUM_API_KEY
import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["LYCEUM_API_KEY"],
base_url="https://api.lyceum.technology/openai/v1",
)
response = client.embeddings.create(
model="qwen/qwen3-embedding-8b",
input="A short text to embed.",
)
print(len(response.data[0].embedding))app.ts · Install openai and set LYCEUM_API_KEY in Node.js
import OpenAI from "openai";
const client = new OpenAI({
apiKey: process.env.LYCEUM_API_KEY,
baseURL: "https://api.lyceum.technology/openai/v1",
});
const response = await client.embeddings.create({
model: "qwen/qwen3-embedding-8b",
input: "A short text to embed.",
});
console.log(response.data[0].embedding.length);Terminal · Set LYCEUM_API_KEY in your terminal
curl https://api.lyceum.technology/openai/v1/embeddings \
-H "Authorization: Bearer $LYCEUM_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen/qwen3-embedding-8b",
"input": "A short text to embed."
}'What it costs
You pay only for the tokens you use.
Token pricing
| USD per million tokens | |
|---|---|
| Input | $0.02 |
| Cached input | $0.00 |
| Output | Input only |
No base fee. Caching applies only where listed. Check the model's processing region before sending data.
The model in detail
Checked against the sources below. Anything they don't state is left out.
Capabilities
- Input and output
- Text in, embedding vectors out
- Context window
- 32K tokens
Architecture and licence
- Total parameters
- 8B
- Architecture
- Dense
- Licence
- Apache 2.0
- Released
- 5 June 2025
Availability
- API model ID
qwen/qwen3-embedding-8b- Provider
- Qwen
- Processing region
- EU-hosted
- Input and output types: official model card, checked 10 September 2026.
- Capabilities and request fields: Lyceum API documentation, checked 9 September 2026.
- Parameters, licence and release date: Qwen’s own sources, including its model card and announcement, checked 29 September 2026.
Common questions about Qwen3 Embedding 8B
Short answers about the model ID, price, limits and behaviour.
What is the model ID for Qwen3 Embedding 8B?
Use qwen/qwen3-embedding-8b as the model value. The OpenAI-compatible base URL is https://api.lyceum.technology/openai/v1. You authenticate with your Lyceum API key as a Bearer token.
How much does Qwen3 Embedding 8B cost?
Qwen3 Embedding 8B costs $0.02 per 1M input tokens. Prices are in US dollars. Billed per token. No base fee.
What is the context window of Qwen3 Embedding 8B?
Qwen3 Embedding 8B has a 32K tokens context window.
Where does Qwen3 Embedding 8B run, and is my data kept?
Qwen3 Embedding 8B runs on EU-hosted infrastructure. Prompts and outputs are processed, not stored, and never used for training.
How do I call Qwen3 Embedding 8B with the OpenAI SDK?
Point the OpenAI SDK at https://api.lyceum.technology/openai/v1, use your Lyceum API key and call client.embeddings.create with model="qwen/qwen3-embedding-8b".
Behaviour notes follow the Lyceum API documentation, checked 6 October 2026.