Qwen3 Embedding 8B

A multilingual text embedding model for search and retrieval, including code, and for classification and clustering.

Start buildingOpen dashboard

Context window
32K tokens
Accepts
Text
Returns
Embedding vectors
Input per 1M tokens
$0.02

Send your first request

Set LYCEUM_API_KEY to your API key, then run this request. Usage is billed to your account.

Base URL
https://api.lyceum.technology/openai/v1
Model ID
qwen/qwen3-embedding-8b
Request fields
model, input

Documented request fields for this endpoint. See the documentation for model-specific controls and limits.

app.py · Install openai and set LYCEUM_API_KEY

import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["LYCEUM_API_KEY"],
    base_url="https://api.lyceum.technology/openai/v1",
)

response = client.embeddings.create(
    model="qwen/qwen3-embedding-8b",
    input="A short text to embed.",
)

print(len(response.data[0].embedding))

What it costs

You pay only for the tokens you use.

Token pricing

Qwen3 Embedding 8B: token prices in US dollars per million tokens
USD per million tokens
Input$0.02
Cached input$0.00
OutputInput only

No base fee. Caching applies only where listed. Check the model's processing region before sending data.

The model in detail

Checked against the sources below. Anything they don't state is left out.

Capabilities

Input and output
Text in, embedding vectors out
Context window
32K tokens

Architecture and licence

Total parameters
8B
Architecture
Dense
Licence
Apache 2.0
Released
5 June 2025

Availability

API model ID
qwen/qwen3-embedding-8b
Provider
Qwen
Processing region
EU-hosted
Sources

Common questions about Qwen3 Embedding 8B

Short answers about the model ID, price, limits and behaviour.

What is the model ID for Qwen3 Embedding 8B?

Use qwen/qwen3-embedding-8b as the model value. The OpenAI-compatible base URL is https://api.lyceum.technology/openai/v1. You authenticate with your Lyceum API key as a Bearer token.

How much does Qwen3 Embedding 8B cost?

Qwen3 Embedding 8B costs $0.02 per 1M input tokens. Prices are in US dollars. Billed per token. No base fee.

What is the context window of Qwen3 Embedding 8B?

Qwen3 Embedding 8B has a 32K tokens context window.

Where does Qwen3 Embedding 8B run, and is my data kept?

Qwen3 Embedding 8B runs on EU-hosted infrastructure. Prompts and outputs are processed, not stored, and never used for training.

How do I call Qwen3 Embedding 8B with the OpenAI SDK?

Point the OpenAI SDK at https://api.lyceum.technology/openai/v1, use your Lyceum API key and call client.embeddings.create with model="qwen/qwen3-embedding-8b".

Behaviour notes follow the Lyceum API documentation, checked 6 October 2026.