All modelsQwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Context 1MReasoningTools
Pricing
Cache read
Reusing a prompt already cached upstream
¥2.40¥2.02/ 1M−16.0%
Cache write
Storing a prompt for later reuse
¥15.00¥12.60/ 1M−16.0%
Struck-through figures are the vendor's published list price.
The API is OpenAI-compatible — point the base URL here and pass the model id.
curl https://api.lulutokens.cn/v1/chat/completions \
-H "Authorization: Bearer $LULU_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen3.7-max",
"messages": [{"role": "user", "content": "Hello"}]
}'
from openai import OpenAI
client = OpenAI(base_url="https://api.lulutokens.cn/v1", api_key="LULU_API_KEY")
response = client.chat.completions.create(
model="qwen3.7-max",
messages=[{"role": "user", "content": "Hello"}],
)
Endpoints:
openaiopenai-responseanthropic