deepseek-v4.1-flash-fast API
Call deepseek-v4.1-flash-fast through a single OpenAI-compatible endpoint — one key, no separate vendor account, no multi-provider wiring.
Pricing
Input$0.15 / 1M tokens (¥1.00)
Output$0.62 / 1M tokens (¥4.00)
Output multiplier×4.00
Prices are per 1M tokens in USD · ¥6.5 = $1 · Billed in CNY
When to use it
At $0.15 per 1M input tokens it sits below the catalogue median of $0.62 — a cost-first choice. Best for general-purpose text workloads. Every model here is reachable with the same API key, so trialling it against a stronger sibling costs one line of code.
Code
from openai import OpenAI
client = OpenAI(
base_url="https://starseaapi.com/v1",
api_key="sk-your-key"
)
r = client.chat.completions.create(
model="deepseek-v4.1-flash-fast",
messages=[{"role": "user", "content": "Explain prefix caching in one paragraph"}]
)
print(r.choices[0].message.content)