Overview
DeepSeek-V4.1-Flash is a multimodal Mixture-of-Experts model with 552B backbone parameters. DeepSeek describes it as the smallest model in its new architecture family, trained from scratch on a 45T-token multimodal corpus, with image understanding built in from the start of pre-training.
DeepSeek's model card says it uses a Causal Encoder-Decoder (CED) layout: about 8B parameters active on input (prefill) and 16B on output (decode). DeepSeek also describes compressed sparse attention and FP4 KV caching that cuts the global KV cache footprint to roughly a quarter of its earlier V4-Flash generation. Vision embeddings are trained jointly with text from the start of pre-training, not bolted on later.
DeepSeek says it scores ahead of its own V4-Pro on the benchmarks in its release notes, and it has replaced DeepSeek's earlier V4-Flash models on DeepSeek's own API.
The id comes from live GET https://api.luv13.ai/v1/models; every other fact comes from the maker's sources below. The context window is the maker's published figure, not a luv13 limit; luv13 hasn't published its own per-model limits.
Price on luv13: see Pricing.
Examples
Set your key first: export LUV13_API_KEY=sk-luv13-... (see Keys and Accounts).
curl:
curl https://api.luv13.ai/v1/chat/completions \
-H "Authorization: Bearer $LUV13_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model": "luv13/deepseek-v4.1-flash", "messages": [{"role": "user", "content": "Hello"}]}'Python (pip install openai):
import os
from openai import OpenAI
client = OpenAI(base_url="https://api.luv13.ai/v1", api_key=os.environ["LUV13_API_KEY"])
reply = client.chat.completions.create(
model="luv13/deepseek-v4.1-flash",
messages=[{"role": "user", "content": "Hello"}],
)
print(reply.choices[0].message.content)JavaScript (npm install openai, Node.js):
import OpenAI from "openai";
const client = new OpenAI({ baseURL: "https://api.luv13.ai/v1", apiKey: process.env.LUV13_API_KEY });
const reply = await client.chat.completions.create({
model: "luv13/deepseek-v4.1-flash",
messages: [{ role: "user", content: "Hello" }],
});
console.log(reply.choices[0].message.content);Without a valid key, all three return HTTP 401 with "type": "invalid_auth" (checked on 2026-09-30). See Errors and Status Codes.
FAQ
What model id do I use on luv13? luv13/deepseek-v4.1-flash, exactly as GET /v1/models lists it. See Model IDs.
Who makes DeepSeek V4.1 Flash? DeepSeek.
Are the weights open? Yes. DeepSeek released them under the MIT license.
Is the context window a luv13 limit? No. It's the maker's published figure. luv13 hasn't published its own per-model limits; see Limits.
How much does it cost on luv13? See Pricing.
