# luv13 > luv13 is an OpenAI-compatible API for open-weight models at base URL https://api.luv13.ai/v1, billed at $0.33 per 1M tokens on every model, input priced the same as output. - Canonical documentation: https://docs.luv13.ai/ (full product docs, tools, models, pricing, A–Z library) - AI index: https://docs.luv13.ai/llms.txt and https://docs.luv13.ai/llms-full.txt - Marketing site: https://luv13.ai/ - OpenAI-compatible: `POST /v1/chat/completions`, `GET /v1/models` - Base URL: https://api.luv13.ai/v1 - Auth: `Authorization: Bearer `; keys start with `sk-luv13-` and are created at https://dash.luv13.ai/ - Pricing: $0.33 per 1M tokens on every model, input priced the same as output; source: https://models.luv13.ai - Model IDs: `luv13/kimi-k3`, `luv13/kimi-k3-fast`, `luv13/glm-5.3`, `luv13/glm-5.3-flash`, `luv13/deepseek-v4.1-flash`, `luv13/deepseek-v4-pro`, `luv13/qwen-3.8-27b` - Support: hi@luv13.ai ## Docs - [Docs](https://docs.luv13.ai/): Base URL, keys, pricing, a first request, and the A–Z library. - [Agents](https://docs.luv13.ai/a/agents): An AI agent is a program that lets a model work toward a goal over several steps, choosing and using tools and checking the results as it goes. - [Aider](https://docs.luv13.ai/a/aider): Aider is an open-source AI pair-programming tool for the terminal that can use luv13 as an OpenAI-compatible endpoint. - [API Key Best Practices](https://docs.luv13.ai/a/api-key-best-practices): API key best practices are the habits that keep your luv13 key, which starts with sk-luv13- and spends your prepaid credit, from leaking or being misused. - [Authentication](https://docs.luv13.ai/a/auth): Authentication on luv13 means sending your sk-luv13- API key as a Bearer token in the Authorization header of each chat completion request. - [Base URL](https://docs.luv13.ai/b/base-url): A base URL is the fixed start of an API's address that every endpoint path is added to. - [Browser Requests](https://docs.luv13.ai/b/browser-requests): Browser requests are calls to luv13 made from JavaScript running in a web page, which luv13 blocks for other sites' origins, so they should go through your own server. - [Chain-of-Thought Prompting](https://docs.luv13.ai/c/chain-of-thought-prompting): Chain-of-thought prompting means asking a model to work through a problem step by step before it gives the final answer. - [Chat Completions](https://docs.luv13.ai/c/chat-completions): Chat completions is the luv13 endpoint that takes a list of messages and returns the model's next reply. - [Chat Templates](https://docs.luv13.ai/c/chat-templates): A chat template is the fixed text format a model uses to turn a list of chat messages into the single token sequence it was trained on. - [Compatible Tools](https://docs.luv13.ai/c/compatible-tools): Compatible tools are apps and coding agents that accept an OpenAI-style base URL and key, and so can use luv13 models once pointed at https://api.luv13.ai/v1. - [Contact and Support](https://docs.luv13.ai/c/contact-and-support): Contact and support covers how to reach the luv13 team, by email at hi@luv13.ai or the form on luv13.ai, and what to include so a problem can be traced. - [Context Window](https://docs.luv13.ai/c/context-window): A context window is the most tokens a model can handle in one request, counting both what you send and what it writes back. - [Conversation History](https://docs.luv13.ai/c/conversation-history): Conversation history is the list of earlier messages you send with each request so a model can follow an ongoing chat. - [curl Examples](https://docs.luv13.ai/c/curl-examples): curl examples are ready-to-run terminal commands for luv13's two endpoints, listing models and sending a chat completion. - [Dashboard](https://docs.luv13.ai/d/dashboard): The luv13 dashboard at luv13.ai/dashboard is where you sign in, create API keys, top up prepaid credit and see your balance and recent usage. - [Data and Privacy](https://docs.luv13.ai/d/data-and-privacy): Data and privacy covers what information luv13 says it handles when you use the API, and where its privacy notice and terms stand. - [DeepSeek V4-Pro](https://docs.luv13.ai/m/deepseek-v4-pro): DeepSeek V4-Pro is DeepSeek's large MIT-licensed Mixture-of-Experts text model, listed on luv13 as luv13/deepseek-v4-pro and temporarily unavailable. - [DeepSeek V4.1 Flash](https://docs.luv13.ai/m/deepseek-v4-1-flash): DeepSeek V4.1 Flash is DeepSeek's MIT-licensed multimodal Mixture-of-Experts model, available on luv13 as luv13/deepseek-v4.1-flash. - [Embeddings](https://docs.luv13.ai/e/embeddings): An embedding is a list of numbers that represents the meaning of a piece of text, so that similar texts get similar numbers. - [Endpoints](https://docs.luv13.ai/e/endpoints): luv13's endpoints are the two URL paths under https://api.luv13.ai/v1 that it serves, one to list models and one to create chat completions. - [Environment Variables](https://docs.luv13.ai/e/environment-variables): An environment variable is a named value set outside your code, such as an API key, that your program reads when it runs. - [Errors and Status Codes](https://docs.luv13.ai/e/errors-and-status-codes): Errors and status codes are the HTTP codes and bodies luv13 returns when a request can't be served, and what each one means you should do. - [Estimating Costs](https://docs.luv13.ai/e/estimating-costs): Estimating costs means turning luv13 token counts into dollars with one multiplication, because every model has the same flat rate and input costs the same as output. - [FAQ](https://docs.luv13.ai/f/faq): The luv13 FAQ answers the questions people ask most about the API, each in a sentence or two, using only facts luv13 has published or that were checked live. - [Few-Shot Prompting](https://docs.luv13.ai/f/few-shot-prompting): Few-shot prompting means showing a model a few worked examples in the prompt so it copies the pattern for a new input. - [Fine-Tuning](https://docs.luv13.ai/f/fine-tuning): Fine-tuning is further training of an existing model on your own examples so it learns a specific task, style or format. - [GLM 5.3](https://docs.luv13.ai/m/glm-5-3): GLM 5.3 is Z.ai's flagship open-weight coding and agent model, available on luv13 as luv13/glm-5.3. - [GLM-5.3 Flash](https://docs.luv13.ai/m/glm-5-3-flash): GLM-5.3 Flash is Z.ai's natively multimodal, MIT-licensed GLM-5 model, available on luv13 as luv13/glm-5.3-flash. - [Glossary](https://docs.luv13.ai/g/glossary): The luv13 glossary defines the terms, ids and fields you meet when using the luv13 API, each in one line. - [Hallucinations](https://docs.luv13.ai/h/hallucinations): A hallucination is when a model states something false or made up as if it were true. - [Health Checks](https://docs.luv13.ai/h/health-checks): A luv13 health check is a quick request, usually GET /v1/models, that tells you whether the API is reachable before you debug your own code. - [HTTP Headers](https://docs.luv13.ai/h/http-headers): HTTP headers are the name-value lines sent with each luv13 request and response; you need two on requests, Authorization and Content-Type. - [Image Input](https://docs.luv13.ai/i/image-input): Image input means sending a picture to a luv13 model that accepts images; five of the seven models list image input, and all return text. - [Input vs. Output Tokens](https://docs.luv13.ai/i/input-vs-output-tokens): Input tokens are the tokens you send to a model, and output tokens are the tokens it writes back. - [JavaScript Example](https://docs.luv13.ai/j/javascript-example): The JavaScript example is a short Node.js script that calls luv13 with the official OpenAI JavaScript SDK. - [JSON Mode](https://docs.luv13.ai/j/json-mode): JSON mode is a request option that tells a model to reply with valid JSON instead of free text. - [Keys and Accounts](https://docs.luv13.ai/k/keys-and-accounts): A luv13 account is where you create API keys, which start with sk-luv13-, and hold the prepaid credit that your requests spend. - [Kimi K3](https://docs.luv13.ai/m/kimi-k3): Kimi K3 is Moonshot AI's open-weight multimodal model for long coding and agent work, available on luv13 as luv13/kimi-k3. - [Kimi K3 Fast](https://docs.luv13.ai/m/kimi-k3-fast): Kimi K3 Fast is a model id luv13 lists as luv13/kimi-k3-fast, named after Moonshot AI's Kimi K3. - [LangChain](https://docs.luv13.ai/l/langchain): LangChain is a framework for building LLM apps whose ChatOpenAI class can call luv13 by setting base_url. - [Latency](https://docs.luv13.ai/l/latency): Latency is how long you wait for a model's response, often measured as the time to the first token and the time to the full reply. - [Listing Models](https://docs.luv13.ai/l/listing-models): Listing models means calling luv13's GET /v1/models endpoint to see every model id you can use right now. - [Migrating from OpenAI](https://docs.luv13.ai/m/migrating-from-openai): Migrating from OpenAI means moving code that calls the OpenAI API over to luv13 by changing the base URL, the API key and the model id. - [Model Fallback](https://docs.luv13.ai/m/model-fallback): Model fallback means trying a second luv13 model id when the first one is unavailable, instead of failing or waiting. - [Model IDs](https://docs.luv13.ai/m/model-ids): A model ID is the exact string, such as luv13/kimi-k3, that you put in the model field of a luv13 request to choose which model answers. - [Model Routing](https://docs.luv13.ai/m/model-routing): Model routing means sending each request to the model best suited for it, based on rules like task type, cost or speed. - [Models](https://docs.luv13.ai/models): Exact model ids, from GET https://api.luv13.ai/v1/models. - [n8n](https://docs.luv13.ai/n/n8n): n8n is a workflow automation tool whose OpenAI credential has a Base URL field, so its OpenAI nodes can call luv13. - [Node.js Fetch](https://docs.luv13.ai/n/nodejs-fetch): Node.js has a built-in fetch function that can call luv13's OpenAI-compatible API with no extra packages. - [Nucleus Sampling](https://docs.luv13.ai/n/nucleus-sampling): Nucleus sampling, set with top_p, makes a model pick each next token only from the smallest group of likely tokens whose probabilities add up to a set share. - [Open-Weight Models](https://docs.luv13.ai/o/open-weight-models): An open-weight model is a language model whose trained weights are published, so anyone allowed by its license can download and run it. - [OpenAI-Compatible APIs](https://docs.luv13.ai/o/openai-compatible-apis): An OpenAI-compatible API accepts the same requests and returns the same response shapes as OpenAI's API, so existing tools and code work with it after changing the base URL and key. - [OpenCode](https://docs.luv13.ai/o/opencode): OpenCode is an open-source AI coding agent for the terminal that can use luv13 as a custom OpenAI-compatible provider. - [Prompt Engineering](https://docs.luv13.ai/p/prompt-engineering): Prompt engineering is the practice of writing and testing the instructions you give a model so it reliably produces the output you want. - [Prompt Injection](https://docs.luv13.ai/p/prompt-injection): Prompt injection is when text from an untrusted source, such as a web page, email or user message, contains instructions that trick a model into ignoring its real ones. - [Python Example](https://docs.luv13.ai/p/python-example): The Python example is a short script that calls luv13 with the official OpenAI Python SDK. - [Python Requests](https://docs.luv13.ai/p/python-requests): The Python requests library can call luv13's OpenAI-compatible API directly with plain HTTP, without an SDK. - [Quantization](https://docs.luv13.ai/q/quantization): Quantization shrinks a model by storing its weights with fewer bits, which saves memory and speeds it up at some cost to accuracy. - [Quickstart](https://docs.luv13.ai/quickstart): Get a luv13 key, set base URL https://api.luv13.ai/v1, and make your first OpenAI-compatible chat completion with one curl. - [Qwen 3.8 27B](https://docs.luv13.ai/m/qwen-3-8-27b): Qwen 3.8 27B is Alibaba's Apache-2.0 dense vision-language model from the Qwen3.8 series, available on luv13 as luv13/qwen-3.8-27b. - [Rate Limiting](https://docs.luv13.ai/r/rate-limiting): Rate limiting is when an API caps how many requests or tokens you can use in a period of time, and rejects extra ones until the window resets. - [Reasoning Models](https://docs.luv13.ai/r/reasoning-models): A reasoning model is a language model trained to work through a problem in intermediate steps before it gives its final answer. - [Retrieval-Augmented Generation](https://docs.luv13.ai/r/retrieval-augmented-generation): Retrieval-augmented generation (RAG) means looking up relevant text first and adding it to the prompt so the model answers from that source. - [Retrying Requests](https://docs.luv13.ai/r/retrying-requests): Retrying requests means sending a failed luv13 call again after a growing wait, and only for errors that can succeed on a second try. - [Roo Code](https://docs.luv13.ai/r/roo-code): Roo Code is an AI coding agent for VS Code that can use luv13 through its OpenAI Compatible provider. - [Server-Sent Events](https://docs.luv13.ai/s/server-sent-events): Server-Sent Events (SSE) is a simple web standard for a server to push a stream of text messages to a client over one open HTTP connection. - [Sources](https://docs.luv13.ai/s/sources): Maker and product URLs cited by public luv13 docs. Gateway catalogs and supplier pages are not listed here. - [Stop Sequences](https://docs.luv13.ai/s/stop-sequences): A stop sequence is a string that tells the model to stop writing as soon as it would produce that text. - [Streaming](https://docs.luv13.ai/s/streaming): Streaming means the API sends a model's reply in small pieces as it is written, instead of all at once at the end. - [System Prompts](https://docs.luv13.ai/s/system-prompts): A system prompt is an instruction at the start of a conversation that sets how the model should behave for every reply that follows. - [Temperature](https://docs.luv13.ai/t/temperature): Temperature is a sampling setting that controls how random a model's word choices are. - [Throughput](https://docs.luv13.ai/t/throughput): Throughput is how much work a model or API gets done over time, usually measured in output tokens per second. - [Timeouts](https://docs.luv13.ai/t/timeouts): A timeout is the longest your code will wait for a request to finish before it gives up. - [Tokenizers](https://docs.luv13.ai/t/tokenizers): A tokenizer is the part of a language model system that splits text into tokens and turns them into numbers the model can read. - [Tool Calling](https://docs.luv13.ai/t/tool-calling): Tool calling lets a model ask your code to run a function you described, then use the result in its reply. - [Tools](https://docs.luv13.ai/tools): Setup in coding tools: where to paste the base URL and key in Cursor, VS Code, Cline, Open WebUI, Codex, Hermes, Kilo Code and the OpenAI SDK. - [Troubleshooting](https://docs.luv13.ai/t/troubleshooting): Troubleshooting is a symptom-to-fix list for the most common problems when calling luv13, based on the responses the API actually returns. - [Usage and Billing](https://docs.luv13.ai/u/usage-and-billing): Usage and billing is how luv13 charges the tokens your requests use against your prepaid credit, at a flat $0.33 per 1M tokens. - [Using Claude Code](https://docs.luv13.ai/u/using-claude-code): Using Claude Code with luv13 isn't possible today, because Claude Code needs an Anthropic-format API and luv13 serves only OpenAI-style chat completions. - [Using Cline](https://docs.luv13.ai/u/using-cline): Cline is an open-source AI coding agent for VS Code and other editors that can use luv13 through its OpenAI Compatible provider. - [Using Codex](https://docs.luv13.ai/u/using-codex): Using Codex with luv13 would mean setting luv13 as a custom model provider in OpenAI's Codex coding agent. - [Using Continue](https://docs.luv13.ai/u/using-continue): Continue is an open-source AI coding assistant for VS Code and JetBrains that can use luv13 through its openai provider with a custom apiBase. - [Using Cursor](https://docs.luv13.ai/u/using-cursor): Using Cursor with luv13 means pointing Cursor's OpenAI API key and base URL override at luv13 so local Chat and Agent run on a luv13 model. - [Using Hermes Agent](https://docs.luv13.ai/u/using-hermes): Using Hermes Agent with luv13 means pointing Nous Research's Hermes Agent at luv13 through its custom OpenAI-compatible provider. - [Using Kilo Code](https://docs.luv13.ai/u/using-kilo-code): Using Kilo Code with luv13 means adding luv13 as an OpenAI Compatible custom provider in the Kilo Code agent. - [Using Open WebUI](https://docs.luv13.ai/u/using-open-webui): Using Open WebUI with luv13 means adding luv13 as an OpenAI API connection so Open WebUI's chat can use luv13 models. - [Using the OpenAI SDKs](https://docs.luv13.ai/u/using-the-openai-sdks): The official OpenAI SDKs for Python and JavaScript can call luv13 by setting their base URL to https://api.luv13.ai/v1 and using a luv13 API key. - [Using VS Code](https://docs.luv13.ai/u/using-vs-code): Using VS Code with luv13 means adding luv13 to VS Code's chat as a Custom Endpoint model that uses the Chat Completions API. - [Vercel AI SDK](https://docs.luv13.ai/v/vercel-ai-sdk): The Vercel AI SDK is a TypeScript library for building AI features that can call luv13 through its OpenAI Compatible provider package. - [Vibe Coding](https://docs.luv13.ai/v/vibe-coding): Vibe coding is building software mostly by describing what you want to an AI coding tool and accepting its changes, with little reading of the code yourself. - [What Is a Token](https://docs.luv13.ai/w/what-is-a-token): A token is a small chunk of text, often a word or part of a word, that a language model reads and writes one at a time. - [What Is luv13](https://docs.luv13.ai/w/what-is-luv13): luv13 is an OpenAI-compatible API that serves seven open-weight models from one base URL and one API key, at one flat price per token. - [XML Prompts](https://docs.luv13.ai/x/xml-prompts): An XML prompt uses simple XML-style tags to separate the parts of a prompt, such as instructions, documents and examples. - [YAML Config](https://docs.luv13.ai/y/yaml-config): A YAML config is a settings file written in YAML, a plain-text format that many AI tools use to store provider, model and key settings. - [Zed Editor](https://docs.luv13.ai/z/zed-editor): Zed is a code editor with built-in AI features that can use luv13 as an OpenAI-compatible provider. - [Zero-Shot Prompting](https://docs.luv13.ai/z/zero-shot-prompting): Zero-shot prompting means asking a model to do a task with instructions only, without giving it any examples. ## Optional - [Full text](https://docs.luv13.ai/llms-full.txt): every page above as plain markdown.