Q

2 pages under Q.

Quantization

Quantization shrinks a model by storing its weights with fewer bits, which saves memory and speeds it up at some cost to accuracy.

Quickstart

Get a luv13 key, set base URL https://api.luv13.ai/v1, and make your first OpenAI-compatible chat completion with one curl.