R

5 pages under R.

Rate Limiting

Rate limiting is when an API caps how many requests or tokens you can use in a period of time, and rejects extra ones until the window resets.

Reasoning Models

A reasoning model is a language model trained to work through a problem in intermediate steps before it gives its final answer.

Retrieval-Augmented Generation

Retrieval-augmented generation (RAG) means looking up relevant text first and adding it to the prompt so the model answers from that source.

Retrying Requests

Retrying requests means sending a failed luv13 call again after a growing wait, and only for errors that can succeed on a second try.

Roo Code

Roo Code is an AI coding agent for VS Code that can use luv13 through its OpenAI Compatible provider.