Models
GET /v1/models lists the models available to your organization right now.
beatrice-flash
The general-purpose model: polished Italian, coding, reasoning.
| Feature | Value |
|---|---|
| Id | beatrice-flash |
| Context | 262,144 tokens |
| Default output | 8,192 tokens |
Maximum output (max_tokens) | 32,768 tokens |
| Reasoning | yes, in reasoning_content |
| Tool calling | yes, parallel too |
- Reasoning: the model thinks before answering. Its reasoning comes in the
reasoning_contentfield (when streaming:delta.reasoning_content) and counts as output tokens. With amax_tokenstoo low the answer may stay empty because reasoning used all the room. - Cached context: when you continue a conversation the model has already seen, the repeated part is read from the cache. It is faster and does not count as input towards your limits (see Limits and plans).
The public id stays stable even when the model behind it changes: updates and replacements are announced in advance.