BEATRICE Docs
Sections

Models

GET /v1/models lists the models available to your organization right now.

beatrice-flash

The general-purpose model: polished Italian, coding, reasoning.

FeatureValue
Idbeatrice-flash
Context262,144 tokens
Default output8,192 tokens
Maximum output (max_tokens)32,768 tokens
Reasoningyes, in reasoning_content
Tool callingyes, parallel too
  • Reasoning: the model thinks before answering. Its reasoning comes in the reasoning_content field (when streaming: delta.reasoning_content) and counts as output tokens. With a max_tokens too low the answer may stay empty because reasoning used all the room.
  • Cached context: when you continue a conversation the model has already seen, the repeated part is read from the cache. It is faster and does not count as input towards your limits (see Limits and plans).

The public id stays stable even when the model behind it changes: updates and replacements are announced in advance.

l'amor che move il sole e l'altre stelle