CocoboxAI models

The curated set of models Cocobox runs on your behalf — what they're good at and what they cost.

Last updated

CocoboxAI is the managed-models option. We pick a small, opinionated lineup across the major frontier providers, expose them under stable aliases, and charge a flat credit price per request.

The lineup

AliasUnderlyingBest forCredits / request
tamborAmazon Nova LiteEveryday SQL, autocomplete, the cheapest tier0.33
truenoMistral 7BFast classifications, simple Q&A1
prismaAmazon Nova ProDay-to-day coding, SQL, reviews3
sonnetAnthropic Claude Sonnet 4.6Complex SQL, multi-step reasoning3
sonnet-5Anthropic Claude Sonnet 5Newest Sonnet, faster and cheaper2
haikuAnthropic Claude Haiku 4.5Snappy answers with strong reasoning, MCP tool loops1
deepseek-r1DeepSeek R1Math-heavy / analytical2
nova-premierAmazon Nova PremierHardest queries, deep reasoning3
opusAnthropic Claude Opus 4.6The most demanding work7
opus-4-7Anthropic Claude Opus 4.7Newest Opus, frontier reasoning7
opus-4-8Anthropic Claude Opus 4.8Latest Opus, frontier reasoning7
fable-5Anthropic Claude Fable 5Highest-tier model, hardest problems10

Aliases are stable. The underlying model can be upgraded — we’ll announce it in the changelog and quality-test before flipping.

Why aliases

Frontier model names change every few months. By coding to prisma instead of claude-sonnet-4-5-20250929, your snippets, scripts, and integrations don’t break when the provider deprecates a checkpoint.

Picking a model

A simple decision tree:

  1. Tight latency budget or burning through credits? → tambor (0.33 credits — cheapest).
  2. Day-to-day SQL editing or chat? → haiku (1 credit) for cheaper, or prisma/sonnet (3 credits) for stronger reasoning.
  3. Tool-heavy MCP loops? → haiku is best for tool calling (also what Cocobox auto-routes to when tools are attached).
  4. Stuck — needs deep thinking on a hard refactor? → opus, opus-4-7, or opus-4-8 (7 credits).
  5. The hardest problem you’ve got? → fable-5 (10 credits).

Defaults in the product:

  • Editor inline suggestions: tambor.
  • “Fix this error”: haiku.
  • Chats: sonnet or prisma.
  • “Plan a migration”: opus.

You can override defaults in Settings → AI → Defaults.

Limits

  • Context window: each alias inherits the underlying model’s context window.
  • Concurrent requests: 8 per workspace by default; raised on Team and Enterprise.
  • RPM: 120 per workspace. Burstable.

For exact request/response shapes, see API → Chat completions and API → Messages.