Skip to content
DecisionNodeDecisionNdedocs
  • Guides
  • API reference
  • Examples
  • Playground

start here

  • QuickstartGet startedGet an API key, send one request with three questions, and branch your code on the typed answers. Plain HTTPS, no SDK to install.
  • POST /v1/decideAPI referenceAnswer typed questions about a state and optional images. One request, one buffered JSON response, one answer per question.
  • QuestionsConceptsQuestions say what to decide. Each one has a type that fixes the shape of its answer: a choice from your options, a score on your scale, a…
  • ConfidenceConceptsProbabilities are calibrated per question type, so a threshold means what it says.
  • ImagesConceptsSend images and text in the same request. The model reads printed and handwritten text, amounts, dates, objects and layout, and answers…
  • Batch jobsPatternsSend up to 10,000 requests in one file and collect the answers later, at half the live price.
  • Pricing and billingYou pay for input tokens only. Output is free because the model generates no text.
↑↓ moveopen7 suggestions
Get API keyGet API key
DecisionNodeDecisionNde

Get started

  • Introduction
  • Quickstart
  • Playground
  • Console and keys
  • With coding agents
  • Examples

Concepts

  • State
  • Questions
  • Choice
  • Score
  • Truth
  • Number
  • Images
  • Confidence
  • Determinism

Models

  • DecisionNode-1.0
  • DecisionNode-1.0 Flash
  • Limits
  • Versions
  • Dedicated capacity

Patterns

  • Confidence-gated routing
  • Fan-out
  • Guardrails
  • Control loops
  • Batch jobs

API reference

  • Overview
  • POST/v1/decide
  • GET/v1/models
  • Errors
  • Safety check
  • Rate limits

Sessions API

  • Sessions overview
  • POSTOpen a session
  • WSStream frames
  • DELEnd a session

Batch API

  • The batch object
  • POSTCreate a batch
  • POSTAdd requests
  • POSTFinalize a batch
  • GETRetrieve a batch
  • GETGet batch results
  • POSTCancel a batch
  • GETList batches

Pricing and billing

  • Pricing and billing

Policies

  • Responsible use
  • Data and privacy

Migrate

  • Coming from a Jev-shaped API
  • Benchmarks
  • Pricing
  • Playground
Get API key
  • Guides
  • API reference
  • Examples
  • Playground

Get started

  • Introduction
  • Quickstart
  • Playground
  • Console and keys
  • With coding agents
  • Examples

Concepts

  • State
  • Questions
  • Choice
  • Score
  • Truth
  • Number
  • Images
  • Confidence
  • Determinism

Models

  • DecisionNode-1.0
  • DecisionNode-1.0 Flash
  • Limits
  • Versions
  • Dedicated capacity

Patterns

  • Confidence-gated routing
  • Fan-out
  • Guardrails
  • Control loops
  • Batch jobs

API reference

  • Overview
  • POST/v1/decide
  • GET/v1/models
  • Errors
  • Safety check
  • Rate limits

Sessions API

  • Sessions overview
  • POSTOpen a session
  • WSStream frames
  • DELEnd a session

Batch API

  • The batch object
  • POSTCreate a batch
  • POSTAdd requests
  • POSTFinalize a batch
  • GETRetrieve a batch
  • GETGet batch results
  • POSTCancel a batch
  • GETList batches

Pricing and billing

  • Pricing and billing

Policies

  • Responsible use
  • Data and privacy

Migrate

  • Coming from a Jev-shaped API
  1. docs
  2. /
  3. Models

Dedicated capacity

Capacity on our own GPUs reserved for your traffic, with limits set for your deployment, or the models on your own premises under the Defence Contract Addendum. The same API, the same answers.

on this page6 sections
  1. When to use it
  2. What stays the same
  3. What changes
  4. On your premises
  5. Health check
  6. How to ask for it

When to use it#

  • Your steady traffic is above what the default rate limits of a few keys cover, and you want headroom that is yours alone.
  • You need capacity that other customers' bursts cannot take, so a 529 never reaches your hot path.
  • You are a defence or security body of an eligible state and need the models inside your own network.

What stays the same#

A dedicated deployment answers the same API: POST /v1/decide, sessions, batch jobs and GET /v1/models, with the same request and response shapes, the same errors and the same models. The same request to the same pinned version returns the same answer as on the shared API, so tests, thresholds and replays carry over unchanged.

What changes#

  • Capacity

    Shared API
    Shared with every customer
    Dedicated
    Reserved for your workspace on our own GPUs
  • Rate limits

    Shared API
    The defaults per key and model
    Dedicated
    Set for your deployment, sized to your traffic
  • Limits

    Shared API
    As on Limits
    Dedicated
    Set per deployment and published at its /v1/models: read them there, not from constants in your code
  • Price

    Shared API
    $0.042 and $0.021 per 1M input tokens
    Dedicated
    Agreed for the capacity you reserve
  • Safety check

    Shared API
    On every request
    Dedicated
    On every request, as on the shared API
Shared API and dedicated deployments compared
Shared APIDedicated
CapacityShared with every customerReserved for your workspace on our own GPUs
Rate limitsThe defaults per key and modelSet for your deployment, sized to your traffic
LimitsAs on LimitsSet per deployment and published at its /v1/models: read them there, not from constants in your code
Price$0.042 and $0.021 per 1M input tokensAgreed for the capacity you reserve
Safety checkOn every requestOn every request, as on the shared API

On your premises#

The models can run inside your own network only under the Defence Contract Addendum, which defence and security bodies of the member states of the North Atlantic Treaty Organization, the member states of the European Union, Australia, Japan, New Zealand and the Republic of Korea may sign. Military use is banned outside it. The addendum's rules are part of the model licence. Under it you:

  • run our safety-check software beside the models, with its self-harm part always on;
  • keep the records the addendum names for 2 years, and allow the audits it sets out;
  • never use the models for what the addendum forbids in every case, whatever the contract says.

Responses carry x-decisionnode-safety wherever the check runs. A self-hosted server that runs without it sends no safety header at all, so code that reads the header should treat it as optional.

Health check#

A dedicated or on-premises deployment also answers GET /healthz, for your load balancer's readiness checks. It needs no key and costs nothing. It returns 200 once the deployment is ready to answer, and 529 with Retry-After while the models are still loading. On the shared API, check a key and your connection with GET /v1/models instead.

curl -i https://<your-deployment>/healthz

How to ask for it#

Write to hello@bynn.com with your model, your peak requests and input tokens a minute, and where it must run. We size the deployment with you.

previousVersions and retirementnextConfidence-gated routing

DecisionNode is built and run by Bynn Intelligence, Inc.

  • Home
  • Playground
  • Examples
  • Console
  • Responsible use
  • Terms
  • Acceptable use
  • Privacy
  • Data processing
  • Defence addendum
  • Cookies

on this page

  1. When to use it
  2. What stays the same
  3. What changes
  4. On your premises
  5. Health check
  6. How to ask for it