Skip to content
DecisionNodeDecisionNdedocs
  • Guides
  • API reference
  • Examples
  • Playground

start here

  • QuickstartGet startedGet an API key, send one request with three questions, and branch your code on the typed answers. Plain HTTPS, no SDK to install.
  • POST /v1/decideAPI referenceAnswer typed questions about a state and optional images. One request, one buffered JSON response, one answer per question.
  • QuestionsConceptsQuestions say what to decide. Each one has a type that fixes the shape of its answer: a choice from your options, a score on your scale, a…
  • ConfidenceConceptsProbabilities are calibrated per question type, so a threshold means what it says.
  • ImagesConceptsSend images and text in the same request. The model reads printed and handwritten text, amounts, dates, objects and layout, and answers…
  • Pricing and billingYou pay for input tokens only. Output is free because the model generates no text.
↑↓ moveopen6 suggestions
Get API keyGet API key
DecisionNodeDecisionNde

Get started

  • Introduction
  • Quickstart
  • With coding agents
  • Examples

Concepts

  • State
  • Questions
  • Choice
  • Score
  • Truth
  • Number
  • Images
  • Confidence
  • Determinism

Models

  • DecisionNode-1.0
  • DecisionNode-1.0 Flash
  • Limits

Patterns

  • Confidence-gated routing
  • Fan-out
  • Guardrails
  • Control loopscomingcoming soon

API reference

  • POST/v1/decide
  • POST/v1/sessionscomingcoming soon
  • GET/v1/models
  • Errors
  • Rate limits

Pricing and billing

  • Pricing and billing

Policies

  • Responsible use

Migrate

  • Coming from a Jev-shaped API
  • Benchmarks
  • Pricing
  • Playground
Get API key
  • Guides
  • API reference
  • Examples
  • Playground

Get started

  • Introduction
  • Quickstart
  • With coding agents
  • Examples

Concepts

  • State
  • Questions
  • Choice
  • Score
  • Truth
  • Number
  • Images
  • Confidence
  • Determinism

Models

  • DecisionNode-1.0
  • DecisionNode-1.0 Flash
  • Limits

Patterns

  • Confidence-gated routing
  • Fan-out
  • Guardrails
  • Control loopscomingcoming soon

API reference

  • POST/v1/decide
  • POST/v1/sessionscomingcoming soon
  • GET/v1/models
  • Errors
  • Rate limits

Pricing and billing

  • Pricing and billing

Policies

  • Responsible use

Migrate

  • Coming from a Jev-shaped API
  1. docs
  2. /
  3. Pricing and billing

Pricing and billing

You pay for input tokens only. Output is free because the model generates no text. DecisionNode-⁠1.0 is $0.042 and DecisionNode-⁠1.0 Flash $0.021 (provisional) per million input tokens, from a prepaid balance. Batch jobs cost half.

on this page7 sections
  1. How a request is billed
  2. Batch jobs
  3. Number questions
  4. Sessionscomingcoming soon
  5. Worked examples
  6. Prepaid balance
  7. Watching spend
  • DecisionNode-⁠1.0

    Model id
    decisionnode-latest
    Input / 1M
    $0.042
    Batch / 1M
    $0.021
    Output
    Free
  • DecisionNode-⁠1.0 Flash

    Model id
    decisionnode-flash-latest
    Input / 1M
    $0.021 (provisional)
    Batch / 1M
    $0.0105 (provisional)
    Output
    Free
  • Dedicated

    Reserved capacity on our own GPUs, custom limits: talk to us

Prices per 1M input tokens
ModelModel idInput / 1MBatch / 1MOutput
DecisionNode-⁠1.0decisionnode-latest$0.042$0.021Free
DecisionNode-⁠1.0 Flashdecisionnode-flash-latest$0.021 (provisional)$0.0105 (provisional)Free
DedicatedReserved capacity on our own GPUs, custom limits: talk to us

How a request is billed#

cost = usage.input_tokens

× price per token

usage.input_tokens
the state, images and every question, as returned in the response
price per token
the model's price per million, divided by 1,000,000

usage.output_tokens is always 0. There are no seats, no minimums and no charge per question: the price depends only on what you send. Requests refused with 400, 401, 402, 422, 429 or 529 are not billed, so you pay only for requests that return answers. A request refused by the safety check (403) is not charged either (pending confirmation).

Batch jobs#

Work that can wait runs as a batch job at half the price per input token, on either model. Output stays free, and a batch answer has the same shape as a real-time one. Batch jobs take every question type, number questions included.

  • Batch results come with a multi-hour delay. There is no set time at which a batch runs or finishes.
  • Poll the batch job for its results. Nothing is pushed to you.
  • Use batches for backfills, nightly scoring and re-checks of past decisions. Anything your software is waiting on belongs in a real-time request.

Number questions#

A number question costs what a choice with the same number of options costs: each value of its grid counts like one option. A count on 0 to 50 is billed like a choice of 51 options, and the default grid of 0 to 255 like one of 256. A tight grid costs less. Images are priced as on every call.

Sessions comingcoming soon#

A session bills its context once, when it opens: the instructions, the state and the questions, at the live price per input token. After that, each frame is billed for its own tokens and the questions it asks, at the same live price. The context is never billed again per frame, and output stays free.

session = opening tokens × price

+ sum(frame tokens × price)

opening tokens
the instructions, the state and the questions, once
frame tokens
each frame's own tokens plus the questions it asks

Worked examples#

Each workload asks three questions per request. Tokens is usage.input_tokens for one request, as the response reports it.

  • Support tickets

    Tokens
    62
    Requests a month
    1,000,000
    DecisionNode-⁠1.0
    $2.60
    DecisionNode-⁠1.0 Flash
    $1.30
  • Comment moderation

    Tokens
    94
    Requests a month
    30,000,000
    DecisionNode-⁠1.0
    $118.44
    DecisionNode-⁠1.0 Flash
    $59.22
  • Receipt photo checks

    Tokens
    1,236
    Requests a month
    200,000
    DecisionNode-⁠1.0
    $10.38
    DecisionNode-⁠1.0 Flash
    $5.19
Monthly cost at these volumes
WorkloadTokensRequests a monthDecisionNode-1.0DecisionNode-1.0 Flash
Support tickets621,000,000$2.60$1.30
Comment moderation9430,000,000$118.44$59.22
Receipt photo checks1,236200,000$10.38$5.19

Prepaid balance#

  • Add credit in the console under Billing. Requests draw it down token by token.
  • Turn on auto reload to top up when the balance drops below an amount you choose.
  • Every top-up gets an invoice in the console.
  • When the balance reaches zero, requests return 402 Out of credit until you add credit, so a runaway job cannot spend money you did not put in. A refused request is not billed.

Watching spend#

The console's Usage page charts spend, tokens and requests by model and by day. In code, read usage.input_tokens from each response; it is exactly what was billed.

Cut cost without cutting quality

Ask every question about an input in one request, trim states to what the decision needs, and route the clear cases to Flash. Each of these lowers tokens, not accuracy.

previousRate limitsnextResponsible use

DecisionNode is built and run by Bynn Intelligence, Inc.

  • Home
  • Playground
  • Examples
  • Console
  • Responsible use
  • Terms
  • Acceptable use
  • Privacy
  • Data processing
  • Defence addendum
  • Cookies

on this page

  1. How a request is billed
  2. Batch jobs
  3. Number questions
  4. Sessionscomingcoming soon
  5. Worked examples
  6. Prepaid balance
  7. Watching spend