modeliddecisionnode-flash-latest(todaydecisionnode-1.0-flash)- Speed
- About 5 ms for a short request, server side (measured October 2026)
- Input
- The same fields as the full model: state, images, questions
- Context
- 64k tokens per request
- Price
- $0.021 per 1M input tokens (provisional), output free
When to use it#
- High volume: moderation of every comment, triage of every event, scoring of every row.
- Inside a latency budget: a check on the hot path of a request, or a gate in front of every agent step.
- Clear decisions: short inputs and well separated options, where the full model's extra capacity is not needed.
Calling it#
curl https://api.decisionnode.com/v1/decide \
-H "Authorization: Bearer $DECISIONNODE_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "decisionnode-flash-latest",
"state": "Customer: I was charged twice and nobody has replied for 3 days.",
"questions": {
"route": {
"type": "choice",
"instructions": "Where should this go?",
"criteria": {
"billing": "money",
"bug": "broken",
"account": "login"
}
},
"urgency": {
"type": "score",
"instructions": "How urgent is this?",
"criteria": ["routine", "today", "urgent", "critical"]
},
"refund": {
"type": "truth",
"instructions": "Refund this automatically?"
}
}
}'Only model changes. At $0.021 per million, the 62-token request above costs $0.0000013, and a million of them about $1.30.