For prototypes & evaluation
Test
125,000,000 free input tokens, renewed monthly.
test_skkeys · every capability- 30 inference requests / minute
- 500,000 input tokens / minute
- No card required. Shared across your organization.
Pricing / USD
$0.04 per million input tokens. Output tokens are free. Start with 125 million free input tokens a month on test keys.
For prototypes & evaluation
125,000,000 free input tokens, renewed monthly.
test_sk keys · every capabilityFor production workloads
prod_sk keys use purchased credits. Free test tokens stay separate.
Images
An image bills as a fixed number of input tokens. The detail you send sets the count. The image bills once per request. Text in the same request bills as usual.
| Capability | low · 512 px | medium · 768 px · default | high · 1024 px |
|---|---|---|---|
| yes-no, classify, classify-tree, rate | 1,000 tokens $0.04 / 1,000 images | 2,000 tokens $0.08 / 1,000 images | 4,000 tokens $0.16 / 1,000 images |
| answer | 1,500 tokens $0.06 / 1,000 images | 3,000 tokens $0.12 / 1,000 images | 6,000 tokens $0.24 / 1,000 images |
| extract, entities, verify | 2,000 tokens $0.08 / 1,000 images | 4,000 tokens $0.16 / 1,000 images | 8,000 tokens $0.32 / 1,000 images |
The input
Send bytes, a data URL, or base64. Remote image URLs aren’t supported.
The timing
answerextractModel time at medium detail.
The extraction
5 fields per extract call
A receipt total. An ID number. A date. Image extract, entities, and verify use a generative pass for focused reading, not full-page transcription.
Capacity that grows with you
Your production tier follows your lifetime paid spend and the time since your first payment. Meet both thresholds to reach the next tier. A paid credit bundle of at least $10 unlocks production access.
| Production tier | Lifetime paid spend | Since first payment | Inference requests / minute | Input tokens / minute |
|---|---|---|---|---|
| Tier 1 | $10 minimum | Immediate | 1,000 | 1 million |
| Tier 2 | $50 minimum | 14 days | 5,000 | 1 million |
| Tier 3 | $250 minimum | 21 days | 5,000 | Unlimited |
| Tier 4 | $500 minimum | 30 days | 10,000 | Unlimited |
Each inference counts toward the rate limit. A batch of 24 texts with 3 statements counts as 72 inference requests.
Need immediate assistance with a higher tier? Reach out to our support team: support@milliseconds.ai.
Estimate production usage
Set your monthly API calls and average billed input tokens per call. Start with an assumption, then use the tokens shown in a real response.
Count every call in a workflow. An invoice processed with three API calls counts as three calls here.
Estimated monthly usage
Estimated production cost
USD / month in purchased token credits
Get startedAn estimate, not an invoice. Production uses purchased credits from the first token; the free monthly test allowance is not deducted. Actual usage depends on the complete request body. See billing details.
The validated request body: your text plus labels, descriptions, hints, questions, types, or schema. Repeated text in separate calls is billed again. Responses include x-input-tokens so you can inspect actual usage.
Test requests pause until the next monthly renewal. Buying credits does not refill the test allowance: use a production key for paid workloads. All test keys in your organization share the same allowance.
No. Production keys use purchased credits only. Test keys use the monthly free allowance only. Both have access to every capability.
No. A paid credit bundle of at least $10 unlocks production access. Auto top-up is optional: enable it if you want to replenish purchased credits automatically.
Yes. All decision capabilities and the OpenAI-compatible endpoint use the same input-token price. Output tokens carry no charge. Batch requests bill the input of the items they contain.
Supplied examples on the demo site are free to try and can use cached responses. Custom inputs require your API key and use your account’s credits. Demos show estimated inference cost to help you plan.
As a fixed number of input tokens, set by the detail you send: low 1,000, medium 2,000, high 4,000. The image bills once per request. Text in the same request bills as usual. Generation capabilities multiply the fixed count.
No. Paperwork parsing, document extraction, and redaction have their own pricing. Your account is shared, but these are separate API products.