Pricing
For full pricing per model see our dedicated pricing page with all models listed. We do not charge any minimum fee per query, nor do we charge any fee on deposits. Pricing examples and model IDs follow the same canonical-ID policy as this documentation: use exact IDs returned byGET /api/v1/models, and do not rely on hidden/internal aliases.
If you are (potentially) going to be a large user of our API reach out to us at support@nano-gpt.com or our Discord for a discount.
Look up a recent request charge
Request Billing returns the original recorded primary charge and token usage using the response’sX-Request-ID and
the same inference API key. It is available for 24 hours from the charge’s
accounting timestamp. The exact cost is a decimal string in USD or XNO; refunds
and separately billed extras are excluded. For aggregate spend and refund
totals, use Usage.