How AI work and inference costs are calculated
On this page
Credits pay for supported AI work. The cost varies with the amount of work, the material the AI must read and the length of the result. A short change and a large conversion should not be expected to cost the same.
OwlMeans usually uses defined AI pipelines. Planning, development and review follow a structured workflow, which supports faster, more predictable progress and a consistent application architecture. Reusable application services and access conventions help security stay part of feature development.

Understand the charge
- Check whether the action has an included allowance.
- Read a displayed estimate or credit confirmation before proceeding.
- Check your current weekly credits and purchased balance.
- Start once and follow the existing request.
- Review recorded usage in Billing after the work.
Cloud inference costs include a markup whose interval is 20%–120%, depending on the type of work. This is a broad pricing interval, not a fixed quote for every request. Credits are reference units; the work’s actual charge can vary with its input and output.
Inference is the computation an AI service performs to produce a result. Longer inputs, longer outputs and more work can increase its cost.
With delegated inference, the connected coding agent performs the requested AI work. Your agent provider’s subscription or usage charges may still apply. Platform allowances and other billable services remain relevant. Delegation does not automatically mean free work or an on-device model.
Use the mode guide to choose the setup. See the Terms and Platform License for credit and subscription conditions.