Pay only for what you use
Usage based pricing that scales with your work.
You are billed only for the usage that goes beyond your monthly free allowance. The allowance is generous and renews at the start of every month, so many smaller projects stay free month after month.
What does Namirasoft Inference pricing cover?
Namirasoft Inference is priced by what you use: the requests you send and the conversation memory you keep. Creating and configuring models and routing adds no separate charges.
Requests
A call Namirasoft applications make through Namirasoft Inference. A Load Balancer, Failover, or Race counts as one request, even when it is processed through multiple targets.
Conversation memory
The conversation context stored through Namirasoft Inference so Namirasoft applications maintain continuity across messages.
Pay for what you useFree every month, then a simple flat rate.
Usage |
Monthly free allowance |
Paid rate |
|---|---|---|
Requests |
30,000 requests per month |
1 USD per 10,000 requests |
Conversation memory |
1 GB per month |
1 USD per GB |
AI model usage is billed separately
Namirasoft Inference manages the configuration of your AI models. What you pay for a model depends on the provider and how you access it, through Namirasoft Squirrel or your own provider key.

Namirasoft Squirrel
Namirasoft Squirrel gives Namirasoft applications the fastest, most cost efficient AI model for each request, intelligently matched to the task.



Bring Your Own Key
Use supported third party AI models through your own key.
Connect your own provider API key and use supported AI models such as Claude, ChatGPT, and DeepSeek through Namirasoft Inference.
Ready to get started with Namirasoft Inference?
Pricing FAQs
Find details about how pricing works in Namirasoft Inference.
1. What does Namirasoft Inference pricing cover?
Namirasoft Inference pricing covers your platform usage: the requests you send and the conversation memory you keep. AI models are billed separately, each with its own pricing. You either bring your own API key for a supported model, or use Namirasoft Squirrel.
2. What counts as a request?
A request is one input you send to Namirasoft Inference. A Load Balancer, Failover, or Race counts as a single request, even when it makes several calls.
3. What is conversation memory?
It is the prompts you send and the responses you receive, retained so your applications keep context across messages. Along with requests, it is what Namirasoft Inference pricing is based on.
4. Is AI model usage included in Namirasoft Inference pricing?
No. The AI models are governed and billed by their own provider, not by Namirasoft Inference. You bring your own API key for a supported model, or use Namirasoft Squirrel. Namirasoft Inference does not set or display these prices, it only links you to them. See provider API pricing or Namirasoft Squirrel pricing.
5. How do I configure my AI models?
You configure your AI models in the Namirasoft Inference app. Use Namirasoft Squirrel, or bring your own provider API key for a supported model.
6. Can I control AI costs?
Yes. You can set spending limits per run, per chat, per day, per week, and per month. When a limit is reached, further usage is restricted to your settings.
7. Can I change the AI models my applications use?
Yes. You can configure and change supported AI models at any time to fit each application.
8. Do I only pay for what I use?
Yes. Namirasoft Inference uses usage based pricing, and no credit card is needed to start. You pay only for requests and conversation memory beyond your monthly free allowance. See the pricing table for the rates.