InferLiBuy AI credits

[ INFERENCE, ON YOUR TERMS ]

Your next idea.Already connected.

Buy credits, use the models you need, and see where every dollar goes. One key. No mandatory subscription.

Text & code · Usage-based billing · Solana USDC or card

inferli / balanceUSD

Buy AI credits

Receive the same amount in usage balance. Payment fees are shown before checkout.

01 / CHOOSE A MODEL   →   02 / ADD CREDITS   →   03 / MAKE YOUR FIRST CALLScroll to explore

[ TRY A MODEL ]

One prompt.
A clear price.

Compare the same task, with the same model and token usage.

Powered by InferLi

DeepSeek V4.1 Flash

Input, output and cache. One clear price.
Example prompt

The answerExample response

Illustrative estimates. Final cost depends on verified usage and the rate accepted for your request.

View an itemized receipt +
InferLi

USAGE RECEIPT

Crypto basics
Workload estimate
DeepSeek V4.1 Flash / DeepInfra
PRICE VERSIONv2

Usage breakdown below uses provider net rates.

Provider net
$1.860000
Platform fee
$0.206667
Total charged$2.066667
Reserved $5.000000Released $2.933333
How this is calculated +

Buyer total = $1.86 ÷ 0.90.
Platform fee = buyer total × 10%.
The platform fee is included in the buyer price. Values shown to six decimal places.

INFERENCE, ACCOUNTED FOR.ILLUSTRATIVE ESTIMATE

Models & pricing

8 models in the catalogue · Output price / 1M tokens

Reference → Through InferLi

Illustrative prices include the platform fee. Reference prices are public retail snapshots for the same route, dated 16 Sep 2026. Savings vary by model; input and cache are billed separately.

Open the model market

Compare models, capabilities and pricing in one place.

[ USAGE PATTERNS ]

Explore usage
patterns.

Requests, tokens and model activity from a fixed public reference dataset.

Reference dataset

Requests by day

Model distribution

Share of all models
Data source & definitions +

Source: Engy public activity ↗. Captured 18 Sep 2026. The chart covers 18 Aug–16 Sep 2026, UTC. The current partial day and projected values are excluded.

Complete-day totals: 23,550,647 requests and 328,943,005,132 tokens. The source displayed 24,292,807 requests and 343,383,362,119 tokens under its own 30d window. This page consistently sums complete daily buckets instead.

Success rate and average total duration use the source snapshot’s separate 24h window. Duration is not time to first token. Reference models do not imply availability in the InferLi catalogue. This dataset never changes your account balance.

02 / From choice to your first call

One key.
You’re ready.

Add credits, create your personal key and start using your chosen model. Advanced controls stay available when you need them.

  1. 01

    Sign in and add credits.

    Review your amount and any payment fees before checkout.

  2. 02

    Create your personal key.

    Copy your key into an OpenAI-compatible tool or app.

  3. 03

    Make your first request.

    Find every call and its cost in your account.

Start using
FIRST REQUEST
HTTPOpenAI-compatible
OpenAI-compatible request format
Model selection stays explicit. Provider selection follows your policy.
REQUEST DETAILS
Request
Preview / not yet run
Selected route
DeepInfra
Price version
v2
Follow this example to its receipt

THE $AI ECOSYSTEM

Stake $AI.
Earn inference credits.

$AI coordinates the ecosystem. Staking rewards are delivered as $CREDIT inference credits, which can be activated into AI usage balance.

Stake $AIEarn $CREDITActivate & use
Buy AI credits ↗

FAQ

Do I need a token to use InferLi?+

No. You can add AI usage balance directly. Holding or staking $AI is a separate way to participate in the ecosystem.

How do I choose an InferLi route?+

Start with the exact model and required capabilities, then check provider terms and capacity. Lowest-cost routing compares the estimated whole request. Performance priority needs provider measurements. You can also pin a named provider; switching to a different model is not the default.

What if a request fails?+

A rate limit or connection failure can try another eligible route within your price and data constraints. Once streaming has started, output from another provider is not spliced in. If the upstream may have executed but the response is lost, the request is reconciled instead of blindly retried. The final buyer bill is deduplicated by request ID.

What happens to my data?+

The gateway processes request content and the selected provider receives it. Review each offer’s retention, training-use and region terms before sending sensitive work. Routine records are designed around model, usage, cost, latency and error metadata; troubleshooting content capture requires a separate opt-in.