Full pricing · one offering

Ctxdex Omni 2 on Runvora Inference

Serverless Inference

Pricing effective

Identity

Model
Ctxdex Omni 2
Provider
Runvora Inference
Channel
Serverless Inference
Provider's name for it
Not recorded
Offering
ctxdex:offering:runvora-inference:ctxdex-omni-2
Currency
USD

Rates

Charged by Runvora Inference on this offering. Effective from 15 Sep 2026. Captured as of 14 Sep 2026.

Processing tier

Applies unless another rate set's condition is met.

Standard rates. Applies unless another rate set's condition is met.
MeterRateAs recorded
Audited base input-token rate$0.36per 1M tokensNo source string recorded
Audited base output-token rate$1.45per 1M tokensNo source string recorded
Audited base cached-input-token rate$0.054per 1M tokensNo source string recorded

Where this sits in the industry

The external offerings this one is compared against, captured as of 2026-09-14. Each names the surface it is sold on and how well that surface matches this one. No price is shown here: most of these providers do not publish one per unit, and the link goes to whatever they do publish.

Mistral API — Mistral Small 4

Mistral AI

Offering directModel directSurface close

Sold as Native API, public self service.

Uses Mistral API — Mistral Small 4 as the ranked 1 current commercial/access reference. Surface fit to CTXDEX role SERVERLESS_INFERENCE: CLOSE.

This provider publishes or documents its pricing. A cost comparison is possible where the provider's own rate covers the same workload.

Provider pricing page

Google Vertex AI / Model Garden — Gemma 4 31B

Google

Offering referenceModel closeSurface reference

Sold as Model Garden, managed or model distribution.

Surface differs: CTXDEX commercial role is SERVERLESS_INFERENCE; external reference surface is MODEL_GARDEN / MANAGED_OR_MODEL_DISTRIBUTION.

Uses Google Vertex AI / Model Garden — Gemma 4 31B as the ranked 2 current commercial/access reference. Surface fit to CTXDEX role SERVERLESS_INFERENCE: REFERENCE.

This provider publishes or documents its pricing. This provider publishes no per-unit rate, so there is nothing to compare a cost against.

Provider pricing page

Alibaba Cloud Model Studio — Qwen3.5-35B-A3B

Alibaba Cloud

Offering closeModel closeSurface close

Sold as Native API, public self service.

Uses Alibaba Cloud Model Studio — Qwen3.5-35B-A3B as the ranked 3 current commercial/access reference. Surface fit to CTXDEX role SERVERLESS_INFERENCE: CLOSE.

This provider publishes or documents its pricing. This provider publishes no per-unit rate, so there is nothing to compare a cost against.

Provider pricing page

Conditions & rules

A headline rate is not what you pay. Everything below changes it, and none of it is reflected in the figures above.

Effective from
15 Sep 2026
Effective to
No end date recorded
Captured as of
14 Sep 2026
Eligibility
Published list pricing — available to anyone on this channel.
Offering status
Active
Capacity
Shared On Demand.
Commitment
No commitment option is recorded on this offering.
Tier coverage
2 tiers, each with a recorded cost.
Separately-priced services
Code Execution, Remote Tool Access. These are charged in addition to the rates above.

Pricing rules

A pricing rule is a recorded Commitment — open the definition that changes what a Rate Schedule — open the definition charges — an Allowance — open the definition, a band, a premium or a Spend Commitment — open the definition.

No pricing rule is recorded on this offering’s current version. The rates above are what is charged.

Pricing versions

Every recorded version of this offering's pricing policy, newest first, with its amounts. A scheduled version is a published future price, not a forecast.

  1. Runvora InferenceIn forceEffective from 15 Sep 2026 — no end date recordedCaptured as of 14 Sep 2026
    • Audited base input-token rate $0.36 per 1M tokens
    • Audited base output-token rate $1.45 per 1M tokens
    • Audited base cached-input-token rate $0.054 per 1M tokens

    Input rate down 74.4% on the previous version.

  2. Runvora InferenceSupersededEffective from 1 Apr 2026 — 15 Sep 2026Captured as of 31 Aug 2026
    • Input tokens $1.404 per 1M tokens
    • Output tokens $5.616 per 1M tokens

    Input rate down 15.3% on the previous version.

  3. Runvora InferenceSupersededEffective from 29 Mar 2026 — 1 Apr 2026Captured as of 1 Apr 2026
    • Input tokens $1.65672 per 1M tokens
    • Output tokens $6.62688 per 1M tokens

    The earliest recorded version, so there is nothing before it to compare with.

Other offerings for this model

All pricing for Ctxdex Omni 2