Overview
Intended useBest suited for speech-to-speech interaction, speech recognition and transcription, and text-to-speech synthesis. |
FamilyCtxdex SpeechSpeech recognition, synthesis and realtime audio family. |
AnnouncedAnnounced |
DeprecatedNot applicableModel is current. |
RetiredNot applicableModel is not retired. |
What it does
Capabilities carry their support status; tasks are labels, because seven descriptions would bury the section.
Capabilities (0)
None recordedNo capability is recorded for this model.Tasks (3)
Inputs & Outputs
Each modality is paired with the representation it is carried in, so a non-token model reads the same way a text one does.
Accepts
- AudioSpeech Audio
- TextChat Messages
Produces
- AudioSpeech Audio
- TextPlain Text
Technical Characteristics
A characteristic with no value states why it has none. An absence here is a recorded fact, not a gap in the page.
Availability
Current offering availableAgent Makers
First Party Managed Api
Serving surfaces
Serving Performance
Measured as ofEvery figure is measured under the conditions beside it. Read without them, a latency number claims more than it can.
| Metric | Typical value | Conditions |
|---|---|---|
| Audio Processing Realtime FactorProcessing time divided by input-audio duration; lower is faster. | 0.245Ratio | Capacity Mode: SHARED_ON_DEMANDDelivery Mode: SYNCHRONOUSProcessing Tier: STANDARDConcurrency Band: C1 |
| First Audio Response MsLatency to first generated audio response. | 190Ms | Capacity Mode: SHARED_ON_DEMANDDelivery Mode: REALTIMEInput Size Band: 5S_AUDIO_TURNProcessing Tier: STANDARDConcurrency Band: C1 |
| First Transcript Response MsLatency to first transcript/recognition result for streaming speech recognition. | 155Ms | Capacity Mode: SHARED_ON_DEMANDDelivery Mode: STREAMINGInput Size Band: 60S_AUDIOProcessing Tier: STANDARDConcurrency Band: C1 |
| Turn Response Latency MsEnd-to-end conversational turn response latency under a defined turn workload. | 400Ms | Capacity Mode: SHARED_ON_DEMANDDelivery Mode: REALTIMEInput Size Band: 5S_AUDIO_TURNProcessing Tier: STANDARDConcurrency Band: C1 |
Pricing Reference
Reference offering
Agent Makers · First Party Managed Api
| Meter | Basis | Rate |
|---|---|---|
| Audio Input | Audited audio-input minute rate | $0.006 / 1Minute |
| Audio Output | Audited audio-output minute rate | $0.045 / 1Minute |
| Input | Audited text-input rate for realtime/speech context | $0.4 / 1000000Token |
Full pricing for a single offering is one step further in, from this model’s pricing page, because it is addressed by offering and this page carries the model. The calculator is disabled because it is not built.
The meters below are on 2 different units and cannot be summed into a single rate. Rates are not comparable across offerings unless meter, unit and conditions match.
Derived chronology
Model Journey
Assembled from this model’s own lifecycle records, its snapshots, its providers’ offering and pricing history, and recorded succession. It is a view over those records, not a second source of truth, and undated lineage is deliberately not on it.
- Effective fromPricingNormalUpcoming
Agent Makers pricing scheduled
Offering context: Agent Makers
Scheduled to take effect on 2026-09-15. Not yet effective.
- Effective fromAvailabilityMajor
Agent Makers · First Party Managed Api: Active
Offering context: Agent Makers
Synthetic commercial activation.
- Effective fromPricingNormal
Agent Makers pricing became effective
Offering context: Agent Makers
Version current, effective from 2026-06-20.
- Effective fromModelMajor
Model released
The first control expands this section in place, without filters or navigation. The second opens the standalone journey, which adds filtering on three axes, year grouping, per-event provenance and a shareable address.
Every date is the source record’s own effective date — a model lifecycle date, a snapshot release, an offering’s effective date, a pricing version’s effective-from. None of them is a load or capture timestamp.
Licence & Use Conditions
LicenceAgent Makers Proprietary Model Terms |
Version1.0 |
SPDX identifierNot applicableSPDX identifies open-source licences; these are proprietary terms. |
Legal textNo public URI recordedThe terms exist and are held internally; no public address has been published.Captured as of |
Use conditions (0)
None recordedNo condition constrains use of this model beyond its licence.Lifecycle & Lineage
Direct lineage
None recordedNo lineage relationship is recorded for this model.Recommended replacement for
None recordedNo replacement recommendation is recorded in either direction.Lifecycle history: Active on 2026-06-16.
Similar Industry Models
Captured as ofA positioning reference, not an equivalence claim. Each row states the axes it was drawn on, and covers only that part of this model's positioning.
GPT-Live-1
OpenAI
Selected as a close current reference for Ctxdex Speech 2 Realtime based on realtime speech, full duplex.
Checked 2026-09-14 · High confidence
Gemini 3.1 Flash Live
Differences recorded: Preview.
Selected as a close current reference for Ctxdex Speech 2 Realtime based on audio to audio, realtime dialogue.
Checked 2026-09-14 · High confidence
Eleven v3 Conversational
ElevenLabs
Selected as a close current reference for Ctxdex Speech 2 Realtime based on low latency speech generation, conversation.
Checked 2026-09-14 · High confidence
Comparison method CTXDEX_COMPARABLES_V1.