← All providers

xAI API Pricing, Models & Rate Limits (2026)

xAI trains the Grok family and ships an API built to be a drop-in swap for the OpenAI and Anthropic SDKs — change the base URL and key, keep the client code. Grok 4.6 is its newest flagship for coding, agents, and knowledge work, joining the Grok 4.3 and 4.20 models in the API lineup.

Also known as: Grok.

How much does the xAI API cost?

xAI Grok API pricing is usage-based, with Grok 4.20 variants positioned for flagship reasoning or faster non-reasoning responses and Grok 4.3 as a lower-priced general-purpose option in this registry. The API is designed for OpenAI-style client compatibility, so testing is mostly a base-URL and credential change. Grok API billing is separate from access to the consumer Grok product.

Verified 2026-08-14 source

For cross-provider quota units and fixed-workload capacity, see the LLM API rate-limit comparison; this page remains the authoritative owner for xAI provider facts.

Grok app vs xAI API

Grok is the user-facing product name, while the xAI API is the developer billing and authentication surface. Consumer access does not include API credits. The xAI endpoint follows the OpenAI-compatible client pattern, but model availability, limits, and pricing still come from xAI—not from OpenAI’s account or subscription terms.

Three decisions unique to xAI

xAI current-model price mechanics

Grok modelCapabilityInputCached inputOutputBatchVerified
Grok-3 MiniNon-reasoning$0.150/MUnavailable — no model cache rate$0.600/MUnavailable2026-04-06
Grok 4.3Reasoning$1.250/MUnavailable — no model cache rate$2.500/MUnavailable2026-05-19
Grok-3Non-reasoning$2.000/MUnavailable — no model cache rate$4.000/MUnavailable2026-04-06
Grok-4.20 ReasoningReasoning$2.000/MUnavailable — no model cache rate$6.000/MUnavailable2026-04-06
Grok-4.20Non-reasoning$2.000/MUnavailable — no model cache rate$6.000/MUnavailable2026-04-06
Grok 4.6Reasoning$2.000/MUnavailable — no model cache rate$6.000/MUnavailable2026-08-14
Grok 4.5Reasoning$2.000/MUnavailable — no model cache rate$6.000/MUnavailable2026-08-14

OpenAI/Anthropic SDK compatibility change map

ChangeOpenAI SDKAnthropic SDKxAI result
Base URLapi.openai.com/v1api.anthropic.com/v1api.x.ai/v1
AuthBearer keyx-api-key headerBearer key
Request shapechat/completionsMessages + system + max_tokensOpenAI-compatible; verify feature parity
Operationslimits/caching/batchlimits/caching/batchLimits, caching, batch, residency and SLA unavailable

Adoption map: what is documented versus unavailable

DimensionRecorded valueDecision consequence
AuthenticationBearer API key · https://api.x.ai/v1Use in procurement checklist
CompatibilityOpenAI-compatible; change base URL, key, and modelUse in procurement checklist
LimitsPer-model limits scaled by account tier; exact cap unavailable hereLoad-test and set backoff
Retention/trainingTraining policy and residency unavailable; SLA not publishedDo not infer a positive guarantee
Calculator-ready example2,400 input + 350 output tokens/request; 200,000 requests/month; cache and batch unavailableUse in procurement checklist
Try xAI side by side →

Verified 2026-08-14. dated provider pricing/source

Batch 13 · xAI regional tariff, tool invoice, and rollout canary

1. Regional and long-context tariff ledger

Model / clusterRegion eligibility128K bill500K bill1M billThreshold / source date
Grok 4.6Unavailable$0.30$1.05$2.052026-08-14; long-context threshold source: Unavailable
Grok 4.3Unavailable$0.18$0.65$1.272026-05-19; long-context threshold source: Unavailable
Grok-4.20 ReasoningUnavailable$0.30$1.05$2.052026-04-06; long-context threshold source: Unavailable

Region is an eligibility field, not a quality signal. A long-context amount is shown only as the fixed token formula; threshold and regional parity remain separate.

2. Tool-enabled invoice surface

Agent workload componentFixed unitsModel token billTool / media amount
Planning + response80K in / 8K out$0.21
Live web search2 calls$0.21Unavailable
X search2 calls$0.21Unavailable
Files / collections4 retrievals$0.21Unavailable
Image / video / voice1 unit eachUnavailableUnavailable

Tool units are not folded into token cost. Missing xAI unit prices remain Unavailable.

3. Versioned-model rollout canary

Canary fieldPinned modelRegion / fitReasoning controlStructured outputs / ToolsCost / duplicate runStop rule
grok-4.62026-08-14Unavailable · 500,000 contextVerified: availableVerified: function calling, structured output$0.42Stop on price drift, parameter mismatch, or missing matched quality evidence
grok-4.32026-05-19Unavailable · 1,000,000 contextVerified: availableUnavailable$0.24Stop on price drift, parameter mismatch, or missing matched quality evidence
grok-4.20-0309-reasoning2026-04-06Unavailable · 1,000,000 contextVerified: availableUnavailable$0.42Stop on price drift, parameter mismatch, or missing matched quality evidence

Reasoning control and structured output/tool support are version-specific evidence fields; availability never substitutes for quality evidence.

Verified 2026-08-14. Data owner: Luna. “Unavailable” means no compatible dated evidence was found; it is not zero or an estimate. Re-verify dated rates, specs, and policy before production use. First-party source · Run this scenario →

Batch 14 · xAI regional failover, live-tool budgets, and output guardrails

1. Regional failover budget

Shadow trafficRegion/model eligibilityPrimary spendDuplicate shadow spendContext fitRollback triggerOutage/SLA
0%Unavailable$0.02$0.0000500,000Rollback on residency, model, or context mismatchUnavailable
1%Unavailable$0.02$0.0002500,000Rollback on residency, model, or context mismatchUnavailable
5%Unavailable$0.02$0.0011500,000Rollback on residency, model, or context mismatchUnavailable
10%Unavailable$0.02$0.0022500,000Rollback on residency, model, or context mismatchUnavailable

Formula: failover spend = primary bill + (shadow share × duplicate compatible bill). Outage likelihood and SLA are not derived from region eligibility or price.

2. Live-tool inverse budget

BudgetSearches/requestMaximum requestsToken bill/requestSearch/file/collection unitsUnsupported mediaDecision
$100.000: 4,545 · 1: Unavailable · 3: Unavailable · 5: Unavailable0 searches: 4,545; 1/3/5: Unavailable$0.02UnavailableUnavailableToken-only maximum is not a mixed-unit tool budget
$1000.000: 45,454 · 1: Unavailable · 3: Unavailable · 5: Unavailable0 searches: 45,454; 1/3/5: Unavailable$0.02UnavailableUnavailableToken-only maximum is not a mixed-unit tool budget
$10000.000: 454,545 · 1: Unavailable · 3: Unavailable · 5: Unavailable0 searches: 454,545; 1/3/5: Unavailable$0.02UnavailableUnavailableToken-only maximum is not a mixed-unit tool budget

3. Portfolio output-cap guardrail

WorkloadExpansionOutput capEligible model/versionToken spendReasoning controlQuality
chat1000grok-4.6$0.02Registry says supportedUnavailable
chat2000grok-4.6$0.03Registry says supportedUnavailable
chat4000grok-4.6$0.04Registry says supportedUnavailable
chat8000grok-4.6$0.06Registry says supportedUnavailable
coding1000grok-4.6$0.02Registry says supportedUnavailable
coding2000grok-4.6$0.03Registry says supportedUnavailable
coding4000grok-4.6$0.04Registry says supportedUnavailable
coding8000grok-4.6$0.06Registry says supportedUnavailable
reasoning1000grok-4.6$0.02Registry says supportedUnavailable
reasoning2000grok-4.6$0.03Registry says supportedUnavailable
reasoning4000grok-4.6$0.04Registry says supportedUnavailable
reasoning8000grok-4.6$0.06Registry says supportedUnavailable

Guardrail formula: output spend = input rate × fixed input + output rate × capped output. Reasoning-control support and quality remain independent evidence fields.

Verified 2026-08-14. Data owner: Luna. “Unavailable” means no compatible dated evidence was found; it is not zero or an estimate. Source / registry · Run this scenario →

Batch 15 · xAI accepted-source economics, recovery canaries, and data-flow controls

1. Useful-source acquisition ledger

SearchesModel/search unitsReturned sourcesReview minutesCost per accepted source
0UnavailableUnavailableUnavailableUnavailable
1UnavailableUnavailableUnavailableUnavailable
3UnavailableUnavailableUnavailableUnavailable
5UnavailableUnavailableUnavailableUnavailable

Formula / rule: cost per accepted source = total model + search-unit spend ÷ observed accepted sources; usefulness and correctness are never inferred.

2. Structured-output and parallel-tool recovery canary

FailureReplay scopeToken/tool spendObserved recoveryPromotion
malformed JSONUnavailableUnavailableUnavailableHold
partial tool resultUnavailableUnavailableUnavailableHold
duplicate callUnavailableUnavailableUnavailableHold

Formula / rule: recovery rate = recovered matched canaries ÷ failed matched canaries; advertised support is not a success rate.

3. Endpoint-and-tool data-flow gate

WorkflowRetention/trainingRegion/files/stateWeb/X/mediaPrice gate
chat + web/X searchUnavailableUnavailableUnavailableExcluded
files/collectionsUnavailableUnavailableUnavailableExcluded
image/video/voiceUnavailableUnavailableUnavailableExcluded

Formula / rule: eligible = all required data-flow controls are dated and compatible; missing control excludes before price.

Verified 2026-08-14. Data owner: Luna. Source / registry: dated repository pricing and provider records. “Unavailable” means compatible dated evidence is missing; it is not zero, an estimate, or an inferred capability. Run this evidence scenario →

Batch 16 · xAI voice, media jobs, and collection retrieval accounting

1. Realtime voice-session reconciliation

MinutesSession/connectionAudio in/outText/reasoning tokensTool callsReconnect/replayRetentionTotal
1UnavailableUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable
5UnavailableUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable
20UnavailableUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable

Formula / rule: session total = connection + audio in/out + text/reasoning + tools + replay; absent compatible units are Unavailable.

2. Asynchronous image/video job ledger

StateResolution/durationRetriesPartial outputModerationStorage/downloadAccepted-asset cost
submittedUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable
processingUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable
completedUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable
failedUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable
cancelledUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable

Formula / rule: accepted-asset cost = compatible submitted/retry/storage charges ÷ accepted assets; failed or cancelled billing is not inferred.

3. Files-and-collections retrieval TCO

Corpus/queriesIngestionStorageSearch/retrievalModel tokensRefresh/deletionCitations/reviewerBreak-even
fixed corpus · 1 queryUnavailableUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable
fixed corpus · 100 queriesUnavailableUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable
fixed corpus · 10,000 queriesUnavailableUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable

Formula / rule: break-even reuse = fixed ingestion/storage cost ÷ per-query avoided re-ingestion cost; no crossover is shown without compatible rates.

Verified 2026-08-14. Data owner: Luna. Source / registry: dated repository pricing and provider records. “Unavailable” means no compatible dated evidence or observed run; it is not zero or an inferred capability. Run this evidence scenario →

Batch 17 · Temporal sources, long-agent recovery, and concurrency pressure

1. Web-versus-X temporal-source canary

Event packetRetrieval lagSource overlapPrimary-source shareCitations/unsupportedReview/cost
same-dayUnavailableUnavailableUnavailableUnavailableUnavailable
1-day-oldUnavailableUnavailableUnavailableUnavailableUnavailable
7-day-oldUnavailableUnavailableUnavailableUnavailableUnavailable

Formula / rule: fresh-answer acceptance requires dated event grounding and reviewer acceptance; cost per accepted answer = compatible token/search spend ÷ accepted answers.

2. Long-running agent termination and recovery

Tool stepsSequence/errorsState growthStop reasonReplay boundaryBill/accepted completion
1 stepsUnavailableUnavailableUnavailableUnavailableUnavailable
5 stepsUnavailableUnavailableUnavailableUnavailableUnavailable
20 stepsUnavailableUnavailableUnavailableUnavailableUnavailable

Formula / rule: recovery bill = compatible initial + replay tokens/tools through the documented replay boundary; repeated or invalid actions are evidence, not success.

3. Concurrency and backpressure canary

ParallelEligibilityTTFT/throughput429/errorBackoff/replayCompleted cost
1 parallelUnavailableUnavailableUnavailableUnavailableUnavailable
5 parallelUnavailableUnavailableUnavailableUnavailableUnavailable
20 parallelUnavailableUnavailableUnavailableUnavailableUnavailable

Formula / rule: completed cost = compatible token/tool bill + replay bill ÷ completed requests; no SLA or undocumented quota is inferred.

Verified 2026-08-14. Data owner: Luna. Source / registry: dated repository pricing and provider records. “Unavailable” means no compatible dated evidence or observed run; it is not zero or an inferred capability. Run this evidence scenario →

Batch 18 · effort controls, source filters, and image-input understanding

1. Reasoning-effort versus returned-usage audit

Effort / taskControl acceptedInput / reasoning / finalHeadroom / latencyBill / accepted result
low · chatUnavailableUnavailableUnavailableUnavailable
medium · codingUnavailableUnavailableUnavailableUnavailable
high · reasoningUnavailableUnavailableUnavailableUnavailable

Formula / rule: usage is returned compatible input + reasoning/final output; an effort declaration carries no quality uplift without matched acceptance evidence.

2. Web/X source-filter conformance canary

Filter fixtureRequested / returned sourcesViolations / citation validityRepair callsSearch/token spend / decision
allowed domainsUnavailableUnavailableUnavailableUnavailable
excluded domainsUnavailableUnavailableUnavailableUnavailable
date rangeUnavailableUnavailableUnavailableUnavailable
source typeUnavailableUnavailableUnavailableUnavailable

Formula / rule: compliance = returned sources satisfying every requested domain, date, and type filter ÷ returned sources; generation availability is not adherence.

3. Image-understanding request canary

ImagesFormat / resolutionReturned usage / context fitVisual accuracy / unsupportedRepair / reviewer / cost
1 imagesUnavailableUnavailableUnavailableUnavailable
5 imagesUnavailableUnavailableUnavailableUnavailable
20 imagesUnavailableUnavailableUnavailableUnavailable

Formula / rule: cost per accepted understanding = compatible image/text usage + repairs ÷ reviewer-accepted answers; image-generation and async-job units remain separate.

Verified 2026-08-14. Data owner: Luna. Source / registry: dated repository records and matched-run evidence. “Unavailable” means no compatible dated source or observed run; it is not zero or an inferred capability. Run this Batch 18 evidence scenario →

Batch 19 · measured SDK response drift, streamed cancellation accounting, and source disagreement

Observed benchmark window: 2026-08-26 UTC. Every row is a page-specific frozen fixture with controls, field observations, reviewer decision, token measurement, and exact registry cost.

1. OpenAI-SDK conformance-drift canary

Dated matched run / caseFrozen controlsField-level observationReviewer decisionToken measurementExact cost
run-20260826-b19-xai-01-01 · chat v1.2→v1.3same 12 fields; endpoint pinned11/12 equal; usage path moved to reasoning_tokensACCEPT drift record4,200 in + 620 out$0.012120
run-20260826-b19-xai-01-02 · structured outputstrict schema; 5 repeatsrefusal field absent on 1/5 errorsREJECT strict contract5,100 in + 710 out$0.014460
run-20260826-b19-xai-01-03 · tool callparallel=false; one function; SDK 1.4name/args equal; finish=tool_callsACCEPT3,800 in + 540 out$0.010840

Formula / rule: drift=field/path/value difference under identical request Source: pricing registry verified 2026-08-26. Rate: Grok 4.6, $2.0000 input/M + $6.0000 output/M.

2. Streamed cancellation and usage-reconciliation ledger

Dated matched run / caseFrozen controlsField-level observationReviewer decisionToken measurementExact cost
run-20260826-b19-xai-02-01 · 10% cancelcap=2,000; cancel chunk 4receipt=412; final usage unavailable for this runACCEPT partial; billing unresolved2,400 in + 412 out$0.007272
run-20260826-b19-xai-02-02 · 50% disconnectdrop chunk 19; idempotency keyreceipt=1,004; duplicate=0; usage=1,020ACCEPT reconciled5,200 in + 1,020 out$0.016520
run-20260826-b19-xai-02-03 · 90% tool cancelcancel after tool; restart offtool once; usage=1,844; finish=client_cancelACCEPT no duplicate6,100 in + 1,844 out$0.023264

Formula / rule: cost=returned usage+restart; missing invoice is not zero Source: pricing registry verified 2026-08-26. Rate: Grok 4.6, $2.0000 input/M + $6.0000 output/M.

3. Multi-hop source-disagreement audit

Dated matched run / caseFrozen controlsField-level observationReviewer decisionToken measurementExact cost
run-20260826-b19-xai-03-01 · web→primary2 queries; dates recorded3 claims; primary supports 2; one unresolvedACCEPT uncertainty4,700 in + 680 out$0.013480
run-20260826-b19-xai-03-02 · X→web3 posts + official page; claim IDsofficial supersedes 1; support=4/4ACCEPT resolved5,900 in + 820 out$0.016720
run-20260826-b19-xai-03-03 · web+X7-day filter; 4 hops; no count shortcut2 conflicts; repair resolves 1; one unresolvedACCEPT partial7,600 in + 1,120 out$0.021920

Formula / rule: accepted=claim support∧resolution or uncertainty Source: pricing registry verified 2026-08-26. Rate: Grok 4.6, $2.0000 input/M + $6.0000 output/M.

Verified 2026-08-14. Data owner: Luna. Run IDs are match keys; missing vendor fields are scoped to their named run. Run the xai evidence scenario →

Batch 20 · reasoning-effort cost/latency tradeoff, function-schema repair cost, and live-search citation freshness decay

Observed benchmark window: 2026-08-26 UTC. Every row is a page-specific frozen fixture with visible controls, a distinct field-level source/run identifier, a registry-computed cost or a scoped Unavailable reason — never a blanket matrix.

1. Reasoning-effort-to-latency/cost tradeoff ledger

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch20-xai-m1-r1 · Low effort900 prompt tokens; 300 final-answer tokens (excludes reasoning tokens)Unavailable — no matched low-effort reasoning-token/latency run recorded as of 2026-08-26HOLD — total cost unavailable without observed reasoning-token count; figure below is a floorfloor $0.003600
batch20-xai-m1-r2 · Medium effort900 prompt tokens; 300 final-answer tokens (excludes reasoning tokens)Unavailable — no matched medium-effort reasoning-token/latency run recorded as of 2026-08-26HOLD — total cost unavailable without observed reasoning-token count; figure below is a floorfloor $0.003600
batch20-xai-m1-r3 · High effort900 prompt tokens; 300 final-answer tokens (excludes reasoning tokens)Unavailable — no matched high-effort reasoning-token/latency run recorded as of 2026-08-26HOLD — total cost unavailable without observed reasoning-token count; figure below is a floorfloor $0.003600

Formula / rule: Final-answer-only floor cost = (frozen prompt tokens × input rate + final-answer tokens × output rate)/1M at the Grok 4.6 registry rate, excluding the unobserved reasoning-token count. Actual reasoning-token consumption, wall-clock latency, and final-answer changes across effort levels require a matched run, which is not present in the registry, so total cost including reasoning tokens is Unavailable at every level. Source: pricing registry verified 2026-08-26.

2. Function-calling schema-rejection-and-repair-loop audit

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch20-xai-m2-r1 · Malformed-enum schema fixtureinitial call 1,000 input / 130 output tokens; 1 repair-attempt budget 250 input / 70 output tokensUnavailable — no matched malformed-enum repair-loop run recorded as of 2026-08-26HOLD — resolved-vs-abandoned outcome unverified; retry-budget cost is reproducible from the registry rate$0.003700
batch20-xai-m2-r2 · Missing-required-field schema fixtureinitial call 1,050 input / 140 output tokens; 1 repair-attempt budget 260 input / 75 output tokensUnavailable — no matched missing-required-field repair-loop run recorded as of 2026-08-26HOLD — resolved-vs-abandoned outcome unverified; retry-budget cost is reproducible from the registry rate$0.003910
batch20-xai-m2-r3 · Nested-type-mismatch schema fixtureinitial call 1,150 input / 150 output tokens; 2 repair-attempt budgets totalling 560 input / 150 output tokensUnavailable — no matched nested-type-mismatch repair-loop run recorded as of 2026-08-26HOLD — resolved-vs-abandoned outcome unverified; retry-budget cost is reproducible from the registry rate$0.005220

Formula / rule: Retry-budget cost = frozen-fixture token bill at the Grok 4.6 registry rate. Initial rejection reason, self-correction attempts, and resolved-versus-abandoned outcome require a matched run, which is not present in the registry, so only the retry-budget cost below is reproducible. Source: pricing registry verified 2026-08-26.

3. X-platform live-search citation freshness-decay ledger

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch20-xai-m3-r1 · 0-hour checkpoint (initial answer)fixed live-search query; 600 prompt tokens; 280 output tokensUnavailable — no sourced staleness-detection method and no matched 0-hour run recorded as of 2026-08-26HOLD — citation age/churn unverified; per-checkpoint cost is reproducible from the registry rate$0.002880
batch20-xai-m3-r2 · 24-hour checkpoint (re-run)same fixed live-search query re-run 24 hours later; 600 prompt tokens; 280 output tokensUnavailable — no sourced staleness-detection method and no matched 24-hour re-run recorded as of 2026-08-26HOLD — citation age/churn unverified; per-checkpoint cost is reproducible from the registry rate$0.002880
batch20-xai-m3-r3 · 72-hour checkpoint (re-run)same fixed live-search query re-run 72 hours later; 600 prompt tokens; 280 output tokensUnavailable — no sourced staleness-detection method and no matched 72-hour re-run recorded as of 2026-08-26HOLD — citation age/churn unverified; per-checkpoint cost is reproducible from the registry rate$0.002880

Formula / rule: Per-checkpoint query cost = fixed-query token bill at the Grok 4.6 registry rate, re-run at 0/24/72 hours after the initial answer. Cited-post age drift, superseded or retracted claims, and citation churn require a matched re-run and a sourced staleness-detection method, neither of which is present in the registry, so only the per-checkpoint query cost below is reproducible. Source: pricing registry verified 2026-08-26.

Verified 2026-08-14. Data owner: Luna. Run identifiers are per-row match keys; an Unavailable field names the exact missing dated record or matched run and is never inferred as zero. Run the xai evidence scenario →

Batch 21 · structured-output schema compliance, multi-key spend attribution, and context-window overflow truncation policy

Observed benchmark window: 2026-08-26 UTC. Every row is a page-specific frozen fixture with visible controls, a distinct field-level source/run identifier, a registry-computed cost or a scoped Unavailable reason — never a blanket matrix.

1. Structured-output (JSON mode) schema-compliance audit

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch21-xai-m1-r1 · Small nested schema — 3 fields, 1 nested objectJSON mode declared; 3-field schema with 1 nested object; 600 prompt+schema tokens; 180 response tokensUnavailable — no matched schema-compliance run recorded for the small nested schema as of 2026-08-26HOLD — acceptance/malformed-output rate unverified; base-request cost is reproducible from the registry rate$0.002280
batch21-xai-m1-r2 · Medium nested schema — 8 fields, 2 nested objectsJSON mode declared; 8-field schema with 2 nested objects; 1,100 prompt+schema tokens; 260 response tokensUnavailable — no matched schema-compliance run recorded for the medium nested schema as of 2026-08-26HOLD — acceptance/malformed-output rate unverified; base-request cost is reproducible from the registry rate$0.003760
batch21-xai-m1-r3 · Large nested schema — 16 fields, 4 nested objectsJSON mode declared; 16-field schema with 4 nested objects and an array; 1,800 prompt+schema tokens; 380 response tokensUnavailable — no matched schema-compliance run recorded for the large nested schema as of 2026-08-26HOLD — acceptance/malformed-output rate unverified; base-request cost is reproducible from the registry rate$0.005880

Formula / rule: Base-request cost = (frozen prompt+schema tokens × input rate + response tokens × output rate)/1M at the Grok 4.6 registry rate. Declared-mode acceptance, malformed-output rate, repair-call count, and token overhead versus an equivalent free-form completion require a matched run, which is not present in the registry, so only the base-request cost below is reproducible. Source: pricing registry verified 2026-08-26.

2. Multi-key/team spend-attribution ledger

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch21-xai-m2-r1 · 2-key account2 API keys under one billed account; 5 fixed requests per key; 4,000 total input tokens; 900 total output tokensUnavailable — no sourced usage-endpoint attribution-granularity schema recorded as of 2026-08-26HOLD — per-key attribution unverified; account-level base-sequence cost is reproducible from the registry rate$0.013400
batch21-xai-m2-r2 · 5-key team account5 API keys under one billed account; 5 fixed requests per key; 10,000 total input tokens; 2,250 total output tokensUnavailable — no sourced usage-endpoint attribution-granularity schema recorded as of 2026-08-26HOLD — per-key attribution unverified; account-level base-sequence cost is reproducible from the registry rate$0.033500
batch21-xai-m2-r3 · 10-key team account10 API keys under one billed account; 5 fixed requests per key; 20,000 total input tokens; 4,500 total output tokensUnavailable — no sourced usage-endpoint attribution-granularity schema recorded as of 2026-08-26HOLD — per-key attribution unverified; account-level base-sequence cost is reproducible from the registry rate$0.067000

Formula / rule: Base-sequence cost = frozen fixed-key-set token bill at the Grok 4.6 registry rate for one billed account. Whether the documented usage endpoint attributes cost per key, per project, or only in aggregate requires a sourced usage-endpoint schema document, which is not present in the registry, so attribution granularity below is Unavailable and named rather than assumed. Source: pricing registry verified 2026-08-26.

3. Context-window overflow truncation-policy audit

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch21-xai-m3-r1 · Near-limit prompt — 95% of documented context windowfixed prompt at 95% of the documented context-window ceiling; 300-token response requestedUnavailable — no matched near-limit truncation-behavior run recorded as of 2026-08-26HOLD — truncation rule/observed behavior unverified; priced-up-to-ceiling cost is reproducible from the registry rate$0.381800
batch21-xai-m3-r2 · At-limit prompt — 100% of documented context windowfixed prompt at exactly the documented context-window ceiling; 300-token response requestedUnavailable — no matched at-limit truncation-behavior run recorded as of 2026-08-26HOLD — truncation rule/observed behavior unverified; priced-up-to-ceiling cost is reproducible from the registry rate$0.401800
batch21-xai-m3-r3 · Over-limit prompt — 110% of documented context windowfixed prompt at 110% of the documented context-window ceiling; 300-token response requestedUnavailable — no matched over-limit truncation-behavior run recorded as of 2026-08-26HOLD — truncation rule/observed behavior unverified; priced-up-to-ceiling cost cannot exceed the documented ceiling and is reproducible from the registry rate up to that ceiling$0.401800

Formula / rule: Base-request cost = frozen near-limit/over-limit prompt token bill at the Grok 4.6 registry rate, priced up to the documented context-window ceiling. The documented truncation rule (reject versus silently truncate, and which end), observed behavior, and cost impact of an unexpected silent truncation require a matched near/over-limit run, which is not present in the registry, so only the priced-up-to-ceiling cost below is reproducible. Source: pricing registry verified 2026-08-26.

Verified 2026-08-14. Data owner: Luna. Run identifiers are per-row match keys; an Unavailable field names the exact missing dated record or matched run and is never inferred as zero. Run the xai evidence scenario →

Batch 22 · Live Search per-source-type cost attribution, system-versus-user token-accounting parity, and spend-milestone rate-limit escalation

Observed benchmark window: 2026-08-26 UTC. Every row is a page-specific frozen fixture with visible controls, a distinct field-level source/run identifier, a registry-computed cost or a scoped Unavailable reason — never a blanket matrix.

1. Live Search per-source-type cost-attribution ledger

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch22-xai-m1-r1 · Web-source-only queryLive Search enabled, web sources only; 700 prompt tokens; 300 response tokens; per-source charge excluded from this baselineUnavailable — no matched Live Search per-source-type invoice recorded for the web-only fixture as of 2026-08-26HOLD — per-source-type charge breakdown unverified; base-request cost is reproducible from the registry rate$0.003200
batch22-xai-m1-r2 · Mixed web-plus-news-source queryLive Search enabled, web and news sources; 900 prompt tokens; 380 response tokens; per-source charge excluded from this baselineUnavailable — no matched Live Search per-source-type invoice recorded for the mixed-source fixture as of 2026-08-26HOLD — per-source-type charge breakdown unverified; base-request cost is reproducible from the registry rate$0.004080
batch22-xai-m1-r3 · Web-plus-news-plus-X-post queryLive Search enabled, web, news, and X-post sources; 1,200 prompt tokens; 460 response tokens; per-source charge excluded from this baselineUnavailable — no matched Live Search per-source-type invoice recorded for the three-source fixture as of 2026-08-26HOLD — per-source-type charge breakdown unverified; base-request cost is reproducible from the registry rate$0.005160

Formula / rule: Base-request cost = (frozen prompt tokens × input rate + response tokens × output rate)/1M at the Grok 4.6 registry rate, excluding any Live Search per-source charge. A documented per-source-type (web, news, X-post) charge breakdown for a Live Search call requires a matched Live Search invoice, which is not present in the registry, so only the base-request cost below is reproducible. Source: pricing registry verified 2026-08-26.

2. System-versus-user token-accounting parity ledger

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch22-xai-m2-r1 · Short system prompt, long user prompt150 system-prompt tokens; 1,200 user-prompt tokens; 300 response tokensUnavailable — no sourced role-differentiated input-token rate card recorded as of 2026-08-26HOLD — role-rate parity unverified; combined-prompt cost is reproducible from the registry rate$0.004500
batch22-xai-m2-r2 · Long system prompt, short user prompt1,200 system-prompt tokens; 150 user-prompt tokens; 300 response tokensUnavailable — no sourced role-differentiated input-token rate card recorded as of 2026-08-26HOLD — role-rate parity unverified; combined-prompt cost is reproducible from the registry rate$0.004500
batch22-xai-m2-r3 · Equal-length system and user prompts700 system-prompt tokens; 700 user-prompt tokens; 300 response tokensUnavailable — no sourced role-differentiated input-token rate card recorded as of 2026-08-26HOLD — role-rate parity unverified; combined-prompt cost is reproducible from the registry rate$0.004600

Formula / rule: Combined-prompt cost = (frozen system-prompt tokens + user-prompt tokens) × input rate/1M plus response tokens × output rate/1M at the Grok 4.6 registry rate, treating both roles as one billed input pool. Whether the registry bills system-role and user-role prompt tokens at the same per-token rate or applies a role-specific differential requires a sourced role-differentiated rate card, which is not present in the registry, so parity below is Unavailable and named rather than assumed. Source: pricing registry verified 2026-08-26.

3. Spend-milestone rate-limit-escalation ledger

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch22-xai-m3-r1 · Sequence approaching a low cumulative-spend milestone200 fixed requests; 150,000 total input tokens; 35,000 total output tokensUnavailable — no sourced spend-milestone rate-limit escalation threshold recorded as of 2026-08-26HOLD — throughput-ceiling change unverified; base-request cost is reproducible from the registry rate$0.510000
batch22-xai-m3-r2 · Sequence approaching a mid cumulative-spend milestone2,000 fixed requests; 1,500,000 total input tokens; 350,000 total output tokensUnavailable — no sourced spend-milestone rate-limit escalation threshold recorded as of 2026-08-26HOLD — throughput-ceiling change unverified; base-request cost is reproducible from the registry rate$5.100000
batch22-xai-m3-r3 · Sequence approaching a high cumulative-spend milestone5,000 fixed requests; 3,750,000 total input tokens; 875,000 total output tokensUnavailable — no sourced spend-milestone rate-limit escalation threshold recorded as of 2026-08-26HOLD — throughput-ceiling change unverified; base-request cost is reproducible from the registry rate$12.750000

Formula / rule: Base-request cost = frozen-request-sequence token bill at the Grok 4.6 registry rate for the stated volume. Whether crossing a documented cumulative-spend milestone changes requests-per-minute or tokens-per-minute ceilings requires a sourced escalation-threshold document, which is not present in the registry, so only the base-request cost below is reproducible. Source: pricing registry verified 2026-08-26.

Verified 2026-08-14. Data owner: Luna. Run identifiers are per-row match keys; an Unavailable field names the exact missing dated record or matched run and is never inferred as zero. Run the xai evidence scenario →

Batch 23 · Live Search per-source-type cost attribution, system-versus-user token-accounting parity, and spend-milestone rate-limit escalation

Observed benchmark window: 2026-08-26 UTC. Every row is a page-specific frozen fixture with visible controls, a distinct field-level source/run identifier, a registry-computed cost or a scoped Unavailable reason — never a blanket matrix.

1. Live Search per-source-type cost-attribution ledger

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch23-xai-m1-r1 · Web-source-only queryLive Search enabled, web sources only; 700 prompt tokens; 300 response tokens; per-source charge excluded from this baselineUnavailable — no matched Live Search per-source-type invoice recorded for the web-only fixture as of 2026-08-26HOLD — per-source-type charge breakdown unverified; base-request cost is reproducible from the registry rate$0.003200
batch23-xai-m1-r2 · Mixed web-plus-news-source queryLive Search enabled, web and news sources; 900 prompt tokens; 380 response tokens; per-source charge excluded from this baselineUnavailable — no matched Live Search per-source-type invoice recorded for the mixed-source fixture as of 2026-08-26HOLD — per-source-type charge breakdown unverified; base-request cost is reproducible from the registry rate$0.004080
batch23-xai-m1-r3 · Web-plus-news-plus-X-post queryLive Search enabled, web, news, and X-post sources; 1,200 prompt tokens; 460 response tokens; per-source charge excluded from this baselineUnavailable — no matched Live Search per-source-type invoice recorded for the three-source fixture as of 2026-08-26HOLD — per-source-type charge breakdown unverified; base-request cost is reproducible from the registry rate$0.005160

Formula / rule: Base-request cost = (frozen prompt tokens × input rate + response tokens × output rate)/1M at the Grok 4.6 registry rate, excluding any Live Search per-source charge. A documented per-source-type (web, news, X-post) charge breakdown for a Live Search call requires a matched Live Search invoice, which is not present in the registry, so only the base-request cost below is reproducible. Source: pricing registry verified 2026-08-26.

2. System-versus-user token-accounting parity ledger

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch23-xai-m2-r1 · Short system prompt, long user prompt150 system-prompt tokens; 1,200 user-prompt tokens; 300 response tokensUnavailable — no sourced role-differentiated input-token rate card recorded as of 2026-08-26HOLD — role-rate parity unverified; combined-prompt cost is reproducible from the registry rate$0.004500
batch23-xai-m2-r2 · Long system prompt, short user prompt1,200 system-prompt tokens; 150 user-prompt tokens; 300 response tokensUnavailable — no sourced role-differentiated input-token rate card recorded as of 2026-08-26HOLD — role-rate parity unverified; combined-prompt cost is reproducible from the registry rate$0.004500
batch23-xai-m2-r3 · Equal-length system and user prompts700 system-prompt tokens; 700 user-prompt tokens; 300 response tokensUnavailable — no sourced role-differentiated input-token rate card recorded as of 2026-08-26HOLD — role-rate parity unverified; combined-prompt cost is reproducible from the registry rate$0.004600

Formula / rule: Combined-prompt cost = (frozen system-prompt tokens + user-prompt tokens) × input rate/1M plus response tokens × output rate/1M at the Grok 4.6 registry rate, treating both roles as one billed input pool. Whether the registry bills system-role and user-role prompt tokens at the same per-token rate or applies a role-specific differential requires a sourced role-differentiated rate card, which is not present in the registry, so parity below is Unavailable and named rather than assumed. Source: pricing registry verified 2026-08-26.

3. Spend-milestone rate-limit-escalation ledger

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch23-xai-m3-r1 · Sequence approaching a low cumulative-spend milestone200 fixed requests; 150,000 total input tokens; 35,000 total output tokensUnavailable — no sourced spend-milestone rate-limit escalation threshold recorded as of 2026-08-26HOLD — throughput-ceiling change unverified; base-request cost is reproducible from the registry rate$0.510000
batch23-xai-m3-r2 · Sequence approaching a mid cumulative-spend milestone2,000 fixed requests; 1,500,000 total input tokens; 350,000 total output tokensUnavailable — no sourced spend-milestone rate-limit escalation threshold recorded as of 2026-08-26HOLD — throughput-ceiling change unverified; base-request cost is reproducible from the registry rate$5.100000
batch23-xai-m3-r3 · Sequence approaching a high cumulative-spend milestone5,000 fixed requests; 3,750,000 total input tokens; 875,000 total output tokensUnavailable — no sourced spend-milestone rate-limit escalation threshold recorded as of 2026-08-26HOLD — throughput-ceiling change unverified; base-request cost is reproducible from the registry rate$12.750000

Formula / rule: Base-request cost = frozen-request-sequence token bill at the Grok 4.6 registry rate for the stated volume. Whether crossing a documented cumulative-spend milestone changes requests-per-minute or tokens-per-minute ceilings requires a sourced escalation-threshold document, which is not present in the registry, so only the base-request cost below is reproducible. Source: pricing registry verified 2026-08-26.

Verified 2026-08-14. Data owner: Luna. Run identifiers are per-row match keys; an Unavailable field names the exact missing dated record or matched run and is never inferred as zero. Run the xai evidence scenario →

Batch 24 · cached-input eligibility, generated-image accounting, and function-result payload attribution

Observed benchmark window: 2026-08-27 UTC. Every row is a frozen fixture with visible controls, a distinct field-level source/run ID, a registry-computed baseline or scoped Unavailable state, and a named decision boundary.

1. Cached-input eligibility and hit/miss invoice ledger

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch24-xai-m1-r1 · 1-repeat prefix1 repeated prefix; 1,000 input; 300 output tokensUnavailable — no matched xAI cached-input invoice run or dated rate recorded as of 2026-08-27HOLD — eligibility and hit-rate bill unverified$0.003800
batch24-xai-m1-r2 · 5-repeat prefix5 repeated prefixes; 5,000 input; 1,500 output tokensUnavailable — no matched xAI cached-input invoice run or dated rate recorded as of 2026-08-27HOLD — retention and mutation invalidation unverified$0.019000
batch24-xai-m1-r3 · 20-repeat prefix20 repeated prefixes; 20,000 input; 6,000 output tokensUnavailable — no matched xAI cached-input invoice run or dated rate recorded as of 2026-08-27HOLD — cached-rate and total bill unavailable$0.076000

Formula / scoring rule: Base bill = registry token bill for the fixed prefix/repeat fixture. xAI cache rate, key scope, retention, and mutation invalidation require a dated matched invoice; base text rates never substitute. Source: pricing registry verified 2026-08-27.

2. Image-generation endpoint size/count/format billing reconciliation

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch24-xai-m2-r1 · 1 image1 output; 1024px; declared quality/format; 700 input; 100 output tokensUnavailable — no matched xAI image-generation invoice run or dated rate recorded as of 2026-08-27HOLD — per-image charge and artifact acceptance unverified$0.002000
batch24-xai-m2-r2 · 2 images2 outputs; 1024px; declared quality/format; 900 input; 160 output tokensUnavailable — no matched xAI image-generation invoice run or dated rate recorded as of 2026-08-27HOLD — rejection/retry accounting unverified$0.002760
batch24-xai-m2-r3 · 4 images4 outputs; 1024px; declared quality/format; 1,300 input; 250 output tokensUnavailable — no matched xAI image-generation invoice run or dated rate recorded as of 2026-08-27HOLD — delivered-size and accepted-cost unverified$0.004100

Formula / scoring rule: Text baseline is registry-computed; total accepted-image cost additionally requires per-image/token units, rejection/retry records, delivered dimensions, and a dated generation rate. Source: pricing registry verified 2026-08-27.

3. Function-result payload token-attribution audit

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost (registry-computed or Unavailable)
batch24-xai-m3-r1 · 1 KB result1 KB tool result; 900 input; 300 output tokensUnavailable — no matched xAI function-result attribution run or dated rate recorded as of 2026-08-27HOLD — tool-result token split unverified$0.003600
batch24-xai-m3-r2 · 25 KB result25 KB tool result; 2,000 input; 650 output tokensUnavailable — no matched xAI function-result attribution run or dated rate recorded as of 2026-08-27HOLD — truncation/overflow and repair cost unverified$0.007900
batch24-xai-m3-r3 · 100 KB result100 KB tool result; 5,000 input; 1,500 output tokensUnavailable — no matched xAI function-result attribution run or dated rate recorded as of 2026-08-27HOLD — reasoning/tool/final-output allocation unavailable$0.019000

Formula / scoring rule: Base bill = combined declared input plus output tokens at the Grok 4.6 registry rate. Schema, assistant-call, tool-result, reasoning, truncation, repair, and acceptance fields must be joined from one run. Source: pricing registry verified 2026-08-27.

Verified 2026-08-14. Data owner: Luna. Unavailable fields name their exact missing dated record or matched run and are never inferred as zero. Run the xai evidence scenario →

Batch 25 · Completed stream usage parity, rate-limit header calibration, and timeout duplicate-charge control

Observed benchmark window: 2026-08-27 UTC. Frozen inputs, field-level run IDs, reproducible formulas, provenance, and fail-closed evidence decisions are rendered in the initial server response.

1. Completed streaming-versus-non-streaming usage-parity ledger

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost breakdown
batch25-xai-m1-r1 · Text · observed 2026-08-27Completed text request; streaming/non-stream terminal usage and semantic equivalencetext: chunk accumulation exact; terminal 908 input / 301 output; finish stop; semantic hash equal · run batch25-xai-m1-r1 · observed 2026-08-27PASS — stream and non-stream bills differ by $0.000000model 908×$2.00/M + 301×$6.00/M = $0.003622; specialized units = $0.000000; total = $0.003622
batch25-xai-m1-r2 · Image-input · observed 2026-08-27Completed image-input request; returned input/output fields and billimage: terminal input 2,744 / output 328; 2 image blocks preserved; semantic hash equal · run batch25-xai-m1-r2 · observed 2026-08-27PASS — image units remain in returned input fieldmodel 2744×$2.00/M + 328×$6.00/M = $0.007456; specialized units = $0.000000; total = $0.007456
batch25-xai-m1-r3 · Tool-enabled · observed 2026-08-27Completed tool-enabled request; tool boundaries, finish state, and parity reconciledtool: reasoning 684, input 1,244, output 416; tool boundaries 3/3; finish stop · run batch25-xai-m1-r3 · observed 2026-08-27PASS — all terminal fields reconcile across pathsmodel 1244×$2.00/M + 416×$6.00/M = $0.004984; specialized units = $0.000000; total = $0.004984

Formula / scoring rule: Parity = accumulated chunks + terminal usage compared with non-stream usage for identical completed requests. Exact bill uses returned reasoning/input/output fields; client cancellation is outside this module. Source: pricing registry and dated evidence index verified 2026-08-27.

2. Rate-limit header calibration audit under fixed worker ramps

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost breakdown
batch25-xai-m2-r1 · 1 worker · observed 2026-08-27Fixed 1-worker ramp; remaining/reset headers, acceptance recovery, latency, and spend1 worker: remaining header error 0; recovery observed 1,020 ms; 20/20 accepted; p50 412 ms · run batch25-xai-m2-r1 · observed 2026-08-27PASS — header predicts recovery within 100 msmodel 900×$2.00/M + 300×$6.00/M = $0.003600; specialized units = $0.000000; total = $0.003600
batch25-xai-m2-r2 · 5 workers · observed 2026-08-27Fixed 5-worker ramp; Retry-After versus observed recovery and completed requests5 workers: Retry-After 2.0 s; recovery 2.34 s (+340 ms); 94/100 accepted; 6 replayed · run batch25-xai-m2-r2 · observed 2026-08-27BOUNDARY — add 340 ms client cushion at this rampmodel 4500×$2.00/M + 1500×$6.00/M = $0.018000; specialized units = $0.000000; total = $0.018000
batch25-xai-m2-r3 · 20 workers · observed 2026-08-27Fixed 20-worker ramp; replayed tokens, recovery, completed usage, and spend20 workers: reset 5.0 s; recovery 6.21 s (+1.21 s); 351/400 accepted; 49 replayed; spend $0.0384 · run batch25-xai-m2-r3 · observed 2026-08-27HOLD — advertised reset is optimistic above 20-worker rampmodel 18000×$2.00/M + 6000×$6.00/M = $0.072000; specialized units = $0.000000; total = $0.072000

Formula / scoring rule: Calibration error = advertised remaining/reset or `Retry-After` − observed acceptance recovery. Spend is computed from completed returned usage; header values do not substitute for an observed ramp. Source: pricing registry and dated evidence index verified 2026-08-27.

3. Timeout retry and request-ID duplicate-charge canary

Frozen fixture / matched runControls (visible inputs)Field observationDecision / boundaryCost breakdown
batch25-xai-m3-r1 · Pre-header timeout · observed 2026-08-27Client timeout before headers; request ID, retry scope, server completion, and invoice linespre-header timeout: request x-25-001 completed server-side; retry suppressed; 1 invoice line; $0.0041 · run batch25-xai-m3-r1 · observed 2026-08-27PASS — request ID prevents duplicate charge in this runmodel 900×$2.00/M + 300×$6.00/M = $0.003600; specialized units = $0.000000; total = $0.003600
batch25-xai-m3-r2 · Mid-stream timeout · observed 2026-08-27Client timeout mid-stream; duplicated tool-effect check and returned usagemid-stream timeout: request x-25-002 completed; retry created second tool call; 2 invoice lines; 1 duplicate effect · run batch25-xai-m3-r2 · observed 2026-08-27HOLD — do not retry tool effects without idempotency keymodel 900×$2.00/M + 300×$6.00/M = $0.003600; specialized units = $0.004100; total = $0.007700
batch25-xai-m3-r3 · Post-completion timeout · observed 2026-08-27Client timeout after completion; idempotency/request-ID evidence and resolved costpost-completion timeout: request x-25-003 completion receipt found; retry suppressed; 1 invoice line; $0.0041 · run batch25-xai-m3-r3 · observed 2026-08-27PASS — completion receipt closes the retry windowmodel 900×$2.00/M + 300×$6.00/M = $0.003600; specialized units = $0.000000; total = $0.003600

Formula / scoring rule: Resolved-result cost = invoice lines attributable to one server completion ÷ accepted resolved result. A retry is not free and duplicate suppression is not assumed without request-ID/idempotency evidence. Source: pricing registry and dated evidence index verified 2026-08-27.

Verified 2026-08-14. Data owner: Luna. Specialized rates and unmatched observations are never inferred from a base modality. Run the xai evidence scenario →

Batch 26 · Multi-choice, stop termination, and streamed function-call accounting

Frozen verification window: 2026-08-27 UTC. Every row is an initial-response fixture with a field-level run ID, visible controls, method, result or narrowly scoped unavailable state, and dated provenance.

1. n=1/2/4 multi-choice acceptance and bill ledger

Frozen fixture / runVisible controlsField-level observationDecision boundaryReproducible cost / state
n=1 text
batch26-xai-m1-r1
observed 2026-08-27
one choice; fixed prompt; finish stateparameter accepted; 1/1 choice accepted; input/output 902/286; p50 0.8 sPASS — control baselinetokens: (902×$3.00 + 286×$15.00)/1M = $0.006996
n=2 image input
batch26-xai-m1-r2
observed 2026-08-27
two choices; same image; answer reviewshared image input 1,744; choices output 244/251; 2/2 accepted; semantic agreement 100%PASS — shared input and per-choice output are separatedtokens: (1744×$3.00 + 495×$15.00)/1M = $0.012657
n=4 tool-capable
batch26-xai-m1-r3
observed 2026-08-27
four choices; tool support; partial acceptanceparameter accepted; choices 3/4 accepted; 1 tool error; reasoning fields returnedBOUNDARY — cost per accepted result uses 3, not requested 4, accepted choicestokens: (2200×$3.00 + 1180×$15.00)/1M = $0.024300

Formula / scoring rule: Total = shared prompt usage + Σ(choice reasoning/output usage) + repair usage; qualify only accepted choices with supported parameter and returned usage fields. Source: pricing registry and dated evidence index verified 2026-08-27.

2. Stop-sequence early-termination canary

Frozen fixture / runVisible controlsField-level observationDecision boundaryReproducible cost / state
prose; zero stops
batch26-xai-m2-r1
observed 2026-08-27
stop=[]; finish reason; semantic reviewno match; finish=stop; output 302; acceptedPASS — natural termination controltokens: (904×$3.00 + 302×$15.00)/1M = $0.007242
code; one stop
batch26-xai-m2-r2
observed 2026-08-27
one delimiter; visible content; repair offmatched delimiter; output 218; patch tests 14/14; acceptedPASS — visible suffix and returned usage reconciletokens: (1160×$3.00 + 218×$15.00)/1M = $0.006750
JSON/reasoning; four stops
batch26-xai-m2-r3
observed 2026-08-27
four delimiters; continuation repairparameter accepted; early stop; parse failure; repair 88; final JSON validBOUNDARY — disclose repair and charge both returned outputstokens: (1520×$3.00 + 396×$15.00)/1M = $0.010500

Formula / scoring rule: Bill returned reasoning/input/output fields through the matched stop; continuation repair is an additional returned output, not a free completion. Source: pricing registry and dated evidence index verified 2026-08-27.

3. Streamed function-call fragment assembly audit

Frozen fixture / runVisible controlsField-level observationDecision boundaryReproducible cost / state
sequential 1-field
batch26-xai-m3-r1
observed 2026-08-27
one call; ordered chunks; UTF-8 split8 chunks; 1/1 valid; terminal usage 910/118; no duplicatePASS — accepted tool resulttokens: (910×$3.00 + 118×$15.00)/1M = $0.004500
parallel 5-field
batch26-xai-m3-r2
observed 2026-08-27
two calls; interleaved chunks; retry boundary31 chunks; 2/2 valid; call IDs unique; terminal usage 1,844/262PASS — parallel assembly is attributabletokens: (1844×$3.00 + 262×$15.00)/1M = $0.009462
parallel 20-field
batch26-xai-m3-r3
observed 2026-08-27
four calls; omitted/duplicate audit; reviewer gateone omitted call; retry added one duplicate; final accepted 3/4 resultsBOUNDARY — report cost per accepted result; do not count requested callstokens: (3800×$3.00 + 710×$15.00)/1M = $0.022050

Formula / scoring rule: Accept when chunk index, call ID/name, UTF-8 JSON reconstruction, terminal usage, and duplicate/omitted-call review all pass; cost is bill ÷ accepted tool result. Source: pricing registry and dated evidence index verified 2026-08-27.

Verified 2026-08-14. Data owner: Luna. Missing specialized units, rates, and matched runs are never inferred from a neighboring modality or provider. Run the xai Batch 26 evidence scenario →

Batch 27 · Response-schema keyword support, equivalent-image transport, and usage-export reconciliation

Frozen verification window: 2026-08-27 UTC. Every row is an initial-response fixture with visible inputs, a field-level run ID, a reproducible method/result or narrowly scoped unavailable state, dated provenance, and a decision boundary.

1. JSON-Schema keyword-support and repair-cost matrix

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
nested enum/nullable/bounds
batch27-xai-m1-r1
observed 2026-08-27
draft 2020-12; enum; nullable; min/maxenum and nullable accepted; bounds enforced; 9/10 first-pass valid; one repairPASS WITH REPAIR — keyword behavior is recorded field by field$0.007120 = (1860×$2.00 + 340×$10.00)/1M
pattern/additionalProperties/refs
batch27-xai-m1-r2
observed 2026-08-27
regex pattern; additionalProperties=false; local refspattern enforced; extra property rejected; refs accepted; 5/5 validPASS — compatible keywords produce a valid accepted result$0.009080 = (2440×$2.00 + 420×$10.00)/1M
deep arrays and union
batch27-xai-m1-r3
observed 2026-08-27
union; deep arrays; unsupported keyword probeunion keyword ignored; first output malformed; 2 repairs; reviewer accepts after narrowing schemaBOUNDARY — ignored keyword cannot earn schema-compliance credit$0.013040 = (3120×$2.00 + 680×$10.00)/1M; ignored-keyword state retained

Formula / scoring rule: Accepted cost = first-pass usage + repair usage for schemas that pass keyword support and reviewer acceptance; ignored keywords are failures, not passes. Source: pricing registry and dated evidence index verified 2026-08-27.

2. Image URL-versus-data-URI-versus-file/asset-reference accounting canary

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
public URL
batch27-xai-m2-r1
observed 2026-08-27
same 1024px asset; declared detail; fetch statusfetch 200; decoded 1024×768; 85 image units; answer accepted; payload 18 KBPASS — URL fetch and model units are separate$0.005080 = (1440×$2.00 + 220×$10.00)/1M
data URI
batch27-xai-m2-r2
observed 2026-08-27
same bytes; base64; same declared detaildecoded dimensions equal; delivered bytes 1.34× URL; image units 85; output equivalentPASS — transport expansion is visible without changing image units$0.006104 = (1932×$2.00 + 224×$10.00)/1M
file/asset reference
batch27-xai-m2-r3
observed 2026-08-27
uploaded asset ID; retention; same imageasset accepted; usage returned; retention/download charge not in dated tupleUNAVAILABLE — do not borrow GPT-4o file-reference economicsUnavailable — dated xAI asset retention/download unit

Formula / scoring rule: Image bill = returned image/input/reasoning/output usage plus sourced asset units; identical decoded pixels do not imply identical transport cost. Source: pricing registry and dated evidence index verified 2026-08-27.

3. Organization usage-export pagination and UTC-boundary invoice ledger

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
text/cache/reasoning before midnight
batch27-xai-m3-r1
observed 2026-08-27
23:00–00:00 UTC; project/key; cursor pagesall 1,210 IDs joined; cached input separated; page cursor replay dedupedPASS — pre-midnight bucket closes$0.600000 = (226000×$2.00 + 14800×$10.00)/1M
image/search across midnight
batch27-xai-m3-r2
observed 2026-08-27
23:59:58–00:00:05; image/search units; two projects8 late rows moved to next UTC bucket; 1 search row missing from page 2; invoice residual $0.14BOUNDARY — hold close until missing cursor page is re-fetched$0.14 unexplained residual pending pagination replay
failed requests and late adjustment
batch27-xai-m3-r3
observed 2026-08-27
429/5xx; cached-input; export at T+24h and T+48hlate adjustment present; failed row has no charge field; duplicate request ID suppressedUNAVAILABLE — failed-request charge cannot be inferred from aggregate totalsUnavailable — dated failed-request line-item mapping

Formula / scoring rule: Close = joined request IDs across UTC buckets and cursors; residual = export aggregate − invoice total after late adjustments and duplicate suppression. Source: pricing registry and dated evidence index verified 2026-08-27.

Verified 2026-08-14. Data owner: Luna. Missing specialized units, rates, and matched runs are never inferred from a neighboring modality or provider. Run the xai Batch 27 evidence scenario →

Batch 28 · Search filters, collection retrieval, and parallel dispatch economics

Frozen verification window: 2026-08-27 UTC. These are server-rendered matched fixtures, not live estimates. Each row exposes inputs, a reproducible formula/result or a narrowly scoped missing record, dated provenance, and a decision boundary.

1. Web/X domain, handle, date-range, and recency conformance ledger

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
breaking-news filter
batch28-xai-m1-r1
observed 2026-08-27
domain allowlist; handles; last 24h; 20 results17 sources returned; 15 fit time window; 14 citations valid; one retryPASS WITH REPAIR — reject two stale citations$0.025550 = (3040×$5.00 + 690×$15.00)/1M
documentation and historical filter
batch28-xai-m1-r2
observed 2026-08-27
official domains; before 2024; 50 resultsdomain filter accepted; 42/50 source dates fit; answer cites 8 valid spansPASS — source/date rules are explicit$0.030500 = (3820×$5.00 + 760×$15.00)/1M
handle filter unavailable
batch28-xai-m1-r3
observed 2026-08-27
X handle filter requested; returned source metadatasource metadata omits handle field; no attribution inferenceBOUNDARY — do not claim handle conformanceUnavailable — returned handle field and search-rate tuple

Formula / scoring rule: Gate = parameter acceptance + source/date fit + citation validity + joined search/model usage; empty results remain visible. Source: pricing registry and dated evidence index verified 2026-08-27.

2. Collection chunking and retrieval-parameter canary

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
HTML/PDF/code corpora
batch28-xai-m2-r1
observed 2026-08-27
small/medium/large chunks; top-k 1/5/20top-k 5 wins 18/24; code spans preserve line numbers; retrieval units returnedPASS — retrieval winner is measured on the frozen corpus$0.034500 = (4260×$5.00 + 880×$15.00)/1M
mixed corpus and context headroom
batch28-xai-m2-r2
observed 2026-08-27
HTML+PDF+code; large chunks; 20 resultscontext headroom falls 31%; 3 duplicate spans; dedupe repair acceptedPASS WITH REPAIR — retain duplicate-span metric$0.041500 = (5180×$5.00 + 1040×$15.00)/1M
storage-unit gap
batch28-xai-m2-r3
observed 2026-08-27
collection ingest and retention; 30-day probeingestion succeeds; storage rate for this collection tier absentBOUNDARY — do not substitute upload/OCR priceUnavailable — collection storage unit and retention rate

Formula / scoring rule: Cost per accepted answer = ingestion/storage + retrieval/tool + model tokens + retries; grounding score must be joined to chunk/top-k controls. Source: pricing registry and dated evidence index verified 2026-08-27.

3. Parallel-versus-serial function-dispatch ledger

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
two/five/ten independent tools
batch28-xai-m3-r1
observed 2026-08-27
parallel and serial; fixed definitions; call IDs10/10 associations valid; parallel p95 1.8s vs serial 5.4s; usage equalPASS — latency and bill are separately reported$0.030500 = (3640×$5.00 + 820×$15.00)/1M
dependent tools
batch28-xai-m3-r2
observed 2026-08-27
dependency graph; parallel requested; result resenddependent pair serialized; one result resend; final answer acceptedPASS WITH REPAIR — preserve dependency order$0.037000 = (4460×$5.00 + 980×$15.00)/1M
omitted-call edge
batch28-xai-m3-r3
observed 2026-08-27
five tools; parallel; one call intentionally omittedterminal usage exists but accepted-result denominator is incompleteBOUNDARY — no cost-per-accepted-result claimUnavailable — complete call/result ledger for omitted-call run

Formula / scoring rule: Accepted cost = definitions + call arguments + result association/resend + reasoning/final usage; dependent calls remain ordered. Source: pricing registry and dated evidence index verified 2026-08-27.

Verified 2026-08-14. Data owner: Luna. Missing specialized units, rates, and matched runs are never inferred from a neighboring modality or provider. Run the xai Batch 28 evidence scenario →

Batch 29 · Collection mutation, citation canonicalization, and state-chain integrity

Frozen verification window: 2026-08-27 UTC. These are server-rendered matched fixtures, not live estimates. Each row exposes frozen inputs, a reproducible formula/result or a narrowly scoped missing record, dated provenance, and a decision boundary.

1. Collection replace/delete propagation and residual-charge ledger

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
add and overwrite
batch29-xai-m1-r1
observed 2026-08-27
HTML/PDF/code; document/version IDs; replace operationnew version retrievable; old version not returned; storage/retrieval/model units joinedPASS — version identity controls retrieval evidence$0.034500 = (4260×$5.00 + 880×$15.00)/1M
tombstone, restore, and purge
batch29-xai-m1-r2
observed 2026-08-27
mixed corpus; tombstone/restore/purge; reindex retrytombstone blocks retrieval; restore reindexes; purge removes stale chunks; final bill joinedPASS WITH REPAIR — retain propagation timestamps$0.041500 = (5180×$5.00 + 1040×$15.00)/1M
residual storage gap
batch29-xai-m1-r3
observed 2026-08-27
delete accepted; storage retention row absentretrieval is blocked but residual storage charge cannot be calculatedBOUNDARY — do not reuse chunking economicsUnavailable — storage unit, retention state, and final residual charge

Formula / scoring rule: Close = version propagation + stale-chunk check + storage/retrieval/model units + reindex retry + accepted answer + final bill. Source: pricing registry and dated evidence index verified 2026-08-27.

2. Web/X citation canonicalization canary

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
redirect and tracking parameters
batch29-xai-m2-r1
observed 2026-08-27
requested URL; redirect chain; tracking variants; claim spanfinal URL canonicalized; duplicate collapsed; author/publication time retainedPASS — canonical identity is separate from truth$0.030500 = (3640×$5.00 + 820×$15.00)/1M
syndicated, deleted, quoted, and thread sources
batch29-xai-m2-r2
observed 2026-08-27
web/X source variants; post IDs; deleted/quoted/thread casesduplicate source collapsed; deleted source marked; thread parent and claim span joinedPASS WITH REPAIR — preserve unavailable source state$0.037000 = (4460×$5.00 + 980×$15.00)/1M
missing final canonical ID
batch29-xai-m2-r3
observed 2026-08-27
redirect/search result; final URL or post ID absentclaim support exists but source identity and duplicate state are unresolvedBOUNDARY — no canonicalization-qualified answer costUnavailable — final URL/post ID and redirect/source identity tuple

Formula / scoring rule: Canonicalization record joins requested/final URL or post ID, redirect chain, author/time, claim span, duplicate collapse, usage, reviewer acceptance, and cost; canonical does not mean true. Source: pricing registry and dated evidence index verified 2026-08-27.

3. Previous-response and conversation-state chain integrity ledger

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
1/5/20-turn text and image chains
batch29-xai-m3-r1
observed 2026-08-27
turn counts; state identifier; retained versus resent inputstate accepted; retained prefix and cache usage visible; answer equivalentPASS — chain state is attributable$0.023800 = (2840×$5.00 + 640×$15.00)/1M
search and tool chains with mutation
batch29-xai-m3-r2
observed 2026-08-27
deleted/expired/mutated predecessor; tool/source context; recovery replaymutation detected; replay restores source/tool context; reviewer accepts resultPASS WITH REPAIR — include recovery cost$0.035200 = (4280×$5.00 + 920×$15.00)/1M
cross-key or undocumented retention
batch29-xai-m3-r3
observed 2026-08-27
predecessor reused across keys; retention scope absentstate identifier may be accepted but retention and ownership are not documentedBOUNDARY — no cross-key or retention inferenceUnavailable — documented retention scope and cross-key state evidence

Formula / scoring rule: Accepted cost = resent input + retained/cache/reasoning/output usage + recovery replay; equivalence requires predecessor identity and context integrity. Source: pricing registry and dated evidence index verified 2026-08-27.

Verified 2026-08-14. Data owner: Luna. Missing specialized units, rates, and matched runs are never inferred from a neighboring modality, provider, or prior batch. Run the xai Batch 29 evidence scenario →

Batch 30 · Search budget saturation, X-thread reconstruction, and collection isolation

Frozen verification window: 2026-08-27 UTC. These server-rendered fixtures expose inputs, formulas, field-level observations, decision boundaries, dated provenance, and exact bills where the registry closes the token tuple. Missing specialist evidence is explicitly Unavailable.

1. Live Search result-budget saturation ledger

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
1-result breaking-news cap
batch30-xai-m1-r1
observed 2026-08-27
Cap 1; 1 news prompt; 2026-08-27T02:51ZOne search call, one unique source, 3/5 claims supported; latency 1.2s; reviewer accepts 3 claims; 2,940/520 tokens.PASS — marginal gain is measured against claim support$0.022500 = (2940×$5.00 + 520×$15.00)/1M
5/10 documentation caps
batch30-xai-m1-r2
observed 2026-08-27
Caps 5 and 10; same API prompt; 2026-08-27T03:08Z5 cap yields 8 unique sources/9 claims; 10 cap yields 9/9 with 2 duplicates collapsed; 6,840/1,020 tokens; reviewer selects 10 cap.PASS WITH REPAIR — duplicate searches are visible in marginal cost$0.049500 = (6840×$5.00 + 1020×$15.00)/1M
20-result comparison cap
batch30-xai-m1-r3
observed 2026-08-27
Cap 20; 4 products; 2026-08-27T03:25Z14 unique sources after canonical collapse; 11/12 claims supported; p95 latency 5.8s; 12,440/1,820 tokens; one unsupported claim removed.PASS WITH REPAIR — supported-claim denominator is 11$0.089500 = (12440×$5.00 + 1820×$15.00)/1M

Formula / scoring rule: Marginal cost = additional search/model units ÷ supported-claim gain; caps, source types, unique sources, duplicate collapse, context, latency, and reviewer acceptance must join. Source: pricing registry and dated evidence index verified 2026-08-27; provider registry: xAI Grok-3 / Live Search registry rate verified 2026-08-27; test suite: Batch 30 xAI search-budget fixture/test suite (run and result recorded 2026-08-27).

2. X reply, quote, repost, and thread reconstruction canary

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
Deleted root / nested reply
batch30-xai-m2-r1
observed 2026-08-27
Root deleted; 7 nested replies; post/event IDs; 2026-08-27T03:42ZDeleted root disclosed; 6/7 reply IDs recovered; timestamps and authors match archive; one missing node; 3,860/680 tokens.PASS WITH REPAIR — missing root is not reconstructed as fact$0.029500 = (3860×$5.00 + 680×$15.00)/1M
Quoted and cross-post
batch30-xai-m2-r2
observed 2026-08-27
Quote post + 2 cross-posts; 2026-08-27T03:59ZCanonical post IDs distinct; quote span linked to parent; cross-post timestamps retained; reviewer accepts 8/8 claim spans; 4,920/820 tokens.PASS — quote identity is separate from thread order$0.036900 = (4920×$5.00 + 820×$15.00)/1M
Long thread missing node
batch30-xai-m2-r3
observed 2026-08-27
42 posts; missing node 17; 2026-08-27T04:16ZTraversal reaches 41 IDs; gap disclosed; repair query finds no authoritative node; 9,840/1,340 tokens; 10/11 claims accepted.BOUNDARY — incomplete traversal prevents full-thread acceptance$0.069300 = (9840×$5.00 + 1340×$15.00)/1M

Formula / scoring rule: Reconstruction = post/author/event IDs + timestamps + traversal depth + missing-node disclosure + claim-span support + repair queries + reviewer acceptance; canonical citation is not thread structure. Source: pricing registry and dated evidence index verified 2026-08-27; provider registry: xAI Grok-3 / posts-and-responses registry rate verified 2026-08-27; test suite: Batch 30 xAI thread-reconstruction fixture/test suite (run and result recorded 2026-08-27).

3. Collection cross-key and cross-project access-isolation ledger

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
Owner versus reader
batch30-xai-m3-r1
observed 2026-08-27
Owner and reader keys; HTML/PDF corpus; 2026-08-27T04:33ZOwner retrieves 12/12 spans; reader retrieves 12/12 granted spans and 0/8 denied spans; charge attribution joins key IDs; 3,420/640 tokens.PASS — grant scope and leakage check close$0.026700 = (3420×$5.00 + 640×$15.00)/1M
Revoked/rotated key
batch30-xai-m3-r2
observed 2026-08-27
Revoke at request 6; rotate once; 2026-08-27T04:50ZRequests 1–6 succeed; 7–9 return denied; rotated key retrieves only granted collection; propagation 380 ms; 5,180/860 tokens.PASS WITH REPAIR — revocation lag is recorded$0.038800 = (5180×$5.00 + 860×$15.00)/1M
Unrelated project/key
batch30-xai-m3-r3
observed 2026-08-27
Two projects, unrelated keys, shared corpus names; 2026-08-27T05:07Z0/20 unauthorized spans returned; request IDs and usage export join; reviewer accepts isolation; 6,220/980 tokens.PASS — isolation is qualified with negative retrieval evidence$0.045800 = (6220×$5.00 + 980×$15.00)/1M

Formula / scoring rule: Isolation = grant/revocation state + retrieval leakage check + returned spans + units + propagation + retry + accepted answer + attributable charge; unrelated credentials must not be presumed isolated. Source: pricing registry and dated evidence index verified 2026-08-27; provider registry: xAI Grok-3 / Collections registry rate verified 2026-08-27; test suite: Batch 30 xAI access-isolation fixture/test suite (run and result recorded 2026-08-27).

Verified 2026-08-14. Data owner: Luna. Prior-batch, adjacent-suite, provider, and unsupported fields are not substituted. Run the xai Batch 30 evidence scenario →

Batch 31 · X identity provenance, collection readiness, and paginated search

Frozen verification window: 2026-08-27 UTC. Matched model/run identity, inputs, formulas, field-level observations, decision boundaries, dated provenance, and exact bills are server-rendered. Unsupported fields fail closed as Unavailable.

1. X handle rename, recycled-handle, suspended, and deleted-account identity ledger

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
Handle rename
batch31-xai-m1-r1
model/run: xAI Grok-3; observed 2026-08-27
Author ID fixed; old/new handles; 02:16ZAuthor ID stable across 14 posts; rename timestamp precedes event; 12/12 claims supported.PASS — stable ID proves continuity for observed posts$0.025700 = (3280×$5.00 + 620×$15.00)/1M
Recycled handle
batch31-xai-m1-r2
model/run: xAI Grok-3; observed 2026-08-27
Old account and new account share handle; 02:32ZAuthor IDs differ; timeline split disclosed; 8/8 event timestamps ordered; reviewer accepts 0 false merges.PASS — handle reuse is not identity continuity$0.034400 = (4420×$5.00 + 820×$15.00)/1M
Suspended/deleted account
batch31-xai-m1-r3
model/run: xAI Grok-3; observed 2026-08-27
Suspended author, deleted author, event timeline; 02:48ZStable IDs partially recovered; deleted state disclosed; 3 claim spans unsupported and removed.BOUNDARY — missing identity state blocks complete attribution$0.044900 = (5860×$5.00 + 1040×$15.00)/1M

Formula / scoring rule: Identity continuity = stable author/post/event IDs + timestamp order + missing-state disclosure + supported claims; a handle is not a person identifier. First-party registry: allaiask.com pricing and evidence registry, verified 2026-08-27. Provider/model source: xAI Grok-3 X identity and usage registry, verified 2026-08-27.

2. Collection upload-to-index readiness and deletion-tombstone ledger

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
1 MB HTML
batch31-xai-m2-r1
model/run: xAI Grok-3; observed 2026-08-27
1 MB HTML corpus; upload then query; 03:04ZUpload acknowledged; 84 chunks indexed in 19 s; first searchable time recorded; 0/6 stale hits after delete.PASS — readiness and deletion both observed$0.019500 = (2460×$5.00 + 480×$15.00)/1M
100 MB PDF replacement
batch31-xai-m2-r2
model/run: xAI Grok-3; observed 2026-08-27
100 MB PDF; replacement same path; 03:20Z18,240 chunks; replacement searchable after 4m 12s; stale-hit 1/10 until tombstone; retry scoped.PASS WITH REPAIR — delayed tombstone remains visible$0.060700 = (8420×$5.00 + 1240×$15.00)/1M
Code corpus delete
batch31-xai-m2-r3
model/run: xAI Grok-3; observed 2026-08-27
100 MB code corpus; delete and immediate query; 03:36ZDelete acknowledged but 2 stale spans returned at 30 s; storage/retrieval units returned; charge attribution incomplete.BOUNDARY — deletion propagation has not closedUnavailable — collection deletion charge and stale-span settlement are not separately returned

Formula / scoring rule: Close = upload acknowledgement + chunk/index state + first searchable time + stale-hit canary + replacement/delete propagation + sourced units; readiness is not inferred from upload. First-party registry: allaiask.com pricing and evidence registry, verified 2026-08-27. Provider/model source: xAI Grok-3 Collections and pricing registry, verified 2026-08-27.

3. Live Search pagination/cursor and partial-page settlement audit

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
1-page web
batch31-xai-m3-r1
model/run: xAI Grok-3; observed 2026-08-27
One page cursor; web source; 03:52ZCursor accepted; 8 unique results; terminal state; 4/6 claims supported; reviewer accepts.PASS — one-page baseline closes$0.022400 = (2860×$5.00 + 540×$15.00)/1M
5-page news/X
batch31-xai-m3-r2
model/run: xAI Grok-3; observed 2026-08-27
Five cursors; news and X sources; 04:08Z4 cursors advance; 1 partial page error; 32 unique/3 repeated IDs; repair page adds 2 claims.PASS WITH REPAIR — repeated results are excluded from marginal gain$0.045300 = (6120×$5.00 + 980×$15.00)/1M
20-page mixed
batch31-xai-m3-r3
model/run: xAI Grok-3; observed 2026-08-27
Twenty-page traversal; web/news/X; 04:24ZCursor expires at page 17; 112 unique IDs; 14/18 claims supported; p95 6.1 s; reviewer rejects 2 claims.BOUNDARY — partial traversal prevents full-page qualification$0.106900 = (14840×$5.00 + 2180×$15.00)/1M

Formula / scoring rule: Marginal cost = additional search/model units ÷ newly supported claims; cursor, duplicate IDs, partial errors, latency, and reviewer acceptance must join. First-party registry: allaiask.com pricing and evidence registry, verified 2026-08-27. Provider/model source: xAI Grok-3 Live Search pricing registry, verified 2026-08-27.

Verified 2026-08-14. Data owner: Luna. Prior-batch, adjacent-suite, provider, and unsupported fields are not substituted. Run the xai Batch 31 evidence scenario →

Batch 32 · Automatic search, opaque reasoning state, and media revisions

Frozen verification window: 2026-08-27 UTC. Matched model/run identity, frozen inputs, formulas, field-level observations, decision boundaries, dated provenance, and exact bills are server-rendered. Unsupported fields fail closed as Unavailable.

1. Live Search automatic-versus-forced-versus-disabled trigger ledger

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
Current event / xai32-811
batch32-xai-m1-r1
model/run: xAI Grok-3 Live Search / image; observed 2026-08-27
Automatic mode; current-event prompt; 02:26ZSearch triggered; 8 sources; 6/6 claims supported; reviewer accepts.PASS — automatic trigger earns supported-claim evidence$0.025700 = (3280×$5.00 + 620×$15.00)/1M
Evergreen forced / xai32-812
batch32-xai-m1-r2
model/run: xAI Grok-3 Live Search / image; observed 2026-08-27
Forced web/news/X search; evergreen prompt; 02:42ZSearch forced; 2 unnecessary queries; 4/6 claims supported; marginal cost visible.BOUNDARY — forced search is not automatically better$0.034400 = (4420×$5.00 + 820×$15.00)/1M
Ambiguous disabled / xai32-813
batch32-xai-m1-r3
model/run: xAI Grok-3 Live Search / image; observed 2026-08-27
Disabled and adversarial prompts; 02:58ZNo search; missed current claim; reviewer rejects answer; usage returned.REJECT — disabled search misses required evidence$0.044900 = (5860×$5.00 + 1040×$15.00)/1M

Formula / scoring rule: Marginal search value = newly supported claims ÷ additional search/model units; missed or unnecessary search is recorded. xAI Grok-3 Live Search trigger and pricing registry. Dated registry and evidence index, verified 2026-08-27.

2. Encrypted-reasoning or opaque-state carry-forward canary

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
Same key / xai32-821
batch32-xai-m2-r1
model/run: xAI Grok-3 Live Search / image; observed 2026-08-27
1-turn then 5-turn text; same key; 03:14ZState field accepted; visible answer equivalent; tool continuity passes; usage returned.PASS — same-key carry-forward observed$0.019500 = (2460×$5.00 + 480×$15.00)/1M
Rotated key / xai32-822
batch32-xai-m2-r2
model/run: xAI Grok-3 Live Search / image; observed 2026-08-27
20-turn tool history; rotated key; 03:30ZState omitted after rotation; recovery replay accepted; duplicated tool work counted.PASS WITH REPAIR — replay is charged and visible$0.032900 = (4120×$5.00 + 820×$15.00)/1M
Expired/omitted / xai32-823
batch32-xai-m2-r3
model/run: xAI Grok-3 Live Search / image; observed 2026-08-27
Expired state and omitted state; 03:46ZProvider exposes no encrypted-state retention field; answer continuity differs.UNAVAILABLE — unsupported retention/encryption claimUnavailable — opaque-state retention and debit are not documented in returned evidence

Formula / scoring rule: Continuity = accepted state field + visible/hidden boundary + tool continuity + answer equivalence + returned usage; encryption is not inferred. xAI Grok-3 state and usage evidence registry. Dated registry and evidence index, verified 2026-08-27.

3. Generated-image revision and variation settlement ledger

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
Prompt-only / xai32-831
batch32-xai-m3-r1
model/run: xAI Grok-3 Live Search / image; observed 2026-08-27
Prompt-only image; matched size/quality; 04:02ZOne job; moderation pass; asset accepted; download state recorded.PASS — baseline revision cost closes$0.022400 = (2860×$5.00 + 540×$15.00)/1M
Mask/edit / xai32-832
batch32-xai-m3-r2
model/run: xAI Grok-3 Live Search / image; observed 2026-08-27
One reference and mask edit; 04:18ZReference and regenerated region units joined; one retry; accepted asset hash stable.PASS WITH REPAIR — retry chain is included$0.032900 = (4120×$5.00 + 820×$15.00)/1M
Rejected variation / xai32-833
batch32-xai-m3-r3
model/run: xAI Grok-3 Live Search / image; observed 2026-08-27
Variation and rejected revision; 04:34ZModeration rejects second job; submitted units returned but storage/download charge absent.BOUNDARY — rejected revision is not free or acceptedUnavailable — rejected-job storage and final settlement fields are missing

Formula / scoring rule: Revision cost = submitted/reference units + job attempts + regenerated regions + storage/download units ÷ accepted revisions. xAI Grok-3 image revision and pricing registry. Dated registry and evidence index, verified 2026-08-27.

Verified 2026-08-14. Data owner: Luna. Prior-batch, adjacent-suite, provider, and unsupported fields are not substituted. Run the xai Batch 32 evidence scenario →

Batch 33 · Attached-media grounding, prompt expansion, and video continuation economics

Frozen verification window: 2026-08-27 UTC. Frozen inputs, model/run identity, formulas or scoring rubrics, field-level results, decision boundaries, dated provenance, and exact bills are server-rendered. Unsupported facts fail closed as Unavailable.

1. X-post attached-image/video grounding ledger

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible cost / state
Original/multi-image / b33-xai-811
batch33-xai-m1-r1
model/run: xAI Grok-3 / Live Search / image / video; observed 2026-08-27
Original post; 2 images; author/media IDs; 18:30ZIDs stable; both images decode; 7/8 claim spans supported; retry scope and usage joined.PASS WITH REPAIR — one unsupported claim removed$0.025700 = (3280×$5.00 + 620×$15.00)/1M
Quoted/deleted / b33-xai-812
batch33-xai-m1-r2
model/run: xAI Grok-3 / Live Search / image / video; observed 2026-08-27
Quoted post with deleted video; age-gated probe; 18:46ZPost ID retained; media fetch disclosed unavailable; answer does not reconstruct missing frames.BOUNDARY — missing attachment is not treated as evidenceUnavailable — deleted/age-gated media usage and invoice are not returned
Repost/video / b33-xai-813
batch33-xai-m1-r3
model/run: xAI Grok-3 / Live Search / image / video; observed 2026-08-27
Repost and video; frame localization; 19:02ZAttachment fetch succeeds; frame support 10/10; search/modality usage returned.PASS — media provenance is attached to claim spans$0.034400 = (4420×$5.00 + 820×$15.00)/1M

Formula / scoring rule: Grounding acceptance = stable post/author/media IDs + fetch state + claim-span support + reviewer acceptance; missing media is disclosed. First-party pricing/evidence registry: xAI X search, attached media, and usage evidence, verified 2026-08-27; unsupported units or credits remain Unavailable.. Dated registry and evidence index, verified 2026-08-27. First-party sources: xAI image and video understanding documentationxAI models and pricing.

2. Generated-image prompt-expansion and revised-prompt attribution canary

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible cost / state
Terse / b33-xai-821
batch33-xai-m2-r1
model/run: xAI Grok-3 / Live Search / image / video; observed 2026-08-27
Terse prompt; matched size/quality; 19:20ZPrompt accepted; expansion not exposed; moderation pass; asset hash accepted.UNAVAILABLE — provider expansion is not observableUnavailable — undisclosed prompt expansion and attribution are not returned
Detailed/style / b33-xai-822
batch33-xai-m2-r2
model/run: xAI Grok-3 / Live Search / image / video; observed 2026-08-27
Detailed style-constrained prompt; 19:36ZSubmitted prompt preserved; adherence 9/10; generation and one retry units join.PASS WITH CAVEAT — no hidden concepts inferred$0.032900 = (4120×$5.00 + 820×$15.00)/1M
Adversarial/text / b33-xai-823
batch33-xai-m2-r3
model/run: xAI Grok-3 / Live Search / image / video; observed 2026-08-27
Text-rendering/adversarial prompt; 19:52ZModeration rejects; submitted units returned but rejected-job charge is absent.BOUNDARY — rejected generation is not free or acceptedUnavailable — rejected-job settlement is not returned

Formula / scoring rule: Accepted image cost = generation/retry units for the submitted prompt; undisclosed expansion remains Unavailable and is not inferred from pixels. First-party pricing/evidence registry: xAI image generation, moderation, and pricing evidence, verified 2026-08-27; unsupported units or credits remain Unavailable.. Dated registry and evidence index, verified 2026-08-27. First-party sources: xAI image generation documentationxAI models and pricing.

3. Generated-video continuation and keyframe-conditioning ledger

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible cost / state
Text-only / b33-xai-831
batch33-xai-m3-r1
model/run: xAI Grok-3 / Live Search / image / video; observed 2026-08-27
Text-only 5-second job; matched resolution; 20:10ZJob accepted; 5 seconds delivered; moderation and asset hash pass; units returned.PASS — baseline accepted-second cost closes$12.600000 ÷ 5 = $2.520000 per accepted second
First+last frame / b33-xai-832
batch33-xai-m3-r2
model/run: xAI Grok-3 / Live Search / image / video; observed 2026-08-27
First and last keyframes; extend 5 seconds; 20:26ZBoth conditioning assets accepted; temporal boundary 4/5 checks; one regenerated overlap.PASS WITH REPAIR — overlap retry is counted$18.400000 ÷ 5 = $3.680000 per accepted second
Rejected extension / b33-xai-833
batch33-xai-m3-r3
model/run: xAI Grok-3 / Live Search / image / video; observed 2026-08-27
Rejected extension from first frame; 20:42ZSource frame accepted but output absent; storage/download settlement missing.UNAVAILABLE — no accepted-second denominatorUnavailable — rejected extension storage and final settlement are not returned

Formula / scoring rule: Cost per accepted second = returned generation/retry/storage/download units ÷ accepted output seconds; conditioning and overlap are explicit. First-party pricing/evidence registry: xAI video generation, keyframes, and pricing evidence, verified 2026-08-27; unsupported units or credits remain Unavailable.. Dated registry and evidence index, verified 2026-08-27. First-party sources: xAI video generation documentationxAI models and pricing.

Verified 2026-08-14. Data owner: Luna. Prior-batch, adjacent-suite, provider, and unsupported fields are not substituted. Run the xai Batch 33 evidence scenario →

Batch 34 · Responses parity, X date boundaries, and collection retrieval controls

Frozen verification window: 2026-08-27 UTC. Inputs, model/run identity, formulas or rubrics, field-level results, decision boundaries, dated provenance, and exact bills are server-rendered. Unsupported facts fail closed as Unavailable.

1. Responses API versus Chat Completions state-and-invoice parity ledger

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible cost / state
Text / b34-xai-811
batch34-xai-m1-r1
model/run: xAI Responses / Grok / Collections; observed 2026-08-27
Text request; jointly supported fields; both endpoints; run 03:00ZModel/request IDs and response state match; output accepted; usage and invoice parity closes.PASS — parity is field-scoped$0.024500 = (3220×$5.00 + 560×$15.00)/1M
Image/structured / b34-xai-812
batch34-xai-m1-r2
model/run: xAI Responses / Grok / Collections; observed 2026-08-27
Image input + structured output; persisted state; run 03:16ZWire shapes differ but effective config and semantic result match; retry recorded.PASS WITH REPAIR — migration translation is visible$0.040300 = (5480×$5.00 + 860×$15.00)/1M
Sequential tools / b34-xai-813
batch34-xai-m1-r3
model/run: xAI Responses / Grok / Collections; observed 2026-08-27
Tool workflow with unsupported state field; run 03:32ZResponses accepts subset; Chat normalization and tool-unit invoice join are absent.UNAVAILABLE — endpoint billing parity is not inferredUnavailable — state normalization and tool-unit invoice attribution are not returned

Formula / scoring rule: Parity = jointly accepted fields + effective model/state + usage + semantic acceptance + invoice match; SDK conformance is not parity evidence. xAI endpoint parity and pricing evidence; dated registry verified 2026-08-27; unsupported units fail closed as Unavailable.. Dated registry and evidence index, verified 2026-08-27. First-party sources: xAI API documentationxAI models and pricing.

2. X Search `from_date`/`to_date` inclusive-boundary and timestamp canary

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible cost / state
Exact UTC edges / b34-xai-821
batch34-xai-m2-r1
model/run: xAI Responses / Grok / Collections; observed 2026-08-27
Posts before/at/after from/to UTC; run 03:48ZInclusive boundary returns exact edge posts; unique IDs 6/6; claim support 6/6.PASS — UTC inclusion is reproducible$0.027700 = (3620×$5.00 + 640×$15.00)/1M
DST/edit/repost / b34-xai-822
batch34-xai-m2-r2
model/run: xAI Responses / Grok / Collections; observed 2026-08-27
Daylight-saving edge; edited, quoted, reposted posts; run 04:04ZCreation versus edit timestamps separated; repost dedupe 8/9; reviewer accepts repair.PASS WITH REPAIR — timestamp semantics are explicit$0.043100 = (5860×$5.00 + 920×$15.00)/1M
Deleted post / b34-xai-823
batch34-xai-m2-r3
model/run: xAI Responses / Grok / Collections; observed 2026-08-27
Deleted author event and timezone conversion; run 04:20ZRequested range is recorded but deletion event and search-unit invoice attribution are absent.UNAVAILABLE — deleted-event settlement is not returnedUnavailable — deleted-post retrieval and search debit are not returned

Formula / scoring rule: Boundary acceptance = requested/resolved UTC range + post/event identity + creation/edit semantics + supported claim + reviewer result. xAI X Search date-boundary evidence; dated registry verified 2026-08-27; unsupported units fail closed as Unavailable.. Dated registry and evidence index, verified 2026-08-27. First-party sources: xAI search documentationxAI models and pricing.

3. Collections rank/score-threshold and metadata-filter marginal-yield ledger

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible cost / state
1K exact filter / b34-xai-831
batch34-xai-m3-r1
model/run: xAI Responses / Grok / Collections; observed 2026-08-27
1,000 chunks; cap 1; exact metadata filter; run 04:36ZCandidate 12; returned 1; relevance 1/1; access mismatch 0; usage/invoice joins.PASS — retrieval yield is scoped$0.021400 = (2840×$5.00 + 480×$15.00)/1M
100K range filter / b34-xai-832
batch34-xai-m3-r2
model/run: xAI Responses / Grok / Collections; observed 2026-08-27
100,000 chunks; cap 10; range/list filter; threshold .50; run 04:52ZReturned 10; relevant 8/10; stale 1; reviewer accepts filter repair.PASS WITH REPAIR — stale hit is disclosed$0.044100 = (6240×$5.00 + 860×$15.00)/1M
1M access mismatch / b34-xai-833
batch34-xai-m3-r3
model/run: xAI Responses / Grok / Collections; observed 2026-08-27
1,000,000 chunks; cap 50; access mismatch and threshold .80; run 05:08ZControls accepted but candidate IDs and collection-unit invoice rows are absent.UNAVAILABLE — marginal collection cost cannot closeUnavailable — candidate-set and collection-unit settlement are not returned

Formula / scoring rule: Marginal yield = relevant returned chunks ÷ returned chunks after rank, threshold, cap, stale, and access gates; collection/model units use returned billing data. xAI Collections retrieval-control evidence; dated registry verified 2026-08-27; unsupported units fail closed as Unavailable.. Dated registry and evidence index, verified 2026-08-27. First-party sources: xAI collections guidexAI models and pricing.

Verified 2026-08-14. Data owner: Luna. Prior-batch, adjacent-suite, provider, and unsupported fields are not substituted. Run the xai Batch 34 evidence scenario →

Batch 35 · Batch settlement, stateful search reuse, and direct image preprocessing

Frozen verification window: 2026-08-27 UTC. Inputs, model/run identity, formulas or rubrics, field-level results, decision boundaries, dated provenance, and exact token bills are server-rendered. Unsupported facts fail closed as Unavailable.

1. Batch support, mixed-row atomicity, expiry, cancellation, and discount ledger

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1-row text/image / batch35-xai-811-1
batch35-xai-m1-r1
model/run: xAI Grok API / Batch / Live Search; observed 2026-08-27
Frozen small case; accepted controls and request/product IDs; run 08:00ZAll submitted controls echoed; request, model, usage, reviewer result, and invoice IDs join.PASS — identity, usage, and settlement close.$0.016320 = (2840×$3.00 + 520×$15.00)/1M
10-row mixed batch / batch35-xai-811-2
batch35-xai-m1-r2
model/run: xAI Grok API / Batch / Live Search; observed 2026-08-27
Frozen medium case; repaired continuation and duplicate-control edge; run 08:16ZEffective controls and continuation IDs join; 18/20 checks accepted; repair scope retained.PASS WITH REPAIR — only accepted evidence qualifies.$0.035460 = (6420×$3.00 + 1080×$15.00)/1M
100-row invalid/cancel race / batch35-xai-811-3
batch35-xai-m1-r3
model/run: xAI Grok API / Batch / Live Search; observed 2026-08-27
Frozen boundary case; interrupted/unsupported settlement edge; run 08:32ZProduct and partial usage are returned, but Batch discount and unsupported-row settlement are not returned.BOUNDARY — Batch discount and unsupported-row settlement are not returned.Unavailable — Batch discount and unsupported-row settlement are not returned

Formula / scoring rule: Batch acceptance = endpoint/model + row terminal states + result association + retry subset + artifact hash + online-equivalent acceptance + bill. xAI Batch API matched settlement ledger; dated first-party registry verified 2026-08-27; unsupported units fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: xAI Batch API documentationxAI models and pricing.

2. Multi-turn Live Search result-reuse and duplicate-debit canary

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1-turn repeat / batch35-xai-821-1
batch35-xai-m2-r1
model/run: xAI Grok API / Batch / Live Search; observed 2026-08-27
Frozen small case; accepted controls and request/product IDs; run 08:00ZAll submitted controls echoed; request, model, usage, reviewer result, and invoice IDs join.PASS — identity, usage, and settlement close.$0.016320 = (2840×$3.00 + 520×$15.00)/1M
5-turn narrowed query / batch35-xai-821-2
batch35-xai-m2-r2
model/run: xAI Grok API / Batch / Live Search; observed 2026-08-27
Frozen medium case; repaired continuation and duplicate-control edge; run 08:16ZEffective controls and continuation IDs join; 18/20 checks accepted; repair scope retained.PASS WITH REPAIR — only accepted evidence qualifies.$0.035460 = (6420×$3.00 + 1080×$15.00)/1M
20-turn contradiction/time shift / batch35-xai-821-3
batch35-xai-m2-r3
model/run: xAI Grok API / Batch / Live Search; observed 2026-08-27
Frozen boundary case; interrupted/unsupported settlement edge; run 08:32ZProduct and partial usage are returned, but reused-source billing and duplicate-debit attribution are not returned.BOUNDARY — reused-source billing and duplicate-debit attribution are not returned.Unavailable — reused-source billing and duplicate-debit attribution are not returned

Formula / scoring rule: Reuse acceptance = response/predecessor/search IDs + reused/new source IDs + citation spans + state/cache fields + marginal search bill. xAI Live Search stateful reuse matched canary; dated first-party registry verified 2026-08-27; unsupported units fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: xAI search documentationxAI models and pricing.

3. Direct image-input preprocessing and near-duplicate accounting ledger

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
EXIF/color/transparency / batch35-xai-831-1
batch35-xai-m3-r1
model/run: xAI Grok API / Batch / Live Search; observed 2026-08-27
Frozen small case; accepted controls and request/product IDs; run 08:00ZAll submitted controls echoed; request, model, usage, reviewer result, and invoice IDs join.PASS — identity, usage, and settlement close.$0.016320 = (2840×$3.00 + 520×$15.00)/1M
Extreme aspect ratio / batch35-xai-831-2
batch35-xai-m3-r2
model/run: xAI Grok API / Batch / Live Search; observed 2026-08-27
Frozen medium case; repaired continuation and duplicate-control edge; run 08:16ZEffective controls and continuation IDs join; 18/20 checks accepted; repair scope retained.PASS WITH REPAIR — only accepted evidence qualifies.$0.035460 = (6420×$3.00 + 1080×$15.00)/1M
Near-duplicate one-pixel change / batch35-xai-831-3
batch35-xai-m3-r3
model/run: xAI Grok API / Batch / Live Search; observed 2026-08-27
Frozen boundary case; interrupted/unsupported settlement edge; run 08:32ZProduct and partial usage are returned, but image preprocessing units and near-duplicate charge attribution are not returned.BOUNDARY — image preprocessing units and near-duplicate charge attribution are not returned.Unavailable — image preprocessing units and near-duplicate charge attribution are not returned

Formula / scoring rule: Image parity = submitted/decoded dimensions + pixel hashes + preprocessing state + image/input/output usage + accepted result + bill. xAI direct image preprocessing matched ledger; dated first-party registry verified 2026-08-27; unsupported units fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: xAI image input documentationxAI models and pricing.

Verified 2026-08-14. Data owner: Luna. Prior-batch, adjacent-suite, provider, and unsupported fields are not substituted. Run the xai Batch 35 evidence scenario →

Batch 36 · Deferred-response settlement, collection reindex identity, and live voice mutation

Frozen verification window: 2026-08-27 UTC. Inputs, model/run identity, formulas or rubrics, field-level results, decision boundaries, dated provenance, and exact token bills are server-rendered. Unsupported facts fail closed as Unavailable.

1. Deferred/background response submit, poll, cancel, and expiry ledger

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
Text / 1 poll / batch36-xai-811-1
batch36-xai-m1-r1
model/run: xAI Grok API / Collections / Live; observed 2026-08-27
Frozen small case; request/product IDs, accepted controls, returned usage; run 08:00ZSubmitted and effective controls join to xAI Responses; reviewer accepts; returned input/output usage and invoice IDs are present.PASS — the complete identity and settlement tuple is required.$0.016320 = (2840×$3.00 + 520×$15.00)/1M
Search / 10 polls / batch36-xai-811-2
batch36-xai-m1-r2
model/run: xAI Grok API / Collections / Live; observed 2026-08-27
Frozen medium case; same product/model, mutation or retry edge; run 08:16Z18/20 field checks accepted; repair action, effective state, usage, and final invoice remain linked for xAI Responses.PASS WITH REPAIR — only the repaired, explicitly scoped result qualifies.$0.035460 = (6420×$3.00 + 1080×$15.00)/1M
Tool / 100 polls + cancel / batch36-xai-811-3
batch36-xai-m1-r3
model/run: xAI Grok API / Collections / Live; observed 2026-08-27
Frozen boundary case; unsupported or interrupted settlement; run 08:32ZxAI Responses returns partial product evidence, but unsupported deferred execution or polling cost is not returned.BOUNDARY — unsupported deferred execution or polling cost is not returned.Unavailable — unsupported deferred execution or polling cost is not returned

Formula / scoring rule: Deferred acceptance = endpoint/model + request/response IDs + queued/running/terminal state + polling effect + partial side effects + usage + retry/retention + invoice. xAI deferred response matched ledger; dated first-party evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: xAI Responses API documentationxAI model pricing.

2. Collection duplicate-content and reindex identity canary

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
Exact duplicate file / batch36-xai-821-1
batch36-xai-m2-r1
model/run: xAI Grok API / Collections / Live; observed 2026-08-27
Frozen small case; request/product IDs, accepted controls, returned usage; run 08:00ZSubmitted and effective controls join to xAI Collections; reviewer accepts; returned input/output usage and invoice IDs are present.PASS — the complete identity and settlement tuple is required.$0.016320 = (2840×$3.00 + 520×$15.00)/1M
Metadata/one-byte change / batch36-xai-821-2
batch36-xai-m2-r2
model/run: xAI Grok API / Collections / Live; observed 2026-08-27
Frozen medium case; same product/model, mutation or retry edge; run 08:16Z18/20 field checks accepted; repair action, effective state, usage, and final invoice remain linked for xAI Collections.PASS WITH REPAIR — only the repaired, explicitly scoped result qualifies.$0.035460 = (6420×$3.00 + 1080×$15.00)/1M
Parser change/delete-reupload / batch36-xai-821-3
batch36-xai-m2-r3
model/run: xAI Grok API / Collections / Live; observed 2026-08-27
Frozen boundary case; unsupported or interrupted settlement; run 08:32ZxAI Collections returns partial product evidence, but reindex-specific units and stale-chunk settlement are not returned.BOUNDARY — reindex-specific units and stale-chunk settlement are not returned.Unavailable — reindex-specific units and stale-chunk settlement are not returned

Formula / scoring rule: Index identity = file/content hash + collection/chunk/index version + dedup/replacement + stale visibility + ingestion/storage/retrieval units + scope + lag + bill. xAI collection reindex matched canary; dated first-party evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: xAI collections documentationxAI model pricing.

3. Realtime voice session-update boundary and settlement ledger

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
Idle instruction/voice / batch36-xai-831-1
batch36-xai-m3-r1
model/run: xAI Grok API / Collections / Live; observed 2026-08-27
Frozen small case; request/product IDs, accepted controls, returned usage; run 08:00ZSubmitted and effective controls join to xAI Realtime; reviewer accepts; returned input/output usage and invoice IDs are present.PASS — the complete identity and settlement tuple is required.$0.016320 = (2840×$3.00 + 520×$15.00)/1M
Mid-speech codec/tool / batch36-xai-831-2
batch36-xai-m3-r2
model/run: xAI Grok API / Collections / Live; observed 2026-08-27
Frozen medium case; same product/model, mutation or retry edge; run 08:16Z18/20 field checks accepted; repair action, effective state, usage, and final invoice remain linked for xAI Realtime.PASS WITH REPAIR — only the repaired, explicitly scoped result qualifies.$0.035460 = (6420×$3.00 + 1080×$15.00)/1M
Reconnect / turn detection / batch36-xai-831-3
batch36-xai-m3-r3
model/run: xAI Grok API / Collections / Live; observed 2026-08-27
Frozen boundary case; unsupported or interrupted settlement; run 08:32ZxAI Realtime returns partial product evidence, but voice mutation boundary and marginal charge attribution are not returned.BOUNDARY — voice mutation boundary and marginal charge attribution are not returned.Unavailable — voice mutation boundary and marginal charge attribution are not returned

Formula / scoring rule: Voice mutation = session/event/response IDs + accepted/effective configuration + buffered audio/context + duplicate effects + modality usage + recovery + latency + charge. xAI Realtime voice mutation matched ledger; dated first-party evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: xAI Realtime documentationxAI model pricing.

Verified 2026-08-14. Data owner: Luna. Prior-batch and adjacent evidence are not substituted. Run the xai Batch 36 evidence scenario →

Batch 37 · Cross-source search arbitration, collection parser integrity, and live-audio loss

Frozen verification window: 2026-08-27 UTC. Inputs, model/run identity, formulas or rubrics, field-level results, decision boundaries, dated provenance, and exact token bills are server-rendered. Unsupported facts fail closed as Unavailable.

1. Simultaneous Web Search and X Search arbitration ledger

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
Web-only/first-party / batch37-xai-811-r1
batch37-xai-m1-r1
model/run: xAI Grok Search / Collections / Voice; observed 2026-08-27
Frozen small case; request/product IDs, accepted controls, returned usage; run 08:00ZSubmitted/effective controls join to xAI Grok Search; reviewer accepts; returned input/output usage and invoice IDs are present.PASS — the complete identity and settlement tuple is required.$0.016320 = (2840×$3.00 + 520×$15.00)/1M
X-only/long thread / batch37-xai-811-r2
batch37-xai-m1-r2
model/run: xAI Grok Search / Collections / Voice; observed 2026-08-27
Frozen medium case; same product/model, mutation or retry edge; run 08:16Z18/20 field checks accepted; repair, effective state, usage, and final invoice remain linked for xAI Grok Search.PASS WITH REPAIR — only the repaired, explicitly scoped result qualifies.$0.035460 = (6420×$3.00 + 1080×$15.00)/1M
Both/contradictory sources / batch37-xai-811-r3
batch37-xai-m1-r3
model/run: xAI Grok Search / Collections / Voice; observed 2026-08-27
Frozen boundary case; unsupported or interrupted settlement; run 08:32ZxAI Grok Search returns partial evidence, but cross-source arbitration or per-source debit is not returned.BOUNDARY — cross-source arbitration or per-source debit is not returned.Unavailable — cross-source arbitration or per-source debit is not returned

Formula / scoring rule: Search arbitration = effective modes + call order + source/post identity + duplicate collapse + supported claim spans + per-source usage + reviewer acceptance + bill. xAI cross-source search matched ledger; matched evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: xAI search documentationxAI model pricing.

2. Collection parser page, section, and code-boundary integrity canary

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
HTML/PDF headers / batch37-xai-821-r1
batch37-xai-m2-r1
model/run: xAI Grok Search / Collections / Voice; observed 2026-08-27
Frozen small case; request/product IDs, accepted controls, returned usage; run 08:00ZSubmitted/effective controls join to xAI Collections; reviewer accepts; returned input/output usage and invoice IDs are present.PASS — the complete identity and settlement tuple is required.$0.016320 = (2840×$3.00 + 520×$15.00)/1M
Markdown/CSV tables / batch37-xai-821-r2
batch37-xai-m2-r2
model/run: xAI Grok Search / Collections / Voice; observed 2026-08-27
Frozen medium case; same product/model, mutation or retry edge; run 08:16Z18/20 field checks accepted; repair, effective state, usage, and final invoice remain linked for xAI Collections.PASS WITH REPAIR — only the repaired, explicitly scoped result qualifies.$0.035460 = (6420×$3.00 + 1080×$15.00)/1M
Repository code/malformed encoding / batch37-xai-821-r3
batch37-xai-m2-r3
model/run: xAI Grok Search / Collections / Voice; observed 2026-08-27
Frozen boundary case; unsupported or interrupted settlement; run 08:32ZxAI Collections returns partial evidence, but parser-specific units or boundary debit is not returned.BOUNDARY — parser-specific units or boundary debit is not returned.Unavailable — parser-specific units or boundary debit is not returned

Formula / scoring rule: Parser integrity = file/content/parser/chunk IDs + extracted span hashes/locations + boundary loss/duplication + searchable state + units + rollback + bill. xAI collection parser matched canary; matched evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: xAI collections documentationxAI model pricing.

3. Realtime voice packet-loss, jitter, resampling, and duplicate-audio ledger

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
Clean/mono stream / batch37-xai-831-r1
batch37-xai-m3-r1
model/run: xAI Grok Search / Collections / Voice; observed 2026-08-27
Frozen small case; request/product IDs, accepted controls, returned usage; run 08:00ZSubmitted/effective controls join to xAI Voice; reviewer accepts; returned input/output usage and invoice IDs are present.PASS — the complete identity and settlement tuple is required.$0.016320 = (2840×$3.00 + 520×$15.00)/1M
1% or 5% loss / batch37-xai-831-r2
batch37-xai-m3-r2
model/run: xAI Grok Search / Collections / Voice; observed 2026-08-27
Frozen medium case; same product/model, mutation or retry edge; run 08:16Z18/20 field checks accepted; repair, effective state, usage, and final invoice remain linked for xAI Voice.PASS WITH REPAIR — only the repaired, explicitly scoped result qualifies.$0.035460 = (6420×$3.00 + 1080×$15.00)/1M
20% loss/reordered audio / batch37-xai-831-r3
batch37-xai-m3-r3
model/run: xAI Grok Search / Collections / Voice; observed 2026-08-27
Frozen boundary case; unsupported or interrupted settlement; run 08:32ZxAI Voice returns partial evidence, but loss concealment and duplicate-audio charge are not returned.BOUNDARY — loss concealment and duplicate-audio charge are not returned.Unavailable — loss concealment and duplicate-audio charge are not returned

Formula / scoring rule: Audio settlement = packet/session IDs + committed/decoded duration + transcript continuity + duplicate effects + modality usage + concealment/reconnect + acceptance + charge. xAI live-audio-loss matched ledger; matched evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: xAI Realtime documentationxAI model pricing.

Verified 2026-08-14. Data owner: Luna. Prior-batch and adjacent evidence are not substituted. Run the xai Batch 37 evidence scenario →

Batch 38 · X identity transitions, web canonicalization, and voice side effects

Frozen verification window: 2026-08-27 UTC. Inputs, model/run identity, formulas or rubrics, field-level results, decision boundaries, dated provenance, and exact token bills are server-rendered. Unsupported facts fail closed as Unavailable.

1. X Search handle filter and account-identity transition ledger

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
Quoted handle / batch38-xai-811-r1
batch38-xai-m1-r1
model/run: xAI Grok Search / Voice; observed 2026-08-27
query `from:OpenAI` plus quoted post ID; account acct_91; run 20:00Zcase-folded handle maps to account ID; 12/12 posts retain author IDs; input 3,220/output 640 tokens; reviewer accepts.PASS — account identity, not display text, is the join key.$0.019260 = (3220×$3.00 + 640×$15.00)/1M
Rename and prior handle / batch38-xai-811-r2
batch38-xai-m1-r2
model/run: xAI Grok Search / Voice; observed 2026-08-27
handle rename event between query and fetch, five result pages; repair run 20:16Zold/new handle relation disclosed; 18/20 result fields accepted; deleted alias excluded; input 6,380/output 1,080 tokens.PASS WITH REPAIR — historical identity is not attributed beyond the event record.$0.035340 = (6380×$3.00 + 1080×$15.00)/1M
Suspended protected repost / batch38-xai-811-r3
batch38-xai-m1-r3
model/run: xAI Grok Search / Voice; observed 2026-08-27
suspended and protected accounts, deleted post, repost expansion; run 20:32Zresult IDs are returned, but identity-transition filtering and handle-specific charge are not returned.UNAVAILABLE — identity-transition filtering or handle-specific charge is not returned.Unavailable — identity-transition filtering or handle-specific charge is not returned

Formula / scoring rule: Identity filter = submitted/effective handle controls + account/post/event IDs + creation/rename state + included results + claim support + usage + reviewer decision + latency + charge. xAI identity-transition matched ledger; matched evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: xAI X Search documentationxAI model pricing.

2. Web Search redirect, canonical-collapse, paywall, and duplicate-source settlement canary

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
Redirect tracking variants / batch38-xai-821-r1
batch38-xai-m2-r1
model/run: xAI Grok Search / Voice; observed 2026-08-27
query q_101, three redirects, UTM variants collapsed; 8 cited spans; run 21:00Zfinal URLs normalize to 4 unique sources; 8/8 claims supported; input 3,440/output 700 tokens; reviewer accepts.PASS — canonical collapse is separately visible from claim acceptance.$0.020820 = (3440×$3.00 + 700×$15.00)/1M
Syndicated disagreement / batch38-xai-821-r2
batch38-xai-m2-r2
model/run: xAI Grok Search / Voice; observed 2026-08-27
five syndicated pages with conflicting canonical headers; one retry mutation; run 21:16Zthree unique sources retained; two duplicate copies disclosed; 19/22 fields accepted; input 6,620/output 1,120 tokens.PASS WITH REPAIR — freshness disagreement remains in the result.$0.036660 = (6620×$3.00 + 1120×$15.00)/1M
Soft-404 and paywall / batch38-xai-821-r3
batch38-xai-m2-r3
model/run: xAI Grok Search / Voice; observed 2026-08-27
soft-404, paywall, mixed freshness, and one blocked source; run 21:32Zsearch results exist, but canonical-collapse and duplicate-source settlement are not returned.UNAVAILABLE — canonical-collapse or duplicate-source settlement is not returned.Unavailable — canonical-collapse or duplicate-source settlement is not returned

Formula / scoring rule: Canonical settlement = query/call/result/final-URL IDs + unique sources + supported claims + citation spans + search/model units + retry mutation + acceptance + marginal bill. xAI Web Search canonical matched canary; matched evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: xAI Web Search documentationxAI model pricing.

3. Realtime voice function-call cancellation, barge-in, and exactly-once side-effect ledger

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
Call cancelled before args / batch38-xai-831-r1
batch38-xai-m3-r1
model/run: xAI Grok Search / Voice; observed 2026-08-27
session s_201, one function call, barge-in before arguments; run 22:00Zcancel event precedes execution; no side effect; transcript commit joins; input 2,980/output 540 tokens; reviewer accepts.PASS — cancellation is distinct from a completed call.$0.017040 = (2980×$3.00 + 540×$15.00)/1M
Barge-in during execution / batch38-xai-831-r2
batch38-xai-m3-r2
model/run: xAI Grok Search / Voice; observed 2026-08-27
five calls, two barge-ins, one compensation action; reconnect repair run 22:16Z4/5 calls exactly once; one cancelled call compensated; 21/24 event fields accepted; input 6,180/output 1,060 tokens.PASS WITH REPAIR — only four effects enter the accepted ledger.$0.034440 = (6180×$3.00 + 1060×$15.00)/1M
Reconnect after result / batch38-xai-831-r3
batch38-xai-m3-r3
model/run: xAI Grok Search / Voice; observed 2026-08-27
reconnect after result event, duplicated audio commit, five side-effect candidates; run 22:32Zsession events remain, but exactly-once voice side-effect settlement is not returned.UNAVAILABLE — exactly-once voice side-effect settlement is not returned.Unavailable — exactly-once voice side-effect settlement is not returned

Formula / scoring rule: Voice side effects = session/response/call/result/event IDs + audio/transcript commit + cancellation acknowledgement + executed effects + modality/model usage + compensation + acceptance + invoice. xAI Realtime voice side-effect matched ledger; matched evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: xAI Realtime voice documentationxAI model pricing.

Verified 2026-08-14. Data owner: Luna. Prior-batch and adjacent evidence are not substituted. Run the xai Batch 38 evidence scenario →

Current models
5
Legacy models
2
Price range /M
$1.56–$3.00
Max context
1M
Median tok/s
98
Next retirement

Grok API model pricing

xAI’s table is a capability-versus-speed choice: use the Grok 4.20 reasoning variant for harder analysis, the non-reasoning variant for faster multimodal responses, or Grok 4.3 when its lower token rates fit the task. Blended prices below normalize the rows; xAI bills actual input and output tokens.

ModelInput /MOutput /MGrok blend /M*
Grok 4.3$1.25$2.50$1.56
Grok-4.20 Reasoning$2.00$6.00$3.00
Grok-4.20$2.00$6.00$3.00
Grok 4.6$2.00$6.00$3.00
Grok 4.5$2.00$6.00$3.00

Pricing values are registry-backed and were most recently verified on 2026-08-14. Sources: https://docs.x.ai/docs/overview#pricing, https://docs.x.ai/docs/models, https://docs.x.ai/developers/models/grok-4.5. Model detail pages preserve each model's own title and verification date.

* Blended comparison assumes 3 input tokens for every output token; it is not the provider's billing unit.

2 legacy xAI models
Grok-3 Mini$0.26/M blended
Grok-3$2.50/M blended

Speed

Fastest measured xAI model is Grok-4.20 at 104 tokens/sec (290ms TTFT), median across measured xAI models is 98 tokens/sec. See the full speed benchmark methodology.

Best for

Muse Spark 1.3 Contributor is our pick for CodingMuse Spark 1.3 Contributor is our pick for Structured Data ExtractionMuse Spark 1.3 Contributor is our pick for Writing & ContentGLM-5.2 is our pick for Math & ReasoningMuse Spark 1.3 Contributor is our pick for Agents & Tool Use
What it will cost →
xAI's 7 priced models, ranked by verbosity-adjusted monthly cost, not list rate.

Related xAI pages

xAI alternatives →All LLM API pricing →xAI speed benchmarks →xAI cost calculator →

Build with Grok

xAI rate limits →

xAI implementation details

Verified 2026-08-14 against source.

xAI’s low migration friction comes from its OpenAI-compatible base URL and bearer-key authentication. Plan for model-specific limits and confirm the current account tier before launch; xAI does not publish the same SLA and prompt-caching commitments represented for some other providers in this dataset.

OpenAI-compatibleYes
API base URLhttps://api.x.ai/v1
Auth modelBearer API key
Prompt cachingNot documented
Batch discountNot documented
Free tierFree starting credits for new accounts
Free-tier limitsPromotional credits are limited to eligible new accounts; amount and expiry vary by account.
Free-tier expiryNot published
Rate-limit modelPer-model rate limits scaled by account tier
Data residencyNot documented
Trains on API dataNot documented
SLA publishedNo
DocsOfficial pricingStatus pageFree-tier terms

Lifecycle

xAI has 2 legacy models still routable. Full dates and successors on the model deprecation tracker.

Switching to and from xAI

The closest parity-aware alternative to Grok-4.20 Reasoning ($3.00/M) outside xAI is GLM-5.2 ($2.15/M, -28.3%) — a config migration. Biggest gap: no vision input.
The closest parity-aware alternative to Grok-4.20 ($3.00/M) outside xAI is Gemini 3.7 Flash ($1.50/M, -50%) — a code-change migration.
Full xAI alternatives comparison →

Calling xAI through All AI Ask

curl https://allaiask.com/api/v1/chat \
  -H "Authorization: Bearer $ALLAIASK_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model": "grok-4.3", "messages": [{"role": "user", "content": "Hello"}]}'

FAQ

Is xAI OpenAI-compatible?

Yes — xAI's API base (https://api.x.ai/v1) accepts the OpenAI SDK request/response shape, so existing OpenAI client code works with a base-URL and key swap.

Does xAI support prompt caching?

Not documented as of 2026-08-14 — we did not find a published prompt-caching feature for xAI. If that changes, this page updates.

Does xAI have a free tier?

Yes — Free starting credits for new accounts. Promotional credits are limited to eligible new accounts; amount and expiry vary by account.

How much does the xAI API cost?

Current xAI models range from $1.56 to $3.00 per million blended tokens (3:1 input:output). Full per-model pricing is in the table below.

Where is xAI API data hosted?

Not documented as of 2026-08-14 — no published data-residency commitment found for xAI.

Try xAI for free

Run real prompts against every current xAI model, and every other provider on this site, in one workspace.

Try It Free