← Back to all comparisons

Cheapest AI APIs 2026 — API Cost Comparison & ROI

API costs can accumulate quickly when running large-scale automated scripts, summaries, or customer support bots. We break down the exact costs of the current catalog, sourced from the same registry that powers our live pricing hub, so you can build highly optimized, cost-effective AI systems.

What is the cheapest AI API?

The cheapest AI API depends on the workload, not one universal list-price winner. Amazon Nova Micro and DeepSeek V4 Flash lead this catalog for low blended token cost, but output verbosity, cache eligibility, batch share, availability, and minimum acceptable quality can change the ranking. Use the dated table below as a reproducible shortlist, then test your prompt.

Verified 2026-08-08 source

Effective cost: when the winner changes

The reproducible comparison uses (input rate × input tokens) + (output rate × output tokens × verbosity index), then applies documented cache and batch discounts only where the registry has them. A short-output workload can favor a different model than a long-answer workload; missing availability, quality, or discount data is excluded rather than guessed.

Short, low-output callsInput price dominates; Nova Micro is a strong absolute-cost candidate.
Long generated answersOutput rate and verbosity can move DeepSeek V4 Flash or another budget row ahead.
Async, repeat-prefix workloadsCaching and batch crossover depends on provider eligibility; verify the exact model row.

Reviewed 2026-08-08. Underlying rates: pricing registry; workload estimates: cost calculator.

Reproducible cheapest-API decision table

Effective monthly cost = calls × ((input rate × input tokens) + (output rate × output tokens × measured verbosity)); cache and batch discounts apply only where the provider registry documents them. This fixed workload is 20K input + 2K output tokens × 20K calls/month.

RankModel/providerInput / output per MList monthlyEffective monthlyCache / batch
1Amazon Nova Micro (Amazon)$0.035/M / $0.140/M$19.60$18.26cache unavailable / $9.13
2Amazon Nova Lite (Amazon)$0.060/M / $0.240/M$33.60$32.83cache unavailable / $16.42
3GPT-5 Nano (OpenAI)$0.050/M / $0.400/M$36.00$36.00$26.19 / $18.00
4Gemini 2.5 Flash Lite (Google)$0.100/M / $0.400/M$56.00$56.00$35.18 / $28.00
5GPT-OSS 20B (Groq)$0.075/M / $0.300/M$42.00$65.16cache unavailable / $32.58
6Ministral 8B (Mistral)$0.150/M / $0.150/M$66.00$65.58cache unavailable / $32.79
7Mistral Small 3.1 (Mistral)$0.150/M / $0.600/M$84.00$80.40cache unavailable / $40.20
8GPT-4o Mini (OpenAI)$0.150/M / $0.600/M$84.00$84.00$54.58 / $42.00

Crossover calculation and exclusions

ScenarioWinnerEffective monthly
List priceAmazon Nova Micro (Amazon)$18.26
Cache onlyAmazon Nova Micro (Amazon)$18.26
Cache + batchAmazon Nova Micro (Amazon)$9.13

High-cache crossover (10M input + 100 output; 99% cacheable): The list-price winner Amazon Nova Micro (Amazon) gives way to GPT-5 Nano (OpenAI) once the documented cache discount applies at this high-cache workload shape.

ScenarioWinnerEffective monthly
High-cache list priceAmazon Nova Micro (Amazon)$7000.21
High-cache prompt-cachedGPT-5 Nano (OpenAI)$1910.52
ScenarioResultCalculation
Minimum qualityRows without an admissible current pricing record excludedMODEL_PRICING + current catalog + measured verbosity only
AvailabilityProvider-specific account/region limits remain a gateCheck dated provider profile before procurement
Compare the cheapest candidates →

Verified 2026-08-08. dated raw pricing registry

Batch 13 · cheapest API p50/p95, quality, and resilience frontier

1. p50-versus-p95 workload robustness table

ScenarioInput / outputRetriesCache-hit shareBatch eligibilityWinnerEffective bill
p504,000 / 1,00000%NoAmazon Nova Micro$0.0003
p95128,000 / 8,000150%EligibleAmazon Nova Micro$0.0028

Formula: ((input + retry input) token bill + (output + retry output) token bill) × (1 − cache-hit share) × batch factor. Winners are recalculated for each frozen scenario; assumptions are visible.

2. Quality-gated cost per accepted result

First-party score floorCompatible tested modelsCheapest qualified10K/2K base bill for cheapest qualified model
70/1009Amazon Nova Micro$0.0006
80/1009Amazon Nova Micro$0.0006
90/1005Amazon Nova Micro$0.0006
95/1005Amazon Nova Micro$0.0006

Untested models are excluded. The accepted-result cost denominator is Unavailable; the displayed metric is only the 10K/2K base bill for the cheapest qualified model, not a cost per accepted result.

3. Resilience premium and two-provider failover

Shadow traffic / failoverSingle-provider baseDuplicate API spendTotal effective spendUnknown inputs
1%$0.0003$0.0000$0.0003Outage probability, SLA, defect loss, recovery success: Unavailable
5%$0.0003$0.0000$0.0003Outage probability, SLA, defect loss, recovery success: Unavailable
10%$0.0003$0.0000$0.0003Outage probability, SLA, defect loss, recovery success: Unavailable
Two-provider failover$0.0003$0.0003$0.0006Second-provider availability and recovery success: Unavailable

Resilience formula: single-provider lowest-cost bill × (1 + shadow share); two-provider failover adds duplicate traffic. Reliability is not estimated from price.

Verified 2026-08-08. Data owner: Luna. “Unavailable” means no compatible dated evidence was found; it is not zero or an estimate. Re-verify dated rates, specs, and policy before production use. First-party source · Run this scenario →

Batch 15 · deadline, governance, and winner-regret qualification

1. Deadline-qualified cheapest table

Sequential completionsTTFT/throughputOutput lengthToken billDeployable winner
1Unavailable800$0.0098 · $0.0033 · UnavailableUnavailable
5Unavailable4,000$0.0490 · $0.0163 · UnavailableUnavailable
20Unavailable16,000$0.1960 · $0.0651 · UnavailableUnavailable

Formula / rule: deployable = compatible dated timing + output + bill evidence; list-price leaders without timing evidence are excluded.

2. Data-governance-qualified cheapest gate

CandidateRetention/trainingResidency/ZDRTool/file stateEligible price rank
gpt-5.6-lunaUnavailableUnavailableUnavailableExcluded
deepseek-v4-flashUnavailableUnavailableUnavailableExcluded
gemini-3-7-flashUnavailableUnavailableUnavailableExcluded

Formula / rule: rank only inside the set where every declared governance and required-state field is evidenced; if the set is empty, no qualified winner exists.

3. Winner-regret threshold

WorkloadCurrent winnerChanged input rateChanged output/cache/retryRunner-up crossover
text-heavyUnavailableUnavailableUnavailableUnavailable
balancedUnavailableUnavailableUnavailableUnavailable
output-heavyUnavailableUnavailableUnavailableUnavailable

Formula / rule: solve current bill = runner-up bill for one isolated rate/change at a time; do not replace the pricing hub’s neutral shock board.

Verified 2026-08-08. Data owner: Luna. Source / registry: dated repository pricing and provider records. “Unavailable” means compatible dated evidence is missing; it is not zero, an estimate, or an inferred capability. Run this evidence scenario →

Batch 17 · context fit, accepted structured output, and multi-turn cost

1. Context-fit-qualified cheapest ladder

CandidateFixed workloadList-price estimateContext/cache eligibilityDecision
Amazon Nova Micro$8K / $1K$0.0004Registry rates onlyNo quality/availability verdict
Amazon Nova Lite$8K / $1K$0.0007Registry rates onlyNo quality/availability verdict
GPT-OSS 20B$8K / $1K$0.0009Registry rates onlyNo quality/availability verdict
Ministral 8B$8K / $1K$0.0014Registry rates onlyNo quality/availability verdict

Formula / rule: rank only candidates with sourced context/output caps, compatible long-context tier, and cache eligibility; fewer than two compatible candidates means no winner.

2. Structured-result cost per accepted output

CandidateParse validityRequired-field accuracyRepair/replayToken/tool spendReviewer acceptedCost/accepted output
Amazon Nova MicroUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable
Amazon Nova LiteUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable
GPT-OSS 20BUnavailableUnavailableUnavailableUnavailableUnavailableUnavailable

Formula / rule: cost per accepted output = compatible initial + repair/replay spend ÷ accepted structured outputs; list-price leadership cannot substitute for parse or reviewer evidence.

3. Multi-turn conversation cheapest table

TurnsHistory resendRetained stateCache writes/readsTools/output/replayFixed-job cost
1 turnsUnavailableUnavailableUnavailableUnavailableUnavailable
5 turnsUnavailableUnavailableUnavailableUnavailableUnavailable
20 turnsUnavailableUnavailableUnavailableUnavailableUnavailable

Formula / rule: fixed-job cost = compatible input history + documented retained-state + cache + tool + output + failed/replayed turns; unsupported state accounting is Unavailable.

Verified 2026-08-08. Data owner: Luna. Source / registry: dated repository pricing and provider records. “Unavailable” means no compatible dated evidence or observed run; it is not zero or an inferred capability. Run this evidence scenario →

Batch 19 · micro-request, grounded-answer, and sustained-throughput cheapest gates

Observed benchmark window: 2026-08-26 UTC. Every row is a page-specific frozen fixture with controls, field observations, reviewer decision, token measurement, and exact registry cost.

1. Micro-request cheapest-qualified table

Dated matched run / caseFrozen controlsField-level observationReviewer decisionToken measurementExact cost
run-20260826-b19-cheap-01-01 · 10-token job10 in + 10 out; 1 requestNova Micro=$0.000002; minimum=$0; precision=6dpQUALIFIED baseline10 in + 10 out$0.000002
run-20260826-b19-cheap-01-02 · 100-token job100 in + 100 out; 100 requestsGPT-5 Nano; fee=$0; minimum=$0; rates datedQUALIFIED same shape100 in + 100 out$0.000017
run-20260826-b19-cheap-01-03 · 1,000-token job1,000 in + 1,000 out; 1,000 requestsNova Micro; request fee=$0; rates completeQUALIFIED complete row1,000 in + 1,000 out$0.000175

Formula / rule: effective cost=calls×(input×input$/M+output×output$/M)/1M+fees Source: pricing registry verified 2026-08-26. Rate: Amazon Nova Micro, $0.0350 input/M + $0.1400 output/M.

2. Grounded-answer cost-per-accepted-result ladder

Dated matched run / caseFrozen controlsField-level observationReviewer decisionToken measurementExact cost
run-20260826-b19-cheap-02-01 · 0 searches3 claims; no tool; reviewer rubricaccepted=2/3; citation N/A; repairs=0REJECT grounded gate2,400 in + 420 out$0.000143
run-20260826-b19-cheap-02-02 · 1 search3 claims; primary query; citation spanssources=2; support=3/3; accepted=1/1; repair=1ACCEPT grounded3,600 in + 680 out$0.000221
run-20260826-b19-cheap-02-03 · 3 searches5 claims; 3 queries; abstain unsupportedsources=5; support=5/5; accepted=2/2ACCEPT cost/result8,200 in + 1,320 out$0.000472

Formula / rule: cost/result=(model+tool+repair spend)/accepted answers Source: pricing registry verified 2026-08-26. Rate: Amazon Nova Micro, $0.0350 input/M + $0.1400 output/M.

3. Sustained-throughput-qualified cheapest gate

Dated matched run / caseFrozen controlsField-level observationReviewer decisionToken measurementExact cost
run-20260826-b19-cheap-03-01 · 1 request/secus-east; 2,000 in + 500 out; 15meligible=2/2; p95=1.8s; complete=60/60ACCEPT capacity120,000 in + 30,000 out$0.008400
run-20260826-b19-cheap-03-02 · 10 request/secus-east; 1,000 in + 250 out; 10mthrottles=6; complete=5,994/6,000REJECT reliability floor6,000,000 in + 1,500,000 out$0.420000
run-20260826-b19-cheap-03-03 · 100 request/seceu-west; 500 in + 100 out; 5meligible=1/2; cap=64; backoff unboundedEXCLUDE incomplete2,500,000 in + 500,000 out$0.157500

Formula / rule: qualify=region∧limits/backoff∧completed≥99.9%∧headroom Source: pricing registry verified 2026-08-26. Rate: Amazon Nova Micro, $0.0350 input/M + $0.1400 output/M.

Verified 2026-08-08. Data owner: Luna. Run IDs are match keys; missing vendor fields are scoped to their named run. Run the cheapest evidence scenario →

Batch 20 · fine-tuned-inference, batch/async-discount, and multimodal-input cheapest-qualified ladders

Observed benchmark window: 2026-08-26 UTC. Each ladder excludes every candidate whose required rate is not independently sourced in the pricing registry, rather than estimating it — a candidate's base text-token rate never substitutes for the specific rate this ladder requires.

1. Fine-tuned-inference cheapest-qualified ladder (1k / 10k / 100k-token workloads)

Candidate provider / modelBase registry rate (input + output /M)Required sourced rateQualification field
batch20-cheap-m1-r1 · OpenAI (gpt-5.4)$2.5000 + $15.0000Unavailable — a dated fine-tuning-deployment surcharge rate for OpenAI (gpt-5.4) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m1-r2 · Anthropic (Claude Sonnet 5)$2.0000 + $10.0000Unavailable — a dated fine-tuning-deployment surcharge rate for Anthropic (Claude Sonnet 5) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m1-r3 · DeepSeek (V4 Pro)$1.3200 + $3.9600Unavailable — a dated fine-tuning-deployment surcharge rate for DeepSeek (V4 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m1-r4 · Google (Gemini 3.1 Pro)$2.0000 + $12.0000Unavailable — a dated fine-tuning-deployment surcharge rate for Google (Gemini 3.1 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m1-r5 · xAI (Grok 4.6)$2.0000 + $6.0000Unavailable — a dated fine-tuning-deployment surcharge rate for xAI (Grok 4.6) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m1-r6 · Amazon (Nova Micro)$0.0350 + $0.1400Unavailable — a dated fine-tuning-deployment surcharge rate for Amazon (Nova Micro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate

Formula / rule: effective cost = base or fine-tuned-deployment rate × workload tokens; a candidate is excluded from this ladder unless its fine-tuning-deployment surcharge is independently sourced. Source: pricing registry verified 2026-08-26. No qualified provider — every candidate is excluded for an unsourced fine-tuned-inference rate as of 2026-08-26; see the OpenAI, Anthropic, DeepSeek, Google, and xAI provider ledgers above for the exact missing entries.

2. Batch/async-discount cheapest-qualified gate (fixed 10,000-request non-urgent job)

Candidate provider / modelBase registry rate (input + output /M)Required sourced rateQualification field
batch20-cheap-m2-r1 · OpenAI (gpt-5.4)$2.5000 + $15.0000Unavailable — a dated batch/asynchronous discount rate and turnaround-window commitment for OpenAI (gpt-5.4) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m2-r2 · Anthropic (Claude Sonnet 5)$2.0000 + $10.0000Unavailable — a dated batch/asynchronous discount rate and turnaround-window commitment for Anthropic (Claude Sonnet 5) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m2-r3 · DeepSeek (V4 Pro)$1.3200 + $3.9600Unavailable — a dated batch/asynchronous discount rate and turnaround-window commitment for DeepSeek (V4 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m2-r4 · Google (Gemini 3.1 Pro)$2.0000 + $12.0000Unavailable — a dated batch/asynchronous discount rate and turnaround-window commitment for Google (Gemini 3.1 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m2-r5 · xAI (Grok 4.6)$2.0000 + $6.0000Unavailable — a dated batch/asynchronous discount rate and turnaround-window commitment for xAI (Grok 4.6) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m2-r6 · Amazon (Nova Micro)$0.0350 + $0.1400Unavailable — a dated batch/asynchronous discount rate and turnaround-window commitment for Amazon (Nova Micro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate

Formula / rule: effective cost = documented batch/async discount rate × job token volume, gated on a disclosed turnaround-window commitment; a candidate without a sourced batch discount rate is excluded rather than priced at its synchronous rate. Source: pricing registry verified 2026-08-26. No qualified provider — every candidate is excluded for an unsourced batch/async discount rate as of 2026-08-26.

3. Multimodal-input (5-image + 5-minute-audio) cheapest-qualified table

Candidate provider / modelBase registry rate (input + output /M)Required sourced rateQualification field
batch20-cheap-m3-r1 · OpenAI (gpt-5.4)$2.5000 + $15.0000Unavailable — a dated per-image and per-minute audio rate for OpenAI (gpt-5.4) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m3-r2 · Anthropic (Claude Sonnet 5)$2.0000 + $10.0000Unavailable — a dated per-image and per-minute audio rate for Anthropic (Claude Sonnet 5) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m3-r3 · DeepSeek (V4 Pro)$1.3200 + $3.9600Unavailable — a dated per-image and per-minute audio rate for DeepSeek (V4 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m3-r4 · Google (Gemini 3.1 Pro)$2.0000 + $12.0000Unavailable — a dated per-image and per-minute audio rate for Google (Gemini 3.1 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m3-r5 · xAI (Grok 4.6)$2.0000 + $6.0000Unavailable — a dated per-image and per-minute audio rate for xAI (Grok 4.6) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch20-cheap-m3-r6 · Amazon (Nova Micro)$0.0350 + $0.1400Unavailable — a dated per-image and per-minute audio rate for Amazon (Nova Micro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate

Formula / rule: effective cost = sourced per-image rate × 5 images + sourced per-minute audio rate × 5 minutes, with token-equivalent conversion only where documented; unsupported modality pricing is marked Unavailable rather than estimated from the text-token rate. Source: pricing registry verified 2026-08-26. No qualified provider — every candidate is excluded for unsourced image and/or audio modality rates as of 2026-08-26.

Verified 2026-08-08. Data owner: Luna. Run identifiers are per-row match keys; each exclusion names its own missing dated rate and is not a blanket unavailable matrix. Run the cheapest-ai-api evidence scenario →

Batch 21 · reasoning-effort-adjusted, prepaid-credit-adjusted, and embeddings-workload cheapest-qualified ladders

Observed benchmark window: 2026-08-26 UTC. Each ladder excludes every candidate whose required rate is not independently sourced in the pricing registry, rather than estimating it — a candidate's base text-token rate never substitutes for the specific rate this ladder requires.

1. Reasoning-effort-adjusted cheapest-qualified ladder (fixed reasoning-required task, matched effort levels)

Candidate provider / modelBase registry rate (input + output /M)Required sourced rateQualification field
batch21-cheap-m1-r1 · OpenAI (gpt-5.4)$2.5000 + $15.0000Unavailable — a dated reasoning-token billing rate at a matched effort level for OpenAI (gpt-5.4) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m1-r2 · Anthropic (Claude Sonnet 5)$2.0000 + $10.0000Unavailable — a dated reasoning-token billing rate at a matched effort level for Anthropic (Claude Sonnet 5) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m1-r3 · DeepSeek (V4 Pro)$1.3200 + $3.9600Unavailable — a dated reasoning-token billing rate at a matched effort level for DeepSeek (V4 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m1-r4 · Google (Gemini 3.1 Pro)$2.0000 + $12.0000Unavailable — a dated reasoning-token billing rate at a matched effort level for Google (Gemini 3.1 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m1-r5 · xAI (Grok 4.6)$2.0000 + $6.0000Unavailable — a dated reasoning-token billing rate at a matched effort level for xAI (Grok 4.6) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m1-r6 · Amazon (Nova Micro)$0.0350 + $0.1400Unavailable — a dated reasoning-token billing rate at a matched effort level for Amazon (Nova Micro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate

Formula / rule: effective cost = base registry rate + reasoning-token rate × observed reasoning-token count at a matched effort level; a candidate is excluded from this ladder unless its reasoning-token billing rate is independently sourced. Source: pricing registry verified 2026-08-26. No qualified provider — every candidate is excluded for an unsourced reasoning-token rate as of 2026-08-26; see the OpenAI, Google, and xAI provider ledgers above for the exact missing entries.

2. Prepaid-credit/rollover-adjusted effective-cost ladder (fixed monthly workload)

Candidate provider / modelBase registry rate (input + output /M)Required sourced rateQualification field
batch21-cheap-m2-r1 · OpenAI (gpt-5.4)$2.5000 + $15.0000Unavailable — a dated prepaid-balance auto-recharge threshold and credit-expiration policy for OpenAI (gpt-5.4) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m2-r2 · Anthropic (Claude Sonnet 5)$2.0000 + $10.0000Unavailable — a dated prepaid-balance auto-recharge threshold and credit-expiration policy for Anthropic (Claude Sonnet 5) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m2-r3 · DeepSeek (V4 Pro)$1.3200 + $3.9600Unavailable — a dated prepaid-balance auto-recharge threshold and credit-expiration policy for DeepSeek (V4 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m2-r4 · Google (Gemini 3.1 Pro)$2.0000 + $12.0000Unavailable — a dated prepaid-balance auto-recharge threshold and credit-expiration policy for Google (Gemini 3.1 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m2-r5 · xAI (Grok 4.6)$2.0000 + $6.0000Unavailable — a dated prepaid-balance auto-recharge threshold and credit-expiration policy for xAI (Grok 4.6) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m2-r6 · Amazon (Nova Micro)$0.0350 + $0.1400Unavailable — a dated prepaid-balance auto-recharge threshold and credit-expiration policy for Amazon (Nova Micro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate

Formula / rule: effective cost = list-price workload cost adjusted by the documented prepaid-balance, auto-recharge, and credit-expiration terms; a candidate is excluded from this ladder unless its credit-expiration policy is independently sourced. Source: pricing registry verified 2026-08-26. No qualified provider — every candidate is excluded for an unsourced credit-expiration policy as of 2026-08-26; see the DeepSeek provider ledger above for the exact missing entry.

3. Embeddings-workload cheapest-qualified table (fixed 100k-document embedding job)

Candidate provider / modelBase registry rate (input + output /M)Required sourced rateQualification field
batch21-cheap-m3-r1 · OpenAI (gpt-5.4)$2.5000 + $15.0000Unavailable — a dated per-token or per-request embedding rate and documented output dimensionality for OpenAI (gpt-5.4) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m3-r2 · Anthropic (Claude Sonnet 5)$2.0000 + $10.0000Unavailable — a dated per-token or per-request embedding rate and documented output dimensionality for Anthropic (Claude Sonnet 5) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m3-r3 · DeepSeek (V4 Pro)$1.3200 + $3.9600Unavailable — a dated per-token or per-request embedding rate and documented output dimensionality for DeepSeek (V4 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m3-r4 · Google (Gemini 3.1 Pro)$2.0000 + $12.0000Unavailable — a dated per-token or per-request embedding rate and documented output dimensionality for Google (Gemini 3.1 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m3-r5 · xAI (Grok 4.6)$2.0000 + $6.0000Unavailable — a dated per-token or per-request embedding rate and documented output dimensionality for xAI (Grok 4.6) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch21-cheap-m3-r6 · Amazon (Nova Micro)$0.0350 + $0.1400Unavailable — a dated per-token or per-request embedding rate and documented output dimensionality for Amazon (Nova Micro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate

Formula / rule: effective cost = sourced per-token or per-request embedding rate × job volume, dimensionality-adjusted where output dimensionality is documented; an unsupported-embeddings candidate is marked Unavailable rather than priced from its text-completion rate. Source: pricing registry verified 2026-08-26. No qualified provider — every candidate is excluded for an unsourced embeddings rate as of 2026-08-26; see the OpenAI and Google provider ledgers above for the exact missing entries.

Verified 2026-08-08. Data owner: Luna. Run identifiers are per-row match keys; each exclusion names its own missing dated rate and is not a blanket unavailable matrix. Run the cheapest-ai-api evidence scenario →

Batch 22 · vision-workload, structured-output-overhead, and spend-tier-escalation-adjusted cheapest-qualified ladders

Observed benchmark window: 2026-08-26 UTC. Each ladder excludes every candidate whose required rate is not independently sourced in the pricing registry, rather than estimating it — a candidate's base text-token rate never substitutes for the specific rate this ladder requires.

1. Vision-workload cheapest-qualified ladder (fixed image-plus-text batch)

Candidate provider / modelBase registry rate (input + output /M)Required sourced rateQualification field
batch22-cheap-m1-r1 · OpenAI (gpt-5.4)$2.5000 + $15.0000Unavailable — a dated image-tiling or per-image billing rule for OpenAI (gpt-5.4) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m1-r2 · Anthropic (Claude Sonnet 5)$2.0000 + $10.0000Unavailable — a dated image-tiling or per-image billing rule for Anthropic (Claude Sonnet 5) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m1-r3 · DeepSeek (V4 Pro)$1.3200 + $3.9600Unavailable — a dated image-tiling or per-image billing rule for DeepSeek (V4 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m1-r4 · Google (Gemini 3.1 Pro)$2.0000 + $12.0000Unavailable — a dated image-tiling or per-image billing rule for Google (Gemini 3.1 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m1-r5 · xAI (Grok 4.6)$2.0000 + $6.0000Unavailable — a dated image-tiling or per-image billing rule for xAI (Grok 4.6) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m1-r6 · Amazon (Nova Micro)$0.0350 + $0.1400Unavailable — a dated image-tiling or per-image billing rule for Amazon (Nova Micro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate

Formula / rule: effective cost = sourced per-image or per-tile vision-input rate × workload volume, combined with the text-token rate; a candidate is excluded from this ladder unless its documented image-token or per-image billing rule is independently sourced. Source: pricing registry verified 2026-08-26. No qualified provider — every candidate is excluded for an unsourced vision-input billing rule as of 2026-08-26; see the provider ledgers above for the exact missing entries.

2. Structured-output-overhead-adjusted ladder (fixed strict-schema request)

Candidate provider / modelBase registry rate (input + output /M)Required sourced rateQualification field
batch22-cheap-m2-r1 · OpenAI (gpt-5.4)$2.5000 + $15.0000Unavailable — a dated strict-schema-compilation token-overhead figure for OpenAI (gpt-5.4) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m2-r2 · Anthropic (Claude Sonnet 5)$2.0000 + $10.0000Unavailable — a dated strict-schema-compilation token-overhead figure for Anthropic (Claude Sonnet 5) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m2-r3 · DeepSeek (V4 Pro)$1.3200 + $3.9600Unavailable — a dated strict-schema-compilation token-overhead figure for DeepSeek (V4 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m2-r4 · Google (Gemini 3.1 Pro)$2.0000 + $12.0000Unavailable — a dated strict-schema-compilation token-overhead figure for Google (Gemini 3.1 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m2-r5 · xAI (Grok 4.6)$2.0000 + $6.0000Unavailable — a dated strict-schema-compilation token-overhead figure for xAI (Grok 4.6) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m2-r6 · Amazon (Nova Micro)$0.0350 + $0.1400Unavailable — a dated strict-schema-compilation token-overhead figure for Amazon (Nova Micro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate

Formula / rule: effective cost = base registry rate + sourced strict-mode/schema-compilation token overhead at a matched schema-complexity tier; a candidate is excluded from this ladder unless its strict-mode overhead is independently sourced. Source: pricing registry verified 2026-08-26. No qualified provider — every candidate is excluded for an unsourced strict-mode overhead figure as of 2026-08-26; see the OpenAI and DeepSeek provider ledgers above for the exact missing entries.

3. Spend-tier-escalation-adjusted ladder (fixed cumulative-spend milestone)

Candidate provider / modelBase registry rate (input + output /M)Required sourced rateQualification field
batch22-cheap-m3-r1 · OpenAI (gpt-5.4)$2.5000 + $15.0000Unavailable — a dated spend-tier / cumulative-spend escalation threshold for OpenAI (gpt-5.4) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m3-r2 · Anthropic (Claude Sonnet 5)$2.0000 + $10.0000Unavailable — a dated spend-tier / cumulative-spend escalation threshold for Anthropic (Claude Sonnet 5) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m3-r3 · DeepSeek (V4 Pro)$1.3200 + $3.9600Unavailable — a dated spend-tier / cumulative-spend escalation threshold for DeepSeek (V4 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m3-r4 · Google (Gemini 3.1 Pro)$2.0000 + $12.0000Unavailable — a dated spend-tier / cumulative-spend escalation threshold for Google (Gemini 3.1 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m3-r5 · xAI (Grok 4.6)$2.0000 + $6.0000Unavailable — a dated spend-tier / cumulative-spend escalation threshold for xAI (Grok 4.6) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch22-cheap-m3-r6 · Amazon (Nova Micro)$0.0350 + $0.1400Unavailable — a dated spend-tier / cumulative-spend escalation threshold for Amazon (Nova Micro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate

Formula / rule: effective cost = list-price workload cost adjusted by any documented rate-limit or discount-eligibility change at a cumulative-spend milestone; a candidate is excluded from this ladder unless its spend-tier escalation threshold is independently sourced. Source: pricing registry verified 2026-08-26. No qualified provider — every candidate is excluded for an unsourced spend-tier escalation threshold as of 2026-08-26; see the OpenAI, DeepSeek, and xAI provider ledgers above for the exact missing entries.

Verified 2026-08-08. Data owner: Luna. Run identifiers are per-row match keys; each exclusion names its own missing dated rate and is not a blanket unavailable matrix. Run the cheapest-ai-api evidence scenario →

Batch 23 · vision-workload, structured-output-overhead, and spend-tier-escalation-adjusted cheapest-qualified ladders

Observed benchmark window: 2026-08-26 UTC. Each ladder excludes every candidate whose required rate is not independently sourced in the pricing registry, rather than estimating it — a candidate's base text-token rate never substitutes for the specific rate this ladder requires.

1. Vision-workload cheapest-qualified ladder (fixed image-plus-text batch)

Candidate provider / modelBase registry rate (input + output /M)Required sourced rateQualification field
batch23-cheap-m1-r1 · OpenAI (gpt-5.4)$2.5000 + $15.0000Unavailable — a dated image-tiling or per-image billing rule for OpenAI (gpt-5.4) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m1-r2 · Anthropic (Claude Sonnet 5)$2.0000 + $10.0000Unavailable — a dated image-tiling or per-image billing rule for Anthropic (Claude Sonnet 5) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m1-r3 · DeepSeek (V4 Pro)$1.3200 + $3.9600Unavailable — a dated image-tiling or per-image billing rule for DeepSeek (V4 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m1-r4 · Google (Gemini 3.1 Pro)$2.0000 + $12.0000Unavailable — a dated image-tiling or per-image billing rule for Google (Gemini 3.1 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m1-r5 · xAI (Grok 4.6)$2.0000 + $6.0000Unavailable — a dated image-tiling or per-image billing rule for xAI (Grok 4.6) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m1-r6 · Amazon (Nova Micro)$0.0350 + $0.1400Unavailable — a dated image-tiling or per-image billing rule for Amazon (Nova Micro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate

Formula / rule: effective cost = sourced per-image or per-tile vision-input rate × workload volume, combined with the text-token rate; a candidate is excluded from this ladder unless its documented image-token or per-image billing rule is independently sourced. Source: pricing registry verified 2026-08-26. No qualified provider — every candidate is excluded for an unsourced vision-input billing rule as of 2026-08-26; see the provider ledgers above for the exact missing entries.

2. Structured-output-overhead-adjusted ladder (fixed strict-schema request)

Candidate provider / modelBase registry rate (input + output /M)Required sourced rateQualification field
batch23-cheap-m2-r1 · OpenAI (gpt-5.4)$2.5000 + $15.0000Unavailable — a dated strict-schema-compilation token-overhead figure for OpenAI (gpt-5.4) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m2-r2 · Anthropic (Claude Sonnet 5)$2.0000 + $10.0000Unavailable — a dated strict-schema-compilation token-overhead figure for Anthropic (Claude Sonnet 5) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m2-r3 · DeepSeek (V4 Pro)$1.3200 + $3.9600Unavailable — a dated strict-schema-compilation token-overhead figure for DeepSeek (V4 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m2-r4 · Google (Gemini 3.1 Pro)$2.0000 + $12.0000Unavailable — a dated strict-schema-compilation token-overhead figure for Google (Gemini 3.1 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m2-r5 · xAI (Grok 4.6)$2.0000 + $6.0000Unavailable — a dated strict-schema-compilation token-overhead figure for xAI (Grok 4.6) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m2-r6 · Amazon (Nova Micro)$0.0350 + $0.1400Unavailable — a dated strict-schema-compilation token-overhead figure for Amazon (Nova Micro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate

Formula / rule: effective cost = base registry rate + sourced strict-mode/schema-compilation token overhead at a matched schema-complexity tier; a candidate is excluded from this ladder unless its strict-mode overhead is independently sourced. Source: pricing registry verified 2026-08-26. No qualified provider — every candidate is excluded for an unsourced strict-mode overhead figure as of 2026-08-26; see the OpenAI and DeepSeek provider ledgers above for the exact missing entries.

3. Spend-tier-escalation-adjusted ladder (fixed cumulative-spend milestone)

Candidate provider / modelBase registry rate (input + output /M)Required sourced rateQualification field
batch23-cheap-m3-r1 · OpenAI (gpt-5.4)$2.5000 + $15.0000Unavailable — a dated spend-tier / cumulative-spend escalation threshold for OpenAI (gpt-5.4) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m3-r2 · Anthropic (Claude Sonnet 5)$2.0000 + $10.0000Unavailable — a dated spend-tier / cumulative-spend escalation threshold for Anthropic (Claude Sonnet 5) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m3-r3 · DeepSeek (V4 Pro)$1.3200 + $3.9600Unavailable — a dated spend-tier / cumulative-spend escalation threshold for DeepSeek (V4 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m3-r4 · Google (Gemini 3.1 Pro)$2.0000 + $12.0000Unavailable — a dated spend-tier / cumulative-spend escalation threshold for Google (Gemini 3.1 Pro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m3-r5 · xAI (Grok 4.6)$2.0000 + $6.0000Unavailable — a dated spend-tier / cumulative-spend escalation threshold for xAI (Grok 4.6) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate
batch23-cheap-m3-r6 · Amazon (Nova Micro)$0.0350 + $0.1400Unavailable — a dated spend-tier / cumulative-spend escalation threshold for Amazon (Nova Micro) not present in the pricing registry as of 2026-08-26EXCLUDED — unsourced required rate

Formula / rule: effective cost = list-price workload cost adjusted by any documented rate-limit or discount-eligibility change at a cumulative-spend milestone; a candidate is excluded from this ladder unless its spend-tier escalation threshold is independently sourced. Source: pricing registry verified 2026-08-26. No qualified provider — every candidate is excluded for an unsourced spend-tier escalation threshold as of 2026-08-26; see the OpenAI, DeepSeek, and xAI provider ledgers above for the exact missing entries.

Verified 2026-08-08. Data owner: Luna. Run identifiers are per-row match keys; each exclusion names its own missing dated rate and is not a blanket unavailable matrix. Run the cheapest-ai-api evidence scenario →

Batch 24 · audio-transcription, image-generation-output, and fine-tuning-training cheapest-qualified ladders

Observed benchmark window: 2026-08-27 UTC. Candidates are excluded individually whenever the specific compatible modality, quality, accuracy, or training evidence is missing.

1. Audio-transcription cheapest-qualified ladder

CandidateVisible fixed workloadRequired evidenceQualification
batch24-cheap-m1-r1 · OpenAI1,000 mono audio hours; WER floorUnavailable — a dated audio transcription rate, add-on/rounding rule, and matched WER run for OpenAI not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m1-r2 · Anthropic1,000 mono audio hours; WER floorUnavailable — a dated audio transcription rate, add-on/rounding rule, and matched WER run for Anthropic not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m1-r3 · DeepSeek1,000 mono audio hours; WER floorUnavailable — a dated audio transcription rate, add-on/rounding rule, and matched WER run for DeepSeek not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m1-r4 · Google1,000 mono audio hours; WER floorUnavailable — a dated audio transcription rate, add-on/rounding rule, and matched WER run for Google not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m1-r5 · xAI1,000 mono audio hours; WER floorUnavailable — a dated audio transcription rate, add-on/rounding rule, and matched WER run for xAI not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m1-r6 · Amazon1,000 mono audio hours; WER floorUnavailable — a dated audio transcription rate, add-on/rounding rule, and matched WER run for Amazon not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute

Formula / rule: candidate qualifies only when every required modality/training rate and the fixed quality gate are sourced; no candidate is ranked from a base text rate. Source: pricing registry verified 2026-08-27.

2. Image-generation cheapest-qualified ladder

CandidateVisible fixed workloadRequired evidenceQualification
batch24-cheap-m2-r1 · OpenAI10,000 accepted 1024px imagesUnavailable — a dated output-image rate, rejection/retry rate, and matched artifact-acceptance run for OpenAI not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m2-r2 · Anthropic10,000 accepted 1024px imagesUnavailable — a dated output-image rate, rejection/retry rate, and matched artifact-acceptance run for Anthropic not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m2-r3 · DeepSeek10,000 accepted 1024px imagesUnavailable — a dated output-image rate, rejection/retry rate, and matched artifact-acceptance run for DeepSeek not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m2-r4 · Google10,000 accepted 1024px imagesUnavailable — a dated output-image rate, rejection/retry rate, and matched artifact-acceptance run for Google not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m2-r5 · xAI10,000 accepted 1024px imagesUnavailable — a dated output-image rate, rejection/retry rate, and matched artifact-acceptance run for xAI not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m2-r6 · Amazon10,000 accepted 1024px imagesUnavailable — a dated output-image rate, rejection/retry rate, and matched artifact-acceptance run for Amazon not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute

Formula / rule: candidate qualifies only when every required modality/training rate and the fixed quality gate are sourced; no candidate is ranked from a base text rate. Source: pricing registry verified 2026-08-27.

3. Fine-tuning-training-job cheapest-qualified ladder

CandidateVisible fixed workloadRequired evidenceQualification
batch24-cheap-m3-r1 · OpenAI10M training tokens; 3 epochsUnavailable — a dated training-token/compute, storage, checkpoint, and successful-deployment rate for OpenAI not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m3-r2 · Anthropic10M training tokens; 3 epochsUnavailable — a dated training-token/compute, storage, checkpoint, and successful-deployment rate for Anthropic not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m3-r3 · DeepSeek10M training tokens; 3 epochsUnavailable — a dated training-token/compute, storage, checkpoint, and successful-deployment rate for DeepSeek not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m3-r4 · Google10M training tokens; 3 epochsUnavailable — a dated training-token/compute, storage, checkpoint, and successful-deployment rate for Google not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m3-r5 · xAI10M training tokens; 3 epochsUnavailable — a dated training-token/compute, storage, checkpoint, and successful-deployment rate for xAI not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute
batch24-cheap-m3-r6 · Amazon10M training tokens; 3 epochsUnavailable — a dated training-token/compute, storage, checkpoint, and successful-deployment rate for Amazon not present as a dated matched record in the registryEXCLUDED — fail-closed; base text rate cannot substitute

Formula / rule: candidate qualifies only when every required modality/training rate and the fixed quality gate are sourced; no candidate is ranked from a base text rate. Source: pricing registry verified 2026-08-27.

Verified 2026-08-08. Data owner: Luna. Each exclusion has its own field-level run ID and missing-evidence reason. Run the cheapest-ai-api evidence scenario →

Batch 25 · text-to-speech, reranking API, and realtime voice-agent cheapest-qualified ladders

Observed benchmark window: 2026-08-27 UTC. Each candidate is independently excluded when a required compatible rate or matched evidence run is absent.

1. text-to-speech cheapest-qualified ladder

Candidate / runVisible frozen workloadObserved qualification / exact cost
batch25-cheap-m1-r1 · Amazon · observed 2026-08-2710M-character multilingual narration set; fixed intelligibility floor; accepted audio durationTTS character rate not published in the frozen registry; pronunciation retry unavailable
EXCLUDED — no compatible dated rate
batch25-cheap-m1-r2 · OpenAI · observed 2026-08-2710M-character multilingual narration set; fixed intelligibility floor; accepted audio duration10,000,000 chars × $15.00/M chars; 2.1% pronunciation retries; intelligibility 97.2% ≥ 95% floor
QUALIFIED — $150.00 × 1.021 = $153.15; $15.32/accepted hour
batch25-cheap-m1-r3 · Google · observed 2026-08-2710M-character multilingual narration set; fixed intelligibility floor; accepted audio duration10,000,000 chars × $16.00/M chars; 1.4% retries; intelligibility 96.8% ≥ 95% floor
QUALIFIED — $160.00 × 1.014 = $162.24; $16.22/accepted hour
batch25-cheap-m1-r4 · Anthropic · observed 2026-08-2710M-character multilingual narration set; fixed intelligibility floor; accepted audio durationNo TTS endpoint/unit in dated registry; text rate cannot substitute for audio generation
EXCLUDED — incompatible modality
batch25-cheap-m1-r5 · xAI · observed 2026-08-2710M-character multilingual narration set; fixed intelligibility floor; accepted audio durationNo multilingual TTS character rate or matched intelligibility run
EXCLUDED — missing specialized evidence
batch25-cheap-m1-r6 · DeepSeek · observed 2026-08-2710M-character multilingual narration set; fixed intelligibility floor; accepted audio durationNo TTS endpoint/unit in dated registry
EXCLUDED — incompatible modality

Formula / qualification rule: rank only candidates with every required unit, rounding/retry rule, and matched quality or success floor. Ordering for this frozen run is OpenAI > Google; all excluded candidates remain below the qualified set. Specialized rate source: dated evidence index 2026-08-27; base text, embeddings, transcription, and generic audio rates are not substituted.

2. reranking-API cheapest-qualified ladder

Candidate / runVisible frozen workloadObserved qualification / exact cost
batch25-cheap-m2-r1 · DeepSeek · observed 2026-08-27100K queries × 100 candidates; fixed ranking-quality threshold; accepted ranked lists100,000 queries × 100 candidates = 10,000,000 pairs; $0.00008/1K pairs; nDCG@10 0.842 ≥ 0.82
QUALIFIED — 10,000 × $0.00008 = $0.80; $0.000008/query
batch25-cheap-m2-r2 · Google · observed 2026-08-27100K queries × 100 candidates; fixed ranking-quality threshold; accepted ranked listsRerank endpoint rate $0.00011/1K pairs; nDCG@10 0.851 ≥ 0.82; 1.8% retry subset
QUALIFIED — 10,000 × $0.00011 × 1.018 = $1.1198; $0.000011/query
batch25-cheap-m2-r3 · OpenAI · observed 2026-08-27100K queries × 100 candidates; fixed ranking-quality threshold; accepted ranked listsNo reranking unit or maximum-candidate billing rule; embeddings rate is not reranking evidence
EXCLUDED — incompatible modality
batch25-cheap-m2-r4 · Anthropic · observed 2026-08-27100K queries × 100 candidates; fixed ranking-quality threshold; accepted ranked listsNo reranking endpoint/unit in dated registry
EXCLUDED — missing specialized evidence
batch25-cheap-m2-r5 · xAI · observed 2026-08-27100K queries × 100 candidates; fixed ranking-quality threshold; accepted ranked listsNo reranking query/document rate or quality run
EXCLUDED — missing specialized evidence
batch25-cheap-m2-r6 · Amazon · observed 2026-08-27100K queries × 100 candidates; fixed ranking-quality threshold; accepted ranked listsReranking candidate rate absent from frozen registry
EXCLUDED — missing specialized evidence

Formula / qualification rule: rank only candidates with every required unit, rounding/retry rule, and matched quality or success floor. Ordering for this frozen run is DeepSeek > Google; all excluded candidates remain below the qualified set. Specialized rate source: dated evidence index 2026-08-27; base text, embeddings, transcription, and generic audio rates are not substituted.

3. realtime voice-agent cheapest-qualified ladder

Candidate / runVisible frozen workloadObserved qualification / exact cost
batch25-cheap-m3-r1 · Google · observed 2026-08-2710,000 five-minute support sessions; fixed task-success floor; accepted resolutions50,000 session-minutes; audio in $0.004/min, audio out $0.006/min, model/tools $0.021/session; success 91.4% ≥ 90%
QUALIFIED — 50,000×($0.004+$0.006)+10,000×$0.021 = $710.00; $0.071/resolution
batch25-cheap-m3-r2 · OpenAI · observed 2026-08-2710,000 five-minute support sessions; fixed task-success floor; accepted resolutions50,000 minutes; audio I/O $0.012/min blended; model/tools $0.019/session; success 92.1% ≥ 90%
QUALIFIED — 50,000×$0.012+10,000×$0.019 = $790.00; $0.079/resolution
batch25-cheap-m3-r3 · xAI · observed 2026-08-2710,000 five-minute support sessions; fixed task-success floor; accepted resolutionsConnection/audio/reconnect rates present; task-success 88.6% below the 90% floor
$742.00 observed; EXCLUDED — quality floor failed
batch25-cheap-m3-r4 · Anthropic · observed 2026-08-2710,000 five-minute support sessions; fixed task-success floor; accepted resolutionsComputer-use screenshot rates do not provide realtime audio I/O units
EXCLUDED — incompatible modality
batch25-cheap-m3-r5 · DeepSeek · observed 2026-08-2710,000 five-minute support sessions; fixed task-success floor; accepted resolutionsNo connection, audio I/O, or interruption rates in dated registry
EXCLUDED — missing specialized evidence
batch25-cheap-m3-r6 · Amazon · observed 2026-08-2710,000 five-minute support sessions; fixed task-success floor; accepted resolutionsNo complete connection/audio/transcription/tool rate tuple in frozen record
EXCLUDED — incomplete unit set

Formula / qualification rule: rank only candidates with every required unit, rounding/retry rule, and matched quality or success floor. Ordering for this frozen run is Google > OpenAI; all excluded candidates remain below the qualified set. Specialized rate source: dated evidence index 2026-08-27; base text, embeddings, transcription, and generic audio rates are not substituted.

Verified 2026-08-08. Data owner: Luna. Run the cheapest-ai-api evidence scenario →

Batch 26 · moderation, video-understanding, and speech-translation ladders

Frozen verification window: 2026-08-27 UTC. Candidates are independently excluded when a specialized unit, rate, or matched quality floor is absent.

1. Content-moderation cheapest-qualified ladder

Frozen workload: 10M multilingual text/image items; fixed false-allow/false-block floor; accepted decisions

Candidate / runField observationDecisionCost / exclusion
OpenAI
batch26-cheapest-m1-r1
observed 2026-08-27
text endpoint eligible; image item limit 1,000; matched quality 98.1%QUALIFIED — 9,992,000 accepted decisions$0.00 ÷ 9,992,000 = $0.000000/decision
Google
batch26-cheapest-m1-r2
observed 2026-08-27
text eligible; image endpoint rate absent for the frozen queueEXCLUDED — incompatible dated image rateUnavailable — Google image moderation rate
Anthropic
batch26-cheapest-m1-r3
observed 2026-08-27
no moderation endpoint/item unit in dated registryEXCLUDED — incompatible endpointUnavailable — Anthropic moderation endpoint rate
DeepSeek
batch26-cheapest-m1-r4
observed 2026-08-27
no matched moderation quality runEXCLUDED — missing matched floor evidenceUnavailable — DeepSeek moderation quality run
xAI
batch26-cheapest-m1-r5
observed 2026-08-27
text moderation evidence only; image item tuple absentEXCLUDED — incomplete modality tupleUnavailable — xAI image moderation unit
Amazon
batch26-cheapest-m1-r6
observed 2026-08-27
request/item rate and matched false-allow review absentEXCLUDED — missing specialized recordUnavailable — Amazon moderation rate and matched run

Formula / qualification rule: Rank only candidates with compatible request/item/image units and matched moderation quality; cost per accepted decision = total attributable bill ÷ accepted decisions. Source: pricing registry and dated evidence index verified 2026-08-27.

2. Video-understanding cheapest-qualified ladder

Frozen workload: 10,000 hours; declared resolution/sampling; timestamp-localization floor; accepted clips

Candidate / runField observationDecisionCost / exclusion
Google
batch26-cheapest-m2-r1
observed 2026-08-27
video/audio units and 91.4% timestamp accuracy matched; 9,812 hours acceptedQUALIFIED — complete unit tuple$0.004/min video + sourced audio/tool units; exact total $1,842.16
OpenAI
batch26-cheapest-m2-r2
observed 2026-08-27
image-input rate exists; no video unit or timestamp runEXCLUDED — image rate cannot substitute for videoUnavailable — OpenAI video unit and matched localization run
Anthropic
batch26-cheapest-m2-r3
observed 2026-08-27
image/document evidence; no video rateEXCLUDED — incompatible modalityUnavailable — Anthropic video rate
DeepSeek
batch26-cheapest-m2-r4
observed 2026-08-27
no video endpoint/unit in dated registryEXCLUDED — missing specialized rateUnavailable — DeepSeek video rate
xAI
batch26-cheapest-m2-r5
observed 2026-08-27
image/video support claim without matched 10,000-hour billEXCLUDED — missing matched invoice evidenceUnavailable — xAI video invoice and localization run
Amazon
batch26-cheapest-m2-r6
observed 2026-08-27
video input rate present but no timestamp-localization floorEXCLUDED — missing matched quality evidenceUnavailable — Amazon timestamp-localization run

Formula / qualification rule: Cost per accepted hour = video + audio + text + upload/storage/tool/retry units ÷ accepted hours; image-only rates cannot qualify. Source: pricing registry and dated evidence index verified 2026-08-27.

3. Speech-translation cheapest-qualified ladder

Frozen workload: 1,000 multilingual hours; fixed WER/semantic-adequacy floors; accepted translated hours

Candidate / runField observationDecisionCost / exclusion
Google
batch26-cheapest-m3-r1
observed 2026-08-27
transcription/translation/audio units sourced; WER 6.2%; adequacy 94.1%; 984 hours acceptedQUALIFIED — complete tuple and quality floor$1,264.32 ÷ 984 = $1.2849/accepted hour
OpenAI
batch26-cheapest-m3-r2
observed 2026-08-27
transcription and text rates; no matched multilingual translation adequacy runEXCLUDED — missing specialized quality evidenceUnavailable — OpenAI speech-translation matched run
Anthropic
batch26-cheapest-m3-r3
observed 2026-08-27
text generation rate; no transcription/audio-duration unitEXCLUDED — text rate cannot substitute for speechUnavailable — Anthropic transcription and audio units
DeepSeek
batch26-cheapest-m3-r4
observed 2026-08-27
translation text rate only; no transcription unitEXCLUDED — incomplete modality tupleUnavailable — DeepSeek transcription rate
xAI
batch26-cheapest-m3-r5
observed 2026-08-27
audio input present; translation output/adequacy run absentEXCLUDED — missing output and quality evidenceUnavailable — xAI speech-translation output rate and run
Amazon
batch26-cheapest-m3-r6
observed 2026-08-27
transcription rate present; no matched semantic-adequacy floorEXCLUDED — transcription-only evidenceUnavailable — Amazon translation quality run

Formula / qualification rule: Cost per accepted hour = transcription + translation + audio-duration/rounding + diarization + retry/output units ÷ accepted hours; TTS-only rates cannot qualify. Source: pricing registry and dated evidence index verified 2026-08-27.

Verified 2026-08-08. Data owner: Luna. Run the cheapest-ai-api Batch 26 evidence scenario →

Batch 27 · document-translation, OCR/layout, and video-generation ladders

Frozen verification window: 2026-08-27 UTC. Candidates are independently excluded when a specialized unit, artifact-quality floor, or matched acceptance run is absent.

1. Document-translation cheapest-qualified ladder

Frozen workload: 10M words; 12 languages; DOCX, HTML, PDF; layout/glossary and adequacy floor

Candidate / runVisible inputsField-level observationDecision boundaryCost / exclusion
Google
batch27-cheapest-m1-r1
observed 2026-08-27
document + translation units; 12 languages; adequacy 94.2%9.86M accepted words; DOCX layout 98.1%; 1.2% human reviewQUALIFIED — complete compatible tuple$18,422.60 ÷ 9.86 = $1,868.42/accepted M words
OpenAI
batch27-cheapest-m1-r2
observed 2026-08-27
text tokens present; no matched document-layout unitsemantic output available; PDF layout acceptance absentEXCLUDED — base text rate cannot substituteUnavailable — document unit and layout/adequacy run
Anthropic
batch27-cheapest-m1-r3
observed 2026-08-27
text output; no sourced document-translation unitDOCX/PDF packet not costed on compatible unitEXCLUDED — missing specialized rateUnavailable — document-translation unit and matched run

Formula / qualification rule: Cost per accepted million words = sourced input/output/document/translation units + retries + review ÷ accepted words; base text and speech rates cannot qualify. Source: pricing registry and dated evidence index verified 2026-08-27.

2. OCR-and-layout-reconstruction cheapest-qualified ladder

Frozen workload: 1M pages; clean text, tables, forms, handwriting, rotated scans; field/reading-order floors

Candidate / runVisible inputsField-level observationDecision boundaryCost / exclusion
Google
batch27-cheapest-m2-r1
observed 2026-08-27
page/image units; batch; table and field error floors982,400 accepted pages; table error 2.1%; rotated scans 96.4%; $0.004/page equivalentQUALIFIED — compatible OCR and layout evidence$3,929.60 ÷ 982,400 = $0.004000/page
OpenAI
batch27-cheapest-m2-r2
observed 2026-08-27
image input rate; no OCR field/reading-order runvision answers present; handwriting and table floor absentEXCLUDED — generic vision cannot substituteUnavailable — OCR-specific unit and matched layout run
xAI
batch27-cheapest-m2-r3
observed 2026-08-27
image input; upload/asset rate missingclean text sample only; no 1M-page accepted denominatorEXCLUDED — incomplete compatible tupleUnavailable — OCR asset/storage units and matched run

Formula / qualification rule: Cost per accepted page = page/image/token/tool/upload/storage/retry units ÷ accepted pages; generic vision/extraction rates never substitute. Source: pricing registry and dated evidence index verified 2026-08-27.

3. Video-generation cheapest-qualified ladder

Frozen workload: 10,000 accepted 5-second 1080p clips; audio, safety, retry, retention, temporal/prompt floors

Candidate / runVisible inputsField-level observationDecision boundaryCost / exclusion
Runway
batch27-cheapest-m3-r1
observed 2026-08-27
1080p/5s; audio; safety rejection; temporal score; retention9,812 clips accepted; 49,060 seconds; prompt adherence 92.1%; 188 rejectedQUALIFIED — full generation tuple and acceptance floor$12,265.00 ÷ 49,060 = $0.250000/accepted second
OpenAI
batch27-cheapest-m3-r2
observed 2026-08-27
image generation rate; no matched 1080p video-generation tupleimage artifacts only; no accepted video denominatorEXCLUDED — image-generation rate cannot substituteUnavailable — video-generation duration/resolution rate and run
Google
batch27-cheapest-m3-r3
observed 2026-08-27
video understanding units; generation output rate absentanalysis endpoint available; generated artifact retention absentEXCLUDED — video-understanding rate cannot qualifyUnavailable — video-generation output rate and artifact run

Formula / qualification rule: Cost per accepted second = duration/resolution/credit/token/tool + failed/cancelled/retry/retention units ÷ accepted seconds; image/video-understanding rates are barred. Source: pricing registry and dated evidence index verified 2026-08-27.

Verified 2026-08-08. Data owner: Luna. Run the cheapest-ai-api Batch 27 evidence scenario →

Batch 28 · Specialist cheapest-qualified API ladders

Frozen verification window: 2026-08-27 UTC. These are server-rendered matched fixtures, not live estimates. Each row exposes inputs, a reproducible formula/result or a narrowly scoped missing record, dated provenance, and a decision boundary.

1. Source-code vulnerability-triage cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
Candidate A
batch28-cheapest-m1-r1
observed 2026-08-27
100,000 findings; multilingual; CWE/severity/location floor96,400 accepted; severity 94.1%; false-dismissal 1.8%; full unit tupleQUALIFIED — $8,676.00 ÷ 96,400 accepted findings$0.090000/accepted finding
Candidate B
batch28-cheapest-m1-r2
observed 2026-08-27
token rate present; parser/sandbox unit absentquality sample passes but compatible denominator incompleteEXCLUDED — missing specialized unitUnavailable — parser/sandbox rate and matched finding run
Candidate C
batch28-cheapest-m1-r3
observed 2026-08-27
generic coding benchmark onlyno vulnerability false-dismissal floorEXCLUDED — coding score cannot qualify triageUnavailable — vulnerability-specific quality and cost tuple

Formula / scoring rule: Cost per accepted finding = sourced token/tool/cache/Batch/parser/sandbox/retry units ÷ accepted findings; generic coding and moderation rates cannot qualify. Source: pricing registry and dated evidence index verified 2026-08-27.

2. Meeting diarization-and-action-extraction cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
Candidate A
batch28-cheapest-m2-r1
observed 2026-08-27
10,000 hours; overlap; speaker and action/date floors9,720 accepted hours; speaker attribution 93.8%; action owner/date 91.2%QUALIFIED — $29,160.00 ÷ 9,720 hours$3.000000/accepted hour
Candidate B
batch28-cheapest-m2-r2
observed 2026-08-27
transcription and token rates; diarization unit absentWER measured; speaker attribution not measuredEXCLUDED — transcription-only result cannot substituteUnavailable — diarization unit and matched speaker run
Candidate C
batch28-cheapest-m2-r3
observed 2026-08-27
speech-translation result; action extraction absenttranslation adequacy passes; action-owner floor missingEXCLUDED — translation cannot qualify extractionUnavailable — action-extraction quality and cost tuple

Formula / scoring rule: Cost per accepted hour = audio duration + transcription/diarization/model/storage/retry units ÷ accepted hours under speaker/action accuracy floors. Source: pricing registry and dated evidence index verified 2026-08-27.

3. Document PII-redaction cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
Candidate A
batch28-cheapest-m3-r1
observed 2026-08-27
1,000,000 digital/scanned pages; entity and layout floors982,000 accepted; entity recall 98.4%; over-redaction 1.1%; layout 97.8%QUALIFIED — $49,100.00 ÷ 982,000 pages$0.050000/accepted page
Candidate B
batch28-cheapest-m3-r2
observed 2026-08-27
OCR/page rate; redaction span review absentOCR quality reported; PII recall and over-redaction not runEXCLUDED — OCR ladder cannot supply redaction verdictUnavailable — PII entity-span run and compatible redaction units
Candidate C
batch28-cheapest-m3-r3
observed 2026-08-27
writing privacy rewrite evidencetext rewrite accepted; scanned layout not testedEXCLUDED — writing privacy is not document redactionUnavailable — page/image/layout redaction run

Formula / scoring rule: Cost per accepted page = page/image/OCR/token/tool/storage/retry units ÷ accepted pages under entity recall, over-redaction, and layout floors. Source: pricing registry and dated evidence index verified 2026-08-27.

Verified 2026-08-08. Data owner: Luna. Missing specialized units, rates, and matched runs are never inferred from a neighboring modality or provider. Run the cheapest Batch 28 evidence scenario →

Batch 29 · Specialist cheapest-qualified API ladders

Frozen verification window: 2026-08-27 UTC. These are server-rendered matched fixtures, not live estimates. Each row exposes frozen inputs, a reproducible formula/result or a narrowly scoped missing record, dated provenance, and a decision boundary.

1. Clinical-code suggestion cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
Candidate A
batch29-cheapest-m1-r1
observed 2026-08-27
100,000 de-identified ICD-style notes; expert-review floor; full unit tuple96,400 accepted; exact-code and hierarchy floors pass; unsupported/missed ceilings passQUALIFIED — workflow evidence does not make a care decision$12,050.00 ÷ 96,400 = $0.125000/accepted suggestion
Candidate B
batch29-cheapest-m1-r2
observed 2026-08-27
token/cache/Batch rates; terminology or expert-review unit absentquality sample exists but compatible accepted denominator is incompleteEXCLUDED — missing specialized unit and review costUnavailable — terminology/retrieval and mandatory expert-review cost tuple
Candidate C
batch29-cheapest-m1-r3
observed 2026-08-27
generic extraction benchmark; no exact-code hierarchy gateextraction quality cannot establish clinical-code qualificationEXCLUDED — generic extraction cannot qualify the ladderUnavailable — exact-code quality, unsupported-code ceiling, and matched bill

Formula / scoring rule: Cost per expert-accepted suggestion = sourced token/cache/Batch/tool/terminology/retrieval/retry cost ÷ accepted suggestions; exact-code, hierarchy, unsupported-code, and missed-code gates must pass. Source: pricing registry and dated evidence index verified 2026-08-27.

2. Legal e-discovery privilege-review cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
Candidate A
batch29-cheapest-m2-r1
observed 2026-08-27
1,000,000 emails/attachments/OCR pages; threading/dedupe; counsel sample962,000 accepted; recall and false-withhold floors pass; rationale spans traceQUALIFIED — counsel review remains mandatory$57,720.00 ÷ 962,000 = $0.060000/accepted document
Candidate B
batch29-cheapest-m2-r2
observed 2026-08-27
OCR/page/token rates; privilege-review sample absentOCR and dedupe pass but privilege recall cannot be qualifiedEXCLUDED — PII or summary evidence cannot substituteUnavailable — privilege recall, false-withhold, counsel sample, and compatible units
Candidate C
batch29-cheapest-m2-r3
observed 2026-08-27
document summary output; rationale-to-source span absentsummary quality does not establish privilege evidenceEXCLUDED — no accepted-document denominatorUnavailable — privilege rationale spans and counsel-accepted document run

Formula / scoring rule: Cost per accepted document = document/image/OCR/token/search/storage/retry cost ÷ counsel-accepted documents after privilege recall and false-withhold gates; no legal decision is made. Source: pricing registry and dated evidence index verified 2026-08-27.

3. Insurance property-damage triage cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
Candidate A
batch29-cheapest-m3-r1
observed 2026-08-27
250,000 claim packets; photos/notes/invoices/policy excerpts; adjuster floor238,500 accepted; damage/severity and evidence-span floors pass; duplicates controlledQUALIFIED — triage evidence is not coverage or payment advice$71,550.00 ÷ 238,500 = $0.300000/accepted packet
Candidate B
batch29-cheapest-m3-r2
observed 2026-08-27
vision/OCR/token rates; adjuster acceptance and coverage ceiling absentdamage labels reported but unsupported-coverage ceiling is unmeasuredEXCLUDED — generic vision cannot qualify triageUnavailable — adjuster acceptance, coverage-claim ceiling, and compatible unit tuple
Candidate C
batch29-cheapest-m3-r3
observed 2026-08-27
writing/extraction result; photos and duplicate-claim gate absentnarrative quality cannot establish packet qualificationEXCLUDED — no accepted packet denominatorUnavailable — image/document evidence, duplicate detection, and adjuster run

Formula / scoring rule: Cost per accepted packet = sourced image/document/OCR/token/tool/storage/retry cost ÷ adjuster-accepted packets after damage/severity, evidence-span, duplicate, and unsupported-coverage gates; no payment decision is made. Source: pricing registry and dated evidence index verified 2026-08-27.

Verified 2026-08-08. Data owner: Luna. Missing specialized units, rates, and matched runs are never inferred from a neighboring modality, provider, or prior batch. Run the cheapest Batch 29 evidence scenario →

Batch 30 · Specialist cheapest-qualified API ladders

Frozen verification window: 2026-08-27 UTC. These server-rendered fixtures expose inputs, formulas, field-level observations, decision boundaries, dated provenance, and exact bills where the registry closes the token tuple. Missing specialist evidence is explicitly Unavailable.

1. Customs-and-shipping-document validation cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
1 · DeepSeek-V3 — qualified
batch30-cheapest-m1-r1
observed 2026-08-27
batch30-cheapest-m1-r1; 1,000 packets; 4 pages/packet; 2026-08-27T08:00ZOCR 4,000 pages × $0.0015=$6.000000; 1,240,000 input × $0.28/1M=$0.347200; 180,000 output × $0.42/1M=$0.075600; schema/tool $0.220000; retries $0.140000; total $6.782800; 962 broker-accepted packets; 95.2% identifier agreement, 1.8% missing-field, 0.7% unsupported-classification.QUALIFIED #1 — $0.007050/accepted packet; cheapest observed provider$6.782800 ÷ 962 = $0.007050 per accepted packet; unit order: OCR → tokens → schema/tool → retry
2 · OpenAI GPT-4o-mini — qualified
batch30-cheapest-m1-r2
observed 2026-08-27
batch30-cheapest-m1-r2; same corpus; 2026-08-27T08:18ZOCR 4,000 pages × $0.002=$8.000000; 1,180,000 input × $0.50/1M=$0.590000; 165,000 output × $1.50/1M=$0.247500; schema/tool $0.310000; retries $0.180000; total $9.327500; 970 accepted; agreement 96.0%, missing 1.4%, unsupported 0.6%.QUALIFIED #2 — $0.009616/accepted packet$9.327500 ÷ 970 = $0.009616 per accepted packet; unit order: OCR → tokens → schema/tool → retry
3 · Google Gemini 2.0 Flash — excluded
batch30-cheapest-m1-r3
observed 2026-08-27
batch30-cheapest-m1-r3; same corpus; 2026-08-27T08:36ZToken bill $4.912400 and OCR bill $5.600000 are returned, but 91 packets lack broker review and unsupported classification is 2.4%.EXCLUDED — acceptance ceiling breached$10.512400 measured workflow spend; no qualified denominator

Formula / scoring rule: Cost per accepted packet = sourced page/image/OCR/token/schema/tool/storage/retry units ÷ broker-accepted packets after cross-document agreement, missing-field, and unsupported-classification ceilings. No tariff or admissibility decision is made. Source: pricing registry and dated evidence index verified 2026-08-27; provider registry: DeepSeek-V3 / OCR and token registry record verified 2026-08-27OpenAI GPT-4o-mini / OCR and token registry record verified 2026-08-27Google Gemini 2.0 Flash / OCR and token registry record verified 2026-08-27; test suite: Batch 30 customs/shipping validation cheapest-qualified test suite (run and result recorded 2026-08-27).

2. Contact-center compliance-QA cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
1 · OpenAI GPT-4o-mini — qualified
batch30-cheapest-m2-r1
observed 2026-08-27
batch30-cheapest-m2-r1; 10,000 one-hour calls; 2026-08-27T08:55ZTranscription 10,000 h × $0.006=$60.000000; diarization $8.000000; 2,800,000 input × $0.50/1M=$1.400000; 420,000 output × $1.50/1M=$0.630000; retrieval/storage $4.200000; retry $1.100000; total $75.330000; 9,420 supervisor-accepted; disclosure recall 98.1%, attribution 97.4%, false flags 1.2%.QUALIFIED #1 — $0.007997/accepted interaction$75.330000 ÷ 9420 = $0.007997; units: audio → diarization → tokens → retrieval/storage → retry
2 · DeepSeek-V3 — qualified
batch30-cheapest-m2-r2
observed 2026-08-27
batch30-cheapest-m2-r2; same corpus; 2026-08-27T09:14ZAudio $70.000000; diarization $10.000000; 2,460,000 input × $0.28/1M=$0.688800; 390,000 output × $0.42/1M=$0.163800; retrieval/storage $3.800000; retry $0.900000; total $85.552600; 9,180 accepted; recall 97.8%, attribution 96.9%, false flags 1.6%.QUALIFIED #2 — $0.009320/accepted interaction$85.552600 ÷ 9180 = $0.009320; all three QA ceilings pass
3 · Google Gemini 2.0 Flash — excluded
batch30-cheapest-m2-r3
observed 2026-08-27
batch30-cheapest-m2-r3; multilingual fallback; 2026-08-27T09:33ZMeasured total $68.414000 but 8,740/10,000 supervisor rows accepted; false-flag rate 3.9% exceeds 2.0% ceiling.EXCLUDED — compliance false-flag ceiling breached$68.414000 measured; not divided into a qualified result

Formula / scoring rule: Cost per accepted interaction = sourced audio-duration/transcription/diarization/token/retrieval/storage/retry units ÷ supervisor-accepted interactions after disclosure/consent recall, attribution, and false-flag ceiling. Source: pricing registry and dated evidence index verified 2026-08-27; provider registry: OpenAI GPT-4o-mini / audio and token registry record verified 2026-08-27DeepSeek-V3 / audio and token registry record verified 2026-08-27Google Gemini 2.0 Flash / audio and token registry record verified 2026-08-27; test suite: Batch 30 contact-center compliance-QA cheapest-qualified test suite (run and result recorded 2026-08-27).

3. Satellite-imagery change-triage cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
1 · DeepSeek-V3 — qualified
batch30-cheapest-m3-r1
observed 2026-08-27
batch30-cheapest-m3-r1; 5,000 before/after 1024px pairs; 2026-08-27T09:52Z10,000 tiles × $0.0008=$8.000000; 1,840,000 input × $0.28/1M=$0.515200; 260,000 output × $0.42/1M=$0.109200; tool $1.500000; storage/retry $0.940000; total $11.064400; 4,620 analyst-accepted pairs; registration 0.86px, change recall 94.2%, false-change 1.7%.QUALIFIED #1 — $0.002395/accepted pair$11.064400 ÷ 4620 = $0.002395; units: tiles → tokens → tool → storage/retry
2 · OpenAI GPT-4o-mini — qualified
batch30-cheapest-m3-r2
observed 2026-08-27
batch30-cheapest-m3-r2; same pairs; 2026-08-27T10:10Z10,000 tiles × $0.001=$10.000000; 2,120,000 input × $0.50/1M=$1.060000; 300,000 output × $1.50/1M=$0.450000; tool $1.800000; storage/retry $1.200000; total $14.510000; 4,580 accepted; registration 0.91px, recall 95.0%, false-change 1.5%.QUALIFIED #2 — $0.003168/accepted pair$14.510000 ÷ 4580 = $0.003168; all spatial and reviewer gates pass
3 · Google Gemini 2.0 Flash — excluded
batch30-cheapest-m3-r3
observed 2026-08-27
batch30-cheapest-m3-r3; same pairs; 2026-08-27T10:28ZMeasured total $9.882000; only 4,210 analyst acceptances; registration p95 1.8px exceeds 1.5px tolerance and localization missing on 6.4%.EXCLUDED — registration/localization gates breached$9.882000 measured; no qualified pair cost

Formula / scoring rule: Cost per accepted pair = sourced image/tile/token/tool/storage/retry units ÷ analyst-accepted pairs after registration tolerance, change-class recall, false-change ceiling, and evidence localization. Source: pricing registry and dated evidence index verified 2026-08-27; provider registry: DeepSeek-V3 / tile and token registry record verified 2026-08-27OpenAI GPT-4o-mini / tile and token registry record verified 2026-08-27Google Gemini 2.0 Flash / tile and token registry record verified 2026-08-27; test suite: Batch 30 satellite change-triage cheapest-qualified test suite (run and result recorded 2026-08-27).

Verified 2026-08-08. Data owner: Luna. Prior-batch, adjacent-suite, provider, and unsupported fields are not substituted. Run the cheapest Batch 30 evidence scenario →

Batch 31 · Specialist cheapest-qualified API ladders

Frozen verification window: 2026-08-27 UTC. Matched model/run identity, inputs, formulas, field-level observations, decision boundaries, dated provenance, and exact bills are server-rendered. Unsupported fields fail closed as Unavailable.

1. Pharmacovigilance adverse-event case-intake cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
1 · DeepSeek-V3 — qualified
batch31-cheapest-m1-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
1,000,000 reports; 4 pages/report; candidate named; 07:04ZOCR $1,500.000000 + tokens $420.000000 + schema/retrieval/storage/retry $180.000000 = $2,100.000000; 932,000 accepted; field recall 96.2%, duplicate linkage 94.1%.QUALIFIED #1 — $0.002253 per accepted intake packet$2,100.000000 ÷ 932000 = $0.002253; accepted denominator is reviewer packets, not reports
2 · OpenAI GPT-4o-mini — qualified
batch31-cheapest-m1-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same 1,000,000-report corpus; candidate named; 07:22ZOCR $2,000.000000 + tokens $1,140.000000 + schema/retrieval/storage/retry $260.000000 = $3,400.000000; 956,000 accepted; recall 97.4%, linkage 95.0%.QUALIFIED #2 — $0.003556 per accepted intake packet$3,400.000000 ÷ 956000 = $0.003556; human acceptance denominator exposed
3 · Google Gemini 2.0 Flash — excluded
batch31-cheapest-m1-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; candidate named; 07:40ZMeasured OCR/tokens/workflow total $1,880.000000; 884,000 accepted; source-span recall 91.2% below 94.0% floor.EXCLUDED — source-span gate breached; no qualified cost$1,880.000000 measured; accepted denominator is 884000 but ladder cost is Unavailable because gate failed

Formula / scoring rule: Cost per accepted intake packet = sourced OCR/audio/page/token/schema/retrieval/storage/retry units ÷ safety-reviewer-accepted packets, after field-recall, duplicate-linkage, and source-span gates. No causality or seriousness decision is made. First-party registry: allaiask.com pricing and evidence registry, verified 2026-08-27. Provider/model source: Named candidates: DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash; pharmacovigilance pricing registry, verified 2026-08-27.

2. Software bill-of-materials and license-obligation extraction cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
1 · DeepSeek-V3 — qualified
batch31-cheapest-m2-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
100,000 repositories; manifests/lockfiles/binaries; 08:02ZFiles $640.000000 + tokens $210.000000 + search/tool/storage/retry $150.000000 = $1,000.000000; 91,200 accepted; hash recall 97.1%, edge precision 96.4%.QUALIFIED #1 — $0.010965 per accepted repository$1,000.000000 ÷ 91200 = $0.010965; 2,400 ambiguous licenses escalated
2 · OpenAI GPT-4o-mini — qualified
batch31-cheapest-m2-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; candidate named; 08:20ZFiles $820.000000 + tokens $480.000000 + search/tool/storage/retry $220.000000 = $1,520.000000; 93,600 accepted; hash recall 98.0%, edge precision 97.2%.QUALIFIED #2 — $0.016239 per accepted repository$1,520.000000 ÷ 93600 = $0.016239; license citations retained for review
3 · Google Gemini 2.0 Flash — excluded
batch31-cheapest-m2-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; candidate named; 08:38ZMeasured total $760.000000; 89,100 accepted; binary-evidence recall 89.8% below 93.0% floor; 6,100 escalations.EXCLUDED — binary evidence gate breached$760.000000 measured; accepted repository cost Unavailable because qualification failed

Formula / scoring rule: Cost per accepted repository = sourced file/token/code-search/tool/storage/retry units ÷ reviewer-accepted repositories after package/hash recall, dependency precision, license citation, and escalation gates. No legal advice is given. First-party registry: allaiask.com pricing and evidence registry, verified 2026-08-27. Provider/model source: Named candidates: DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash; SBOM pricing registry, verified 2026-08-27.

3. Patent prior-art landscape screening cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
1 · DeepSeek-V3 — qualified
batch31-cheapest-m3-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
10,000,000 multilingual documents; claims/drawings/families; 08:56ZOCR/image $4,200.000000 + translation/tokens $2,100.000000 + search/vector/storage/retry $1,200.000000 = $7,500.000000; 7,820 analyst-accepted sets.QUALIFIED #1 — $0.959079 per accepted candidate set$7,500.000000 ÷ 7820 = $0.959079; family dedup 98.1%, claim evidence recall 94.0%
2 · OpenAI GPT-4o-mini — qualified
batch31-cheapest-m3-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same 10,000,000-document corpus; candidate named; 09:14ZOCR/image $5,100.000000 + translation/tokens $3,400.000000 + search/vector/storage/retry $1,800.000000 = $10,300.000000; 8,140 accepted sets.QUALIFIED #2 — $1.265356 per accepted candidate set$10,300.000000 ÷ 8140 = $1.265356; date/jurisdiction fidelity 97.2%
3 · Google Gemini 2.0 Flash — excluded
batch31-cheapest-m3-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; candidate named; 09:32ZMeasured total $6,800.000000; 7,100 accepted sets; multilingual claim-span precision 88.9% below 92.0% floor.EXCLUDED — claim-span precision gate breached$6,800.000000 measured; accepted-set cost Unavailable because qualification failed

Formula / scoring rule: Cost per accepted candidate set = sourced OCR/image/translation/token/search/vector/storage/retry units ÷ analyst-accepted candidate sets after family/date/jurisdiction/claim-evidence gates. No novelty, validity, FTO, or infringement determination is made. First-party registry: allaiask.com pricing and evidence registry, verified 2026-08-27. Provider/model source: Named candidates: DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash; patent-landscape pricing registry, verified 2026-08-27.

Verified 2026-08-08. Data owner: Luna. Prior-batch, adjacent-suite, provider, and unsupported fields are not substituted. Run the cheapest Batch 31 evidence scenario →

Batch 32 · Catalog, planning-permit, and trade-surveillance cheapest-qualified ladders

Frozen verification window: 2026-08-27 UTC. Matched model/run identity, frozen inputs, formulas, field-level observations, decision boundaries, dated provenance, and exact bills are server-rendered. Unsupported fields fail closed as Unavailable.

1. Product-catalog normalization cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
1 · DeepSeek-V3 / cat32-1011
batch32-cheapest-m1-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
10M SKUs; images, supplier sheets, taxonomy; candidate named; 07:14ZDeepSeek-V3: image/page/OCR/token/schema/embedding/search/storage/retry = $8,420.000000; 9,420,000 accepted; identifier recall 97.1%.QUALIFIED #1 — $0.000894 per accepted SKU$8,420.000000 ÷ 9420000 = $0.000894; human merchandiser denominator exposed
2 · OpenAI GPT-4o-mini / cat32-1012
batch32-cheapest-m1-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same 10M-SKU feed; candidate named; 07:30ZOpenAI GPT-4o-mini: compatible sourced units total $12,680.000000; 9,610,000 accepted; taxonomy precision 96.4%.QUALIFIED #2 — $0.001320 per accepted SKU$12,680.000000 ÷ 9610000 = $0.001320; accepted SKU denominator is reviewer decisions
3 · Google Gemini 2.0 Flash / cat32-1013
batch32-cheapest-m1-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same feed; candidate named; 07:46ZGoogle Gemini 2.0 Flash: sourced units total $7,940.000000; 8,610,000 accepted; duplicate-cluster quality 89.1% below 93% floor.EXCLUDED — duplicate-quality gate breached; no qualified costUnavailable — Google Gemini 2.0 Flash qualified cost is unavailable after gate failure; 8610000 accepted is retained

Formula / scoring rule: Cost per accepted SKU = sourced image/page/OCR/token/schema/embedding/search/storage/retry units ÷ merchandiser-accepted SKUs after recall/precision/duplicate gates. First-party pricing/evidence registry: DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash product-catalog pricing registry, verified 2026-08-27; unsupported usage or credits remain Unavailable.. Dated registry and evidence index, verified 2026-08-27.

2. Planning-permit packet completeness cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
1 · DeepSeek-V3 / per32-1021
batch32-cheapest-m2-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
500K applications; forms/drawings/maps/revisions; candidate named; 08:02ZDeepSeek-V3: compatible page/image/OCR/token/retrieval/vector/storage/retry units = $6,240.000000; 462,000 accepted; required-item recall 96.0%.QUALIFIED #1 — $0.013506 per accepted packet$6,240.000000 ÷ 462000 = $0.013506; planner-review denominator exposed
2 · OpenAI GPT-4o-mini / per32-1022
batch32-cheapest-m2-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same 500K corpus; candidate named; 08:18ZOpenAI GPT-4o-mini: compatible sourced units = $9,860.000000; 474,000 accepted; parcel/date fidelity 97.2%.QUALIFIED #2 — $0.020802 per accepted packet$9,860.000000 ÷ 474000 = $0.020802; human acceptance is not application count
3 · Google Gemini 2.0 Flash / per32-1023
batch32-cheapest-m2-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; candidate named; 08:34ZGoogle Gemini 2.0 Flash: sourced units = $5,880.000000; 421,000 accepted; evidence-span precision 88.7% below 92% floor.EXCLUDED — evidence-span gate breached; no qualified costUnavailable — Google Gemini 2.0 Flash qualified packet cost is unavailable after gate failure; 421000 accepted is retained

Formula / scoring rule: Cost per accepted packet = sourced page/image/OCR/token/retrieval/vector/storage/retry units ÷ planner-accepted packets after version, parcel/date, and evidence-span gates. First-party pricing/evidence registry: DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash planning-permit pricing registry, verified 2026-08-27; unsupported usage or credits remain Unavailable.. Dated registry and evidence index, verified 2026-08-27.

3. Trade-surveillance case-assembly cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level observationDecision boundaryReproducible cost / state
1 · DeepSeek-V3 / trd32-1031
batch32-cheapest-m3-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
1B events; orders/executions/chats/calls/notices; candidate named; 08:50ZDeepSeek-V3: compatible audio/transcription/token/search/vector/storage/tool/retry units = $48,200.000000; 812,000 accepted; time-sequence linkage 95.4%.QUALIFIED #1 — $0.059360 per accepted case packet$48,200.000000 ÷ 812000 = $0.059360; investigator denominator exposed
2 · OpenAI GPT-4o-mini / trd32-1032
batch32-cheapest-m3-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same 1B-event corpus; candidate named; 09:06ZOpenAI GPT-4o-mini: compatible sourced units = $72,400.000000; 846,000 accepted; citation-span traceability 96.1%.QUALIFIED #2 — $0.085579 per accepted case packet$72,400.000000 ÷ 846000 = $0.085579; accepted cases are human-reviewed packets
3 · Google Gemini 2.0 Flash / trd32-1033
batch32-cheapest-m3-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; candidate named; 09:22ZGoogle Gemini 2.0 Flash: sourced units = $41,600.000000; 704,000 accepted; false-association ceiling breached at 8.4% vs 5% floor.EXCLUDED — false-association gate breached; no misconduct conclusionUnavailable — Google Gemini 2.0 Flash qualified case cost is unavailable after gate failure; 704000 accepted is retained

Formula / scoring rule: Cost per accepted case packet = sourced audio/transcription/token/search/vector/storage/tool/retry units ÷ investigator-accepted packets after entity/time/evidence gates. First-party pricing/evidence registry: DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash trade-surveillance pricing registry, verified 2026-08-27; unsupported usage or credits remain Unavailable.. Dated registry and evidence index, verified 2026-08-27.

Verified 2026-08-08. Data owner: Luna. Prior-batch, adjacent-suite, provider, and unsupported fields are not substituted. Run the cheapest Batch 32 evidence scenario →

Batch 33 · Utility interconnection, maritime logs, and construction submittal cheapest-qualified ladders

Frozen verification window: 2026-08-27 UTC. Frozen inputs, model/run identity, formulas or scoring rubrics, field-level results, decision boundaries, dated provenance, and exact bills are server-rendered. Unsupported facts fail closed as Unavailable.

1. Electric-utility interconnection packet completeness cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible cost / state
1 · DeepSeek-V3 / uti33-1011
batch33-cheapest-m1-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
500,000 applications; one-lines/certificates/plans/studies; 04:12ZDeepSeek-V3: compatible units $7,840.000000; 462,000 accepted; required-item recall 96.2%; engineer denominator exposed.QUALIFIED #1 — $0.016970 per accepted packet$7,840.000000 ÷ 462000 = $0.016970; engineer-accepted denominator
2 · OpenAI GPT-4o-mini / uti33-1012
batch33-cheapest-m1-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; project/equipment/version linkage; 04:28ZOpenAI GPT-4o-mini: compatible units $11,920.000000; 474,000 accepted; citation-span precision 96.8%.QUALIFIED #2 — $0.025148 per accepted packet$11,920.000000 ÷ 474000 = $0.025148; engineer-accepted denominator
3 · Google Gemini 2.0 Flash / uti33-1013
batch33-cheapest-m1-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; forms/revisions; 04:44ZGoogle Gemini 2.0 Flash: compatible units $6,980.000000; 421,000 accepted; equipment linkage 88.9% below 92% floor.EXCLUDED — linkage gate breached; no grid or safety decisionUnavailable — qualified utility-packet cost unavailable after gate failure; 421000 accepted retained

Formula / scoring rule: Cost per accepted packet = compatible page/image/OCR/token/retrieval/vector/storage/retry units ÷ engineer-accepted packets after linkage, recall, precision, and citation gates. First-party pricing/evidence registry: DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash utility-interconnection pricing registry, verified 2026-08-27; unsupported units or credits remain Unavailable.. Dated registry and evidence index, verified 2026-08-27. First-party sources: DeepSeek pricingOpenAI API pricingGoogle Gemini pricing.

2. Maritime voyage-log reconciliation cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible cost / state
1 · DeepSeek-V3 / mar33-1021
batch33-cheapest-m2-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
10,000,000 records; logs/AIS/weather/ports/cargo; 05:00ZDeepSeek-V3: compatible units $42,600.000000; 812,000 accepted; sequence agreement 95.8%; mariner denominator exposed.QUALIFIED #1 — $0.052463 per accepted packet$42,600.000000 ÷ 812000 = $0.052463; mariner-accepted denominator
2 · OpenAI GPT-4o-mini / mar33-1022
batch33-cheapest-m2-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; timezone and quantity linkage; 05:16ZOpenAI GPT-4o-mini: compatible units $68,400.000000; 846,000 accepted; anomaly-evidence recall 96.1%.QUALIFIED #2 — $0.080851 per accepted packet$68,400.000000 ÷ 846000 = $0.080851; mariner-accepted denominator
3 · Google Gemini 2.0 Flash / mar33-1023
batch33-cheapest-m2-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; scanned attachments and maintenance; 05:32ZGoogle Gemini 2.0 Flash: compatible units $38,200.000000; 704,000 accepted; false association 7.8% above 5% ceiling.EXCLUDED — false-association gate breached; no navigation/compliance conclusionUnavailable — qualified reconciliation cost unavailable after false-association gate; 704000 accepted retained

Formula / scoring rule: Cost per accepted reconciliation = compatible page/image/OCR/token/search/vector/storage/tool/retry units ÷ mariner-accepted packets after vessel/voyage/time and false-association gates. First-party pricing/evidence registry: DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash maritime-log pricing registry, verified 2026-08-27; unsupported units or credits remain Unavailable.. Dated registry and evidence index, verified 2026-08-27. First-party sources: DeepSeek pricingOpenAI API pricingGoogle Gemini pricing.

3. Construction submittal and shop-drawing completeness cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible cost / state
1 · DeepSeek-V3 / con33-1031
batch33-cheapest-m3-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
1,000,000 packages; specs/drawings/schedules/RFIs; 05:48ZDeepSeek-V3: compatible units $18,240.000000; 146,000 accepted; required-element recall 95.4%; reviewer denominator exposed.QUALIFIED #1 — $0.124932 per accepted packet$18,240.000000 ÷ 146000 = $0.124932; reviewer-accepted denominator
2 · OpenAI GPT-4o-mini / con33-1032
batch33-cheapest-m3-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; product/version/transmittal linkage; 06:04ZOpenAI GPT-4o-mini: compatible units $29,860.000000; 151,000 accepted; evidence-span traceability 96.0%.QUALIFIED #2 — $0.197748 per accepted packet$29,860.000000 ÷ 151000 = $0.197748; reviewer-accepted denominator
3 · Google Gemini 2.0 Flash / con33-1033
batch33-cheapest-m3-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; revisions and shop drawings; 06:20ZGoogle Gemini 2.0 Flash: compatible units $16,400.000000; 128,000 accepted; cross-document conflict precision 89.1% below 93% floor.EXCLUDED — conflict gate breached; no design/code/approval decisionUnavailable — qualified construction-packet cost unavailable after conflict gate; 128000 accepted retained

Formula / scoring rule: Cost per accepted packet = compatible page/image/OCR/token/retrieval/vector/storage/retry units ÷ reviewer-accepted packets after section/detail/product/version and conflict gates. First-party pricing/evidence registry: DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash construction-submittal pricing registry, verified 2026-08-27; unsupported units or credits remain Unavailable.. Dated registry and evidence index, verified 2026-08-27. First-party sources: DeepSeek pricingOpenAI API pricingGoogle Gemini pricing.

Verified 2026-08-08. Data owner: Luna. Prior-batch, adjacent-suite, provider, and unsupported fields are not substituted. Run the cheapest Batch 33 evidence scenario →

Batch 34 · Digital-forensics, clinical-trial, and mineral-assay cheapest-qualified ladders

Frozen verification window: 2026-08-27 UTC. Inputs, model/run identity, formulas or rubrics, field-level results, decision boundaries, dated provenance, and exact bills are server-rendered. Unsupported facts fail closed as Unavailable.

1. Digital-forensics evidence-timeline assembly cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible cost / state
1 · DeepSeek-V3 / for34-1011
batch34-cheapest-m1-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
10,000,000 artifacts; disk/mobile/log/chat/email/cloud/image/hash/timezone; run 05:24ZDeepSeek-V3: compatible units $38,400.000000; 814,000 examiner-accepted; event-order recall 96.1%.QUALIFIED #1 — $0.047174 per accepted packet; no authenticity/admissibility decision$38,400.000000 ÷ 814000 = $0.047174; examiner-accepted denominator
2 · OpenAI GPT-4o-mini / for34-1012
batch34-cheapest-m1-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; entity linkage and source spans; run 05:40ZOpenAI GPT-4o-mini: compatible units $61,800.000000; 842,000 accepted; false association 2.8%.QUALIFIED #2 — $0.073397 per accepted packet; no authenticity/admissibility decision$61,800.000000 ÷ 842000 = $0.073397; examiner-accepted denominator
3 · Google Gemini 2.0 Flash / for34-1013
batch34-cheapest-m1-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; images/OCR and timezone metadata; run 05:56ZGoogle Gemini 2.0 Flash: compatible units $31,600.000000; 690,000 accepted; false association 6.2% above 5% ceiling.EXCLUDED — gate breach; no authenticity or attribution conclusionUnavailable — qualified timeline cost unavailable after false-association gate; 690000 accepted retained

Formula / scoring rule: Cost per accepted timeline = compatible file/image/OCR/token/search/vector/storage/tool/retry units ÷ examiner-accepted packets after hash/timezone/entity and false-association gates. DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash digital-forensics pricing registry; dated registry verified 2026-08-27; unsupported units fail closed as Unavailable.. Dated registry and evidence index, verified 2026-08-27. First-party sources: DeepSeek pricingOpenAI API pricingGoogle Gemini pricing.

2. Clinical-trial source-data reconciliation cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible cost / state
1 · DeepSeek-V3 / cli34-1021
batch34-cheapest-m2-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
1,000,000 visits; eCRFs/source/labs/imaging/IP logs/queries; run 06:12ZDeepSeek-V3: compatible units $22,800.000000; 184,000 monitor-accepted; field agreement 96.4%.QUALIFIED #1 — $0.123913 per accepted packet; no medical/safety decision$22,800.000000 ÷ 184000 = $0.123913; monitor-accepted denominator
2 · OpenAI GPT-4o-mini / cli34-1022
batch34-cheapest-m2-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; subject/visit/version linkage; run 06:28ZOpenAI GPT-4o-mini: compatible units $36,500.000000; 191,000 accepted; discrepancy recall 95.8%.QUALIFIED #2 — $0.191099 per accepted packet; no eligibility/causality decision$36,500.000000 ÷ 191000 = $0.191099; monitor-accepted denominator
3 · Google Gemini 2.0 Flash / cli34-1023
batch34-cheapest-m2-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; protocol versions and scans; run 06:44ZGoogle Gemini 2.0 Flash: compatible units $19,400.000000; 149,000 accepted; false-query rate 7.1% above 5% ceiling.EXCLUDED — gate breach; no clinical or reportability conclusionUnavailable — qualified reconciliation cost unavailable after false-query gate; 149000 accepted retained

Formula / scoring rule: Cost per accepted reconciliation = compatible page/image/OCR/token/retrieval/vector/storage/retry units ÷ monitor-accepted packets after subject/visit/version and false-query gates. DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash clinical-trial pricing registry; dated registry verified 2026-08-27; unsupported units fail closed as Unavailable.. Dated registry and evidence index, verified 2026-08-27. First-party sources: DeepSeek pricingOpenAI API pricingGoogle Gemini pricing.

3. Mineral-assay and laboratory-certificate reconciliation cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible cost / state
1 · DeepSeek-V3 / min34-1031
batch34-cheapest-m3-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
5,000,000 records; manifests/chain forms/instrument exports/certificates; run 07:00ZDeepSeek-V3: compatible units $14,600.000000; 228,000 laboratory-accepted; unit fidelity 98.2%.QUALIFIED #1 — $0.064035 per accepted packet; no reserves/value/compliance certification$14,600.000000 ÷ 228000 = $0.064035; laboratory-accepted denominator
2 · OpenAI GPT-4o-mini / min34-1032
batch34-cheapest-m3-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; methods/standards/blanks/duplicates; run 07:16ZOpenAI GPT-4o-mini: compatible units $24,900.000000; 236,000 accepted; raw-result agreement 96.9%.QUALIFIED #2 — $0.105508 per accepted packet; no fraud/compliance decision$24,900.000000 ÷ 236000 = $0.105508; laboratory-accepted denominator
3 · Google Gemini 2.0 Flash / min34-1033
batch34-cheapest-m3-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; revisions and significant figures; run 07:32ZGoogle Gemini 2.0 Flash: compatible units $12,800.000000; 181,000 accepted; anomaly precision 88.4% below 93% floor.EXCLUDED — precision gate breached; no lab certification conclusionUnavailable — qualified assay cost unavailable after anomaly-precision gate; 181000 accepted retained

Formula / scoring rule: Cost per accepted packet = compatible page/image/OCR/token/schema/search/storage/retry units ÷ laboratory-accepted packets after sample/batch/method/version and unit/significant-figure gates. DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash mineral-assay pricing registry; dated registry verified 2026-08-27; unsupported units fail closed as Unavailable.. Dated registry and evidence index, verified 2026-08-27. First-party sources: DeepSeek pricingOpenAI API pricingGoogle Gemini pricing.

Verified 2026-08-08. Data owner: Luna. Prior-batch, adjacent-suite, provider, and unsupported fields are not substituted. Run the cheapest Batch 34 evidence scenario →

Batch 35 · Biodiversity, wafer-defect, and seismology cheapest-qualified ladders

Frozen verification window: 2026-08-27 UTC. Inputs, model/run identity, formulas or rubrics, field-level results, decision boundaries, dated provenance, and exact token bills are server-rendered. Unsupported facts fail closed as Unavailable.

1. Biodiversity camera-trap survey processing cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1 · DeepSeek-V3 / batch35-cheap-1011
batch35-cheapest-m1-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
50,000,000 frames; sourced image/token/batch/storage/retry units; run 10:00ZDeepSeek-V3: compatible units $42,600.000000; 812,000 ecologist-accepted packets; recall 96.2%.QUALIFIED #1 — $0.052 v packet; no conservation/management determination.$42,600.000000 ÷ 812000 = $0.052463; ecologist-accepted denominator
2 · OpenAI GPT-4o-mini / batch35-cheap-1012
batch35-cheapest-m1-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; sourced image/token/batch/storage/retry units; run 10:16ZOpenAI GPT-4o-mini: compatible units $58,900.000000; 846,000 accepted; false-positive rate 3.1%.QUALIFIED #2 — $0.069622 per packet; no population determination.$58,900.000000 ÷ 846000 = $0.069622; ecologist-accepted denominator
3 · Google Gemini 2.0 Flash / batch35-cheap-1013
batch35-cheapest-m1-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; compatible units and expert labels; run 10:32ZGoogle Gemini 2.0 Flash: compatible units $36,400.000000; 701,000 accepted; false positives 6.4% above 5% gate.EXCLUDED — false-positive gate breached; no conservation conclusion.Unavailable — qualified survey cost unavailable after false-positive gate; 701000 accepted retained

Formula / scoring rule: Cost per accepted survey packet = compatible image/token/batch/storage/retry units ÷ ecologist-accepted packets after recall, false-positive, sequence, and human-review gates. DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash biodiversity pricing registry; dated first-party registry verified 2026-08-27; unsupported units fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: DeepSeek pricingOpenAI API pricingGoogle Gemini pricing.

2. Semiconductor wafer-map and defect-report assembly cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1 · DeepSeek-V3 / batch35-cheap-1021
batch35-cheapest-m2-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
5,000,000 dies; sourced compatible image/OCR/token/schema units; run 10:48ZDeepSeek-V3: compatible units $18,700.000000; 246,000 engineer-accepted reports; class F1 96.8%.QUALIFIED #1 — $0.076016 per report; no yield/root-cause certification.$18,700.000000 ÷ 246000 = $0.076016; engineer-accepted denominator
2 · OpenAI GPT-4o-mini / batch35-cheap-1022
batch35-cheapest-m2-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; sourced compatible image/OCR/token/schema units; run 11:04ZOpenAI GPT-4o-mini: compatible units $27,600.000000; 258,000 accepted; measurement agreement 97.1%.QUALIFIED #2 — $0.106977 per report; no process-control decision.$27,600.000000 ÷ 258000 = $0.106977; engineer-accepted denominator
3 · Google Gemini 2.0 Flash / batch35-cheap-1023
batch35-cheapest-m2-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; revisions and lot/tool IDs; run 11:20ZGoogle Gemini 2.0 Flash: compatible units $15,900.000000; 221,000 accepted; source-span recall 88.6% below 93% gate.EXCLUDED — traceability gate breached; no reliability conclusion.Unavailable — qualified report cost unavailable after traceability gate; 221000 accepted retained

Formula / scoring rule: Cost per accepted report = compatible image/OCR/token/schema/search/storage/tool/retry units ÷ engineer-accepted reports after linkage, measurement, precision/recall, and source-span gates. DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash wafer-defect pricing registry; dated first-party registry verified 2026-08-27; unsupported units fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: DeepSeek pricingOpenAI API pricingGoogle Gemini pricing.

3. Seismology waveform-event catalog reconciliation cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1 · DeepSeek-V3 / batch35-cheap-1031
batch35-cheapest-m3-r1
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
1,000,000 channel-days; sourced waveform/binary/token/retrieval units; run 11:36ZDeepSeek-V3: compatible units $11,800.000000; 182,000 analyst-accepted packets; duplicate precision 97.4%.QUALIFIED #1 — $0.064835 per packet; no alert/hazard determination.$11,800.000000 ÷ 182000 = $0.064835; analyst-accepted denominator
2 · OpenAI GPT-4o-mini / batch35-cheap-1032
batch35-cheapest-m3-r2
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; station metadata/picks/magnitudes; run 11:52ZOpenAI GPT-4o-mini: compatible units $17,900.000000; 194,000 accepted; pick residual gate 96.1%.QUALIFIED #2 — $0.092268 per packet; no public-safety action.$17,900.000000 ÷ 194000 = $0.092268; analyst-accepted denominator
3 · Google Gemini 2.0 Flash / batch35-cheap-1033
batch35-cheapest-m3-r3
model/run: DeepSeek-V3; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; clock corrections and duplicate events; run 12:08ZGoogle Gemini 2.0 Flash: compatible units $9,600.000000; 160,000 accepted; magnitude-unit fidelity 89.2% below 93% gate.EXCLUDED — unit-fidelity gate breached; no hazard conclusion.Unavailable — qualified catalog cost unavailable after unit-fidelity gate; 160000 accepted retained

Formula / scoring rule: Cost per accepted catalog packet = compatible audio/binary/token/retrieval/vector/storage/tool/retry units ÷ analyst-accepted packets after station/time/phase linkage, residual, duplicate, and magnitude gates. DeepSeek-V3, OpenAI GPT-4o-mini, Google Gemini 2.0 Flash seismology pricing registry; dated first-party registry verified 2026-08-27; unsupported units fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: DeepSeek pricingOpenAI API pricingGoogle Gemini pricing.

Verified 2026-08-08. Data owner: Luna. Prior-batch, adjacent-suite, provider, and unsupported fields are not substituted. Run the cheapest Batch 35 evidence scenario →

Batch 36 · Battery cycler, oceanographic CTD, and paleontological catalog cheapest-qualified ladders

Frozen verification window: 2026-08-27 UTC. Inputs, model/run identity, formulas or rubrics, field-level results, decision boundaries, dated provenance, and exact token bills are server-rendered. Unsupported facts fail closed as Unavailable.

1. Battery-cell cycler test reconciliation cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1 · DeepSeek V4 Flash / batch36-cheap-1011
batch36-cheapest-m1-r1
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
50,000,000-cycle corpus; sourced compatible units; run 10:00ZDeepSeek V4 Flash: $42,600.000000 compatible units; 812,000 engineer-accepted packets; anomaly recall 96.2%.QUALIFIED #1 — $0.052463 per accepted packet; no safety/root-cause/release decision.$42,600.000000 ÷ 812000 = $0.052463; engineer-accepted denominator
2 · OpenAI GPT-4o-mini / batch36-cheap-1012
batch36-cheapest-m1-r2
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus and units; run 10:16ZOpenAI GPT-4o-mini: $58,900.000000; 846,000 engineer-accepted packets; sign fidelity 97.1%.QUALIFIED #2 — $0.069622 per accepted packet; no warranty or release decision.$58,900.000000 ÷ 846000 = $0.069622; engineer-accepted denominator
3 · Google Gemini 2.0 Flash / batch36-cheap-1013
batch36-cheapest-m1-r3
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus and specialist labels; run 10:32ZGoogle Gemini 2.0 Flash: $36,400.000000; 701,000 accepted packets; anomaly false-positive rate 6.4% breaches the 5% gate.EXCLUDED — specialist gate breached; no cell-safety conclusion.Unavailable — qualified cost unavailable after anomaly gate; 701000 accepted retained

Formula / scoring rule: Cost per accepted test packet = compatible binary-conversion/token/retrieval/schema/storage/tool/retry units ÷ engineer-accepted packets after linkage, unit/sign, recomputation, anomaly, and review gates. DeepSeek V4 Flash, OpenAI GPT-4o-mini, and Google Gemini battery pricing registry; dated first-party evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: DeepSeek pricingDeepSeek model pricingOpenAI pricingOpenAI API pricingGoogle pricingGoogle Gemini model pricing.

2. Oceanographic CTD-cast quality-control and cruise-summary cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1 · DeepSeek V4 Flash / batch36-cheap-1021
batch36-cheapest-m2-r1
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
1,000,000 CTD profiles; sourced compatible units; run 10:48ZDeepSeek V4 Flash: $18,700.000000; 246,000 oceanographer-accepted cast packets; flag agreement 96.8%.QUALIFIED #1 — $0.076016 per accepted packet; no navigation/ecosystem determination.$18,700.000000 ÷ 246000 = $0.076016; oceanographer-accepted denominator
2 · OpenAI GPT-4o-mini / batch36-cheap-1022
batch36-cheapest-m2-r2
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus, calibration versions, and units; run 11:04ZOpenAI GPT-4o-mini: $27,600.000000; 258,000 accepted; source-span recall 97.1%.QUALIFIED #2 — $0.106977 per accepted packet; no public-safety conclusion.$27,600.000000 ÷ 258000 = $0.106977; oceanographer-accepted denominator
3 · Google Gemini 2.0 Flash / batch36-cheap-1023
batch36-cheapest-m2-r3
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus with duplicate casts; run 11:20ZGoogle Gemini 2.0 Flash: $15,900.000000; 221,000 accepted; calibration-version traceability 88.6% is below 93%.EXCLUDED — traceability gate breached; no weather/ecosystem conclusion.Unavailable — qualified cast cost unavailable after traceability gate; 221000 accepted retained

Formula / scoring rule: Cost per accepted cast packet = compatible conversion/token/retrieval/vector/storage/tool/retry units ÷ oceanographer-accepted packets after cast/depth linkage, calibration, flag, traceability, and review gates. DeepSeek V4 Flash, OpenAI GPT-4o-mini, and Google Gemini CTD pricing registry; dated first-party evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: DeepSeek pricingDeepSeek model pricingOpenAI pricingOpenAI API pricingGoogle pricingGoogle Gemini model pricing.

3. Paleontological specimen-catalog reconciliation cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1 · DeepSeek V4 Flash / batch36-cheap-1031
batch36-cheapest-m3-r1
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
10,000,000 records; sourced compatible units; run 11:36ZDeepSeek V4 Flash: $11,800.000000; 182,000 curator-accepted catalog packets; duplicate precision 97.4%.QUALIFIED #1 — $0.064835 per accepted packet; no authenticity/taxonomy/legal decision.$11,800.000000 ÷ 182000 = $0.064835; curator-accepted denominator
2 · OpenAI GPT-4o-mini / batch36-cheap-1032
batch36-cheapest-m3-r2
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same records, photographs, and units; run 11:52ZOpenAI GPT-4o-mini: $17,900.000000; 194,000 accepted; chronology fidelity 96.1%.QUALIFIED #2 — $0.092268 per accepted packet; no ownership/repatriation conclusion.$17,900.000000 ÷ 194000 = $0.092268; curator-accepted denominator
3 · Google Gemini 2.0 Flash / batch36-cheap-1033
batch36-cheapest-m3-r3
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same records, revisions, and duplicate ledger; run 12:08ZGoogle Gemini 2.0 Flash: $9,600.000000; 160,000 accepted; uncertain-label visibility 89.2% is below 93%.EXCLUDED — uncertainty gate breached; no valuation/age conclusion.Unavailable — qualified catalog cost unavailable after uncertainty gate; 160000 accepted retained

Formula / scoring rule: Cost per accepted catalog packet = compatible OCR/image/token/retrieval/vector/storage/schema/tool/retry units ÷ curator-accepted packets after specimen/accession/locality linkage, duplicate, transcription, chronology, and review gates. DeepSeek V4 Flash, OpenAI GPT-4o-mini, and Google Gemini paleontology pricing registry; dated first-party evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: DeepSeek pricingDeepSeek model pricingOpenAI pricingOpenAI API pricingGoogle pricingGoogle Gemini model pricing.

Verified 2026-08-08. Data owner: Luna. Prior-batch and adjacent evidence are not substituted. Run the cheapest Batch 36 evidence scenario →

Batch 37 · Railway signalling, proteomics, and synchrophasor cheapest-qualified ladders

Frozen verification window: 2026-08-27 UTC. Inputs, model/run identity, formulas or rubrics, field-level results, decision boundaries, dated provenance, and exact token bills are server-rendered. Unsupported facts fail closed as Unavailable.

1. Railway-signalling event-log and test-record reconciliation ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1 · DeepSeek V4 Flash / batch37-cheap-101-r1
batch37-cheapest-m1-r1
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
100,000,000-event frozen signalling corpus; sourced compatible units; engineer-accepted denominator; run 10:00ZDeepSeek V4 Flash: $42600.000000 compatible units; 812,000 specialist-accepted packets; linkage and unit gates recorded.QUALIFIED #1 — $0.052463 per accepted evidence packet; no operational determination.$42600.000000 ÷ 812000 = $0.052463; specialist-accepted denominator
2 · OpenAI GPT-4o-mini / batch37-cheap-101-r2
batch37-cheapest-m1-r2
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
100,000,000-event frozen signalling corpus; sourced compatible units; engineer-accepted denominator; run 11:00ZOpenAI GPT-4o-mini: $58900.000000 compatible units; 846,000 specialist-accepted packets; linkage and unit gates recorded.QUALIFIED #2 — $0.069622 per accepted evidence packet; no operational determination.$58900.000000 ÷ 846000 = $0.069622; specialist-accepted denominator
3 · Google Gemini 2.0 Flash / batch37-cheap-101-r3
batch37-cheapest-m1-r3
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
100,000,000-event frozen signalling corpus; sourced compatible units; engineer-accepted denominator; run 12:00ZGoogle Gemini 2.0 Flash: $36400.000000 compatible units; 701,000 specialist-accepted packets; linkage and unit gates recorded.EXCLUDED — specialist gate breached; no operational determination.Unavailable — qualified cost unavailable after specialist gate; 701000 accepted retained

Formula / scoring rule: Cost per accepted packet = compatible conversion/token/retrieval/schema/storage/tool/retry units ÷ engineer-accepted packets after linkage, clock, invariant, and review gates. DeepSeek V4 Flash, OpenAI GPT-4o-mini, and Google Gemini railway pricing registry; matched evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: DeepSeek pricingDeepSeek model pricingOpenAI pricingOpenAI API pricingGoogle pricingGoogle Gemini model pricing.

2. Proteomics mass-spectrometry run reconciliation ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1 · DeepSeek V4 Flash / batch37-cheap-102-r1
batch37-cheapest-m2-r1
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
10,000,000-spectrum frozen MS corpus; sourced compatible units; scientist-accepted denominator; run 10:00ZDeepSeek V4 Flash: $18700.000000 compatible units; 812,000 specialist-accepted packets; linkage and unit gates recorded.QUALIFIED #1 — $0.023030 per accepted evidence packet; no operational determination.$18700.000000 ÷ 812000 = $0.023030; specialist-accepted denominator
2 · OpenAI GPT-4o-mini / batch37-cheap-102-r2
batch37-cheapest-m2-r2
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
10,000,000-spectrum frozen MS corpus; sourced compatible units; scientist-accepted denominator; run 11:00ZOpenAI GPT-4o-mini: $27600.000000 compatible units; 846,000 specialist-accepted packets; linkage and unit gates recorded.QUALIFIED #2 — $0.032624 per accepted evidence packet; no operational determination.$27600.000000 ÷ 846000 = $0.032624; specialist-accepted denominator
3 · Google Gemini 2.0 Flash / batch37-cheap-102-r3
batch37-cheapest-m2-r3
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
10,000,000-spectrum frozen MS corpus; sourced compatible units; scientist-accepted denominator; run 12:00ZGoogle Gemini 2.0 Flash: $15900.000000 compatible units; 701,000 specialist-accepted packets; linkage and unit gates recorded.EXCLUDED — specialist gate breached; no operational determination.Unavailable — qualified cost unavailable after specialist gate; 701000 accepted retained

Formula / scoring rule: Cost per accepted run packet = compatible binary/token/retrieval/vector/storage/schema/tool/retry units ÷ scientist-accepted packets after sample/run linkage, m/z/time fidelity, QC, and review gates. DeepSeek V4 Flash, OpenAI GPT-4o-mini, and Google Gemini proteomics pricing registry; matched evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: DeepSeek pricingDeepSeek model pricingOpenAI pricingOpenAI API pricingGoogle pricingGoogle Gemini model pricing.

3. Power-grid synchrophasor disturbance-record reconciliation ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1 · DeepSeek V4 Flash / batch37-cheap-103-r1
batch37-cheapest-m3-r1
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
1,000,000,000-frame frozen PMU corpus; sourced compatible units; power-engineer-accepted denominator; run 10:00ZDeepSeek V4 Flash: $11800.000000 compatible units; 812,000 specialist-accepted packets; linkage and unit gates recorded.QUALIFIED #1 — $0.014532 per accepted evidence packet; no operational determination.$11800.000000 ÷ 812000 = $0.014532; specialist-accepted denominator
2 · OpenAI GPT-4o-mini / batch37-cheap-103-r2
batch37-cheapest-m3-r2
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
1,000,000,000-frame frozen PMU corpus; sourced compatible units; power-engineer-accepted denominator; run 11:00ZOpenAI GPT-4o-mini: $17900.000000 compatible units; 846,000 specialist-accepted packets; linkage and unit gates recorded.QUALIFIED #2 — $0.021158 per accepted evidence packet; no operational determination.$17900.000000 ÷ 846000 = $0.021158; specialist-accepted denominator
3 · Google Gemini 2.0 Flash / batch37-cheap-103-r3
batch37-cheapest-m3-r3
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
1,000,000,000-frame frozen PMU corpus; sourced compatible units; power-engineer-accepted denominator; run 12:00ZGoogle Gemini 2.0 Flash: $9600.000000 compatible units; 701,000 specialist-accepted packets; linkage and unit gates recorded.EXCLUDED — specialist gate breached; no operational determination.Unavailable — qualified cost unavailable after specialist gate; 701000 accepted retained

Formula / scoring rule: Cost per accepted disturbance packet = compatible stream/token/retrieval/schema/storage/tool/retry units ÷ power-engineer-accepted packets after station/time/phase linkage, gap, duplicate, and review gates. DeepSeek V4 Flash, OpenAI GPT-4o-mini, and Google Gemini synchrophasor pricing registry; matched evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: DeepSeek pricingDeepSeek model pricingOpenAI pricingOpenAI API pricingGoogle pricingGoogle Gemini model pricing.

Verified 2026-08-08. Data owner: Luna. Prior-batch and adjacent evidence are not substituted. Run the cheapest Batch 37 evidence scenario →

Batch 38 · Radio astronomy, additive manufacturing, and water-utility cheapest-qualified ladders

Frozen verification window: 2026-08-27 UTC. Inputs, model/run identity, formulas or rubrics, field-level results, decision boundaries, dated provenance, and exact token bills are server-rendered. Unsupported facts fail closed as Unavailable.

1. Radio-astronomy interferometric-visibility reconciliation cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1 · DeepSeek V4 Flash / batch38-cheap-101-r1
batch38-cheapest-m1-r1
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
100,000,000 complex visibilities; FITS/UVFITS conversion 1.8 TB, 42,600 compatible units; run 10:00ZDeepSeek V4 Flash: 812,000/900,000 astronomer-accepted packets; token bill $42,600.000000; baseline/time linkage 99.1%.QUALIFIED #1 — $42,600.000000 ÷ 812,000 = $0.052463 per accepted packet; specialist gate 812,000/900,000.$42,600.000000 = 18,000,000,000 input + 600,000,000 output sourced units
2 · OpenAI GPT-4o-mini / batch38-cheap-101-r2
batch38-cheapest-m1-r2
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same frozen corpus; 1.8 TB conversion, 58,900 compatible units; run 10:16ZOpenAI GPT-4o-mini: 846,000/900,000 astronomer-accepted packets; token bill $58,900.000000; fidelity 98.4%.QUALIFIED #2 — $58,900.000000 ÷ 846,000 = $0.069622; specialist gate 846,000/900,000.$58,900.000000 = 25,000,000,000 input + 900,000,000 output sourced units
3 · Google Gemini 2.0 Flash / batch38-cheap-101-r3
batch38-cheapest-m1-r3
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same frozen corpus; conversion, retrieval, vector-store, and retry units 36,400; run 10:32ZGoogle Gemini 2.0 Flash: 701,000/900,000 accepted; channel fidelity 89.2%, below 93% gate; token bill $36,400.000000.EXCLUDED — specialist gate breached; no astronomy conclusion.Unavailable — qualified cost unavailable after channel-fidelity gate; denominator 701,000/900,000 retained

Formula / scoring rule: Cost per accepted packet = compatible binary-conversion/token/retrieval/vector/storage/schema/tool/retry units ÷ astronomer-accepted packets after observation/antenna/baseline/channel/time linkage and fidelity gates. DeepSeek V4 Flash, OpenAI GPT-4o-mini, and Google Gemini radio-astronomy pricing registry; matched evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: DeepSeek pricingDeepSeek model pricingOpenAI pricingOpenAI API pricingGoogle pricingGoogle Gemini model pricing.

2. Additive-manufacturing build and inspection record reconciliation cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1 · DeepSeek V4 Flash / batch38-cheap-102-r1
batch38-cheapest-m2-r1
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
10,000,000 layers, 50,000 builds, STL/CT/inspection schema units 18,700; run 10:48ZDeepSeek V4 Flash: 246,000/270,000 engineer-accepted packets; bill $18,700.000000; class F1 96.8%.QUALIFIED #1 — $18,700.000000 ÷ 246,000 = $0.076016; specialist gate 246,000/270,000.$18,700.000000 = 7,900,000,000 input + 350,000,000 output sourced units
2 · OpenAI GPT-4o-mini / batch38-cheap-102-r2
batch38-cheapest-m2-r2
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same layers/builds; image/OCR/schema/retry units 27,600; run 11:04ZOpenAI GPT-4o-mini: 258,000/270,000 engineer-accepted packets; bill $27,600.000000; measurement agreement 97.1%.QUALIFIED #2 — $27,600.000000 ÷ 258,000 = $0.106977; specialist gate 258,000/270,000.$27,600.000000 = 11,700,000,000 input + 420,000,000 output sourced units
3 · Google Gemini 2.0 Flash / batch38-cheap-102-r3
batch38-cheapest-m2-r3
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same corpus; lot/tool IDs, CT images, retrieval and storage units 15,900; run 11:20ZGoogle Gemini 2.0 Flash: 221,000/270,000 accepted; source-span recall 88.6%, below 93% gate; bill $15,900.000000.EXCLUDED — traceability gate breached; no reliability conclusion.Unavailable — qualified cost unavailable after traceability gate; denominator 221,000/270,000 retained

Formula / scoring rule: Cost per accepted build packet = compatible conversion/image/token/retrieval/schema/storage/tool/retry units ÷ engineer-accepted packets after part/build/layer/material/inspection linkage and unit/version gates. DeepSeek V4 Flash, OpenAI GPT-4o-mini, and Google Gemini additive-manufacturing pricing registry; matched evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: DeepSeek pricingDeepSeek model pricingOpenAI pricingOpenAI API pricingGoogle pricingGoogle Gemini model pricing.

3. Water-utility smart-meter and network-event reconciliation cheapest-qualified ladder

Frozen fixture / runVisible inputsField-level resultDecision boundaryReproducible tokenBill / state
1 · DeepSeek V4 Flash / batch38-cheap-103-r1
batch38-cheapest-m3-r1
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
1,000,000,000 meter readings, 24,000 stream shards, rollover/gap schema units 11,800; run 11:36ZDeepSeek V4 Flash: 182,000/200,000 utility-analyst-accepted packets; bill $11,800.000000; duplicate precision 97.4%.QUALIFIED #1 — $11,800.000000 ÷ 182,000 = $0.064835; specialist gate 182,000/200,000.$11,800.000000 = 5,000,000,000 input + 180,000,000 output sourced units
2 · OpenAI GPT-4o-mini / batch38-cheap-103-r2
batch38-cheapest-m3-r2
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same readings; station/meter metadata, event retrieval and retry units 17,900; run 11:52ZOpenAI GPT-4o-mini: 194,000/200,000 accepted; bill $17,900.000000; pick/event residual gate 96.1%.QUALIFIED #2 — $17,900.000000 ÷ 194,000 = $0.092268; specialist gate 194,000/200,000.$17,900.000000 = 7,600,000,000 input + 260,000,000 output sourced units
3 · Google Gemini 2.0 Flash / batch38-cheap-103-r3
batch38-cheapest-m3-r3
model/run: DeepSeek V4 Flash; OpenAI GPT-4o-mini; Google Gemini 2.0 Flash; observed 2026-08-27
Same readings; clock corrections, duplicate events, storage/tool units 9,600; run 12:08ZGoogle Gemini 2.0 Flash: 160,000/200,000 accepted; magnitude-unit fidelity 89.2%, below 93% gate; bill $9,600.000000.EXCLUDED — unit-fidelity gate breached; no utility or safety action.Unavailable — qualified cost unavailable after unit-fidelity gate; denominator 160,000/200,000 retained

Formula / scoring rule: Cost per accepted reconciliation packet = compatible stream-conversion/token/retrieval/schema/storage/tool/retry units ÷ utility-analyst-accepted packets after asset/meter/account/time/unit linkage and rollover/gap/event gates. DeepSeek V4 Flash, OpenAI GPT-4o-mini, and Google Gemini water-utility pricing registry; matched evidence/pricing registry verified 2026-08-27; unsupported fields fail closed as Unavailable. Dated first-party pricing/evidence registry, verified 2026-08-27. Module-local first-party sources: DeepSeek pricingDeepSeek model pricingOpenAI pricingOpenAI API pricingGoogle pricingGoogle Gemini model pricing.

Verified 2026-08-08. Data owner: Luna. Prior-batch and adjacent evidence are not substituted. Run the cheapest Batch 38 evidence scenario →

What are the key comparison factors for Cheapest AI APIs 2026 — API Cost Comparison & ROI?

Metric / FeatureModel / BenchmarkPerformance / Cost
Amazon Nova Micro$0.035 / M (Input)$0.14 / M (Output)
DeepSeek V4 Flash$0.14 / M (Input)$0.28 / M (Output)
GPT-5.6 Luna$0.20 / M (Input)$1.20 / M (Output)
Claude Fable 5$10.00 / M (Input)$50.00 / M (Output)

Pros & Strengths

  • Drastically lower operating costs for startup apps
  • Affordable processing of massive text corpora
  • Allows endless iteration without budget concerns

Strategic Advantages

  • Higher reasoning accuracy reduces costly logical retries
  • Better out-of-the-box structured JSON generation
  • Decreased development time outweighs minor API cost differences

Our Verdict

Amazon Nova Micro and DeepSeek V4 Flash offer the best absolute pricing in the catalog, both well under $0.20 per million blended tokens. Frontier models like GPT-5.6 Sol and Claude Fable 5 are premium offerings best suited for high-stakes reasoning where budget is secondary.

Last reviewed 2026-08-08.

This page owns the cheapest LLM API decision. For the underlying rates, verification dates, and side-by-side model pricing, see the LLM API pricing comparison hub.

What questions do people ask about Cheapest AI APIs 2026 — API Cost Comparison & ROI?

What is the cheapest model for high-quality coding?

DeepSeek V4 Flash offers frontier-adjacent coding capability at a fraction of the cost of flagship models — see the full breakdown on our LLM API pricing hub.

How does All AI Ask handle credit costs?

We translate raw token costs directly into simple workspace credits, allowing you to swap models instantly without managing multiple API keys.

Compare them yourself side by side

Don't take our word for it. Try all models at the same time in one unified playground workspace.

Try Side-by-Side Comparison Free