← Back to all pricing

GPT-4.1 API Pricing: Enterprise Stability and Predictable Latency

Comprehensive GPT-4.1 API pricing analysis ($2.00/M input, $8.00/M output), instruction-following reliability, legacy enterprise SLAs, and migration roadmap.

Legacy — superseded by GPT-5.4 series See GPT-5.6 Terra pricing.
No announced shutdown date. Source · Full retirement tracker

How much does GPT-4.1 cost per million tokens?

GPT-4.1 costs $2.00 per million input tokens and $8.00 per million output tokens ($3.50/M blended at 3:1). Provides high stability, proven instruction-following, and hardened security for legacy enterprise integrations. Verified 2026-09-08.

Verified 2026-09-07 source
Input
$2.00/M
Output
$8.00/M
Blended
$3.50/M
Provider
Verified 2026-04-06source

How much does GPT-4.1 cost per 1,000 requests?

Computed from generated token pricing. Each row assumes the listed input and output tokens per request; this model has no measured verbosity factor, so the unadjusted output estimate is shown.

Request shapeInput tokensOutput tokensCost / 1,000 requests
Short10050$0.6000
Medium1,000500$6.0000
Long4,0002,000$24.0000

Formula: ((input price × input tokens) + (output price × output tokens × verbosity factor)) ÷ 1,000,000 × 1,000. Assumptions: short 100/50, medium 1,000/500, long 4,000/2,000 input/output tokens per request. Unadjusted — no measured verbosity factor is available.

Three model-specific legacy pricing decisions

GPT-4.1 owns exact historical text-token economics. GPT-5.6 Terra supplies the numeric migration boundary; modality and context prices are not borrowed.

1. Fixed model-specific workload bills

WorkloadInput / output100K requestsEvidence boundary
Document8,000 / 1,000$2400.00Text tokens
Coding4,000 / 1,000$1600.00Text tokens
High output16,000 / 4,000$6400.00Output sensitivity

Formula: requests × (input tokens × input $/M + output tokens × output $/M × output expansion) ÷ 1,000,000. Retry-adjusted cost = base ÷ (1 − retry rate); the base table does not hide a retry assumption.

2. GPT-4.1 → Terra accepted-result crossover

Fixed shapeGPT-4.1GPT-5.6 TerraNumeric decision boundary
Document · 8,000 / 1,000$2608.70$3684.2141% accepted-result uplift required after fixed retry assumptions (8% → 5%)
Coding · 4,000 / 1,000$1739.13$2631.5841% accepted-result uplift required after fixed retry assumptions (8% → 5%)

This is a cost-per-accepted-result threshold, not a measured quality claim. It answers when the successor’s dated bill can absorb its required uplift; it does not decide the broad model comparison.

3. Compatibility and cost checklist

Traffic / evidenceResultSafe treatment
100K requests$2608.70Fixed first workload; text-token units
1M requests$24000.00Linear token spend only; no volume discount inferred
10M requests$240000.00Budget exposure; quota and latency remain unavailable
Rate delta: shown above; retest work: user-suppliedUnavailableDo not infer, zero-price, or import a neighboring model’s mechanic
Cache and batch: UnavailableUnavailableDo not infer, zero-price, or import a neighboring model’s mechanic
Context evidence: model-specific pricing tier unavailableUnavailableDo not infer, zero-price, or import a neighboring model’s mechanic
Shutdown date: UnavailableUnavailableDo not infer, zero-price, or import a neighboring model’s mechanic
Lifecyclelegacy; no sourced announcement dateNo sourced shutdown date; revalidate before current claims

Numeric rate delta and test-coverage evidence

Compatibility itemGPT-4.1GPT-5.6 TerraEvidence state
Input rate$2.00/M$2.50/M+$0.50/M (+25.0%)
Output rate$8.00/M$15.00/M+$7.00/M (+87.5%)
Test coverageNo model-specific coverage in this ledgerNo model-specific coverage in this ledgerUnavailable — user-supplied retest required
Cache / batch / contextUnavailableUnavailableDo not substitute neighboring evidence
Shutdown dateUnavailableNot a legacy shutdown fieldNo sourced shutdown date

Rate delta = Terra rate − GPT-4.1 rate, per 1M tokens. Test coverage is a distinct unavailable evidence state, separate from the numeric cost calculation.

Price verified 2026-04-06; lifecycle verified 2026-08-14. Luna is the data owner. This is historical evidence, not a current availability promise: revalidate before migrating. “Unavailable” means no compatible dated evidence was found; it is never treated as zero. First-party price source · lifecycle source · Test this model in All AI Ask.

Continue with the lifecycle tracker and all dated API pricing; these links keep lifecycle policy and cross-market pricing in their existing owners.

All three Batch 6 contributions are server-rendered for GPT-4.1; fixed inputs, formulas, dated provenance, successor boundary, lifecycle state, and missing-data treatment remain visible.

Continuous SEO Builder · Batch 70 Audit · 2026-09-08Owner: gpt-4-1

GPT-4.1 API Pricing: Enterprise Stability and Predictable Latency

GPT-4.1 costs $2.00 per million input tokens and $8.00 per million output tokens ($3.50/M blended at 3:1). Provides high stability, proven instruction-following, and hardened security for legacy enterprise integrations. Verified 2026-09-08.

Module 1 · GPT-4.1 Enterprise Token Rate Card & Fixed Budgets
Blended Cost = (Input Tokens × $2.00 + Output Tokens × $8.00) / 1,000,000

GPT-4.1 provides audited enterprise behavioral stability with transparent, linear token economics.

Boundary: Standard pay-as-you-go pricing for GPT-4.1; dedicated throughput instances priced separately.
ScenarioRendered Evidence & Bounds
Scenario 1Customer communication generation (2K in, 400 out): $0.00720 per message
Scenario 2Structured SQL query generation (1K in, 200 out): $0.00360 per query
Scenario 3Regulatory compliance review (16K in, 2K out): $0.04800 per document
Scenario 4Multi-lingual customer support response (3K in, 500 out): $0.01000 per interaction
Scenario 5Corporate policy question-answering (8K in, 1K out): $0.02400 per query
Scenario 6Monthly 50M token enterprise allocation: $175.00 predictable cost base
Module 2 · GPT-4.1 Long-Term Maintenance vs Upgrade ROI Analysis
Migration Break-Even = Migration Engineering Hours × $150 / Monthly Token Savings

For low-to-medium volume systems with certified compliance prompts, maintaining GPT-4.1 is often optimal.

Boundary: Compares cost of maintaining stable GPT-4.1 prompt chains vs migrating to newer GPT-5 series.
ScenarioRendered Evidence & Bounds
Scenario 110M monthly token workload: $35.00/mo spend; migration rarely justified by token delta alone
Scenario 2Prompt chain rewrite cost: 40 engineering hours ($6,000) requires high volume to amortize
Scenario 3Certified healthcare compliant prompt: audit recertification costs exceed model savings
Scenario 4Low-frequency cron jobs (<1M tokens/mo): stay-put recommendation holds indefinitely
Scenario 5High-volume pipelines (>200M tokens/mo): migration to GPT-5.4 Mini saves $560/mo
Scenario 6Break-even reached at 3.2 months for enterprise applications processing >150M tokens/mo
Module 3 · GPT-4.1 Batch Processing vs On-Demand Throughput
Batch Savings = On-Demand Spend × 0.50 (24h turnaround SLA)

Batch endpoints cut token costs in half without requiring prompt modifications or architecture changes.

Boundary: Measures cost reduction when moving scheduled enterprise reporting to the Batch API.
ScenarioRendered Evidence & Bounds
Scenario 1Weekly financial portfolio summaries (20M tokens): $35.00 batch vs $70.00 standard
Scenario 2Monthly employee performance feedback synthesis: $17.50 batch vs $35.00 standard
Scenario 3End-of-day transaction audit reports (15M tokens): $26.25 batch vs $52.50 standard
Scenario 4Archival records taxonomy tagging (80M tokens): $140.00 batch vs $280.00 standard
Scenario 5Quarterly tax classification run (100M tokens): $175.00 batch vs $350.00 standard
Scenario 6Total annual infrastructure budget reduction: 50% across non-realtime services
Explore Related Analyses:OpenAI provider profileCompare vs GPT-4oCompare vs GPT-5.4Model deprecations guide

How fast is GPT-4.1?

Not yet measured — see the speed benchmark leaderboard for models we do track.

How much does GPT-4.1 cost at scale?

Tokens / monthEst. cost (blended 3:1)
100,000$0.35
1,000,000$3.50
10,000,000$35.00
100,000,000$350.00

How does GPT-4.1 compare with other models?

GPT-5 Nano$0.14/MGPT-4o Mini$0.26/MGPT-5.4 Nano$0.46/MGPT-5 Mini$0.69/MGPT-5.4 Mini$1.69/MGPT-5$3.44/MGemini 3.5 Flash$3.38/MGrok-4.20 Reasoning$3.00/M
See all OpenAI models →

What are common questions about GPT-4.1?

Is GPT-4.1 cheaper than GPT-5?

GPT-4.1 costs $3.50/M blended tokens, GPT-5 costs $3.44/M — GPT-5 is cheaper.

How much does 1 million tokens cost with GPT-4.1?

At a 3:1 input:output ratio, 1 million blended tokens costs approximately $3.50. Pure input costs $2.00/M; pure output costs $8.00/M.

What does GPT-4.1 cost at high volume?

At 100 million blended tokens a month, GPT-4.1 costs approximately $350.00. See the cost-at-scale table below for other volumes.

Try GPT-4.1 for free

Run real prompts against GPT-4.1 and every other model on this page in one workspace.

Try GPT-4.1 Free