State of Indiana — Office of Technology
Indiana Office of Technology

Copilot AI Credits — Model Pricing Guide

Understand how model selection impacts your organization's AI credit consumption

How Copilot AI Credits Work

Every Copilot interaction (Chat, agent mode, coding agents) consumes tokens. Tokens are converted to AI credits based on the model used. 1 AI credit = $0.01 USD.

The cost of a single interaction depends on:

Code completions are free. Inline code suggestions (the gray text that appears as you type) do not consume AI credits. Only Chat, agent mode, and coding agent interactions are billed.

Included Credit Pools (Business & Enterprise Plans)

Credits are pooled at the organization/enterprise level — all seats share a combined pool.

PlanCredits/Seat/Month10-Seat Pool50-Seat Pool
Copilot Business ($19/seat)1,90019,00095,000
Copilot Enterprise ($39/seat)3,90039,000195,000

Once the pool is exhausted, additional usage is billed at overage rates (1 credit = $0.01) — unless a hard budget cap is configured.

Model Pricing (Per 1M Tokens → AI Credits)

All prices below are per 1 million tokens. To convert to AI credits: multiply the dollar rate by 100 (since 1 credit = $0.01).

OpenAI Models

ModelCategoryInputCachedOutput
GPT-5 miniLightweight$0.25$0.025$2.00
GPT-5.4 nanoLightweight$0.20$0.02$1.25
GPT-5.4 miniLightweight$0.75$0.075$4.50
GPT-5.3-CodexPowerful$1.75$0.175$14.00
GPT-5.4Versatile$2.50$0.25$15.00
GPT-5.5Powerful$5.00$0.50$30.00
GPT-5.6 LunaLightweight$1.00$0.10$6.00
GPT-5.6 TerraVersatile$2.50$0.25$15.00
GPT-5.6 SolPowerful$5.00$0.50$30.00

Anthropic (Claude) Models

Anthropic models include an additional cache write cost.

ModelCategoryInputCachedCache WriteOutput
Claude Haiku 4.5Versatile$1.00$0.10$1.25$5.00
Claude Sonnet 4 / 4.5 / 4.6Versatile$3.00$0.30$3.75$15.00
Claude Sonnet 5 *Versatile$2.00$0.20$2.50$10.00
Claude Opus 4.5–4.8Powerful$5.00$0.50$6.25$25.00
Claude Fable 5Powerful$10.00$1.00$12.50$50.00

* Claude Sonnet 5 promotional pricing through August 31, 2026.

Google Models

ModelCategoryInputCachedOutput
Gemini 3 FlashLightweight$0.50$0.05$3.00
Gemini 2.5 ProPowerful$1.25$0.125$10.00
Gemini 3.5 FlashLightweight$1.50$0.15$9.00
Gemini 3.6 FlashVersatile$1.50$0.15$7.50
Gemini 3.1 ProPowerful$2.00$0.20$12.00

Other Models

ModelProviderCategoryInputCachedOutput
MAI-Code-1-FlashMicrosoftLightweight$0.75$0.075$4.50
Kimi K2.7 CodeMoonshot AIVersatile$0.95$0.19$4.00
Raptor miniGitHubVersatile$0.25$0.025$2.00

Practical Cost Examples

A typical Chat interaction uses roughly 2,000 input tokens + 1,000 output tokens. Here's what that costs across models:

ModelCategoryCost per ChatCredits UsedChats per 1,900 credits
GPT-5.4 nanoLightweight~$0.0017~0.17~11,176
GPT-5 miniLightweight~$0.0025~0.25~7,600
Claude Sonnet 4.6Versatile~$0.021~2.1~905
GPT-5.4Versatile~$0.020~2.0~950
Claude Opus 4.6Powerful~$0.035~3.5~543
GPT-5.5Powerful~$0.040~4.0~475
Claude Fable 5Powerful~$0.070~7.0~271

Based on ~2K input + ~1K output tokens per interaction. Agent mode sessions use significantly more tokens (10K–100K+ per session).

Agent Mode Example

A typical agent session (multi-step code generation, file edits, terminal commands) might consume 50,000 input + 10,000 output tokens:

GPT-5.4 nano: (50K × $0.20/1M) + (10K × $1.25/1M) = $0.0225 → ~2.25 credits
Claude Sonnet: (50K × $3.00/1M) + (10K × $15.00/1M) = $0.30 → ~30 credits
Claude Opus: (50K × $5.00/1M) + (10K × $25.00/1M) = $0.50 → ~50 credits
GPT-5.5: (50K × $5.00/1M) + (10K × $30.00/1M) = $0.55 → ~55 credits
Claude Fable: (50K × $10.00/1M) + (10K × $50.00/1M) = $1.00 → ~100 credits
Key takeaway: A developer using premium models (Opus, GPT-5.5, Fable) in agent mode can consume their share of the credit pool 10–20× faster than one using lightweight models for Chat. Model choice is the biggest lever for cost control.

Cost Control Recommendations

Enterprise control: Organization owners can restrict which models are available to developers. Limiting model selection to lightweight/versatile tiers is the most effective cost control mechanism.

References