BounceGrip

LLM API cost calculator for teams

Compare what your team would spend across OpenAI, Anthropic, Google Gemini, OpenRouter, Kimi, DeepSeek, Qwen, MiniMax, GLM, xAI (Grok), Meta (Muse Spark), Krea, Kling, and Runway — using each provider's published API prices.

Pricing snapshot:

How to use this tool

  1. Choose the response length your team usually needs.
  2. Enter your team size and average daily chats per person.
  3. Add any typical image or PDF uploads.
  4. Choose the billing period you want to budget for.
  5. Select one or more models to compare their estimated costs.

Worked example: Example: five people sending 15 medium chats each workday for a month can be compared across any selected provider models in one view.

GPT-5.5

OpenAI

Context
1.1M
Input /1M
$5.00
Output /1M
$30.00

GPT-5.5 Pro

OpenAI

Context
1.1M
Input /1M
$30.00
Output /1M
$180.00

GPT-5.4

OpenAI

Context
1.1M
Input /1M
$2.50
Output /1M
$15.00

GPT-5.4 mini

OpenAI

Context
400k
Input /1M
$0.75
Output /1M
$4.50

GPT-5.4 nano

OpenAI

Context
400k
Input /1M
$0.20
Output /1M
$1.25

GPT-5.1

OpenAI

Context
400k
Input /1M
$1.25
Output /1M
$10.00

Claude Opus 4.7

Anthropic

Context
1M
Input /1M
$5.00
Output /1M
$25.00

Claude Sonnet 4.6

Anthropic

Context
1M
Input /1M
$3.00
Output /1M
$15.00

Claude Opus 4.6

Anthropic

Context
1M
Input /1M
$5.00
Output /1M
$25.00

Gemini 3.1 Pro Preview

Google Gemini

Context
1M
Input /1M
$2.00
Output /1M
$12.00

Base rate through 200k input tokens; longer prompts use Google's higher context tier.

Gemini 3.1 Flash-Lite

Google Gemini

Context
1M
Input /1M
$0.25
Output /1M
$1.50

DeepSeek V4 Pro (OpenRouter)

OpenRouter

Context
1M
Input /1M
$0.44
Output /1M
$0.87

Qwen3.6 Plus (OpenRouter)

OpenRouter

Context
1M
Input /1M
$0.33
Output /1M
$1.95

GLM-5.1 (OpenRouter)

OpenRouter

Context
203k
Input /1M
$0.97
Output /1M
$3.04

MiMo-V2.5 (OpenRouter)

OpenRouter

Context
1M
Input /1M
$0.11
Output /1M
$0.28

Kimi K2.7 Code

Kimi

Context
262k
Input /1M
$0.95
Output /1M
$4.00

Kimi K2.7 Code Highspeed

Kimi

Context
262k
Input /1M
$1.90
Output /1M
$8.00

Kimi K2.6

Kimi

Context
262k
Input /1M
$0.95
Output /1M
$4.00

DeepSeek V4 Flash

DeepSeek

Context
1M
Input /1M
$0.14
Output /1M
$0.28

DeepSeek V4 Pro

DeepSeek

Context
1M
Input /1M
$0.44
Output /1M
$0.87

Qwen3.6 Plus

Qwen

Context
1M
Input /1M
$0.50
Output /1M
$3.00

International deployment base tier through 256k input tokens; longer prompts use Alibaba's $2 / $6 tier.

MiniMax M3

MiniMax

Context
1M
Input /1M
$0.30
Output /1M
$1.20

Base rate through 512k input tokens; longer prompts use MiniMax's higher tier.

MiniMax M2.7

MiniMax

Context
200k
Input /1M
$0.30
Output /1M
$1.20

MiniMax M2.5

MiniMax

Context
200k
Input /1M
$0.30
Output /1M
$1.20

GLM-5.1

GLM

Context
203k
Input /1M
$1.40
Output /1M
$4.40

GLM-4.6

GLM

Context
200k
Input /1M
$0.60
Output /1M
$2.20

GLM-4.5 Air

GLM

Context
128k
Input /1M
$0.20
Output /1M
$1.10

Grok 4.5

xAI (Grok)

Context
500k
Input /1M
$2.00
Output /1M
$6.00

Base rate through 200k input tokens; longer prompts use xAI's $4 / $12 tier.

Grok 4.3

xAI (Grok)

Context
1M
Input /1M
$1.25
Output /1M
$2.50

Base rate through 200k input tokens; longer prompts use xAI's $2.50 / $5 tier.

Muse Spark 1.1

Meta (Muse Spark)

Context
1M
Input /1M
$1.25
Output /1M
$4.25

Public preview: US-only and waitlist-gated at the time of this snapshot.

Krea 2 Medium

Krea

Context
0k
Input /1M
$0.00
Output /1M
$0.00

Image generation only. From $0.03 per image (Medium); Large $0.06, Medium Turbo $0.015. Not token-metered.

Kling v3 Standard

Kling

Context
0k
Input /1M
$0.00
Output /1M
$0.00

Video generation only. From ~$0.084 per second. Not token-metered.

Runway Gen 4.5

Runway

Context
0k
Input /1M
$0.00
Output /1M
$0.00

Video generation only. $0.12 per second (12 credits × $0.01). Not token-metered.

How this estimate works

We assume conversations split roughly 40% input / 60% output tokens. Response length sets output tokens per chat (300 / 800 / 1,600 / 3,200), each image adds ~1,100 input tokens, and each PDF adds ~3,000. Selecting several models splits usage evenly across them. Real bills vary with prompt caching, batching, and provider billing rules — use this for budgeting, not invoicing.

Want a practical cost-control framework? Read our guide to cutting LLM costs with a model-agnostic AI workspace.

Pricing sources and assumptions

Rates were checked against first-party provider documentation on . Cards show standard or base per-token rates; regional, long-context, batch, cache, and provider routing tiers can change the final bill.

Get your team on one AI workspace.