AI on the bill

See what Bedrock and Vertex AI cost, model by model

GetFinOps reads Bedrock and Vertex AI spend from your own bill, flags spikes and idle provisioned capacity, and records the AI calls GetFinOps makes for you.

It doesn’t collect spend your own applications send directly to Anthropic, OpenAI, Azure OpenAI or other model APIs.

The problem

Managed AI spend lands on the cloud bill as long usage-type and SKU strings, mixed in with everything else. It is hard to tell which model a charge came from, whether yesterday was a spike, or whether provisioned capacity sat idle.

How it works

  1. Collect

    Each scan reads Bedrock spend from Cost Explorer in your connected AWS account and Vertex AI spend from your BigQuery billing export, for the last 30 days by default.

  2. Break down

    Spend rolls up by model, provider and day, week or month, next to the cost of the AI calls GetFinOps makes in your workspace.

  3. Flag

    Spike and idle-capacity rules run on the collected spend, and anything they find appears on the AI Spend page and in your findings.

  • Daily totals

    for Bedrock and Vertex AI: no individual calls, nothing credited to an agent

  • 1,000 tokens

    over seven days: capacity that costs at least $1 but serves fewer is idle

  • Per call

    record of each Claude call GetFinOps makes: model, tokens, cost, stop reason

  • 90 days

    the most an admin can backfill for a provider, once every 15 minutes

Capabilities

Bedrock and Vertex AI spend by model, plus a call-by-call record of GetFinOps’ own AI use.

  1. Spend by model, from your bill

    Bedrock spend comes from your AWS Cost Explorer and Vertex AI spend from your BigQuery billing export, day by day. Each line is matched to a model and region; anything unrecognised stays in the total as “unknown”.

  2. Spikes and idle capacity

    A finding is raised when a model’s daily spend jumps well above its recent average, or when Bedrock Provisioned Throughput or a Vertex AI endpoint costs money while serving almost no tokens.

  3. Every GetFinOps AI call, itemised

    Reports, the copilot and the daily brief run on Claude. Each call is recorded with its model, input, output and cache tokens, cost and stop reason, and grouped by agent when it ran inside one.

  4. Gaps you can see

    For Bedrock and Vertex AI you see the days collected, the dates missing and when data last arrived, and a banner lists any gaps so an admin can backfill them.

In depth

Where the numbers come from

Two kinds of AI spend appear side by side, and they are measured differently.

  • Amazon Bedrock: daily unblended cost from Cost Explorer in your AWS account, grouped by usage type and region. The model and whether the tokens were input or output are read from the usage type.
  • Google Vertex AI: daily cost from your BigQuery billing export for services whose name starts with “Vertex AI”, grouped by SKU, region and project.
  • GetFinOps’ own AI: each Claude call made for your workspace, such as report narratives, copilot answers and the daily brief, recorded as it runs and priced from a per-model rate table.
  • A usage type or SKU that doesn’t match a known model is kept as “unknown” rather than dropped.

Spend by model and over time

The AI Spend page shows month-to-date spend against last month, total tokens, and the change from the previous seven days. A chart plots cost by day, week or month over windows from 24 hours to a year, and a table lists tokens and cost for each model.

Tabs filter the views to one provider: GetFinOps’ own calls, Bedrock or Vertex AI. The same breakdowns are available to AI assistants as read-only MCP tools.

Findings for spikes and idle capacity

Two rules run on the collected Bedrock and Vertex AI spend in every scan, and their findings are listed on the AI Spend page as well as with your other findings.

  • Cost spike: a model’s latest daily spend in a region is above its prior seven-day average plus two standard deviations, or above 1.25 times the average when those earlier days are all equal. The flagged day must cost at least $1, and at least three earlier days are needed. Three standard deviations or more is marked high severity.
  • Idle provisioned capacity: over seven days, Bedrock Provisioned Throughput or a Vertex AI deployed endpoint costs at least $1 while serving fewer than 1,000 tokens. The finding shows the capacity cost and the tokens actually served.

Knowing when data is missing

Each collection records, per provider, the days it covered, the dates missing from the window, and whether it succeeded, came back empty or failed. Bedrock and Vertex AI each get a freshness badge, and a banner lists missing dates.

An admin can backfill up to 90 days for a provider from that banner, once every 15 minutes.

What GetFinOps’ own AI costs you

Every Claude call GetFinOps makes for your workspace is stored with its model, input and output tokens, cache writes and reads, cost, duration, stop reason and, where Anthropic returns one, its request ID. Cache tokens are priced separately because Anthropic bills them separately from input.

Click a model to open its individual calls, with totals for the whole filter rather than just the page. Calls made inside an agent run are grouped by agent, showing runs, calls per run and cost per run, plus the share of spend that could be tied to an agent at all.

What it does not do

  • It doesn’t collect spend your own applications send directly to Anthropic, OpenAI, Azure OpenAI or other model APIs.
  • It doesn’t show individual requests, cache tokens or agents for Bedrock and Vertex AI, which arrive as daily totals.
  • Vertex AI figures are the billing export’s cost column, before credits.
  • AI spend views cover the whole workspace; they don’t follow the cloud account selector.
  • It doesn’t change provisioned capacity, endpoints or model settings.

FAQ

Questions

Which AI spend does it cover?

Amazon Bedrock, read from AWS Cost Explorer; Google Vertex AI, read from your BigQuery billing export; and the Claude calls GetFinOps makes for your workspace. Spend your own applications send straight to Anthropic, OpenAI or other model APIs is not collected.

Can I see individual requests?

For GetFinOps’ own calls, yes: click a model to list its calls with tokens, cost, duration, stop reason and Anthropic request ID. Bedrock and Vertex AI arrive as daily totals from your bill, so there are no individual requests to show.

What do I need to set up?

For Bedrock, a connected AWS account that GetFinOps can read Cost Explorer from. For Vertex AI, Cloud Billing export to BigQuery and a connected GCP project with access to that dataset.

Will it change anything in my AI setup?

No. Spike and idle-capacity findings tell you what to look at, such as removing idle Provisioned Throughput or undeploying an endpoint. GetFinOps does not make those changes.

Try free — no credit card

Create a workspace and connect an AWS or GCP account, or book a demo first.