All news

Industry news September 15, 2026 1 min read

Gemma 4 31B now runs on AWS three different ways

Gemma 4 31B now has three routes onto AWS and Google — Bedrock, SageMaker, and Google Cloud — and each bills differently.

Google DeepMind's Gemma 4 31B now reaches AWS three different ways, each billed differently.

Google's own Cloud program covers it directly — credits apply to Gemini and Gemma with no extra setup. On AWS, Bedrock already lists Gemma 4 31B as a managed, pay-per-token model, and Activate credits cover third-party models via Bedrock the same way they cover any other model there.

What's new, as of 14 September, is a third option: SageMaker JumpStart now deploys Gemma 4 31B — including an NVFP4-quantized build AWS says is 2.5x faster and uses two-thirds less memory — onto an instance you pay for by the hour, not by the token. That only beats Bedrock's metered rate once inference volume is high enough to keep the instance busy.

Three routes, three different bills for the same model — pick by how much of it you will actually run.

Sources

  1. Gemma-4-31B-it-assistant and Gemma-4-31B-IT-NVFP4 models now available on Amazon SageMaker JumpStart — Amazon Web Services, September 14, 2026
  2. Models at a glance — Amazon Web Services, September 15, 2026
  3. AWS Activate credits — Amazon Web Services, September 15, 2026
  4. Google for Startups Cloud Program — Google Cloud, September 15, 2026

Find out what you qualify for today

Tell us four facts and we will come back with an ordered list — or look through the database yourself. No account, no charge.